AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,877 models Compare
🤖
Gemma-SEA-LION-v4-4B-VL
aisingapore

SEA-LION is a collection of Large Language Models (LLMs) which have been pretrained and instruct-tuned for the Southeast Asia (SEA) region.

Multimodal 4.0B ↓ 323
🤖
Nous-Capybara-34B
NousResearch

This is trained on the Yi-34B model with 200K context length, for 3 epochs on the Capybara dataset!

Open Source 34.0B ↓ 323
🤖
youri-7b
rinna

Overview We conduct continual pre-training of llama2-7b on 40B tokens from a mixture of Japanese and English datasets. The continual pre-training significantly improves the model's performance on Japanese tasks.

Open Source 7.0B ↓ 323
🤖
Llama-3.1-Swallow-70B-Instruct-v0.3
tokyotech-llm

Llama 3.1 Swallow is a series of large language models (8B, 70B) that were built by continual pre-training on the Meta Llama 3.1 models. Llama 3.1 Swallow enhanced the Japanese language capabilities of the original Llama 3.1 while retaining the English language capabilities. We u…

Open Source 70.0B ↓ 322
🤖
Step-3.5-Flash-Base
stepfun-ai

Step 3.5 Flash (visit website) is our most capable open-source foundation model, engineered to deliver frontier reasoning and agentic capabilities with exceptional efficiency. We also open-sourced the training codebase (SteptronOss), with support for continue pretrain, SFT, RL (W…

Open Source ↓ 321
🤖
Gemma-SEA-LION-v4-27B-IT-NVFP4
aisingapore

Model Card for Gemma-SEA-LION-v4-27B-IT-NVFP4

Open Source 27.0B ↓ 321
🤖
codegen25-7b-multi_P
Salesforce

Authors: Erik Nijkamp\ , Hiroaki Hayashi\ , Yingbo Zhou, Caiming Xiong

Code 7.0B ↓ 319
🤖
Ring-1T
inclusionAI

🤗 Hugging Face      🤖 ModelScope      🐙 Experience Now

Open Source ↓ 313
🤖
Medical-Qwen3-Swallow-30B-A3B
tokyotech-llm

Medical-Qwen3-Swallow-30B-A3B is a medical-domain language model based on tokyotech-llm/Qwen3-Swallow-30B-A3B-RL-v0.2 . It is designed to support research and development toward safe and trustworthy AI for Japanese clinical settings.

Open Source 30.0B ↓ 310
🤖
Intern-S2-Preview-397B
internlm

💻Github Repo • 🤗HF Model Collections • 🤖ModelScope Collections • 💬Online Chat

Open Source 397.0B ↓ 309
🤖
Llama-2-70b-instruct
upstage

Developed by : Upstage Backbone Model : LLaMA-2 Language(s) : English Library : HuggingFace Transformers License : Fine-tuned checkpoints is licensed under the Non-Commercial Creative Commons license (CC BY-NC-4.0) Where to send comments : Instructions on how to provide feedback…

Open Source 70.0B ↓ 307
🤖
Ring-mini-2.0
inclusionAI

🤗 Hugging Face &nbsp&nbsp &nbsp&nbsp🤖 ModelScope   🐙 Experience Now

Open Source ↓ 306