AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

440 models for "Base Model" Compare
Medical-GPT-OSS-Swallow-120B logo
Medical-GPT-OSS-Swallow-120B
tokyotech-llm

Medical-GPT-OSS-Swallow-120B is a medical-domain language model based on tokyotech-llm/GPT-OSS-Swallow-120B-RL-v0.1. It is designed to support research and development toward safe and trustworthy AI for Japanese clinical settings.

Open Source 120.0B ↓ 282
SynLogic-Mix-3-32B logo
SynLogic-Mix-3-32B
MiniMaxAI

SynLogic Zero-Mix-3: Large-Scale Multi-Domain Reasoning Model

Open Source 32.0B ↓ 281
Aquila2-7B logo
Aquila2-7B
BAAI

We opensource our Aquila2 series, now including Aquila2 , the base language models, namely Aquila2-7B and Aquila2-34B , as well as AquilaChat2 , the chat models, namely AquilaChat2-7B and AquilaChat2-34B , as well as the long-text chat models, namely AquilaChat2-7B-16k and Aquila…

Open Source 7.0B ↓ 280
Medical-Qwen3-Swallow-30B-A3B logo
Medical-Qwen3-Swallow-30B-A3B
tokyotech-llm

Medical-Qwen3-Swallow-30B-A3B is a medical-domain language model based on tokyotech-llm/Qwen3-Swallow-30B-A3B-RL-v0.2 . It is designed to support research and development toward safe and trustworthy AI for Japanese clinical settings.

Open Source 30.0B ↓ 279
SEA-LION-v1-7B-IT logo
SEA-LION-v1-7B-IT
aisingapore

SEA-LION is a collection of Large Language Models (LLMs) which has been pretrained and instruct-tuned for the Southeast Asia (SEA) region. The sizes of the models range from 3 billion to 7 billion parameters.

Open Source 7.5B ↓ 276
Swallow-7b-plus-hf logo
Swallow-7b-plus-hf
tokyotech-llm

Our Swallow model has undergone continual pre-training from the Llama 2 family, primarily with the addition of Japanese language data. The tuned versions use supervised fine-tuning (SFT). Links to other models can be found in the index.

Open Source 7.0B ↓ 274
WeDLM-8B-Instruct logo
WeDLM-8B-Instruct
tencent

WeDLM-8B-Instruct is our flagship instruction-tuned diffusion language model that performs parallel decoding under standard causal attention, fine-tuned from WeDLM-8B.

Open Source 8.0B ↓ 274
Falcon3-10B-Base-1.58bit logo
Falcon3-10B-Base-1.58bit
tiiuae

--- library name: transformers tags: - bitnet - falcon3 base model: tiiuae/Falcon3-10B-Base license: other license name: falcon-llm-license license link: https://falconllm.tii.ae/falcon-terms-and-conditions.html ---

Open Source 10.0B ↓ 270
Gemma-2-Llama-Swallow-2b-pt-v0.1 logo
Gemma-2-Llama-Swallow-2b-pt-v0.1
tokyotech-llm

Gemma-2-Llama-Swallow series was built by continual pre-training on the gemma-2 models. Gemma 2 Swallow enhanced the Japanese language capabilities of the original Gemma 2 while retaining the English language capabilities. We use approximately 200 billion tokens that were sampled…

Open Source 2.0B ↓ 268
Yi-34B-Chat-8bits logo
Yi-34B-Chat-8bits
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 34.0B ↓ 267
AquilaCode-py logo
AquilaCode-py
BAAI

Aquila Language Model is the first open source language model that supports both Chinese and English knowledge, commercial license agreements, and compliance with domestic data regulations.

Code ↓ 266
Yi-6B-Chat-8bits logo
Yi-6B-Chat-8bits
01-ai

Building the Next Generation of Open-Source and Bilingual LLMs

Open Source 6.0B ↓ 261