AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,114 models for "Chat" Compare
🤖
Ling-lite
inclusionAI

Ling is a MoE LLM provided and open-sourced by InclusionAI. We introduce two different sizes, which are Ling-Lite and Ling-Plus. Ling-Lite has 16.8 billion parameters with 2.75 billion activated parameters, while Ling-Plus has 290 billion parameters with 28.8 billion activated pa…

Open Source ↓ 1.7K
🤖
GPT-OSS-Swallow-20B-RL-v0.1
tokyotech-llm

GPT-OSS-Swallow v0.1 is a family of large language models available in 20B and 120B parameter sizes. Built as bilingual Japanese-English models, they were developed through Continual Pre-Training (CPT), Supervised Fine-Tuning (SFT), and Reinforcement Learning with Verifiable Rewa…

Open Source 20.0B ↓ 1.7K
🤖
Qwen3-Swallow-8B-SFT-v0.2
tokyotech-llm

Qwen3-Swallow v0.2 is a family of large language models available in 8B , 30B-A3B , and 32B parameter sizes. Built as bilingual Japanese-English models, they were developed through Continual Pre-Training (CPT), Supervised Fine-Tuning (SFT), and Reinforcement Learning with Verifia…

Open Source 8.0B ↓ 1.7K
🤖
deepseek-coder-33b-base
deepseek-ai

[🏠Homepage] [🤖 Chat with DeepSeek Coder] [Discord] [Wechat(微信)]

Code 33.0B ↓ 1.7K
🤖
Baichuan-M2-32B
baichuan-inc

This repository contains the model presented in Baichuan-M2: Scaling Medical Capability with Large Verifier System.

Open Source 32.0B ↓ 1.7K
🤖
Mistral-Nemo-Japanese-Instruct-2408
cyberagent

This is a Japanese continually pre-trained model based on mistralai/Mistral-Nemo-Instruct-2407.

Open Source ↓ 1.7K
🤖
SingGuard-NSFA-0.8B
inclusionAI

SingGuard-NSFA: Extensible Guardrails for Agentic AI via Generative Reasoning and Real-Time Classification

Open Source 0.8B ↓ 1.6K
🤖
LLaDA2.2-flash
inclusionAI

LLaDA2.2-flash is an agent-oriented diffusion language model in the LLaDA2 series. By introducing Levenshtein Editing (with DELETE and INSERT control tokens) to diffusion language modeling, it represents the LLaDA2 series' first step in agentic applications, including long-contex…

Open Source ↓ 1.6K
🤖
Falcon-H1-Tiny-90M-Instruct
tiiuae

0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation

Open Source ↓ 1.6K
🤖
InternVL3_5-2B-Instruct
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 2.0B ↓ 1.6K
🤖
InternVL3_5-4B-Pretrained
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 4.0B ↓ 1.6K
🤖
SEA-LION-v1-7B-IT
aisingapore

SEA-LION is a collection of Large Language Models (LLMs) which has been pretrained and instruct-tuned for the Southeast Asia (SEA) region. The sizes of the models range from 3 billion to 7 billion parameters.

Open Source 7.0B ↓ 1.6K