AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,115 models for "Chat" Compare
🤖
LFM2-VL-450M
LiquidAI

LFM2‑VL is Liquid AI's first series of multimodal models, designed to process text and images with variable resolutions. Built on the LFM2 backbone, it is optimized for low-latency and edge AI applications.

Multimodal 0.45B ↓ 37.5K
🤖
internlm2_5-7b-chat
internlm

💻Github Repo • 🤔Reporting Issues • 📜Technical Report

Open Source 7.0B ↓ 37.3K
🤖
deepseek-llm-7b-base
deepseek-ai

[🏠Homepage] [🤖 Chat with DeepSeek LLM] [Discord] [Wechat(微信)]

Open Source 7.0B ↓ 37K
🤖
DeepSeek-V2
deepseek-ai

Model Download Evaluation Results Model Architecture API Platform License Citation

Open Source ↓ 37K
🤖
deepseek-coder-1.3b-instruct
deepseek-ai

[🏠Homepage] [🤖 Chat with DeepSeek Coder] [Discord] [Wechat(微信)]

Code 1.3B ↓ 36.1K
🤖
EXAONE-3.5-32B-Instruct
LGAI-EXAONE

We introduce EXAONE 3.5, a collection of instruction-tuned bilingual (English and Korean) generative models ranging from 2.4B to 32B parameters, developed and released by LG AI Research. EXAONE 3.5 language models include: 1) 2.4B model optimized for deployment on small or resour…

Open Source 32.0B ↓ 36.1K
🤖
internlm-chat-7b
internlm

InternLM has open-sourced a 7 billion parameter base model and a chat model tailored for practical scenarios. The model has the following characteristics: - It leverages trillions of high-quality tokens for training to establish a powerful knowledge base. - It supports an 8k cont…

Open Source 7.0B ↓ 35.3K
🤖
Seed-OSS-36B-Instruct
ByteDance-Seed

👋 Hi, everyone! We are ByteDance Seed Team.

Open Source 36.0B ↓ 35.1K
🤖
deepseek-moe-16b-chat
deepseek-ai

[🏠Homepage] [🤖 Chat with DeepSeek LLM] [Discord] [Wechat(微信)]

Open Source 16.4B ↓ 34.4K
🤖
DeepSeek-V3.1-Base
deepseek-ai

DeepSeek-V3.1 is a hybrid model that supports both thinking mode and non-thinking mode. Compared to the previous version, this upgrade brings improvements in multiple aspects:

Open Source ↓ 34.1K
🤖
LongCat-Flash-Chat
meituan-longcat

Model Introduction We introduce LongCat-Flash, a powerful and efficient language model with 560 billion total parameters, featuring an innovative Mixture-of-Experts (MoE) architecture. The model incorporates a dynamic computation mechanism that activates 18.6B∼31.3B parameters (a…

Open Source ↓ 34K
🤖
OLMo-2-1124-7B-SFT
allenai

Upon the initial release of OLMo-2 models, we realized the post-trained models did not share the pre-tokenization logic that the base models use. As a result, we have trained new post-trained models. The new models are available under the same names as the original models, but we…

Open Source 7.0B ↓ 33.8K