AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

362 models for "Fine-tuned" Compare
InternVL3-38B-AWQ logo
InternVL3-38B-AWQ
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 38.0B ↓ 792
Qwen-SEA-Guard-8B-2602 logo
Qwen-SEA-Guard-8B-2602
aisingapore

SEA-Safeguard is a collection of safety-focused Large Language Models (LLMs) built upon the SEA-LION family, designed specifically for the Southeast Asia (SEA) region.

Open Source 8.0B ↓ 787
open-calm-1b logo
open-calm-1b
cyberagent

OpenCALM is a suite of decoder-only language models pre-trained on Japanese datasets, developed by CyberAgent, Inc.

Open Source 1.0B ↓ 733
RedPajama-INCITE-7B-Instruct logo
RedPajama-INCITE-7B-Instruct
togethercomputer

RedPajama-INCITE-7B-Instruct was developed by Together and leaders from the open-source AI community including Ontocord.ai, ETH DS3Lab, AAI CERC, Université de Montréal, MILA - Québec AI Institute, Stanford Center for Research on Foundation Models (CRFM), Stanford Hazy Research r…

Open Source 7.0B ↓ 729
Falcon-E-3B-Base logo
Falcon-E-3B-Base
tiiuae

0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation

Open Source 3.0B ↓ 722
open-calm-medium logo
open-calm-medium
cyberagent

OpenCALM is a suite of decoder-only language models pre-trained on Japanese datasets, developed by CyberAgent, Inc.

Open Source ↓ 719
Qwen-SEA-LION-v4-8B-VL logo
Qwen-SEA-LION-v4-8B-VL
aisingapore

SEA-LION is a collection of Large Language Models (LLMs) which have been pretrained and instruct-tuned for the Southeast Asia (SEA) region.

Multimodal 8.0B ↓ 715
StableBeluga2 logo
StableBeluga2
stabilityai

Use Stable Chat (Research Preview) to test Stability AI's best language models for free

Open Source ↓ 705
InternVL3-78B-AWQ logo
InternVL3-78B-AWQ
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 78.0B ↓ 702
Falcon-E-1B-Base logo
Falcon-E-1B-Base
tiiuae

0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation

Open Source 1.0B ↓ 689
calm2-7b-chat logo
calm2-7b-chat
cyberagent

CyberAgentLM2-Chat is a fine-tuned model of CyberAgentLM2 for dialogue use cases.

Open Source 7.0B ↓ 688
ERNIE-4.5-VL-424B-A47B-PT logo
ERNIE-4.5-VL-424B-A47B-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Multimodal 424.0B ↓ 685