AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

362 models for "Fine-tuned" Compare
Qwen-SEA-LION-v4.5-27B-IT logo
Qwen-SEA-LION-v4.5-27B-IT
aisingapore

SEA-LION is a collection of Large Language Models (LLMs) which have been pretrained and instruct-tuned for the Southeast Asia (SEA) region.

Multimodal 27.0B ↓ 1.1K
blip2-opt-6.7b-coco logo
blip2-opt-6.7b-coco
Salesforce

BLIP-2 model, leveraging OPT-6.7b (a large language model with 6.7 billion parameters). It was introduced in the paper BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models by Li et al. and first released in this repository.

Open Source 6.7B ↓ 999
blip2-flan-t5-xxl logo
blip2-flan-t5-xxl
Salesforce

BLIP-2 model, leveraging Flan T5-xxl (a large language model). It was introduced in the paper BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models by Li et al. and first released in this repository.

Open Source ↓ 977
RedPajama-INCITE-Chat-3B-v1 logo
RedPajama-INCITE-Chat-3B-v1
togethercomputer

RedPajama-INCITE-Chat-3B-v1 was developed by Together and leaders from the open-source AI community including Ontocord.ai, ETH DS3Lab, AAI CERC, Université de Montréal, MILA - Québec AI Institute, Stanford Center for Research on Foundation Models (CRFM), Stanford Hazy Research re…

Open Source 3.0B ↓ 976
Llama-2-7B-32K-Instruct logo
Llama-2-7B-32K-Instruct
togethercomputer

Llama-2-7B-32K-Instruct is an open-source, long-context chat model finetuned from Llama-2-7B-32K, over high-quality instruction and chat data. We built Llama-2-7B-32K-Instruct with less than 200 lines of Python script using Together API, and we also make the recipe fully availabl…

Open Source 7.0B ↓ 905
Llama-2-70b-hf logo
Llama-2-70b-hf
NousResearch

Llama 2 Llama 2 is a collection of pretrained and fine-tuned generative text models ranging in scale from 7 billion to 70 billion parameters. This is the repository for the 70B pretrained model, converted for the Hugging Face Transformers format. Links to other models can be foun…

Open Source 70.0B ↓ 899
open-calm-7b logo
open-calm-7b
cyberagent

OpenCALM is a suite of decoder-only language models pre-trained on Japanese datasets, developed by CyberAgent, Inc.

Open Source 7.0B ↓ 895
RedPajama-INCITE-7B-Chat logo
RedPajama-INCITE-7B-Chat
togethercomputer

RedPajama-INCITE-7B-Chat was developed by Together and leaders from the open-source AI community including Ontocord.ai, ETH DS3Lab, AAI CERC, Université de Montréal, MILA - Québec AI Institute, Stanford Center for Research on Foundation Models (CRFM), Stanford Hazy Research resea…

Open Source 7.0B ↓ 885
RedPajama-INCITE-Instruct-3B-v1 logo
RedPajama-INCITE-Instruct-3B-v1
togethercomputer

RedPajama-INCITE-Instruct-3B-v1 was developed by Together and leaders from the open-source AI community including Ontocord.ai, ETH DS3Lab, AAI CERC, Université de Montréal, MILA - Québec AI Institute, Stanford Center for Research on Foundation Models (CRFM), Stanford Hazy Researc…

Open Source 3.0B ↓ 849
ERNIE-4.5-VL-424B-A47B-Base-PT logo
ERNIE-4.5-VL-424B-A47B-Base-PT
baidu

[!NOTE] Note: " -Paddle " models use PaddlePaddle weights, while " -PT " models use Transformer-style PyTorch weights.

Multimodal 424.0B ↓ 840
Qwen-SEA-Guard-4B-2602 logo
Qwen-SEA-Guard-4B-2602
aisingapore

SEA-Guard is a collection of safety-focused Large Language Models (LLMs) built upon the SEA-LION family, designed specifically for the Southeast Asia (SEA) region.

Open Source 4.0B ↓ 822
Falcon-E-3B-Instruct logo
Falcon-E-3B-Instruct
tiiuae

0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation

Open Source 3.0B ↓ 821