AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

362 models for "Fine-tuned" Compare
UltraLM-65b logo
UltraLM-65b
openbmb

This is UltraLM-65b delta weights, a chat language model trained upon UltraChat

Open Source 65.0B ↓ 347
Falcon-E-3B-Base-prequantized logo
Falcon-E-3B-Base-prequantized
tiiuae

0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation

Open Source 3.0B ↓ 346
Nous-Puffin-70B logo
Nous-Puffin-70B
NousResearch

Based off Puffin 13B which was the first commercially available language model released by Nous Research!

Open Source 70.0B ↓ 344
Pythia-Chat-Base-7B logo
Pythia-Chat-Base-7B
togethercomputer

Feel free to try out our OpenChatKit feedback app!

Open Source 7.0B ↓ 342
llama-30b-instruct logo
llama-30b-instruct
upstage

Developed by : Upstage Backbone Model : LLaMA Variations : It has different model parameter sizes and sequence lengths: 30B/1024, 30B/2048, 65B/1024 Language(s) : English Library : HuggingFace Transformers License : This model is under a Non-commercial Bespoke License and governe…

Open Source 30.0B ↓ 340
Nous-Capybara-7B-V1 logo
Nous-Capybara-7B-V1
NousResearch

MUCH BETTER MISTRAL BASED VERSION IS OUT NOW AS CAPYBARA V1.9

Open Source 7.0B ↓ 319
StableBeluga1-Delta logo
StableBeluga1-Delta
stabilityai

Stable Beluga 1 is a Llama65B model fine-tuned on an Orca style Dataset

Open Source ↓ 307
llama-65b-instruct logo
llama-65b-instruct
upstage

Developed by : Upstage Backbone Model : LLaMA Variations : It has different model parameter sizes and sequence lengths: 30B/1024, 30B/2048, 65B/1024 Language(s) : English Library : HuggingFace Transformers License : This model is under a Non-commercial Bespoke License and governe…

Open Source 65.0B ↓ 299
SmolVLM-500M-Base logo
SmolVLM-500M-Base
HuggingFaceTB

This is the model card of a 🤗 transformers model that has been pushed on the Hub. This model card has been automatically generated.

Multimodal ↓ 291
GPT-NeoXT-Chat-Base-20B logo
GPT-NeoXT-Chat-Base-20B
togethercomputer

Feel free to try out our OpenChatKit feedback app!

Open Source 20.0B ↓ 287
SEA-LION-v1-7B-IT logo
SEA-LION-v1-7B-IT
aisingapore

SEA-LION is a collection of Large Language Models (LLMs) which has been pretrained and instruct-tuned for the Southeast Asia (SEA) region. The sizes of the models range from 3 billion to 7 billion parameters.

Open Source 7.5B ↓ 276
WeDLM-8B-Instruct logo
WeDLM-8B-Instruct
tencent

WeDLM-8B-Instruct is our flagship instruction-tuned diffusion language model that performs parallel decoding under standard causal attention, fine-tuned from WeDLM-8B.

Open Source 8.0B ↓ 274