AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

362 models for "Fine-tuned" Compare
mistral-coreml logo
mistral-coreml
apple

[!IMPORTANT] ❗ This repo requires the use of the macOS Sequoia (15) Developer Beta to utilize the latest and greatest CoreML has to offer! Sign up for the Apple Beta Software Program here to get access. Check out the companion blog post to learn more about what's new in iOS 18 &…

Open Source ↓ 93
distill-bloom-1b3 logo
distill-bloom-1b3
bigscience

WARNING: This is an intermediary checkpoint and WIP project. It is not fully trained yet. You might want to use Bloom-1B3 if you want a model that has completed training. This model is a distilled version of Bloom-1B3

Open Source ↓ 90
distill-bloom-1b3-10x logo
distill-bloom-1b3-10x
bigscience

WARNING: This is an intermediary checkpoint and WIP project. It is not fully trained yet. You might want to use Bloom-1B3 if you want a model that has completed training. This model is a distilled version of Bloom-1B3 (10x distillation)

Open Source ↓ 85
BaichuanMed-OCR-72B logo
BaichuanMed-OCR-72B
baichuan-inc

BaichuanMed-OCR-72B is a model fine-tuned from the Qwen2.5-VL-72B-Instruct with our constructed and curated medical report datasets consists of medical report images and related questions and answers (QAs). It has been specifically adapted to perform Optical Character Recognition…

Open Source 72.0B ↓ 81
Gemma-SEA-LION-v4-27B-IT-NVFP4 logo
Gemma-SEA-LION-v4-27B-IT-NVFP4
aisingapore

Model Card for Gemma-SEA-LION-v4-27B-IT-NVFP4

Open Source 27.0B ↓ 75
opencole-typographylmm-llava-v1.5-7b-lora logo
opencole-typographylmm-llava-v1.5-7b-lora
cyberagent

This model is based on LLaVA1.5-7b. The model is finetuned with LoRA on OpenCOLE1.0 dataset to generate text layouts.

Open Source 7.0B ↓ 65
Qwen-SEA-LION-v4-32B-IT-8BIT logo
Qwen-SEA-LION-v4-32B-IT-8BIT
aisingapore

Qwen-SEA-LION-v4-32B-IT-8BIT (GPTQ model)

Open Source 32.0B ↓ 65
Gemma-SEA-LION-v4-27B-IT-FP8-Dynamic logo
Gemma-SEA-LION-v4-27B-IT-FP8-Dynamic
aisingapore

Model Card for Gemma-SEA-LION-v4-27B-IT-FP8-Dynamic

Open Source 27.0B ↓ 59
WangchanLION-v3-IT logo
WangchanLION-v3-IT
aisingapore

WangchanLION is a joint effort between VISTEC and AI Singapore to develop a Thai-specific collection of Large Language Models (LLMs), pre-trained for Southeast Asian (SEA) languages, and instruct-tuned specifically for the Thai language.

Open Source ↓ 55
Gemma2-9b-WangchanLIONv2-instruct logo
Gemma2-9b-WangchanLIONv2-instruct
aisingapore

WangchanLION is a joint effort between VISTEC and AI Singapore to develop a Thai-specific collection of Large Language Models (LLMs), pre-trained for Southeast Asian (SEA) languages, and instruct-tuned specifically for the Thai language.

Open Source 9.0B ↓ 54
tiny-aya-en-thinker logo
tiny-aya-en-thinker
CohereLabs

Reasoning 3.35B ↓ 40
SEA-LION-v1-7B-IT-GPTQ logo
SEA-LION-v1-7B-IT-GPTQ
aisingapore

SEA-LION is a collection of Large Language Models (LLMs) which has been pretrained and instruct-tuned for the Southeast Asia (SEA) region. The sizes of the models range from 3 billion to 7 billion parameters.

Open Source 7.0B ↓ 29