AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

362 models for "Fine-tuned" Compare
Kimi-K2.6-NVFP4 logo
Kimi-K2.6-NVFP4
nvidia

Description: The NVIDIA Kimi-K2.6-NVFP4 model is the quantized version of the Moonshot AI's Kimi-K2.6 model, which is an auto-regressive language model that uses an optimized transformer architecture. For more information, please check here. The NVIDIA Kimi-K2.6 NVFP4 model is qu…

Open Source ↓ 153.2K
Mistral-7B-Instruct-v0.1 logo
Mistral-7B-Instruct-v0.1
mistralai

py from mistral common.tokens.tokenizers.mistral import MistralTokenizer from mistral common.protocol.instruct.messages import UserMessage from mistral common.protocol.instruct.request import ChatCompletionRequest

Open Source 7.0B ↓ 152.3K
NVIDIA-Nemotron-Parse-v1.1 logo
NVIDIA-Nemotron-Parse-v1.1
nvidia

NVIDIA Nemotron Parse v1.1 is designed to understand document semantics and extract text and tables elements with spatial grounding. Given an image, NVIDIA Nemotron Parse v1.1 produces structured annotations, including formatted text, bounding-boxes and the corresponding semantic…

Multimodal 0.885B ↓ 148.9K
opt-350m logo
opt-350m
facebook

OPT : Open Pre-trained Transformer Language Models

Open Source ↓ 139.2K
Llama-2-7b-hf logo
Llama-2-7b-hf
NousResearch

Llama 2 Llama 2 is a collection of pretrained and fine-tuned generative text models ranging in scale from 7 billion to 70 billion parameters. This is the repository for the 7B pretrained model, converted for the Hugging Face Transformers format. Links to other models can be found…

Open Source 7.0B ↓ 138.6K
OLMoE-1B-7B-0125-Instruct logo
OLMoE-1B-7B-0125-Instruct
allenai

OLMoE-1B-7B-0125-Instruct January 2025 is post-trained variant of the OLMoE-1B-7B January 2025 model, which has undergone supervised finetuning on an OLMo-specific variant of the Tülu 3 dataset and further DPO training on this dataset, and finally RLVR training using this data. T…

Open Source 1.0B ↓ 134.8K
deepseek-coder-6.7b-instruct logo
deepseek-coder-6.7b-instruct
deepseek-ai

[🏠Homepage] [🤖 Chat with DeepSeek Coder] [Discord] [Wechat(微信)]

Code 6.7B ↓ 119.5K
Mage-VL logo
Mage-VL
microsoft

Mage-VL An Efficient Codec-Native Streaming Multimodal Foundation Model

Multimodal 4.0B ↓ 117.8K
InternVL3-8B logo
InternVL3-8B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 7.94B ↓ 110.4K
paligemma-3b-ft-cococap-448 logo
paligemma-3b-ft-cococap-448
google

Multimodal 3.0B ↓ 106.3K
NVIDIA-Nemotron-Parse-2.0 logo
NVIDIA-Nemotron-Parse-2.0
nvidia

Description: NVIDIA Nemotron Parse 2.0 transforms document images into structured, machine-readable representations with text, layout classes, bounding boxes, and reading-order information. Given a Red, Green, Blue (RGB) document image and a task prompt, the model produces format…

Multimodal 0.905B ↓ 101.7K
InternVL3-1B logo
InternVL3-1B
OpenGVLab

[\[📂 GitHub\]](https://github.com/OpenGVLab/InternVL) [\[📜 InternVL 1.0\]](https://huggingface.co/papers/2312.14238) [\[📜 InternVL 1.5\]](https://huggingface.co/papers/2404.16821) [\[📜 InternVL 2.5\]](https://huggingface.co/papers/2412.05271) [\[📜 InternVL2.5-MPO\]](https://huggi…

Multimodal 0.94B ↓ 99.9K