LLM Models
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
For transformers versions v4.40.0 or newer, we suggest using OLMo 7B HF instead.
Model Summary: Granite-4.0-1B-Base is a lightweight decoder-only language model designed for scenarios where efficiency and speed are critical. They can run on resource-constrained devices such as smartphones or IoT hardware, enabling offline and privacy-preserving applications.…
Model Summary: Granite-4.0-1B is a lightweight instruct model finetuned from Granite-4.0-1B-Base using a combination of open source instruction datasets with permissive license and internally collected synthetic datasets. This model is developed using a diverse set of techniques…
📣 Update [10-07-2025]: Added a default system prompt to the chat template to guide the model towards more professional, accurate, and safe responses.
LFM2.5-VL-1.6B-Extract extracts user-defined fields from images and returns them as JSON . It is Liquid AI's first vision model in the Liquid Nanos collection—compact, task-specific models built for production workflows—and extends the Extract family alongside LFM2-1.2B-Extract f…
Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.
--- license: apache-2.0 language: - en pipeline tag: image-text-to-text tags: - multimodal library name: transformers ---
Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.
Florence-2: Advancing a Unified Representation for a Variety of Vision Tasks
Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.
--- license name: qwen-research license link: https://huggingface.co/Qwen/Qwen2.5-VL-3B-Instruct/blob/main/LICENSE language: - en pipeline tag: image-text-to-text tags: - multimodal library name: transformers ---