AI Agent Hub

LLM Models · allenai

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

53 models from allenai Compare
OLMo-2-0425-1B-Instruct logo
OLMo-2-0425-1B-Instruct
allenai

OLMo 2 1B Instruct April 2025 is post-trained variant of the allenai/OLMo-2-0425-1B-RLVR1 model, which has undergone supervised finetuning on an OLMo-specific variant of the Tülu 3 dataset, further DPO training on this dataset, and final RLVR training on this dataset. Tülu 3 is d…

open-source 1.0B ↓ 73.4K
OLMo-2-1124-7B-Instruct logo
OLMo-2-1124-7B-Instruct
allenai

Upon the initial release of OLMo-2 models, we realized the post-trained models did not share the pre-tokenization logic that the base models use. As a result, we have trained new post-trained models. The new models are available under the same names as the original models, but we…

open-source 7.0B ↓ 63.6K
OLMo-1B-hf logo
OLMo-1B-hf
allenai

OLMo is a series of O pen L anguage Mo dels designed to enable the science of language models. The OLMo models are trained on the Dolma dataset. We release all code, checkpoints, logs (coming soon), and details involved in training these models. This model has been converted from…

code 1.0B ↓ 51.5K
Molmo2-O-7B logo
Molmo2-O-7B
allenai

Molmo2 is a family of open vision-language models developed by the Allen Institute for AI (Ai2) that support image, video and multi-image understanding and grounding. Molmo2 models are trained on publicly available third party datasets as referenced in our technical report and Mo…

multimodal 7.0B ↓ 43.5K
OLMo-2-1124-7B-SFT logo
OLMo-2-1124-7B-SFT
allenai

Upon the initial release of OLMo-2 models, we realized the post-trained models did not share the pre-tokenization logic that the base models use. As a result, we have trained new post-trained models. The new models are available under the same names as the original models, but we…

open-source 7.0B ↓ 33.8K
truthfulqa-truth-judge-llama2-7B logo
truthfulqa-truth-judge-llama2-7B
allenai

This model is built based on LLaMa2 7B in replacement of the truthfulness/informativeness judge models that were originally introduced in the TruthfulQA paper. That model is based on OpenAI's Curie engine using their finetuning API. However, as of February 08, 2024, OpenAI has ta…

open-source 7.0B ↓ 26.7K
truthfulqa-info-judge-llama2-7B logo
truthfulqa-info-judge-llama2-7B
allenai

This model is built based on LLaMa2 7B in replacement of the truthfulness/informativeness judge models that were originally introduced in the TruthfulQA paper. That model is based on OpenAI's Curie engine using their finetuning API. However, as of February 08, 2024, OpenAI has ta…

open-source 7.0B ↓ 24.5K
Olmo-Hybrid-7B logo
Olmo-Hybrid-7B
allenai

We expand on our Olmo model series by introducing Olmo Hybrid, a new 7B hybrid RNN model in the Olmo family. Olmo Hybrid dramatically outperforms Olmo 3 in final performance, consistently showing roughly 2x data efficiency on core evals over the course of our pretraining run. We…

open-source 7.0B ↓ 21.3K
olmOCR-7B-0825-FP8 logo
olmOCR-7B-0825-FP8
allenai

Quantized to FP8 Version of olmOCR-7B-0825, using llmcompressor.

multimodal 7.0B ↓ 18.6K
Llama-3.1-Tulu-3-8B-SFT logo
Llama-3.1-Tulu-3-8B-SFT
allenai

Tülu3 is a leading instruction following model family, offering fully open-source data, code, and recipes designed to serve as a comprehensive guide for modern post-training techniques. Tülu3 is designed for state-of-the-art performance on a diversity of tasks in addition to chat…

open-source 8.0B ↓ 17.4K
OLMoE-1B-7B-0924-Instruct logo
OLMoE-1B-7B-0924-Instruct
allenai

OLMoE-1B-7B-Instruct is a Mixture-of-Experts LLM with 1B active and 7B total parameters released in September 2024 (0924) that has been adapted via SFT and DPO from OLMoE-1B-7B. It yields state-of-the-art performance among models with a similar cost (1B) and is competitive with m…

open-source 1.0B ↓ 16.7K
OLMoE-1B-7B-0125 logo
OLMoE-1B-7B-0125
allenai

OLMoE-1B-7B is a Mixture-of-Experts LLM with 1B active and 7B total parameters released in January 2025 (0125) that is 100% open-source. It is an improved version of OLMoE-09-24, see the paper appendix for details.

open-source 6.92B ↓ 14.9K