AI Agent Hub

LLM Models · NousResearch

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

67 models from NousResearch Compare
Yarn-Mistral-7b-128k logo
Yarn-Mistral-7b-128k
NousResearch

Nous-Yarn-Mistral-7b-128k is a state-of-the-art language model for long context, further pretrained on long context data for 1500 steps using the YaRN extension method. It is an extension of Mistral-7B-v0.1 and supports a 128k token context window.

Open Source 7.0B ↓ 1K
DeepHermes-3-Llama-3-8B-Preview logo
DeepHermes-3-Llama-3-8B-Preview
NousResearch

DeepHermes 3 Preview is the latest version of our flagship Hermes series of LLMs by Nous Research, and one of the first models in the world to unify Reasoning (long chains of thought that improve answer accuracy) and normal LLM response modes into one model. We have also improved…

Open Source 8.0B ↓ 931
Llama-2-70b-hf logo
Llama-2-70b-hf
NousResearch

Llama 2 Llama 2 is a collection of pretrained and fine-tuned generative text models ranging in scale from 7 billion to 70 billion parameters. This is the repository for the 70B pretrained model, converted for the Hugging Face Transformers format. Links to other models can be foun…

Open Source 70.0B ↓ 899
Nous-Hermes-2-Mistral-7B-DPO logo
Nous-Hermes-2-Mistral-7B-DPO
NousResearch

Nous Hermes 2 on Mistral 7B DPO is the new flagship 7B Hermes! This model was DPO'd from Teknium/OpenHermes-2.5-Mistral-7B and has improved across the board on all benchmarks tested - AGIEval, BigBench Reasoning, GPT4All, and TruthfulQA.

Open Source 7.0B ↓ 848
k2-merged-3.5T-bf16 logo
k2-merged-3.5T-bf16
NousResearch

Experimental frankenmerge of Kimi K2-07, 09 and Base

Open Source ↓ 678
DeepHermes-3-Llama-3-3B-Preview logo
DeepHermes-3-Llama-3-3B-Preview
NousResearch

DeepHermes 3 Preview is the latest version of our flagship Hermes series of LLMs by Nous Research, and one of the first models in the world to unify Reasoning (long chains of thought that improve answer accuracy) and normal LLM response modes into one model. We have also improved…

Open Source 3.0B ↓ 638
Nous-Hermes-13b logo
Nous-Hermes-13b
NousResearch

Nous-Hermes-13b is a state-of-the-art language model fine-tuned on over 300,000 instructions. This model was fine-tuned by Nous Research, with Teknium and Karan4D leading the fine tuning process and dataset curation, Redmond AI sponsoring the compute, and several other contributo…

Open Source 13.0B ↓ 614
Nous-Capybara-7B-V1.9 logo
Nous-Capybara-7B-V1.9
NousResearch

This is currently the best 7B version of Capybara to use

Open Source 7.0B ↓ 469
Yarn-Llama-2-13b-128k logo
Yarn-Llama-2-13b-128k
NousResearch

Nous-Yarn-Llama-2-13b-128k is a state-of-the-art language model for long context, further pretrained on long context data for 600 steps. This model is the Flash Attention 2 patched version of the original model: https://huggingface.co/conceptofmind/Yarn-Llama-2-13b-128k

Open Source 13.0B ↓ 460
Yarn-Llama-2-7b-64k logo
Yarn-Llama-2-7b-64k
NousResearch

Nous-Yarn-Llama-2-7b-64k is a state-of-the-art language model for long context, further pretrained on long context data for 400 steps. This model is the Flash Attention 2 patched version of the original model: https://huggingface.co/conceptofmind/Yarn-Llama-2-7b-64k

Open Source 7.0B ↓ 459
NousCoder-14B logo
NousCoder-14B
NousResearch

We introduce NousCoder-14B , a competitive programming model post-trained on Qwen3-14B via reinforcement learning. On LiveCodeBench v6 (08/01/2024 - 05/01/2025), we achieve a Pass@1 accuracy of 67.87\%, up 7.08\% from the baseline Pass@1 accuracy of 60.79\% of Qwen3-14B. We train…

Code 14.0B ↓ 458
Nous-Capybara-34B logo
Nous-Capybara-34B
NousResearch

This is trained on the Yi-34B model with 200K context length, for 3 epochs on the Capybara dataset!

Open Source 34.0B ↓ 454