AI Agent Hub

LLM Models

Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.

1,877 models Compare
🤖
StepFun-Prover-Preview-32B
stepfun-ai

StepFun-Prover-Preview-32B is a theorem proving model developed by StepFun Team. It can iteratively refine the proof sketch via interacting with Lean4, and achieve 70.0% accuracy with Pass@1 on MiniF2F-test. Advanced usage examples can be seen in github.

Open Source 32.0B ↓ 131
🤖
JanusCoderV-8B
internlm

💻Github Repo • 🤗Model Collections • 📜Technical Report

Code 8.0B ↓ 129
🤖
Intern-S2-Preview-397B-FP8
internlm

💻Github Repo • 🤗HF Model Collections • 🤖ModelScope Collections • 💬Online Chat

Open Source 397.0B ↓ 129
🤖
StepFun-Prover-Preview-7B
stepfun-ai

StepFun-Prover-Preview-7B is a theorem proving model developed by StepFun Team. It can iteratively refine the proof sketch via interacting with Lean4, and achieve 66.0% accuracy with Pass@1 on MiniF2F-test. Advanced usage examples can be seen in github.

Open Source 7.0B ↓ 128
🤖
OREAL-DeepSeek-R1-Distill-Qwen-7B
internlm

- Arxiv - Github - Model Collection - Data

Reasoning 7.0B ↓ 127
🤖
GPT-JT-6B-v0
togethercomputer

python from transformers import pipeline

Open Source 6.0B ↓ 123
🤖
santacoder-cf
bigcode

Code ↓ 122
🤖
internlm2-wqx-20b
internlm

InternLM2-WQX-20B 🤗 | InternLM2-WQX-VL-20B 🤗

Open Source 20.0B ↓ 121
🤖
Swallow-70b-instruct-hf
tokyotech-llm

Our Swallow model has undergone continual pre-training from the Llama 2 family, primarily with the addition of Japanese language data. The tuned versions use supervised fine-tuning (SFT). Links to other models can be found in the index.

Open Source 70.0B ↓ 121
🤖
santacoder-ldf
bigcode

This is SantaCoder finetuned using the Line Diff Format introduced in OctoPack.

Code ↓ 121
🤖
internlm2-math-base-20b
internlm

State-of-the-art bilingual open-sourced Math reasoning LLMs. A solver , prover , verifier , augmentor .

Open Source 20.0B ↓ 119
🤖
Swallow-13b-instruct-v0.1
tokyotech-llm

Our Swallow model has undergone continual pre-training from the Llama 2 family, primarily with the addition of Japanese language data. The tuned versions use supervised fine-tuning (SFT). Links to other models can be found in the index.

Open Source 13.0B ↓ 119