LLM Models
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.
We introduce K-EXAONE 2.0 , a frontier-scale multilingual language model developed by LG AI Research. K-EXAONE 2.0 was scaled to more than three times the size of its predecessor through upcycling, followed by continual pretraining, difficulty-focused mid-training, and post-train…
We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.
We introduce K-EXAONE 2.0 , a frontier-scale multilingual language model developed by LG AI Research. K-EXAONE 2.0 was scaled to more than three times the size of its predecessor through upcycling, followed by continual pretraining, difficulty-focused mid-training, and post-train…
We introduce K-EXAONE 2.0 , a frontier-scale multilingual language model developed by LG AI Research. K-EXAONE 2.0 was scaled to more than three times the size of its predecessor through upcycling, followed by continual pretraining, difficulty-focused mid-training, and post-train…
We introduce K-EXAONE 2.0 , a frontier-scale multilingual language model developed by LG AI Research. K-EXAONE 2.0 was scaled to more than three times the size of its predecessor through upcycling, followed by continual pretraining, difficulty-focused mid-training, and post-train…
Qwen2.5 Bakeneko 32B Instruct (rinna/qwen2.5-bakeneko-32b-instruct)
DeepSeek R1 Distill Qwen2.5 Bakeneko 32B (rinna/deepseek-r1-distill-qwen2.5-bakeneko-32b)
QwQ Bakeneko 32B (rinna/qwq-bakeneko-32b)
Qwen2.5 Bakeneko 32B (rinna/qwen2.5-bakeneko-32b)
DeepSeek R1 Distill Qwen2.5 Bakeneko 32B AWQ (rinna/deepseek-r1-distill-qwen2.5-bakeneko-32b-awq)
Qwen2.5 Bakeneko 32B Instruct AWQ (rinna/qwen2.5-bakeneko-32b-instruct-awq)