LLM Models
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation
AHN: Artificial Hippocampus Networks for Efficient Long-Context Modeling
📖 Paper 🏠 Github 🤗 Spatial-SSRL-7B Model 🤗 Spatial-SSRL-3B Model 🤗 Spatial-SSRL-Qwen3VL-4B Model 🤗 Spatial-SSRL-81k Dataset 📰 Daily Paper
Authors: Erik Nijkamp\ , Hiroaki Hayashi\ , Yingbo Zhou, Caiming Xiong
🦉GitHub 💬WeChat 百川API支持搜索增强和192K长窗口,新增百川搜索增强知识库、限时免费! 🚀 百川大模型在线对话平台 已正式向公众开放 🎉
An ablation of OctoCoder released for research purposes. Generally use OctoCoder, which performs better. Steps: 80
Qianfan-VL: Domain-Enhanced Universal Vision-Language Models
An ablation of OctoCoder released for research purposes. Generally use OctoCoder, which performs better. Steps: 50
An ablation of OctoCoder released for research purposes. Generally use OctoCoder, which performs better. Steps: 45
This model is a version of bigscience/bloom-7b1 post-processed to be run at home using the Petals swarm.
Reinforcement learning (RL) (e.g., GRPO) helps with grounding because of its inherent objective alignment—rewarding successful clicks—rather than encouraging long textual Chain-of-Thought (CoT) reasoning. Unlike approaches that rely heavily on verbose CoT reasoning, GRPO directly…