LLM Models
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
⚠️ This project is intended for research and educational purposes only . Any use for illegal data access, system interference, or unlawful activities is strictly prohibited. Please review our Terms of Use carefully.
Our Swallow model has undergone continual pre-training from the Llama 2 family, primarily with the addition of Japanese language data. The tuned versions use supervised fine-tuning (SFT). Links to other models can be found in the index.
Sachin Mehta, Mohammad Hossein Sekhavat, Qingqing Cao, Maxwell Horton, Yanzi Jin, Chenfan Sun, Iman Mirzadeh, Mahyar Najibi, Dmitry Belenko, Peter Zatloukal, Mohammad Rastegari
SEA-LION is a collection of Large Language Models (LLMs) which have been pretrained and instruct-tuned for the Southeast Asia (SEA) region.
--- inference: false language: - en tags: - instruction-finetuning pretty name: JudgeLM-100K task categories: - text-generation ---
From Inquiry to Decision: Building Trustworthy Medical AI
💻Github Repo • 🤗Model Collections • 💬Online Chat
OLMo-Bitnet-1B is a 1B parameter model trained using the method described in The Era of 1-bit LLMs: All Large Language Models are in 1.58 Bits.
Use Stable Chat (Research Preview) to test Stability AI's best language models for free
0. TL;DR 1. Model Details 2. Training Details 3. Usage 4. Evaluation 5. Citation
State-of-the-art bilingual open-sourced Math reasoning LLMs. A solver , prover , verifier , augmentor .
Falcon3 family of Open Foundation Models is a set of pretrained and instruct LLMs ranging from 1B to 10B.