LLM Models
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen). As of now, Qwen2.5-Coder has covered six mainstream model sizes, 0.5, 1.5, 3, 7, 14, 32 billion parameters, to meet the needs of different developers. Qwen2.5-Coder brings…
1. Model Summary 2. How to use 3. Evaluation 4. Training 5. Limitations 6. License
DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
General OCR Theory: Towards OCR-2.0 via a Unified End-to-end Model
Phi-3.5-vision is a lightweight, state-of-the-art open multimodal model built upon datasets which include - synthetic data and filtered publicly available websites - with a focus on very high-quality, reasoning dense data both on text and vision. The model belongs to the Phi-3 mo…
HunyuanOCR-1.5: Making Lightweight OCR VLMs Faster and Better
--- license: apache-2.0 language: - en pipeline tag: image-text-to-text tags: - multimodal - gui library name: transformers ---
🤗 huggingchat 📰 Tech Blog
Florence-2: Advancing a Unified Representation for a Variety of Vision Tasks
------------------------- ------------------------------------------------------------------------------- Developers Microsoft Research Description phi-4 is a state-of-the-art open model built upon a blend of synthetic datasets, data from filtered public domain websites, and acqu…