DeepSeek-R1-Distill-Llama-70B
About this model
DeepSeek R1 Distill Llama 70B is a distilled large language model based on Llama-3.3-70B-Instruct, using outputs from DeepSeek R1. The model combines advanced distillation techniques to achieve high performance across...
Technical Specs
- Architecture: Transformer
- Context Window: 8,192 tokens
- Input Modalities: text
Hardware Requirements
- API-only (no local hardware needed)
Pricing
| Input | Output | Currency |
|---|---|---|
| 0.7999999999999999 | 0.7999999999999999 | USD |
Related Models
DeepSeek-R1-Distill-Qwen-7B
DeepSeek-R1-Distill-Qwen-7B
Reasoning
↓ 369.5K
DeepSeek-R1-Distill-Qwen-1.5B
DeepSeek-R1-Distill-Qwen-1.5B
Reasoning
↓ 486.6K
DeepSeek-R1-Distill-Llama-8B
DeepSeek-R1-Distill-Llama-8B
Reasoning
↓ 389.2K
DeepSeek-R1-Distill-Qwen-14B
DeepSeek-R1-Distill-Qwen-14B
Reasoning
↓ 486.8K