DeepSeek: DeepSeek V4 Flash 0731 (batch)
About this model
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
Technical Specs
- Architecture: text->text
- Context Window: 1,048,576 tokens
- Input Modalities: text
Hardware Requirements
- API-only (no local hardware needed)
Pricing
| Input | Output | Currency |
|---|---|---|
| 0.14 | 0.28 | USD |
Related Models
OpenAI's flagship model, GPT-4 is a large-scale multimodal language model capable of solving difficult problems with greater accuracy than previous models due to its broader general knowledge and advanced reasoning...
GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.
One of the highest performing and most popular fine-tunes of Llama 2 13B, with rich descriptions and roleplay. #merge
A recreation trial of the original MythoMax-L2-B13 but with updated models. #merge