Qwen: Qwen3 235B A22B Thinking 2507
8.6 / 10
131.1K context
Proprietary
About this model
Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...
Benchmark Scores
BBH
88.87
DROP
88.7
Aider
59.6
GSM8K
96.4
MATH-500
98.0
AIME-2024
85.7
Technical Specs
- Architecture: text->text
- Context Window: 131,072 tokens
- Input Modalities: text
Hardware Requirements
- API-only (no local hardware needed)
Pricing
| Input | Output | Currency |
|---|---|---|
| 0.23 / 1M tokens | 2.3 / 1M tokens | USD |