Google: Gemma 3 12B
8.4 / 10
131.1K context
Proprietary
About this model
Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...
Benchmark Scores
MATH
83.8
Technical Specs
- Architecture: text+image->text
- Context Window: 131,072 tokens
- Input Modalities: text
Hardware Requirements
- API-only (no local hardware needed)
Pricing
| Input | Output | Currency |
|---|---|---|
| 0.05 / 1M tokens | 0.15 / 1M tokens | USD |