AI Agent Hub

Compare LLM Models

Select 2-5 models to compare side by side.

Select models to compare 0 / 5 selected
Select 2-5 models
Attribute
farbodtavakkoli
Company farbodtavakkoli
Release Date 2026-07-24
Parameters (B) 31
Architecture Transformer (Gemma 4 hybrid sliding-window and global attention)
Context Window 256000
Input Modalities text
Open Source Yes
License
Score 8.4
VRAM (GB) 70
Compute FP16 inference needs ~70 GB VRAM (e.g. A100 80GB); INT4 quantization runs on ~17–24 GB (RTX 4090). Training used on-prem AMD MI355X GPUs with Dell infrastructure; data processing used ~530 AMD MI300X GPUs on Microsoft Azure.
Benchmark: GSM8K 82
Benchmark: HumanEval 80
Benchmark: MMLU 85.2
Pricing (Input)
Pricing (Output)