Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
| Provider | Input $/M | Output $/M | Blended $/M ↑ | vs Median |
|---|---|---|---|---|
| ↓Reka | $0.06 | $0.20 | $0.14 | -36.8% |
| Darkbloom | $0.04 | $0.22 | $0.15 | -34.7% |
| NextBit | $0.08 | $0.26 | $0.18 | -19.5% |
| Cloudflare | $0.10 | $0.30 | $0.22 | -3.5% |
| CoreWeave | $0.10 | $0.30 | $0.22 | -3.5% |
| DekaLLM | $0.06 | $0.33 | $0.22 | -2.6% |
| Makora | $0.08 | $0.32 | $0.22 | -1.8% |
| DeepInfra | $0.07 | $0.34 | $0.23 | +1.8% |
| Novita | $0.13 | $0.40 | $0.29 | +28.1% |
| Parasail | $0.13 | $0.40 | $0.29 | +28.1% |
| SiliconFlow | $0.14 | $0.40 | $0.30 | +29.8% |
| Io Net | $0.15 | $0.50 | $0.36 | +57.9% |
| $0.15 | $0.60 | $0.42 | +84.2% | |
| Venice | $0.19 | $0.88 | $0.60 | +164.9% |