Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
What Reroute charges per unit. No markup on inference — you pay the carrier's price, metered from the carrier's own usage numbers.
Each carrier sets its own price. The router weighs these against uptime when it picks one.
| Provider | Input /M | Output /M | Cache read /M | Cache write /M | Discount |
|---|---|---|---|---|---|
| $0.042 | $0.22 | $0.021 | — | — | |
| $0.0517 | $0.225 | $0.0217 | — | 25% | |
| $0.06 | $0.33 | $0.04 | — | — | |
| $0.07 | $0.34 | — | — | — | |
| $0.08 | $0.32 | $0.032 | — | — | |
| $0.09 | $0.30 | $0.05 | — | — | |
| $0.10 | $0.30 | $0.05 | — | — | |
| $0.10 | $0.30 | $0.05 | — | — | |
| $0.13 | $0.40 | $0.05 | — | — | |
| $0.13 | $0.40 | $0.05 | — | — | |
| $0.13 | $0.40 | — | — | — | |
| $0.14 | $0.40 | $0.05 | — | — | |
| $0.15 | $0.60 | — | — | — |