Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
What Reroute charges per unit. No markup on inference — you pay the carrier's price, metered from the carrier's own usage numbers.
Each carrier sets its own price. The router weighs these against uptime when it picks one.
| Provider | Input /M | Output /M | Cache read /M | Cache write /M | Discount |
|---|---|---|---|---|---|
| $0.09 | $0.34 | $0.05 | — | — | |
| $0.10 | $0.34 | $0.10 | — | — | |
| $0.12 | $0.36 | $0.09 | — | — | |
| $0.12 | $0.37 | $0.012 | — | — | |
| $0.14 | $0.40 | $0.14 | — | — | |
| $0.14 | $0.40 | — | — | — | |
| $0.14 | $0.40 | — | — | — | |
| $0.15 | $0.40 | $0.06 | — | — | |
| $0.20 | $0.40 | — | — | — | |
| $0.361 | $1.09 | $0.181 | — | 5% | |
| $0.38 | $1.15 | — | — | — | |
| $0.75 | $1 | $0.20 | — | — | |
| $0.75 | $1 | $0.25 | — | — |