reroute

Google: Gemma 4 31B

google/gemma-4-31b-it

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Modalities
In / Out Price
$0.09 / $0.34 per 1M
Context
262K
Released
Apr 2, 2026
Knowledge Cutoff
—

Pricing

What Reroute charges per unit. No markup on inference — you pay the carrier's price, metered from the carrier's own usage numbers.

Input tokens
$0.09 / M
Output tokens
$0.34 / M
Cache read
$0.05 / M

By provider

Each carrier sets its own price. The router weighs these against uptime when it picks one.

ProviderInput /MOutput /MCache read /MCache write /MDiscount
DeepInfra$0.09$0.34$0.05——
CoreWeave$0.10$0.34$0.10——
Venice$0.12$0.36$0.09——
Chutes$0.12$0.37$0.012——
Crusoe$0.14$0.40$0.14——
Friendli$0.14$0.40———
Novita$0.14$0.40———
Parasail$0.15$0.40$0.06——
DeepInfra$0.20$0.40———
Io Net$0.361$1.09$0.181—5%
SambaNova$0.38$1.15———
ModelRun$0.75$1$0.20——
SiliconFlow$0.75$1$0.25——