reroute

Z.ai: GLM 5.2

z-ai/glm-5.2

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

Modalities
In / Out Price
$0.07 / $4.30 per 1M
Context
1M
Released
Jun 16, 2026
Knowledge Cutoff
—

Pricing

What Reroute charges per unit. No markup on inference — you pay the carrier's price, metered from the carrier's own usage numbers.

Input tokens
$0.07 / M
Output tokens
$4.30 / M
Cache read
$0.07 / M

By provider

Each carrier sets its own price. The router weighs these against uptime when it picks one.

ProviderInput /MOutput /MCache read /MCache write /MDiscount
Relace$0.07$4.30$0.07——
Wafer$0.084$4.30$0.083——
Baidu$0.189$0.594$0.0351—87%
Wafer$0.19$4.30$0.18——
InferenceNet$0.20$4.40$0.13——
Decart$0.279$1.72$0.107—57%
StreamLake$0.56$1.76$0.104—60%
Baidu$0.56$1.96$0.139—75%
DeepInfra$0.563$1.80$0.105—25%
Morph$0.619$3.74$0.137——
Novita$0.65$2.04$0.121—54%
DigitalOcean$0.70$2.20$0.105——
CoreWeave$0.76$2.42$0.14——
AtlasCloud$0.938$2.95$0.174—33%
Alibaba$0.966$3.04$0.193——
Cloudflare$1.18$4.40$0.26——
SiliconFlow$1.19$3.74$0.221—15%
Phala$1.26$3$0.22——
Inceptron$1.39$4.39$0.25——
Nebius$1.40$4.40———
Mistral$1.40$4.40$0.14——
BaseTen$1.40$4.40$0.14——
Together$1.40$4.40$0.26——
BaseTen$1.40$4.40$0.14——
Venice$1.40$4.40$0.26——
GMICloud$1.40$4.40$0.26——
Parasail$1.40$4.40$0.26——
Friendli$1.40$4.40$0.26——
Z.AI$1.40$4.40$0.26——
Mistral$1.54$4.84$0.154——
BaseTen$2.10$6.60$0.21——
BaseTen$2.10$6.60$0.21——
Alibaba$2.31$7.26$0.462——
Decart$2.75$10$0.60——