reroute

Z.ai: GLM 4.6

z-ai/glm-4.6

Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...

Modalities
In / Out Price
$0.43 / $1.75 per 1M
Context
205K
Released
Sep 30, 2025
Knowledge Cutoff
Mar 2025

Pricing

What Reroute charges per unit. No markup on inference — you pay the carrier's price, metered from the carrier's own usage numbers.

Input tokens
$0.43 / M
Output tokens
$1.75 / M
Cache read
$0.08 / M

By provider

Each carrier sets its own price. The router weighs these against uptime when it picks one.

ProviderInput /MOutput /MCache read /MCache write /MDiscount
Venice$0.43$1.75$0.08——
DeepInfra$0.50$2$0.10——
Novita$0.55$2.20$0.11——
Z.AI$0.60$2.20$0.11——