reroute

Z.ai: GLM 4.6

z-ai/glm-4.6

Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...

Modalities
In / Out Price
$0.43 / $1.75 per 1M
Context
205K
Released
Sep 30, 2025
Knowledge Cutoff
Mar 2025

Providers

Different carriers host the same model. Reroute routes each request to the best one on price and uptime, and fails over to the next if a carrier goes down.

ProviderContextInput /MOutput /MCache read /MLatencyThroughputUptime
Venicefp4198K$0.43$1.75$0.08——96.67%
DeepInfrafp4203K$0.50$2$0.10——100.00%
Novitabf16205K$0.55$2.20$0.11——100.00%
Z.AIfp4203K$0.60$2.20$0.11———