reroute

Z.ai: GLM 5.3

z-ai/glm-5.3

GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...

Modalities
In / Out Price
$0.049 / $3.39 per 1M
Context
1M
Released
Aug 18, 2026
Knowledge Cutoff
—

Pricing

What Reroute charges per unit. No markup on inference — you pay the carrier's price, metered from the carrier's own usage numbers.

Input tokens
$0.049 / M
Output tokens
$3.39 / M
Cache read
$0.048 / M

By provider

Each carrier sets its own price. The router weighs these against uptime when it picks one.

ProviderInput /MOutput /MCache read /MCache write /MDiscount
Relace$0.0392$12$0.0365——
Wafer$0.049$3.39$0.048——
Reka$0.14$4.20$0.105—30%
InferenceNet$0.14$4.40$0.07——
Makora$0.14$4.40$0.10——
Wafer$0.17$3.39$0.16——
AkashML$0.19$4.40$0.19——
Sail Research$0.20$3.40$0.15——
Sail Research$0.20$3.40$0.15——
OpenInference$0.26$3.96$0.26——
Morph$0.476$3.74$0.108——
DeepInfra$0.563$2.50$0.125—38%
Inceptron$0.60$3.39$0.20——
Novita$0.70$2.20$0.13—50%
Phala$0.84$2.64$0.156—40%
Decart$0.842$2.65$0.196——
DigitalOcean$0.91$2.86$0.169——
GMICloud$0.98$3.08$0.182—30%
Baidu$1.12$3.52$0.208—20%
SiliconFlow$1.12$3.52$0.208—20%
Alibaba$1.19$3.74$0.238——
Io Net$1.25$4.40$0.46——
Friendli$1.26$3.96$0.234—10%
Mistral$1.40$4.40$0.14——
BaseTen$1.40$4.40$0.14——
Mistral$1.40$4.40$0.14——
Nebius$1.40$4.40———
Crusoe$1.40$4.40$0.26——
PrimeIntellect$1.40$4.40$0.26——
Venice$1.40$4.40$0.26—20%
Together$1.40$4.40$0.26——
Parasail$1.40$4.40$0.26——
Modal$1.40$4.40$0.26——
BaseTen$1.40$4.40$0.14——
Fireworks$1.40$4.40$0.26——
Cloudflare$1.40$4.40$0.26——
AtlasCloud$1.40$4.40$0.26——
Z.AI$1.40$4.40$0.26——
Mistral$1.54$4.84$0.154——
Fireworks$2.10$6.60$0.39——
BaseTen$2.10$6.60$0.21——
BaseTen$2.10$6.60$0.21——
Alibaba$2.80$8.80$0.56——