GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...
Different carriers host the same model. Reroute routes each request to the best one on price and uptime, and fails over to the next if a carrier goes down.
| Provider | Context | Input /M | Output /M | Cache read /M | Latency | Throughput | Uptime |
|---|---|---|---|---|---|---|---|
| 198K | $0.60 | $1.92 | $0.12 | — | — | 100.00% | |
| 203K | $0.60 | $1.92 | $0.12 | — | — | 100.00% | |
| 203K | $0.70 | $2.24 | $0.14 | — | — | 100.00% | |
| 205K | $0.95 | $2.55 | $0.20 | — | — | 100.00% | |
| 203K | $1 | $3.20 | — | — | — | — | |
| 198K | $1 | $3.20 | $0.20 | — | — | — | |
| 203K | $1 | $3.20 | $0.20 | — | — | 100.00% | |
| 203K | $1 | $3.20 | $0.20 | — | — | 100.00% |