GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...
Different carriers host the same model. Reroute routes each request to the best one on price and uptime, and fails over to the next if a carrier goes down.
| Provider | Context | Input /M | Output /M | Cache read /M | Latency | Throughput | Uptime |
|---|---|---|---|---|---|---|---|
| 203K | $0.965 | $3.03 | $0.179 | — | — | 100.00% | |
| 200K | $0.966 | $3.04 | $0.179 | — | — | 100.00% | |
| 203K | $0.98 | $3.08 | $0.098 | — | — | 0.00% | |
| 205K | $1.19 | $3.74 | $0.60 | — | — | 100.00% | |
| 203K | $1.21 | $4.20 | $0.60 | — | — | — | |
| 203K | $1.26 | $3.96 | $0.234 | — | — | 100.00% | |
| 203K | $1.33 | $4.18 | $0.247 | — | — | 100.00% | |
| 205K | $1.38 | $4.40 | $0.26 | — | — | — | |
| 203K | $1.40 | $4.40 | — | — | — | 100.00% | |
| 203K | $1.40 | $4.40 | $0.26 | — | — | — | |
| 203K | $1.40 | $4.40 | $0.26 | — | — | 100.00% | |
| 203K | $1.40 | $4.40 | $0.26 | — | — | 100.00% | |
| 200K | $1.4 | $4.4 | $0.26 | — | — | — |