gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
Different carriers host the same model. Reroute routes each request to the best one on price and uptime, and fails over to the next if a carrier goes down.
| Provider | Context | Input /M | Output /M | Cache read /M | Latency | Throughput | Uptime |
|---|---|---|---|---|---|---|---|
| 131K | $0.018 | $0.09 | $0.009 | — | — | 100.00% | |
| 131K | $0.02 | $0.10 | — | — | — | 100.00% | |
| 131K | $0.029 | $0.14 | $0.029 | — | — | 100.00% | |
| 131K | $0.03 | $0.13 | $0.03 | — | — | 100.00% | |
| 131K | $0.03 | $0.14 | — | — | — | 100.00% | |
| 131K | $0.03 | $0.15 | $0.02 | — | — | 100.00% | |
| 131K | $0.04 | $0.18 | — | — | — | 100.00% | |
| 131K | $0.07 | $0.15 | — | — | — | 100.00% | |
| 131K | $0.07 | $0.15 | — | — | — | 100.00% | |
| 131K | $0.07 | $0.25 | — | — | — | 100.00% | |
| 131K | $0.075 | $0.30 | $0.0375 | — | — | 100.00% |