GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...
Different carriers host the same model. Reroute routes each request to the best one on price and uptime, and fails over to the next if a carrier goes down.
| Provider | Context | Input /M | Output /M | Cache read /M | Latency | Throughput | Uptime |
|---|---|---|---|---|---|---|---|
| 400K | $0.10 | $0.625 | $0.01 | — | — | — | |
| 400K | $0.20 | $1.25 | $0.02 | — | — | 100.00% | |
| 400K | $0.20 | $1.25 | $0.02 | — | — | 100.00% | |
| 400K | $0.22 | $1.38 | $0.022 | — | — | — |