reroute

inclusionAI: Ling 3.0 Flash

inclusionai/ling-3.0-flash

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

Modalities
In / Out Price
$0.021 / $0.063 per 1M
Context
262K
Released
Jul 23, 2026
Knowledge Cutoff
—

Providers

Different carriers host the same model. Reroute routes each request to the best one on price and uptime, and fails over to the next if a carrier goes down.

ProviderContextInput /MOutput /MCache read /MLatencyThroughputUptime
Novita262K$0.021$0.063$0.0042——100.00%
DeepInfrabf16131K$0.06$0.18$0.012——100.00%