reroute

inclusionAI: Ling 3.0 Flash

inclusionai/ling-3.0-flash

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

Modalities
In / Out Price
$0.021 / $0.063 per 1M
Context
262K
Released
Jul 23, 2026
Knowledge Cutoff
—

Pricing

What Reroute charges per unit. No markup on inference — you pay the carrier's price, metered from the carrier's own usage numbers.

Input tokens
$0.021 / M
Output tokens
$0.063 / M
Cache read
$0.0042 / M

By provider

Each carrier sets its own price. The router weighs these against uptime when it picks one.

ProviderInput /MOutput /MCache read /MCache write /MDiscount
Novita$0.021$0.063$0.0042—65%
DeepInfra$0.06$0.18$0.012——