*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
Top apps by tokens routed through Reroute in the last 30 days. Apps appear here when they send an X-Title header.
Once apps call this model with an X-Title header, they're ranked here by tokens.