reroute

Inception: Mercury 2.5

inception/mercury-2.5

Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

Modalities
In / Out Price
$0.04 / $0.15 per 1M
Context
260K
Released
Sep 8, 2026
Knowledge Cutoff
—

Providers

Different carriers host the same model. Reroute routes each request to the best one on price and uptime, and fails over to the next if a carrier goes down.

ProviderContextInput /MOutput /MCache read /MLatencyThroughputUptime
Inception260K$0.04$0.15$0.004——100.00%