reroute

DeepSeek: DeepSeek V4 Flash 0423

deepseek/deepseek-v4-flash

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

Modalities
In / Out Price
$0.0075 / $1.28 per 1M
Context
1M
Released
Apr 24, 2026
Knowledge Cutoff
—

Pricing

What Reroute charges per unit. No markup on inference — you pay the carrier's price, metered from the carrier's own usage numbers.

Input tokens
$0.0075 / M
Output tokens
$1.28 / M
Cache read
$0.0075 / M

By provider

Each carrier sets its own price. The router weighs these against uptime when it picks one.

ProviderInput /MOutput /MCache read /MCache write /MDiscount
OpenInference$0.006$0.489$0.006——
Relace$0.0075$1.28$0.0075——
Baidu$0.0182$0.0364$0.00364—87%
Wafer$0.055$0.17$0.013——
StreamLake$0.07$0.14$0.014—50%
DeepInfra$0.09$0.18$0.018——
GMICloud$0.091$0.182$0.0182—35%
Venice$0.0966$0.193$0.0196—30%
DigitalOcean$0.098$0.196$0.0196——
SiliconFlow$0.13$0.28$0.028——
Alibaba$0.134$0.268$0.0268——
Novita$0.14$0.28$0.028——
AtlasCloud$0.14$0.28$0.028——
Parasail$0.14$0.28$0.07——
Mancer 2$0.19$0.50———
Azure$0.21$0.56$0.031——
Cloudflare$0.44$1.32$0.014——