Ways to buy 26
| Way to buy | Input | Output | In / out per 1M | Cached | Vs list |
|---|---|---|---|---|---|
| Morph Default region, cheapest · cached $0.0015 | $0.075 | $0.3 | $0.075 / $0.3-50% vs list | $0.0015 | -50% vs list |
| DekaLLM Default region · cached $0.01 | $0.04 | $0.49 | $0.04 / $0.49-42% vs list | $0.01 | -42% vs list |
| OpenInference fp4 · cached $0.01 | $0.1 | $0.5 | $0.1 / $0.5-24% vs list | $0.01 | -24% vs list |
| Relace Default region · cached $0.01 | $0.1 | $0.5 | $0.1 / $0.5-24% vs list | $0.01 | -24% vs list |
fp8 · cached $0.0042 | $0.14 | $0.42 | $0.14 / $0.42-20% vs list | $0.0042 | -20% vs list |
| Wafer Default region · cached $0.06 | $0.099 | $0.6 | $0.099 / $0.6-15% vs list | $0.06 | -15% vs list |
The lab's own API · cached $0.003 | $0.15 | $0.6 | $0.15 / $0.6List price | $0.003 | List price |
| Sail Research fp4 · cached $0.01 | $0.13 | $0.75 | $0.13 / $0.75+9% vs list | $0.01 | +9% vs list |
| StreamLake fp8 · cached $0.0033 | $0.165 | $0.66 | $0.165 / $0.66+10% vs list | $0.0033 | +10% vs list |
fp8 · cached $0.03 | $0.2 | $0.65 | $0.2 / $0.65+19% vs list | $0.03 | +19% vs list |
Default region · cached $0.007 | $0.22 | $0.66 | $0.22 / $0.66+26% vs list | $0.007 | +26% vs list |
| GMICloud fp8 · cached $0.0045 | $0.225 | $0.9 | $0.225 / $0.9+50% vs list | $0.0045 | +50% vs list |
| Krea fp8 · cached $0.006 | $0.225 | $0.9 | $0.225 / $0.9+50% vs list | $0.006 | +50% vs list |
| Phala Default region · cached $0.0055 | $0.276 | $1.1 | $0.276 / $1.1+84% vs list | $0.0055 | +84% vs list |
fp8 · cached $0.0057 | $0.285 | $1.14 | $0.285 / $1.14+90% vs list | $0.0057 | +90% vs list |
| AtlasCloud fp8 · cached $0.03 | $0.3 | $1.2 | $0.3 / $1.2+100% vs list | $0.03 | +100% vs list |
fp8 · cached $0.03 | $0.3 | $1.2 | $0.3 / $1.2+100% vs list | $0.03 | +100% vs list |
Default region · cached $0.006 | $0.3 | $1.2 | $0.3 / $1.2+100% vs list | $0.006 | +100% vs list |
| Makora fp8 · cached $0.006 | $0.3 | $1.2 | $0.3 / $1.2+100% vs list | $0.006 | +100% vs list |
Default region · cached $0.03 | $0.3 | $1.2 | $0.3 / $1.2+100% vs list | $0.03 | +100% vs list |
| NextBit fp8 · cached $0.006 | $0.3 | $1.2 | $0.3 / $1.2+100% vs list | $0.006 | +100% vs list |
fp8 · cached $0.006 | $0.3 | $1.2 | $0.3 / $1.2+100% vs list | $0.006 | +100% vs list |
Default region · cached $0.03 | $0.3 | $1.2 | $0.3 / $1.2+100% vs list | $0.03 | +100% vs list |
fp8 · cached $0.006 | $0.3 | $1.2 | $0.3 / $1.2+100% vs list | $0.006 | +100% vs list |
Default region · cached $0.006 | $0.3 | $1.2 | $0.3 / $1.2+100% vs list | $0.006 | +100% vs list |
fp8 · cached $0.0075 | $0.375 | $1.5 | $0.375 / $1.5+150% vs list | $0.0075 | +150% vs list |
USD per 1M tokens; batch $0.112 in, $0.336 out on the lab's API. Regions are separate routes.
Through a gateway
Your own numbers100M input and 20M output tokens a month: $27.00 direct.
Helicone$27.00+0%
Kong AI Gateway$27.00+0%
LiteLLM$27.00+0%
Vercel AI Gateway$27.00+0%
Cloudflare AI Gateway$28.35+5%
Requesty$28.35+5%
Eden AI$28.48+5.5%
OpenRouter$28.48+5.5%
No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.
Related models
| In / out per 1M | Cheapest route | ||||||
|---|---|---|---|---|---|---|---|
DeepSeek · 1M | $0.216 | $0.647 | $0.216 / $0.647 | $0.323 | 1M | on DeepInfra (fp8) | · |
DeepSeek · 1MCheapest: $0.525 on Baidu (fp8) | $0.66 | $1.98 | $0.66 / $1.98 | $0.99 | 1M | $0.525 on Baidu (fp8) | · |
DeepSeek · 1.31M | $0.053 | $0.158 | $0.053 / $0.158 | $0.079 | 1.31M | on StreamLake (fp8) | · |
DeepSeek · 1M | $0.049 | $0.098 | $0.049 / $0.098 | $0.061 | 1M | on Baidu (fp8) | · |
Google · 1MCheapest: $0.281 on Google Vertex AI (global-flex) | $0.25 | $1.5 | $0.25 / $1.5 | $0.563 | 1M | $0.281 on Google Vertex AI (global-flex) | · |
Google · 1MCheapest: $0.281 on Google Gemini API (flex) | $0.25 | $1.5 | $0.25 / $1.5 | $0.563 | 1M | $0.281 on Google Gemini API (flex) | · |
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex) | $0.1 | $0.5 | $0.1 / $0.5 | $0.2 | 1.05M | $0.1 on OpenAI (flex) | · |
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex) | $0.1 | $0.5 | $0.1 / $0.5 | $0.2 | 1.05M | $0.1 on OpenAI (flex) | · |
About DeepSeek V4.1 Flash
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture.
Takes text, image in and returns text. First listed Sep 23, 2026. Up to 393K output tokens. Blended price $0.262 per 1M tokens at 3 input to 1 output. Overall score 3 on leadersboard.fru.dev (rank 33). Model id deepseek-v4-1-flash.
Price history
List price since Sep 23, 2026, from daily reads of the OpenRouter catalog.
Across fru.devCompany profile
Questions
How much does DeepSeek V4.1 Flash cost?
$0.15 per million input tokens and $0.6 per million output tokens, with cached input at $0.003, or $0.112 and $0.336 through the Batch API.
What is the context window of DeepSeek V4.1 Flash?
1M tokens, with up to 393K tokens of output.
Where is DeepSeek V4.1 Flash cheapest?
Of 26 routes tracked, Morph is cheapest at $0.075 input and $0.3 output per million tokens.
Has DeepSeek V4.1 Flash's price changed?
Yes. The output price went from $0.6 to $1.2, first seen Sep 25, 2026.
Sources: OpenRouter's public model catalog and each provider's own pricing page, read every morning at 05:00 UTC; gateway fees re-read every Monday. List prices: your contract may differ. Logos via logo.dev; trademarks belong to their owners.