DeepSeek
DeepSeek V4.1 Flash
deepseek/deepseek-v4.1-flashHugging Face ↗
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...
Only tracked on Vercel AI Gateway.
DeepSeek V4.1 Flash usage over time
Vercel AI Gateway share of tokens, by day.
DeepSeek V4.1 Flash holds 1.01% today, down from 13.82% at the start of the visible window.
Rank history
Adoption curve vs the pantheon
DeepSeek V4.1 Flash's share plotted against days since first tracked, overlaid on other current leaders on the same axis. The question every new-model page asks: is this one going to make it?
Citable stat
As of 2026-09-11, DeepSeek V4.1 Flash holds 39.9% of Vercel AI Gateway token share, up 39.92pp over 30 days. Source: Tidelines (Vercel AI Gateway).
Where to run it
DeepSeek V4.1 Flash pricing by provider
Per-provider pricing, context, and uptime via OpenRouter. Sorted by output price.
DeepSeek V4.1 Flash is served by 8 tracked endpoints on OpenRouter. The cheapest output price is $0.60 per million tokens from DeepInfra, and the most expensive is $1.20 from DeepSeek. The largest context window offered is 1049k tokens. Prices are list rates and refresh daily.
| Provider | Input $/M | Output $/M | Context | Uptime |
|---|---|---|---|---|
| DeepInfra | $0.20 | $0.60 | 1049k | 95.64% |
| Fireworks | $0.22 | $0.66 | 1049k | 99.18% |
| SiliconFlow | $0.30 | $1.20 | 1049k | 99.70% |
| GMICloud | $0.30 | $1.20 | 1049k | 99.92% |
| Morph | $0.30 | $1.20 | 1049k | 99.08% |
| Io Net | $0.30 | $1.20 | 262k | 98.31% |
| Novita | $0.30 | $1.20 | 1049k | 99.96% |
| DeepSeek | $0.30 | $1.20 | 1049k | 99.96% |