DeepSeek

DeepSeek V4 Flash

deepseek/deepseek-v4-flash-vision-expHugging Face ↗

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...

toolsstructuredreasoningimage in
Data source

Only tracked on Vercel AI Gateway.

DeepSeek V4 Flash usage over time

Vercel AI Gateway share of tokens, by day.

View
Tidelines — LLM usage and market share

DeepSeek V4 Flash holds 29.71% today, up from 13.82% at the start of the visible window.

Data: Vercel AI Gateway as of 2026-09-02 · Source: tidelines.ai/models/deepseek-deepseek-v4-flash-vision-exp

Rank history

Tidelines — LLM usage and market share
Data: Vercel AI Gateway as of 2026-09-02 · Source: tidelines.ai/models/deepseek-deepseek-v4-flash-vision-exp

Adoption curve vs the pantheon

DeepSeek V4 Flash's share plotted against days since first tracked, overlaid on other current leaders on the same axis. The question every new-model page asks: is this one going to make it?

Tidelines — LLM usage and market share
Data: Vercel AI Gateway as of 2026-09-02 · Source: tidelines.ai/models/deepseek-deepseek-v4-flash-vision-exp

Citable stat

As of 2026-09-02, DeepSeek V4 Flash holds 29.7% of Vercel AI Gateway token share, up 8.73pp over 30 days. Source: Tidelines (Vercel AI Gateway).

Cost per session

Median dollars spent on one DeepSeek V4 Flash session, by coding agent and session length.

Tidelines — LLM usage and market share

Token rate sets the price; session length and reasoning set the bill.

Agent2–9 turns10–49 turns
Hermes Agent$0.0093$0.045

Compare every model's session cost →

Data: OpenRouter session-cost dataset as of 2026-08-30 · Source: tidelines.ai/models/deepseek-deepseek-v4-flash-vision-exp

Where to run it

DeepSeek V4 Flash pricing by provider

Tidelines — LLM usage and market share

Per-provider pricing, context, and uptime via OpenRouter. Sorted by output price.

DeepSeek V4 Flash is served by 3 tracked endpoints on OpenRouter. The cheapest output price is $0.66 per million tokens from Fireworks, and the most expensive is $1.32 from DeepSeek. The largest context window offered is 1049k tokens. Prices are list rates and refresh daily.

ProviderInput $/MOutput $/MContextUptime
Fireworks$0.22$0.661049k99.98%
Novita$0.44$1.321049k100.00%
DeepSeek$0.44$1.321049k100.00%