Pricing · 2026-07-30

Cheapest AI models

The ten cheapest models that developers are genuinely running, by list price per million output tokens. Free demo endpoints and models with no measurable traffic are excluded.

What cheap costs in July 2026

Across 50 tracked models, gateway traffic of roughly 7.7T tokens a day implies about $23,939,776 of daily list-price spend. Most of that concentrates in a handful of flagships, which is exactly why the cheap tier matters: swapping a high-volume, low-complexity workload onto a model in this table changes the bill by an order of magnitude.

The cheapest model with real usage is inclusionAI: Ling-2.6-flash at $0.03 per million output tokens, holding 0.44% of tracked tokens (+0.08pp over 30 days).

Top 10 cheapest AI models in use

List price per million output tokens, models with real gateway traffic.

Tidelines — LLM usage and market share
#ModelLab$/M outShare %
1inclusionAI: Ling-2.6-flashOther$0.030.44
2Poolside: Laguna XS 2.1Other$0.120.34
3Qwen: Qwen3.7 FlashAlibaba$0.130.32
4OpenAI: gpt-oss-120bOpenAI$0.170.97
5Poolside: Laguna S 2.1Other$0.181.43
6DeepSeek: DeepSeek V4 FlashDeepSeek$0.2813.96
7Xiaomi: MiMo-V2.5Xiaomi$0.288.68
8Google: Gemma 4 31BGoogle$0.340.72
9Google: Gemma 4 26B A4B Google$0.340.67
10Google: Gemini 2.5 Flash LiteGoogle$0.401.12
Data: OpenRouter as of 2026-07-30 · Source: tidelines.ai/cheapest-ai-models

Cheapest option per lab

The lowest-priced model with real usage from each lab.

Tidelines — LLM usage and market share
#ModelLab$/M outShare %
1inclusionAI: Ling-2.6-flashOther$0.030.44
2Qwen: Qwen3.7 FlashAlibaba$0.130.32
3OpenAI: gpt-oss-120bOpenAI$0.170.97
4DeepSeek: DeepSeek V4 FlashDeepSeek$0.2813.96
5Xiaomi: MiMo-V2.5Xiaomi$0.288.68
6Google: Gemma 4 31BGoogle$0.340.72
7NVIDIA: Nemotron 3 SuperNvidia$0.400.59
8Tencent: Hy3Tencent$0.537.36
9MiniMax: MiniMax M2.7MiniMax$1.000.28
10StepFun: Step 3.7 FlashStepFun$1.153.39
Data: OpenRouter as of 2026-07-30 · Source: tidelines.ai/cheapest-ai-models

Cheapest AI models FAQ

What is the cheapest AI model that people actually use?
inclusionAI: Ling-2.6-flash from Other at $0.03 per million output tokens, while still carrying 0.44% of tracked gateway traffic on 2026-07-30.
What are the top 10 cheapest AI models?
Ranked by output list price among models with real usage, the cheapest start with inclusionAI: Ling-2.6-flash, Poolside: Laguna XS 2.1, Qwen: Qwen3.7 Flash. The full top 10 is in the table above and is recomputed daily.
Why exclude free and zero-traffic models?
A $0 price on a model nobody routes traffic to is not a usable option. This page only ranks models with a real list price and measurable gateway usage, so the cheapest entries are ones you can actually deploy.
Is the cheapest model the best value?
Not always. Cheap models can need more tokens or more retries for the same task. Pair this page with the usage-per-dollar table on the best LLMs page to weigh price against adoption.
How often do prices update?
Prices and usage are re-read from the gateway catalogues every day. The figures on this page are from 2026-07-30.