The Magic of AI

LLM API · DeepSeek

deepseek-flash pricing

Input

$0.300 /Mtok

Output

$1.20 /Mtok

Cached input

$0.006 /Mtok

Context window

1,000k tokens

Max output

384k tokens

Tier

budget

Ranked 26 of 87 by price among 87 LLM APIs in the index — the cheapest is Z.AI GLM-4.5-Flash at $0.0000. Figures are USD list prices excluding tax.

Standard (peak) rate shown. Off-peak is 50% lower and applies outside 01:00-04:00 and 06:00-10:00 UTC, Mon-Fri.

Off-peak rate: input $0.15, output $0.6, cached input $0.003 per million tokens.

Price history

Every weekly reading since this item entered the index on 15 August 2026. Dots mark readings where the figure changed.

deepseek-flash price history$0.0000$0.369$0.739$1.11$1.4815 Aug24 Aug31 Aug05 Sep09 Sep14 Sep21 SepInput $/Mtok · 2026-08-15 · $0.440Input $/Mtok · 2026-09-10 · $0.300Input $/Mtok · 2026-09-21 · $0.300Output $/Mtok · 2026-08-15 · $1.32Output $/Mtok · 2026-09-10 · $1.20Output $/Mtok · 2026-09-21 · $1.20Cached input $/Mtok · 2026-08-15 · $0.014Cached input $/Mtok · 2026-09-10 · $0.0060Cached input $/Mtok · 2026-09-21 · $0.0060Input $/MtokOutput $/MtokCached input $/Mtok
DateFigureReadingChange
2026-08-15Cached input $/Mtok$0.014—
2026-08-15Input $/Mtok$0.440—
2026-08-15Output $/Mtok$1.32—
2026-09-10Cached input $/Mtok$0.0060-57.1%
2026-09-10Input $/Mtok$0.300-31.8%
2026-09-10Output $/Mtok$1.20-9.1%
2026-09-21Cached input $/Mtok$0.0060—
2026-09-21Input $/Mtok$0.300—
2026-09-21Output $/Mtok$1.20—

Machine-readable: history.json · Every movement also appears in the change log.

What it costs in practice

Computed from the current figures above. Change the assumptions in the LLM API cost calculator.

WorkloadCost
One million input tokens$0.300
One million output tokens$1.20
Blended $/Mtok at 4:1 input:output$0.480
1,000 chatbot replies (1,500 in / 300 out)$0.810
Summarising 1,000 ten-page documents (5,000 in / 400 out)$1.98
A month of 10,000 requests a day (2,000 in / 500 out)$360.00
1,000 replies with a 1,200-token cached prefix (300 fresh in / 300 out)$0.457

Compared with other models in the same tier

Modelblended $/Mtok (4:1)vs this
Z.AI GLM-4.5-Flash$0.0000-100%
Z.AI GLM-4.6V-Flash$0.0000-100%
Z.AI GLM-4.7-Flash$0.0000-100%
Mistral AI Ministral 3 (3B)$0.100-79%
Z.AI GLM-4.6V-FlashX$0.112-77%
Groq GPT-OSS 20B$0.120-75%
Z.AI GLM-4.7-FlashX$0.136-72%
Mistral AI Ministral 3 (8B)$0.150-69%

How to read this page

The figures are the list prices published on the vendor's own pricing page, re-read every week by an automated check that never invents a number: if the page cannot be read, the previous reading is kept and the item is flagged after three missed weeks. Negotiated rates, volume tiers, free credits and regional pricing are not modelled. The methodology describes the process and its limits; the price index shows every llm api figure side by side.