The Magic of AI

LLM API · DeepSeek

DeepSeek V4 Flash 0731 pricing

Input

$0.140 /Mtok

Output

$0.280 /Mtok

Batch discount

−67%

Tier

budget

Ranked 10 of 81 by price among 81 LLM APIs in the index — the cheapest is Z.AI GLM-4.5-Flash at $0.0000. Figures are USD list prices excluding tax.

Price history

Tracked since 12 September 2026. A chart appears once there are at least two weekly readings.

DateFigureReadingChange
2026-09-12Input $/Mtok$0.140
2026-09-12Output $/Mtok$0.280

Machine-readable: history.json · Every movement also appears in the change log.

What it costs in practice

Computed from the current figures above. Change the assumptions in the LLM API cost calculator.

WorkloadCost
One million input tokens$0.140
One million output tokens$0.280
Blended $/Mtok at 4:1 input:output$0.168
1,000 chatbot replies (1,500 in / 300 out)$0.294
Summarising 1,000 ten-page documents (5,000 in / 400 out)$0.812
A month of 10,000 requests a day (2,000 in / 500 out)$126.00
Those 1,000 document summaries via the batch API (−67%)$0.268

Compared with other models in the same tier

Modelblended $/Mtok (4:1)vs this
Z.AI GLM-4.5-Flash$0.0000-100%
Z.AI GLM-4.6V-Flash$0.0000-100%
Mistral AI Ministral 3 (3B)$0.100-40%
Z.AI GLM-4.6V-FlashX$0.112-33%
Groq GPT-OSS 20B$0.120-29%
Z.AI GLM-4.7-FlashX$0.136-19%
Mistral AI Ministral 3 (8B)$0.150-11%
Google Gemini 2.5 Flash-Lite$0.160-5%

How to read this page

The figures are the list prices published on the vendor's own pricing page, re-read every week by an automated check that never invents a number: if the page cannot be read, the previous reading is kept and the item is flagged after three missed weeks. Negotiated rates, volume tiers, free credits and regional pricing are not modelled. The methodology describes the process and its limits; the price index shows every llm api figure side by side.