The Magic of AI

AI price index

273 individually sourced prices covering LLM APIs, embedding models, image and video generation, speech, and GPU rental. Each one is re-read from the vendor's own pricing page every week; when a figure moves, the movement is recorded in the change log.

Every model name links to its own page with weekly price history and worked examples — or browse by model and by vendor.

Every figure links to its source. Prices are USD list rates excluding tax, read from each vendor's own pricing page. Negotiated rates, volume commitments and credits are not modelled. Click a column header to sort.

Per-million-token list prices. Cached input applies to a repeated prompt prefix; batch applies to asynchronous processing.

ModelTierInput $/MtokOutput $/MtokCached inputBatchContextSource
Z.AI GLM-4.5-Flashnewbudget$0.00$0.00source
Z.AI GLM-4.6V-Flashnewbudget$0.00$0.00source
Z.AI GLM-4.7-Flashnewbudget$0.00$0.00source
Mistral AI Ministral 3 (3B)budget$0.10$0.10source
Z.AI GLM-4-32B-0414-128Knewsmall$0.10$0.10source
Z.AI GLM-4.6V-FlashXnewbudget$0.04$0.40source
Groq GPT-OSS 20Bbudget$0.07$0.30131ksource
Z.AI GLM-4.7-FlashXnewbudget$0.07$0.40source
Mistral AI Ministral 3 (8B)newbudget$0.15$0.15source
DeepSeek V4 Flash 0731newbudget$0.14$0.28−67%source
Mistral AI Ministral 3 (14B)newbudget$0.20$0.20source
Z.AI GLM-5.3-Flashnewbudget$0.15$0.50source
Meta via Fireworks AI Llama 4 Scout Instructsmall$0.15$0.601048ksource
Mistral AI Mistral Small 4small$0.15$0.60source
Groq GPT-OSS 120Bsmall$0.15$0.60131ksource
Meta via Fireworks AI Llama 4 Scout Instruct (Basic)newsmall$0.15$0.60source
Qwen via Together AI Qwen3 235B A22B Instructsmall$0.20$0.60source
Qwen via Together AI Qwen2.5 7B Instruct Turbobudget$0.30$0.30source
Cohere Command-lightbudget$0.30$0.60source
Z.AI GLM-4.5-Airnewbudget$0.20$1.10source
OpenAI GPT-5.6 Lunasmall$0.20$1.20$0.020−50%1050ksource
Mistral AI Codestralsmall$0.30$0.90source
Z.AI GLM-4.6Vnewmid$0.30$0.90source
deepseek-flashbudget$0.30off-peak $0.15$1.20off-peak $0.60$0.006off-peak $0.0031000ksource
DeepSeek V4.1 Flashnewsmall$0.30off-peak $0.15$1.20off-peak $0.60source
MiniMax M3newbudget$0.30$1.20−50%source
Alibaba Qwen3.7-Plusnewmid$0.32$1.28source
Mistral AI Mistral Large 3mid$0.50$1.50source
Cohere Command R (03-2024)small$0.50$1.50source
Cohere Aya Expansenewsmall$0.50$1.50source
Google Gemini 3.5 Flash-Litesmall$0.30$2.50$0.030source
Google Gemini 2.5 Flash Live APInewmid$0.50$2.00source
Z.AI GLM-4.5Vnewmid$0.60$1.80source
Z.AI GLM-4.6newsmall$0.60$2.20source
Z.AI GLM-4.5newsmall$0.60$2.20source
Z.AI GLM-4.7newmid$0.60$2.20source
kimi-k2.7-codenewmid$0.19$4.00source
OpenAI gpt-realtime-2.1-mininewbudget$0.60$2.40source
Perplexity Sonarnewsmall$1.00$1.00source
RunPod IBM Granite 4.0 H Smallnewsmall$1.00$1.00source
Meta via Together AI Llama 3.3 70Bmid$1.04$1.04source
Alibaba Qwen3.5-397B-A17Bnewfrontier$0.60$3.60−83%source
Google Gemini 3.8 Flashnewsmall$0.75$3.75source
Google Gemini 3.7 Flashnewsmall$0.75$3.75source
Google Gemini 3.6 Flashnewsmall$0.75$3.75source
Z.AI GLM-5newmid$1.00$3.20source
Baseten Kimi K2.6newmid$0.95$4.00source
Z.AI GLM-4.5-AirXnewsmall$1.10$4.50source
Anthropic Claude Haiku 4.5small$1.00$5.00$0.100−50%200ksource
DeepSeek V4 Prosmall$1.32off-peak $0.66$3.96off-peak $1.98$0.044off-peak $0.0221000ksource
DeepSeek V4 Pro 0813newfrontier$1.32$3.96−67%source
kimi-k2.7-code-highspeednewmid$0.38$8.00source
Mistral AI GLM 5.2newmid$1.40$4.40source
Z.AI GLM-5.3newfrontier$1.40$4.40source
Z.AI GLM-5.1newfrontier$1.40$4.40source
Mistral AI Mistral Medium 3.5mid$1.50$7.50source
xAI Grok 4.6mid$2.00$6.00500ksource
Alibaba Qwen3.7-Maxnewfrontier$2.00$6.00source
Alibaba Qwen3.8-2.4T-A95Bnewfrontier$2.00$6.00−67%source
Google Gemini 3.5 Flashmid$1.50$9.00$0.150−50%source
Perplexity Sonar Reasoning Pronewmid$2.00$8.00source
Perplexity Sonar Deep Researchnewfrontier$2.00$8.00source
Z.AI GLM-4.5-Xnewfrontier$2.20$8.90source
Anthropic Claude Sonnet 5mid$2.00$10.00$0.200−50%1000ksource
OpenAI GPT-5.6 Terramid$2.00$12.00$0.200−50%1050ksource
Google Gemini 3.1 Profrontier$2.00$12.00−50%source
Cohere Command R+ (08-2024)mid$2.50$10.00source
OpenAI gpt-5.3-codexnewsmall$1.75$14.00source
DeepSeek R1newfrontier$3.75$10.00source
Perplexity Sonar Pronewmid$3.00$15.00source
Baseten Kimi K3newfrontier$3.00$15.00source
OpenAI GPT-5.6 Solfrontier$4.00$20.00$0.400−50%1050ksource
OpenAI gpt-realtime-2.1newmid$4.00$24.00source
Anthropic Claude Opus 5frontier$5.00$25.00$0.500−50%1000ksource
OpenAI chat-latestnewmid$5.00$30.00source
RunPod Qwen3 32B AWQnewsmall$10.00$10.00source
OpenAI gpt-6-astranewfrontier$10.00$50.00source
Anthropic Claude Fable 5.1newfrontier$10.00$50.00$0.250source
Anthropic Claude Fable 5newfrontier$10.00$50.00$1.000−50%source
OpenAI gpt-5.6-cybernewfrontier$12.50$75.00source

Embedding models

Price your workload →

Embedding prices are small enough that retrieval quality, not price, should drive the choice.

Image generation

Price your workload →

Per-image list prices. Effective cost is this figure multiplied by how many attempts a usable asset takes.

ModelTier$/imageSource
RunPod FLUX.1 Schnellnewstandard$0.002source
Black Forest Labs FLUX Schnellnewfast$0.003source
Z.AI CogView-4newstandard$0.010source
Black Forest Labs FLUX.2 Klein 4B1MP$0.014source
Kling AI Kling Image 2.11K/2K$0.014source
Black Forest Labs FLUX.2 Klein 9Bnew9B$0.015source
Z.AI GLM-Imagenewstandard$0.015source
xAI grok-imagine-imagenew1K / 2K$0.020source
RunPod FLUX.1 [dev]newstandard$0.020source
RunPod Qwen Imagenewstandard$0.020source
Stability AI Stable Diffusion 3.5 Flashflash$0.025source
Black Forest Labs FLUX Devnewstandard$0.025source
RunPod bytedance / Seedream 4.0 Editnewstandard$0.027source
Kling AI Kling Image 3.01K/2K$0.028source
Black Forest Labs FLUX.2 Pro1MP$0.030source
Stability AI Stable Image Corecost-optimised$0.030source
fal Seedream V4new1MP$0.030source
fal Nanobanananew1MP$0.040source
Black Forest Labs FLUX 1.1 [pro]standard$0.040source
xAI Grok Imagine Image 2.0new1K$0.040source
Black Forest Labs FLUX.1 Kontext Pronewpro$0.040source
fal Flux Kontext Pronew1MP$0.040source
Recraft AI Recraft v3newstandard$0.040source
Runway Gen-4 Image Turbonew1080p$0.048source
Black Forest Labs FLUX.2 Flexnewflex$0.050source
Black Forest Labs FLUX.1 Fill Pronewinpainting$0.050source
xAI grok-imagine-image-qualitynew1K / 2K$0.050source
Black Forest Labs FLUX1.1 Pro Ultranewultra$0.060source
Black Forest Labs FLUX.1 Raw Pronewtext-to-image$0.060source
Stability AI Stable Diffusion 3.5 Largestandard$0.065source
Black Forest Labs FLUX.2 Max1MP$0.070source
Black Forest Labs FLUX Kontext [max]max$0.080source
Stability AI Stable Image Ultraflagship$0.080source
Ideogram AI Ideogram v3 Qualitynewquality$0.090source
Mistral AI Agent API - Imagesnewstandard$0.100source
RunPod google / Nano Banana Pro Editnewpro$0.140source
Runway Gen-4 Imagenew1080p$0.192source
Runway Nano Banana Pro 2new2K$0.264source
Runway Nano Banana Pronew2K$0.480source

Video generation

Price your workload →

Per-second list prices, with a 30-second clip shown because that is where the numbers stop feeling small.

ModelTier$/second$/30s clipSource
Black Forest Labs FLUX Video Editnewfast$0.030$0.90source
Kling AI Kling 2.5 Turbo720p$0.042$1.26source
xAI grok-imagine-videonew480p / 720p$0.050$1.50source
fal Wannewstandard$0.050$1.50source
Luma AI Ray3.2720p$0.060$1.80source
Black Forest Labs FLUX Video UpscalenewPrecise$0.070$2.10source
fal Klingnew2.5 Turbo Pro$0.070$2.10source
xAI Grok Imagine Video 1.5new480p$0.080$2.40source
Kling AI Kling 3.0720p$0.084$2.52source
OpenAI sora-2new720p$0.100$3.00source
Kling AI Kling 3.0 Turbo720p + audio$0.112$3.36source
Black Forest Labs FLUX 3 VideoHD full render$0.170$5.10source
Z.AI CogVideoX-3newstandard$0.200$6.00source
Luma AI Ray3.21080p$0.240$7.20source
WaveSpeed AI WAN 2.1 I2Vnew720p$0.250$7.50source
Runway Gen-4.5new1080p$0.288$8.64source
Black Forest Labs FLUX 3newfhd t2v full render$0.290$8.70source
RunPod OpenAI / SORA 2 Pro I2Vnew4s$0.300$9.00source
fal Veonew3$0.400$12.00source
OpenAI Sora 2 Pronew1080p$0.500$15.00source
Runway Aleph 2.0new1080p$0.672$20.16source
Runway Seedance 2.0 Fastnewstandard$0.696$20.88source
Runway Seedance 2.0 Pronew1080p$0.960$28.80source

Speech to text

Price your workload →

Billed on audio duration including silence, so trimming dead air at ingest is a direct saving.

Model$/minute$/audio hourSource
Groq Whisper Large V3 Turbonew$0.0007$0.04source
Groq Whisper Large V3new$0.0019$0.11source
Z.AI GLM-ASR-2512new$0.0024$0.14source
AssemblyAI Universal-2$0.0025$0.15source
AssemblyAI Universal-Streamingnew$0.0025$0.15source
AssemblyAI Universal-Streaming Multilingualnew$0.0025$0.15source
OpenAI gpt-4o-mini-transcribe$0.0030$0.18source
Google Speech-to-Text V2 Dynamic Batch$0.0030$0.18source
Google Gemini 3.5 Transcribenew$0.0030$0.18source
Google Speech-to-Text V2 Dynamic Batch Standardnew$0.0030$0.18source
Mistral AI Voxtral Mini Transcribe 2new$0.0030$0.18source
AssemblyAI Universal-3.5 Pro$0.0035$0.21source
Mistral AI Voxtral Smallnew$0.0040$0.24source
Deepgram Nova-3 monolingual$0.0043$0.26source
OpenAI gpt-transcribe$0.0045$0.27source
Deepgram Whisper Largenew$0.0048$0.29source
Google Gemini 3.5 Transcribe Livenew$0.0050$0.30source
Deepgram Nova-3 multilingual$0.0052$0.31source
OpenAI gpt-4o-transcribe$0.0060$0.36source
Mistral AI Voxtral Mini Transcribe Realtimenew$0.0060$0.36source
ElevenLabs Scribe v2new$0.0067$0.40source
AssemblyAI Universal-3.5 Pro Realtimenew$0.0075$0.45source
Deepgram Flux Englishnew$0.0077$0.46source
Deepgram Flux Multilingualnew$0.0078$0.47source
Google Speech-to-Text V2 Standard$0.0160$0.96source
Google Speech-to-Text V1new$0.0160$0.96source
Google Speech-to-Text V2new$0.0160$0.96source
OpenAI gpt-live-transcribenew$0.0170$1.02source
OpenAI gpt-realtime-whispernew$0.0170$1.02source
Google Speech-to-Text V1 (without data logging)new$0.0240$1.44source
OpenAI gpt-realtime-translatenew$0.0340$2.04source
Google Gemini 3.5 Live Translatenew$0.0368$2.21source
ElevenLabs Scribe v2 (Pro rate)$0.0400$2.40source
AssemblyAI Voice Agent APInew$0.0750$4.50source
Google Speech-to-Text V1 Medicalnew$0.0780$4.68source
Google Speech-to-Text Medical Dictationnew$0.0780$4.68source
Google Speech-to-Text Medical Conversationnew$0.0780$4.68source
RunPod Whisper V3 Largenew$3.0000$180.00source

Text to speech

Price your workload →

Note the two different billing units — per character and per minute are not directly comparable without converting.

ModelList priceUnitSource
Google TTS Standard/WaveNet$4.001M charssource
Deepgram Aura-1$15.001M charssource
xAI Grok Voice APInew$15.001M charssource
xAI grok-voicenew$15.001M charssource
Google TTS Neural2$16.001M charssource
Mistral AI Voxtral TTSnew$16.001M charssource
Google Google TTS Polyglotnew$16.001M charssource
Google TTS Chirp 3 HD$30.001M charssource
Deepgram Aura-2$30.001M charssource
Deepgram Flux TTSnew$45.001M charssource
RunPod Minimax Speech 02 HDnew$50.001M charssource
Google Google TTS Instant custom voicenew$60.001M charssource
Cartesia Sonic-3new$65.001M charssource
Cartesia Sonic-2new$65.001M charssource
Google Google TTS Studio voicesnew$160.001M charssource
Runway Text to Speechnew$480.001M charssource

On-demand hourly rates. The monthly column assumes 730 hours — a rented GPU costs that whether you use it or not.

VendorGPUVRAM$/hour$/month at 100%TypeSource
Vast.aiL40S48 GB$0.40$292marketplacesource
Vast.aiA100 SXM4 80GB80 GB$0.52$380marketplacesource
AWSAWS Trainiumnew32 GB$0.60$435reservedsource
LambdaQuadro RTX 6000new24 GB$0.69$504on-demandsource
Google CloudL4 (G2, per GPU)24 GB$0.71$518on-demandsource
LambdaV100new16 GB$0.79$577on-demandsource
RunPodL40S48 GB$1.09$796on-demandsource
LambdaA6000new48 GB$1.09$796on-demandsource
CoreWeaveL40new48 GB$1.25$912on-demandsource
LambdaA10new24 GB$1.29$942on-demandsource
AWSA100new80 GB$1.48$1,077reservedsource
RunPodA100 80GB PCIe80 GB$1.59$1,161on-demandsource
Vast.aiH100 SXM80 GB$1.60$1,168marketplacesource
LambdaA100 SXMnew40 GB$1.99$1,453on-demandsource
LambdaA100 PCIenew40 GB$1.99$1,453on-demandsource
AWSAWS Trainium2new192 GB$2.23$1,632reservedsource
LambdaGH200new96 GB$2.29$1,672on-demandsource
CoreWeaveRTX PRO 6000 Blackwellnew96 GB$2.50$1,825on-demandsource
Lambda LabsA100 80GB SXM80 GB$2.79$2,037on-demandsource
RunPodH100 PCIe80 GB$2.89$2,110on-demandsource
falRTX PRO 6000new96 GB$2.99$2,183on-demandsource
RunPodH100 NVLnew94 GB$3.19$2,329on-demandsource
RunPodH100 SXM80 GB$3.49$2,548on-demandsource
Vast.aiH200141 GB$3.82$2,789marketplacesource
Lambda LabsH100 SXM80 GB$3.99$2,913on-demandsource
Together AIH10080 GB$3.99$2,913on-demandsource
Together AIHGX H100new80 GB$3.99$2,913on-demandsource
Vast.aiB200192 GB$4.13$3,015marketplacesource
CoreWeaveHGX B300new270 GB$4.48$3,270spotsource
RunPodH200141 GB$4.59$3,351on-demandsource
AWSH100 (p5.48xlarge, per GPU)80 GB$5.19$3,789capacity blocksource
Together AIHGX H200new141 GB$5.99$4,373on-demandsource
CoreWeaveH100 HGX (per GPU)80 GB$6.16$4,497on-demandsource
CoreWeaveH200 HGXnew141 GB$6.31$4,606on-demandsource
Lambda LabsB200 SXM6180 GB$6.69$4,884on-demandsource
RunPodB200180 GB$6.79$4,957on-demandsource
Google CloudB200 (A4 High, per GPU)180 GB$8.06$5,884on-demandsource
Together AIB200180 GB$8.19$5,979on-demandsource
CoreWeaveB200 HGX (per GPU)180 GB$8.60$6,278on-demandsource
CoreWeaveGB200 NVL72new186 GB$10.50$7,665on-demandsource
AWSB200 (p6-b200, per GPU)180 GB$12.36$9,019capacity blocksource
AWSB300new192 GB$14.04$10,249reservedsource

Using these figures

Every number here is a published list price in US dollars, excluding tax. If your organisation has a volume commitment or a negotiated agreement, your real rate is lower and none of these figures apply directly.

To turn a price into a monthly bill you need a workload, which is what the calculators are for. To understand why the input and output columns differ by a factor of five or more, see why output tokens cost more.

The methodology page documents how this index is built, how failures to verify are handled, and which vendors publish prices in a form that cannot be read reliably.