AI API pricing, compared

Compare API prices for text, TTS, image, and video models across every major provider and cloud channel — preserving each provider's native billing unit. Click any model for the full breakdown.

Headline prices today

Snapshots from our verified table below. Every row links to its official source and last-checked date.

Model Input /1M Output /1M Cached /1M Context Details
GPT-5.5 OpenAI
$5.00 $30.00 $0.50 1,050,000 View →
Claude Opus 4.8 Anthropic
$5.00 $25.00 $0.50 1M View →
Gemini 3.1 Pro Google Gemini
$2.00 $12.00 $0.20 1,048,576 View →
DeepSeek V4 Pro DeepSeek
$0.435 $0.87 $0.0036 1M View →
Mistral Large 3 Mistral
$0.50 $1.50 256K View →

Budget picks in the same families:

  • GPT-5.4 nano

    OpenAI · Classification, Low latency, High volume

    $0.20 /1M in · $1.25 /1M out

  • Claude Haiku 4.5

    Anthropic · Fast, Extraction, Routing

    $1.00 /1M in · $5.00 /1M out

  • Gemini 3.1 Flash-Lite

    Google Gemini · Low cost, High volume, Fast

    $0.25 /1M in · $1.50 /1M out

Full price table (filterable)

42 modelsPrices per 1M tokens · Checked 2026-06-18
Model & channelInputOutputCached inContextBest for
GPT-5.5OpenAI
$5.00/1M$30.00/1M$0.501,050,000
ReasoningCodingAgentic
3 channels
Official APIDirectVerified
$5.00/1M$30.00/1M$0.501,050,000
Azure OpenAICloudVerified
$5.00/1M$30.00/1M$0.501,050,000
OpenRouterRouterVerified
$5.00/1M$30.00/1M1,050,000
$30.00/1M$180.00/1M1,050,000
ReasoningHigh-stakes
2 channels
Official APIDirectVerified
$30.00/1M$180.00/1M1,050,000
OpenRouterRouterVerified
$30.00/1M$180.00/1M1,050,000
GPT-5.4OpenAI
$2.50/1M$15.00/1M$0.251,050,000
ReasoningCodingGeneral
3 channels
Official APIDirectVerified
$2.50/1M$15.00/1M$0.251,050,000
Azure OpenAICloudVerified
$2.50/1M$15.00/1M$0.251,050,000
OpenRouterRouterVerified
$2.50/1M$15.00/1M1,050,000
$0.75/1M$4.50/1M$0.075400,000
Low latencyCodingAgents at scale
3 channels
Official APIDirectVerified
$0.75/1M$4.50/1M$0.075400,000
Azure OpenAICloudVerified
$0.75/1M$4.50/1M$0.075400,000
OpenRouterRouterVerified
$0.75/1M$4.50/1M400,000
$0.20/1M$1.25/1M$0.02400,000
ClassificationLow latencyHigh volume
3 channels
Official APIDirectVerified
$0.20/1M$1.25/1M$0.02400,000
Azure OpenAICloudAuto
$0.20/1M$1.25/1M$0.02400,000
OpenRouterRouterVerified
$0.20/1M$1.25/1M400,000
$1.75/1M$14.00/1M$0.175400,000
Agentic codingSoftware engineeringTool use
3 channels
Official APIDirectVerified
$1.75/1M$14.00/1M$0.175400,000
Azure OpenAICloudAuto
$1.75/1M$14.00/1M$0.175400,000
OpenRouterRouterVerified
$1.75/1M$14.00/1M400,000
$5.00/1M$25.00/1M$0.501M
Hard reasoningAgentsCoding
4 channels
Claude APIDirectVerified
$5.00/1M$25.00/1M$0.501M
AWS BedrockCloudAuto
$5.00/1M$25.00/1M$0.501M
$5.00/1M$25.00/1M$0.501M
OpenRouterRouterVerified
$5.00/1M$25.00/1M1M
$5.00/1M$25.00/1M$0.501M
Hard reasoningAgents
4 channels
Claude APIDirectVerified
$5.00/1M$25.00/1M$0.501M
AWS BedrockCloudAuto
$5.00/1M$25.00/1M$0.501M
$5.00/1M$25.00/1M$0.501M
OpenRouterRouterVerified
$5.00/1M$25.00/1M1M
$3.00/1M$15.00/1M$0.301M
CodingAgentsProduction
3 channels
Claude APIDirectVerified
$3.00/1M$15.00/1M$0.301M
$3.00/1M$15.00/1M$0.301M
OpenRouterRouterVerified
$3.00/1M$15.00/1M1M
$3.00/1M$15.00/1M$0.301M
CodingAgents
3 channels
Claude APIDirectVerified
$3.00/1M$15.00/1M$0.301M
$3.00/1M$15.00/1M$0.301M
OpenRouterRouterVerified
$3.00/1M$15.00/1M1M
$1.00/1M$5.00/1M$0.10200K
FastExtractionRouting
4 channels
Claude APIDirectVerified
$1.00/1M$5.00/1M$0.10200K
AWS BedrockCloudAuto
$1.00/1M$5.00/1M$0.10200K
$1.00/1M$5.00/1M$0.10200K
OpenRouterRouterVerified
$1.00/1M$5.00/1M200K
Gemini 3.1 ProGoogle Gemini
$2.00/1M$12.00/1M$0.201,048,576
ReasoningAgenticLong context
3 channels
Gemini Developer APIDirectVerified
$2.00/1M$12.00/1M$0.201,048,576
Google Vertex AICloudVerified
$2.00/1M$12.00/1M$0.201,048,576
OpenRouterRouterAuto
$2.00/1M$12.00/1M1,048,576
Gemini 3.5 FlashGoogle Gemini
$1.50/1M$9.00/1M$0.151,048,576
CodingAgenticFast
3 channels
Gemini Developer APIDirectVerified
$1.50/1M$9.00/1M$0.151,048,576
Google Vertex AICloudVerified
$1.50/1M$9.00/1M$0.151,048,576
OpenRouterRouterVerified
$1.50/1M$9.00/1M1,048,576
$0.25/1M$1.50/1M$0.0251,048,576
Low costHigh volumeFast
3 channels
Gemini Developer APIDirectVerified
$0.25/1M$1.50/1M$0.0251,048,576
Google Vertex AICloudVerified
$0.25/1M$1.50/1M$0.0251,048,576
OpenRouterRouterVerified
$0.25/1M$1.50/1M1,048,576
Gemini 2.5 ProGoogle Gemini
$1.25/1M$10.00/1M$0.1251,048,576
ReasoningLong contextMultimodal
3 channels
Gemini Developer APIDirectVerified
$1.25/1M$10.00/1M$0.1251,048,576
Google Vertex AICloudVerified
$1.25/1M$10.00/1M$0.131,048,576
OpenRouterRouterVerified
$1.25/1M$10.00/1M1,048,576
Gemini 2.5 FlashGoogle Gemini
$0.30/1M$2.50/1M$0.031,048,576
Low latencyHigh volumeMultimodal
3 channels
Gemini Developer APIDirectVerified
$0.30/1M$2.50/1M$0.031,048,576
Google Vertex AICloudVerified
$0.30/1M$2.50/1M$0.031,048,576
OpenRouterRouterVerified
$0.30/1M$2.50/1M1,048,576
$0.10/1M$0.40/1M$0.011,048,576
Lowest costHigh volume
3 channels
Gemini Developer APIDirectVerified
$0.10/1M$0.40/1M$0.011,048,576
Google Vertex AICloudVerified
$0.10/1M$0.40/1M$0.011,048,576
OpenRouterRouterVerified
$0.10/1M$0.40/1M1,048,576
$1.25/1M$2.50/1M$0.201M
ReasoningTool callingAgents
3 channels
Official APIDirectVerified
$1.25/1M$2.50/1M$0.201M
OpenRouterRouterVerified
$1.25/1M$2.50/1M1M
$1.25/1M$2.50/1M200K (Foundry cap)
DeepSeek V4 FlashOpenDeepSeek
$0.14/1M$0.28/1M$0.00281M
Open weightLow costHigh volume
2 channels
Official APIDirectVerified
$0.14/1M$0.28/1M$0.00281M
OpenRouterRouterAuto
$0.09/1M$0.18/1M1M
DeepSeek V4 ProOpenDeepSeek
$0.435/1M$0.87/1M$0.00361M
Open weightReasoningCoding
4 channels
Official APIDirectVerified
$0.435/1M$0.87/1M$0.00361M
OpenRouterRouterAuto
$0.435/1M$0.87/1M1M
Together AIHostVerified
$2.10/1M$4.40/1M$0.20512K
FireworksHostAuto
$1.74/1M$3.48/1M1M
Mistral Large 3OpenMistral
$0.50/1M$1.50/1M256K
Open weightAgenticMultilingual
2 channels
La PlateformeDirectVerified
$0.50/1M$1.50/1M256K
AWS BedrockCloudAuto
$0.50/1M$1.50/1M128K
Mistral Medium 3.5Mistral · La Plateforme
$1.50/1M$7.50/1M128K
BalancedEnterprise
Mistral Small 4OpenMistral · La Plateforme
$0.10/1M$0.30/1M128K
Open weightLow costEdge
Devstral 2OpenMistral · La Plateforme
$0.40/1M$2.00/1M256K
Open weightCodingAgentic
$0.10/1M$0.32/1M131K
Open weightGeneralCheap
5 channels
OpenRouterRouterVerified
$0.10/1M$0.32/1M131K
GroqHostVerified
$0.59/1M$0.79/1M128K
AWS BedrockCloudAuto
$0.72/1M$0.72/1M128K
FireworksHostAuto
$0.90/1M$0.90/1M$0.45131K
Together AIHostVerified
$1.04/1M$1.04/1M131K
$0.10/1M$0.30/1M10M
Open weightUltra-long contextMultimodal
3 channels
OpenRouterRouterVerified
$0.10/1M$0.30/1M10M
GroqHostVerified
$0.11/1M$0.34/1M128K
FireworksHostAuto
$0.15/1M$0.60/1M$0.075131K
$0.15/1M$0.60/1M1M
Open weightReasoningMultimodal
4 channels
OpenRouterRouterVerified
$0.15/1M$0.60/1M1M
FireworksHostAuto
$0.22/1M$0.88/1M$0.11131K
AWS BedrockCloudAuto
$0.24/1M$0.97/1M1M
Together AIHostVerified
$0.27/1M$0.85/1M1M
Qwen3 235B A22BOpenAlibaba
$0.09/1M$0.10/1M262K
Open weightReasoningMoE
3 channels
OpenRouterRouterVerified
$0.09/1M$0.10/1M262K
Together AIHostNeeds review
$0.20/1M$0.60/1M131K
FireworksHostAuto
$0.22/1M$0.88/1M$0.11131K
Qwen3 Coder 480BOpenAlibaba · OpenRouter
$0.22/1M$1.80/1M1M
Open weightCodingAgentic
Qwen3 32BOpenAlibaba · Groq
$0.29/1M$0.59/1M131K
Open weightFastDense
Gemma 3 27BOpenGoogle · OpenRouter
$0.08/1M$0.16/1M131K
Open weightSingle-GPUMultimodal
Gemma 3 12BOpenGoogle · OpenRouter
$0.05/1M$0.15/1M131K
Open weightCheapSmall
Phi-4 (14B)OpenMicrosoft · OpenRouter
$0.065/1M$0.14/1M16K
Open weightSmallEfficient
Kimi K2.6OpenMoonshot
$0.95/1M$4.00/1M$0.16262K
Open weightReasoningMultimodal
4 channels
Official APIDirectVerified
$0.95/1M$4.00/1M$0.16262K
OpenRouterRouterVerified
$0.68/1M$3.41/1M262K
Together AIHostAuto
$1.20/1M$4.50/1M$0.20262K
FireworksHostAuto
$0.95/1M$4.00/1M262K
Kimi K2.7 CodeOpenMoonshot
$0.95/1M$4.00/1M$0.19262K
Open weightCodingAgentic+1
3 channels
OpenRouterRouterAuto
$0.74/1M$3.50/1M262K
Official APIDirectVerified
$0.95/1M$4.00/1M$0.19262K
/1M/1M262K
MiniMax M3OpenMiniMax
$0.30/1M$1.20/1M$0.061,048,576
Open weightAgenticCoding+1
3 channels
Official APIDirectVerified
$0.30/1M$1.20/1M$0.061,048,576
$0.60/1M$2.40/1M$0.121,048,576
OpenRouterRouterVerified
$0.30/1M$1.20/1M1,048,576
MiniMax M2.7OpenMiniMax
$0.30/1M$1.20/1M$0.06Needs review
Open weightAgenticAffordable
3 channels
Official APIDirectVerified
$0.30/1M$1.20/1M$0.06Needs review
Official HighSpeedHostVerified
$0.60/1M$2.40/1M$0.06Needs review
OpenRouterRouterAuto
$0.30/1M$1.20/1MNeeds review
MiniMax M2.5OpenMiniMax
$0.30/1M$1.20/1M$0.03Needs review
Open weightLegacyAffordable
2 channels
Official APIDirectVerified
$0.30/1M$1.20/1M$0.03Needs review
Official HighSpeedHostVerified
$0.60/1M$2.40/1M$0.03Needs review
GLM-5.2OpenZhipu
$1.40/1M$4.40/1M$0.261,048,576
Open weightCodingReasoning+1
3 channels
Official API (Z.ai)DirectVerified
$1.40/1M$4.40/1M$0.261,048,576
DeepInfraHostVerified
$1.40/1M$4.40/1M$0.251,048,576
OpenRouterRouterVerified
$1.40/1M$4.40/1M1,048,576
GLM-5.1OpenZhipu
$1.40/1M$4.40/1M$0.26Needs review
Open weightCodingAgentic
2 channels
OpenRouterRouterAuto
$0.98/1M$3.08/1MNeeds review
Official API (Z.ai)DirectVerified
$1.40/1M$4.40/1M$0.26Needs review
Qwen3.7 MaxOpenAlibaba
$1.25/1M$3.75/1M$0.13Needs review
Open weightReasoningAgentic
2 channels
Together AIHostVerified
$1.25/1M$3.75/1M$0.13Needs review
OpenRouterRouterAuto
$1.25/1M$3.75/1MNeeds review
Qwen3.7 PlusAlibaba · Requesty
$0.32/1M$1.28/1MNeeds review
AffordableGeneralAlibaba

Frequently asked questions

What is the cheapest AI API in 2026?
As of 2026-06-18, the cheapest verified text API is Gemini 2.5 Flash-Lite at $0.10 per 1M input tokens and $0.40 per 1M output tokens, with cached input at $0.01 per 1M; DeepSeek V4 Flash is cheaper on output at $0.14/$0.28, and $0.0028/1M for cache hits. For image generation, MiniMax image-01 is the lowest official row we track at $0.0035/image; OpenAI's cheapest image tier is GPT Image 1 Mini from about $0.005/image. For video, Google's Veo 3.1 Lite is the lowest official row we verified at $0.05/sec for 720p.
How much does GPT-5 cost per 1M tokens?
As of 2026-06-17, GPT-5.5 (OpenAI direct) costs $5.00 per 1M input tokens and $30.00 per 1M output tokens, with cached input at $0.50 per 1M. GPT-5.4 is $2.50/$15.00. GPT-5.4 mini is $0.75/$4.50. GPT-5.4 nano is $0.20/$1.25. Prices are identical on Azure OpenAI for the same models.
How much does Claude API cost per 1M tokens?
As of 2026-06-17, Claude Opus 4.8 costs $5.00 per 1M input tokens and $25.00 per 1M output tokens (cached read $0.50). Claude Sonnet 4.6 is $3.00/$15.00 (cached read $0.30). Claude Haiku 4.5 is $1.00/$5.00 (cached read $0.10).
What is prompt caching and how much does it save?
Prompt caching reuses a previously seen input prefix at a discounted rate. On GPT-5.5 cached input is $0.50 per 1M versus $5.00 fresh — a 90% discount. On Claude Sonnet 4.6 cached read is $0.30 versus $3.00 fresh — also 90% off. Anthropic additionally separates 5-minute and 1-hour cache writes.
Is OpenRouter the same price as the official API?
Often yes. As of 2026-06-17, OpenRouter lists Claude Sonnet 4.6 at $3.00/$15.00 per 1M — identical to Anthropic direct. The trade-off: OpenRouter hides fine-grained cache read/write tiers, so if your workload depends on cache economics, the direct API exposes the cleaner price.
Which AI coding tool costs $20/month?
As of 2026-06-17, Cursor Pro is $20/month, ChatGPT Plus is $20/month, Claude Pro is $20/month (with confirmed $200/year annual), and Perplexity Pro is $20/month. GitHub Copilot Pro is the cheapest at $10/user/month.