AI API Cost Calculator — Real-Time LLM Pricing

Free calculator for GPT, Claude, Gemini, DeepSeek and 500+ other LLMs. Paste your prompt or enter token counts to see the exact cost across every model on the market. Live data — updated daily from OpenRouter pricing. No signup, no ads on the calculator itself.

What does a typical prompt cost? (live example, August 2026)

Live cost for each major model at common prompt sizes. Numbers update with current OpenRouter pricing.

Prompt typeTokens (in / out)GoogleOpenAIDeepSeekAnthropic
Short Q&A500 / 200$0.00$0.000041$0.0002$0.0004
Code review (1K)1,000 / 500$0.00$0.000095$0.0005$0.0009
Long doc summarization8,000 / 1,000$0.00$0.0004$0.0026$0.0033
Agent task (typical)4,000 / 2,000$0.00$0.0004$0.0019$0.0035
Heavy agent run50,000 / 10,000$0.00$0.0028$0.017$0.025
1M in / 1M out (scale)1,000,000 / 1,000,000$0.00$0.160$0.669$1.50

Assumes each model's flagship tier (GPT-5.4, Claude Opus 4.6, Gemini 2.5 Pro, DeepSeek V3). Switch to cheaper variants like nano or mini in the table below to see dramatic cost drops (often 10–30×).

Top 5 cheapest LLM APIs right now

Lowest-cost models with non-zero pricing, sorted by blended (3:1) input/output price. Use the calculator below to compare for your exact workload.

  1. 1.Ling-2.6-flashinclusionai$0.010 / $0.030 per 1M
  2. 2.Mistral NemoMistral$0.019 / $0.030 per 1M
  3. 3.Ling-3.0-flashinclusionai$0.021 / $0.063 per 1M
  4. 4.Llama 3 8B Lunarissao10k$0.040 / $0.050 per 1M
  5. 5.MythoMax 13Bgryphe$0.060 / $0.060 per 1M

AI API cost FAQ

How is API cost calculated?

Cost = (input_tokens × input_price + output_tokens × output_price) ÷ 1,000,000. Both prices are in USD per 1M tokens. We pull live pricing from OpenRouter's public API, which mirrors what providers charge.

Why do prices vary between providers?

Each provider sets its own rates based on model capability, compute cost, and competitive positioning. Premium reasoning models (Claude Opus, GPT high-effort) cost more than fast variants (nano, mini, flash) because they spend more compute per token.

Which model is cheapest for production?

For most workloads, look at the fast tier: Claude Haiku, GPT-5 mini, Gemini Flash, DeepSeek V3. They're typically 10–30× cheaper than flagship models while staying good enough for chat, classification, and routing. Reserve flagship models for reasoning-heavy tasks.

Does this include caching discounts?

No. This shows list prices. Most providers offer prompt caching (50–90% off cached input) and batch discounts (50% off async jobs). Check each provider's docs for the exact numbers.

Get API access

Found the model that fits your budget? Sign up via one of these providers to start using it. Some links earn us a small commission at no extra cost to you.

See our affiliate disclosure for the full list and how we choose what to recommend.

406 models
★ Best
liquid
Free
LFM2.5-2.6B (free)
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 128K
nvidia
Free
Nemotron 3.5 Lightning (free)
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 1.0M
poolside
Free
Laguna S 2.1 (free)
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 262K
poolside
Free
Laguna XS 2.1 (free)
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 262K
North Mini Code (free)
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 256K
nvidia
Free
Nemotron 3.5 Content Safety (free)
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 128K
nvidia
Free
Nemotron 3 Ultra (free)
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 1.0M
nvidia
Free
Nemotron 3 Nano Omni (free)
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 256K
Gemma 4 26B A4B (free)
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 262K
Gemma 4 31B (free)
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 262K
Lyria 3 Pro Preview
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 1.0M
Lyria 3 Clip Preview
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 1.0M
nvidia
Free
Nemotron 3 Super (free)
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 262K
Free Models Router
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 200K
nvidia
Free
Nemotron 3 Nano 30B A3B (free)
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 256K
nvidia
Free
Nemotron Nano 12B 2 VL (free)
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 128K
nvidia
Free
Nemotron Nano 9B V2 (free)
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 128K
gpt-oss-20b (free)
Input
Free
Output
Free
Est. Cost
<$0.0001
Context: 131K
inclusionai
Fast
Ling-2.6-flash
Input
$0.01
Output
$0.03
Est. Cost
<$0.0001
Context: 262K
Mistral Nemo
Input
$0.02
Output
$0.03
Est. Cost
<$0.0001
Context: 131K
inclusionai
Fast
Ling-3.0-flash
Input
$0.02
Output
$0.06
Est. Cost
<$0.0001
Context: 262K
sao10k
Reasoning
Llama 3 8B Lunaris
Input
$0.04
Output
$0.05
Est. Cost
<$0.0001
Context: 8K
ibm-granite
Flagship
Granite 4.0 Micro
Input
$0.02
Output
$0.11
Est. Cost
<$0.0001
Context: 131K
nex-agi
Fast
Nex-N2-Mini
Input
$0.02
Output
$0.10
Est. Cost
<$0.0001
Context: 262K
Mistral Small 3
Input
$0.05
Output
$0.08
Est. Cost
<$0.0001
Context: 33K
meta-llama
Flagship
Llama 3.1 8B Instruct
Input
$0.05
Output
$0.08
Est. Cost
<$0.0001
Context: 131K
gryphe
Flagship
MythoMax 13B
Input
$0.06
Output
$0.06
Est. Cost
<$0.0001
Context: 8K
upstage
Reasoning
Solar Pro 4
Input
$0.03
Output
$0.12
Est. Cost
<$0.0001
Context: 524K
Qwen
Fast
Qwen3.7 Flash
Input
$0.03
Output
$0.13
Est. Cost
<$0.0001
Context: 1.0M
gpt-oss-20b
Input
$0.03
Output
$0.13
Est. Cost
<$0.0001
Context: 131K
ibm-granite
Flagship
Granite 4.1 8B
Input
$0.05
Output
$0.10
Est. Cost
<$0.0001
Context: 131K
Gemma 3 4B
Input
$0.05
Output
$0.10
Est. Cost
<$0.0001
Context: 131K
amazon
Flagship
Nova Micro 1.0
Input
$0.04
Output
$0.14
Est. Cost
$0.0001
Context: 128K
Command R7B (12-2024)
Input
$0.04
Output
$0.15
Est. Cost
$0.0001
Context: 128K
gpt-oss-120b
Input
$0.03
Output
$0.17
Est. Cost
$0.0001
Context: 131K
poolside
Flagship
Laguna XS 2.1
Input
$0.06
Output
$0.12
Est. Cost
$0.0001
Context: 262K
Gemma 3n 4B
Input
$0.06
Output
$0.12
Est. Cost
$0.0001
Context: 33K
Gemma 3 12B
Input
$0.05
Output
$0.15
Est. Cost
$0.0001
Context: 131K
GPT-5 Nano (batch)
Input
$0.02
Output
$0.20
Est. Cost
$0.0001
Context: 400K
meta-llama
Flagship
Llama 3.2 1B Instruct
Input
$0.03
Output
$0.20
Est. Cost
$0.0001
Context: 60K
microsoft
Flagship
Phi 4
Input
$0.07
Output
$0.14
Est. Cost
$0.0001
Context: 16K
Qwen
Flagship
Qwen3 30B A3B Instruct 2507
Input
$0.05
Output
$0.19
Est. Cost
$0.0001
Context: 262K
rekaai
Flagship
Reka Edge
Input
$0.10
Output
$0.10
Est. Cost
$0.0001
Context: 16K
Ministral 3 3B 2512
Input
$0.10
Output
$0.10
Est. Cost
$0.0001
Context: 131K
nvidia
Fast
Nemotron 3 Nano 30B A3B
Input
$0.05
Output
$0.20
Est. Cost
$0.0001
Context: 262K
Gemini 2.5 Flash Lite (batch)
Input
$0.05
Output
$0.20
Est. Cost
$0.0001
Context: 1.0M
GPT-4.1 Nano (batch)
Input
$0.05
Output
$0.20
Est. Cost
$0.0001
Context: 1.0M
tencent
Flagship
Hy3 preview
Input
$0.06
Output
$0.21
Est. Cost
$0.0002
Context: 262K
DeepSeek V4 Flash 0731
Input
$0.08
Output
$0.18
Est. Cost
$0.0002
Context: 1.0M
Qwen
Flagship
Qwen3.5-9B
Input
$0.10
Output
$0.15
Est. Cost
$0.0002
Context: 262K
poolside
Flagship
Laguna S 2.1
Input
$0.09
Output
$0.18
Est. Cost
$0.0002
Context: 1.0M
amazon
Flagship
Nova Lite 1.0
Input
$0.06
Output
$0.24
Est. Cost
$0.0002
Context: 300K
Qwen
Fast
Qwen3.5-Flash
Input
$0.07
Output
$0.26
Est. Cost
$0.0002
Context: 1.0M
bytedance
Flagship
UI-TARS 7B
Input
$0.10
Output
$0.20
Est. Cost
$0.0002
Context: 128K
rekaai
Fast
Reka Flash 3
Input
$0.10
Output
$0.20
Est. Cost
$0.0002
Context: 66K
Qwen
Flagship
Qwen2.5 7B Instruct
Input
$0.10
Output
$0.20
Est. Cost
$0.0002
Context: 33K
DeepSeek V4 Flash Latest
Input
$0.08
Output
$0.25
Est. Cost
$0.0002
Context: 1.0M
Qwen
Flagship
Qwen3 Coder 30B A3B Instruct
Input
$0.07
Output
$0.28
Est. Cost
$0.0002
Context: 262K
meta-llama
Flagship
Llama 3.2 3B Instruct
Input
$0.05
Output
$0.33
Est. Cost
$0.0002
Context: 131K
Mistral Small 3.2 24B
Input
$0.09
Output
$0.25
Est. Cost
$0.0002
Context: 256K
Qwen
Flagship
Qwen3 32B
Input
$0.08
Output
$0.28
Est. Cost
$0.0002
Context: 131K
Ministral 3 8B 2512
Input
$0.15
Output
$0.15
Est. Cost
$0.0002
Context: 262K
nvidia
Flagship
Nemotron 3.5 Lightning
Input
$0.10
Output
$0.25
Est. Cost
$0.0002
Context: 1.0M
bytedance-seed
Fast
Seed 1.6 Flash
Input
$0.07
Output
$0.30
Est. Cost
$0.0002
Context: 262K
gpt-oss-safeguard-20b
Input
$0.07
Output
$0.30
Est. Cost
$0.0002
Context: 131K
GPT-4o-mini (batch)
Input
$0.07
Output
$0.30
Est. Cost
$0.0002
Context: 128K
Qwen
Flagship
Qwen3 14B
Input
$0.12
Output
$0.24
Est. Cost
$0.0002
Context: 131K
stepfun
Fast
Step 3.5 Flash
Input
$0.10
Output
$0.30
Est. Cost
$0.0003
Context: 262K
Voxtral Small 24B 2507
Input
$0.10
Output
$0.30
Est. Cost
$0.0003
Context: 32K
meta-llama
Flagship
Llama 4 Scout
Input
$0.10
Output
$0.30
Est. Cost
$0.0003
Context: 1.3M
GPT-5 Nano
Input
$0.05
Output
$0.40
Est. Cost
$0.0003
Context: 400K
z-ai
Fast
GLM 4.7 Flash
Input
$0.06
Output
$0.40
Est. Cost
$0.0003
Context: 203K
meta-llama
Flagship
Llama 3.3 70B Instruct
Input
$0.10
Output
$0.32
Est. Cost
$0.0003
Context: 131K
Gemma 4 31B
Input
$0.10
Output
$0.34
Est. Cost
$0.0003
Context: 262K
meta-llama
Flagship
Llama Guard 4 12B
Input
$0.18
Output
$0.18
Est. Cost
$0.0003
Context: 1.0M
DeepSeek V4 Flash 0423
Input
$0.14
Output
$0.28
Est. Cost
$0.0003
Context: 1.0M
xiaomi
Flagship
MiMo-V2.5
Input
$0.14
Output
$0.28
Est. Cost
$0.0003
Context: 1.1M
nvidia
Flagship
Nemotron 3 Super
Input
$0.08
Output
$0.40
Est. Cost
$0.0003
Context: 1.0M
Ministral 3 14B 2512
Input
$0.20
Output
$0.20
Est. Cost
$0.0003
Context: 262K
bytedance-seed
Fast
Seed-2.0-Mini
Input
$0.10
Output
$0.40
Est. Cost
$0.0003
Context: 262K
Gemini 2.5 Flash Lite
Input
$0.10
Output
$0.40
Est. Cost
$0.0003
Context: 1.0M
GPT-4.1 Nano
Input
$0.10
Output
$0.40
Est. Cost
$0.0003
Context: 1.0M
Gemma 3 27B
Input
$0.08
Output
$0.45
Est. Cost
$0.0003
Context: 262K
Qwen
Flagship
Qwen3 VL 32B Instruct
Input
$0.10
Output
$0.42
Est. Cost
$0.0003
Context: 131K
Gemma 4 26B A4B
Input
$0.12
Output
$0.40
Est. Cost
$0.0003
Context: 262K
Nous
Flagship
Hermes 4 70B
Input
$0.13
Output
$0.40
Est. Cost
$0.0003
Context: 131K
Qwen
Flagship
Qwen3 VL 8B Instruct
Input
$0.12
Output
$0.45
Est. Cost
$0.0003
Context: 262K
Qwen
Flagship
Qwen3 8B
Input
$0.12
Output
$0.45
Est. Cost
$0.0003
Context: 131K
Qwen
Flagship
Qwen3 235B A22B Instruct 2507
Input
$0.09
Output
$0.55
Est. Cost
$0.0004
Context: 262K
Qwen
Flagship
Qwen3 30B A3B
Input
$0.12
Output
$0.50
Est. Cost
$0.0004
Context: 131K
inclusionai
Flagship
Ring-2.6-1T
Input
$0.07
Output
$0.63
Est. Cost
$0.0004
Context: 262K
inclusionai
Flagship
Ling-2.6-1T
Input
$0.07
Output
$0.63
Est. Cost
$0.0004
Context: 262K
tencent
Flagship
Hy3
Input
$0.13
Output
$0.53
Est. Cost
$0.0004
Context: 262K
allenai
Flagship
Olmo 3 32B Think
Input
$0.15
Output
$0.50
Est. Cost
$0.0004
Context: 66K
GPT-5.6 Luna Pro
Input
$0.10
Output
$0.60
Est. Cost
$0.0004
Context: 1.1M
GPT-5.6 Luna Pro (batch)
Input
$0.10
Output
$0.60
Est. Cost
$0.0004
Context: 1.1M
GPT-5.6 Luna
Input
$0.10
Output
$0.60
Est. Cost
$0.0004
Context: 1.1M
GPT-5.6 Luna (batch)
Input
$0.10
Output
$0.60
Est. Cost
$0.0004
Context: 1.1M
GPT-5.4 Nano (batch)
Input
$0.10
Output
$0.63
Est. Cost
$0.0004
Context: 400K
tencent
Flagship
Hunyuan A13B Instruct
Input
$0.14
Output
$0.57
Est. Cost
$0.0004
Context: 131K