OpenAI

GPT-4 Turbo

AI model by OpenAI. Real-time pricing and benchmark data.

Pricing (per 1M tokens)

Input$10.00
Output$30.00
Blended (3:1)$15.00

Source: Artificial Analysis

Performance

Output Speed29 tok/s
Time to First Token4040ms

Median values from Artificial Analysis

Editorial Analysis

Profile. GPT-4 Turbo is a premium-tier model from OpenAI at $30.00/M output tokens. Benchmark scores (e.g. Coding Index 21.5) suggest it targets workloads where correctness matters more than cost — long-running agents, code generation at scale, and complex multi-step reasoning.

Latency. On latency, GPT-4 Turbo reports 29 tok/s output speed and 4040ms time-to-first-token. Output speed is on the slower end; for real-time chat consider pairing with a streaming front-end or routing latency-sensitive requests to a faster model.

Cost model. For a workload of 1M input tokens + 100K output tokens, GPT-4 Turbo costs $13.00 per million input-equivalent requests. Multiply by your monthly request volume to project the bill, and compare against the speed/correctness profile above to decide if the trade-off is worth it for your workload.

Editorial summary generated from public pricing and benchmark data (Artificial Analysis API). All numbers are sourced — see the Artificial Analysis methodology for benchmark definitions. Last refreshed 2026-08-24.

Compare with similar models

ModelInputOutputSpeed
GPT-4 TurboCurrent
$10.00$30.0029 tok/s
Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)
$5.00$25.0054 tok/s
Claude Opus 5 (Adaptive Reasoning, Low Effort)
$5.00$25.0049 tok/s
Claude Opus 4.5 (Reasoning)
$5.00$25.0046 tok/s
Claude Opus 4.7 (Non-reasoning, High Effort)
$5.00$25.0044 tok/s
Claude Opus 4.5 (Non-reasoning)
$5.00$25.0045 tok/s

Example Costs

Single Request
$0.0250
1.0K in / 500 out
1K Requests/day
$25.00
1.0M in / 500.0K out
10K Requests/day
$250.00
10.0M in / 5.0M out