OpenAI · Pricing

GPT-4 Turbo pricing

Complete GPT-4 Turbo API cost reference: $10/M input, $30/M output, 128K token context window, $5/M cached reads. Verified against official sources.

Prices verified August 20, 2026
GPT-4 Turbo
Input / M
$10
Output / M
$30
Cache read
$5
Cache write
$15
Context
128K
Batch
−50%
Vision
Yes
Reasoning

What GPT-4 Turbo costs in practice

WorkloadCost / callCost @ 1K req/day
Light · 1K in / 300 out$0.0190$578.31/mo
Standard · 10K in / 2K out$0.1600$4870.00/mo
Heavy · 100K in / 20K out$1.60$48700.00/mo

Estimates assume medium reasoning effort, no batch discounts and no prompt caching. Turn on prompt caching and the input side typically drops 50–90% — run your own mix in the free calculator preloaded with GPT-4 Turbo.

About this page

Prices were last verified on August 20, 2026 against OpenAI's official pricing page. Providers reprice frequently — see our methodology for exactly how estimates are computed, and always confirm final rates with the provider before budgeting.

GPT-4 Turbo FAQ

How much does GPT-4 Turbo cost?

GPT-4 Turbo charges $10 per million input tokens and $30 per million output tokens, with cached input reads at $5/M. Batch mode cuts both rates by 50% for asynchronous workloads.

What is the context window of GPT-4 Turbo?

GPT-4 Turbo supports a context window of 128,000 tokens (128K) — roughly 96,000 English words or about 192 pages of text in a single request.

Does GPT-4 Turbo support images, audio or reasoning?

GPT-4 Turbo offers vision & image understanding. Legacy flagship. Prefer GPT-4o for new work.

How do I estimate my GPT-4 Turbo bill?

A typical workload of 10K input and 2K output tokens costs about $0.1600 per call with GPT-4 Turbo. At 1,000 requests per day that is roughly $4870.00 per month. Use the free AITokenCalculator to run your exact numbers.