OpenAI · Pricing

GPT-4o mini pricing

Complete GPT-4o mini API cost reference: $0.15/M input, $0.60/M output, 128K token context window, $0.07/M cached reads. Verified against official sources.

Prices verified August 20, 2026
GPT-4o mini
Input / M
$0.15
Output / M
$0.60
Cache read
$0.07
Cache write
$0.30
Context
128K
Batch
−50%
Vision
Yes
Reasoning

What GPT-4o mini costs in practice

WorkloadCost / callCost @ 1K req/day
Light · 1K in / 300 out$0.0003$10.04/mo
Standard · 10K in / 2K out$0.0027$82.18/mo
Heavy · 100K in / 20K out$0.0270$821.81/mo

Estimates assume medium reasoning effort, no batch discounts and no prompt caching. Turn on prompt caching and the input side typically drops 50–90% — run your own mix in the free calculator preloaded with GPT-4o mini.

How GPT-4o mini compares

About this page

Prices were last verified on August 20, 2026 against OpenAI's official pricing page. Providers reprice frequently — see our methodology for exactly how estimates are computed, and always confirm final rates with the provider before budgeting.

GPT-4o mini FAQ

How much does GPT-4o mini cost?

GPT-4o mini charges $0.15 per million input tokens and $0.60 per million output tokens, with cached input reads at $0.07/M. Batch mode cuts both rates by 50% for asynchronous workloads.

What is the context window of GPT-4o mini?

GPT-4o mini supports a context window of 128,000 tokens (128K) — roughly 96,000 English words or about 192 pages of text in a single request.

Does GPT-4o mini support images, audio or reasoning?

GPT-4o mini offers vision & image understanding, audio input. Cheapest OpenAI multimodal option.

How do I estimate my GPT-4o mini bill?

A typical workload of 10K input and 2K output tokens costs about $0.0027 per call with GPT-4o mini. At 1,000 requests per day that is roughly $82.18 per month. Use the free AITokenCalculator to run your exact numbers.