Estimates assume medium reasoning effort, no batch discounts and no prompt caching. Turn on
prompt caching and the input side typically drops 50–90% —
run your own mix in the free calculator preloaded with GPT-4.1.
About this page
Prices were last verified on August 20, 2026 against OpenAI's official pricing page. Providers reprice
frequently — see our methodology for exactly how estimates are computed,
and always confirm final rates with the provider before budgeting.
GPT-4.1 FAQ
How much does GPT-4.1 cost?
GPT-4.1 charges $2 per million input tokens and $8 per million output tokens, with cached input reads at $1/M. Batch mode cuts both rates by 50% for asynchronous workloads.
What is the context window of GPT-4.1?
GPT-4.1 supports a context window of 1,047,576 tokens (1.0M) — roughly 785,682 English words or about 1,571 pages of text in a single request.
Does GPT-4.1 support images, audio or reasoning?
GPT-4.1 offers vision & image understanding. Current production recommendation. 1M context, 50% batch, cached input up to 90% off.
How do I estimate my GPT-4.1 bill?
A typical workload of 10K input and 2K output tokens costs about $0.0360 per call with GPT-4.1. At 1,000 requests per day that is roughly $1095.75 per month. Use the free AITokenCalculator to run your exact numbers.