OpenAI o-series API pricing
OpenAI's earlier reasoning models. o1 shuts down on 23 October 2026 and the o3-2025-04-16 snapshot on 11 December 2026.
o4-mini
PreviousPrevious generation. The current replacement is GPT-6 Luna at $0.10 / $0.50 per 1M tokens.
What it costs in practice
o4-mini at list prices, no batch discount.
| 1,000 chat replies1,500 tokens in and 300 out each | $2.97 |
|---|---|
| Summarising 100 long reports20,000 tokens in and 800 out each | $2.55 |
| A month of a coding agent400 tasks, 85% of input cached | $334 |
The coding-agent example assumes 85% of input is served from cache, typical for agents that resend the same context each step. Model your own workload
Details
o4-miniChecked 13 Jun 2026 against OpenAI pricing .
Other OpenAI o-series versions
GPT-6 models reason natively; OpenAI points o-series users to them.
Version details
o1
Shuts down 23 Oct$15.00 / $60.00 per 1M tokenscached $7.50 · batch 50% off · 200K context · 100K max output
- Shuts down on 23 October 2026.
Newer option: GPT-6 Astra at $10.00 / $50.00. Checked 27 Sep 2026 against OpenAI deprecations .
o3
Shuts down 11 Dec$2.00 / $8.00 per 1M tokenscached $0.50 · batch 50% off · 200K context · 100K max output
- OpenAI cut the price by 80% on 10 June 2025.
- The o3-2025-04-16 snapshot shuts down on 11 December 2026.
Newer option: GPT-6 Sol at $2.00 / $10.00. Checked 27 Sep 2026 against OpenAI pricing and OpenAI deprecations .
Compare o4-mini
No ready-made comparisons for this family yet.
Price history
All changesNo price changes recorded since we began tracking in June 2026.
Coming up
OpenAI pricing rules
- Batch and Flex processing cost 50% less.
- For GPT-5.4 and later, prompts longer than 272K tokens are billed at 2× the input price and 1.5× the output price.
- Cached input is billed at about a tenth of the input price.
Questions
How much does o4-mini cost?
As of 27 Sep 2026, o4-mini costs $1.10 per million input tokens and $4.40 per million output tokens, and $0.275 per million cached input tokens. Batch requests cost 50% less. Prices checked 13 Jun 2026.
What is o4-mini's context window?
200K tokens, with up to 100K tokens of output.
Does o4-mini support prompt caching?
Yes. Repeated input read from the cache costs $0.275 per million tokens, 75% less than normal input.