AI Token Price
OpenAI

OpenAI · previous generation

OpenAI o-series API pricing

OpenAI's earlier reasoning models. o1 shuts down on 23 October 2026 and the o3-2025-04-16 snapshot on 11 December 2026.

o4-mini

Previous
Input
$1.10
per 1M tokens
Output
$4.40
per 1M tokens
Cached input
$0.28
75% off input
Batch
50% off
for jobs that can wait
Context window
200K
100K max output
Verified 13 Jun 2026OpenAI pricing

Previous generation. The current replacement is GPT-6 Luna at $0.10 / $0.50 per 1M tokens.

What it costs in practice

o4-mini at list prices, no batch discount.

What o4-mini costs for three workloads
1,000 chat replies1,500 tokens in and 300 out each$2.97
Summarising 100 long reports20,000 tokens in and 800 out each$2.55
A month of a coding agent400 tasks, 85% of input cached$334

The coding-agent example assumes 85% of input is served from cache, typical for agents that resend the same context each step. Model your own workload

Details

API model IDo4-mini
Released16 Apr 2025
StatusPrevious generation, still available

Checked 13 Jun 2026 against OpenAI pricing .

Other OpenAI o-series versions

GPT-6 models reason natively; OpenAI points o-series users to them.

Other OpenAI o-series models and prices per million tokens
Model Input / 1M Output / 1M Cached input Context Released
o1Shuts down 23 Oct
$15.00 $60.00 $7.50 200K 17 Dec 2024
o3Shuts down 11 Dec
$2.00 $8.00 $0.50 200K 16 Apr 2025

Version details

o1

Shuts down 23 Oct

$15.00 / $60.00 per 1M tokenscached $7.50 · batch 50% off · 200K context · 100K max output

Released 17 Dec 2024 · o1

  • Shuts down on 23 October 2026.

Newer option: GPT-6 Astra at $10.00 / $50.00. Checked 27 Sep 2026 against OpenAI deprecations .

o3

Shuts down 11 Dec

$2.00 / $8.00 per 1M tokenscached $0.50 · batch 50% off · 200K context · 100K max output

Released 16 Apr 2025 · o3

  • OpenAI cut the price by 80% on 10 June 2025.
  • The o3-2025-04-16 snapshot shuts down on 11 December 2026.

Newer option: GPT-6 Sol at $2.00 / $10.00. Checked 27 Sep 2026 against OpenAI pricing and OpenAI deprecations .

Compare o4-mini

No ready-made comparisons for this family yet.

Compare with any model

Price history

All changes

No price changes recorded since we began tracking in June 2026.

Coming up

  • OpenAI
    OpenAI shuts down o1 and the gpt-4o-2024-05-13 snapshot. Source
  • OpenAI
    OpenAI shuts down the o3-2025-04-16 snapshot. Source

OpenAI pricing rules

  • Batch and Flex processing cost 50% less.
  • For GPT-5.4 and later, prompts longer than 272K tokens are billed at 2× the input price and 1.5× the output price.
  • Cached input is billed at about a tenth of the input price.

All OpenAI models

Questions

How much does o4-mini cost?

As of 27 Sep 2026, o4-mini costs $1.10 per million input tokens and $4.40 per million output tokens, and $0.275 per million cached input tokens. Batch requests cost 50% less. Prices checked 13 Jun 2026.

What is o4-mini's context window?

200K tokens, with up to 100K tokens of output.

Does o4-mini support prompt caching?

Yes. Repeated input read from the cache costs $0.275 per million tokens, 75% less than normal input.