AI Token Price
OpenAI

OpenAI · previous generation

GPT-4o API pricing

OpenAI's 2024 multimodal model, still available at $2.50 / $10 per 1M tokens as a legacy option. The gpt-4o-2024-05-13 snapshot shuts down on 23 October 2026.

GPT-4o

Previous
Input
$2.50
per 1M tokens
Output
$10.00
per 1M tokens
Cached input
$1.25
50% off input
Batch
50% off
for jobs that can wait
Context window
128K
16K max output
Verified 27 Sep 2026OpenAI pricing

Previous generation. The current replacement is GPT-6 Sol at $2.00 / $10.00 per 1M tokens.

What it costs in practice

GPT-4o at list prices, no batch discount.

What GPT-4o costs for three workloads
1,000 chat replies1,500 tokens in and 300 out each$6.75
Summarising 100 long reports20,000 tokens in and 800 out each$5.80
A month of a coding agent400 tasks, 85% of input cached$1,158

The coding-agent example assumes 85% of input is served from cache, typical for agents that resend the same context each step. Model your own workload

Details

API model IDgpt-4o
Released13 May 2024
StatusPrevious generation, still available
  • The gpt-4o-2024-05-13 snapshot shuts down on 23 October 2026. Later snapshots have no announced shutdown.

Checked 27 Sep 2026 against OpenAI pricing and OpenAI deprecations .

Compare GPT-4o

No ready-made comparisons for this family yet.

Compare with any model

Price history

All changes

No price changes recorded since we began tracking in June 2026.

Coming up

  • OpenAI
    OpenAI shuts down o1 and the gpt-4o-2024-05-13 snapshot. Source

OpenAI pricing rules

  • Batch and Flex processing cost 50% less.
  • For GPT-5.4 and later, prompts longer than 272K tokens are billed at 2× the input price and 1.5× the output price.
  • Cached input is billed at about a tenth of the input price.

All OpenAI models

Questions

How much does GPT-4o cost?

As of 27 Sep 2026, GPT-4o costs $2.50 per million input tokens and $10 per million output tokens, and $1.25 per million cached input tokens. Batch requests cost 50% less. Prices checked 27 Sep 2026.

What is GPT-4o's context window?

128K tokens, with up to 16K tokens of output.

Does GPT-4o support prompt caching?

Yes. Repeated input read from the cache costs $1.25 per million tokens, 50% less than normal input.