OpenAI

o3 API Pricing

"Well-rounded and powerful model across domains" for complex reasoning over text, code, and images. Page notes it has been "succeeded by GPT-5." Knowledge cutoff Jun 1 2024.

Verified against OpenAI’s pricing page on 2026-07-02.

View source ↗

Facts

Legacy

Specs

  • Familyo-series
  • API IDo3
  • ModalityMultimodal
  • Context window200K tokens
  • Max output100K tokens
  • Release dateNot published

Pricing (per 1M tokens)

  • Input$2.00 / 1M tokens
  • Output$8.00 / 1M tokens
  • Cached input$0.50 / 1M tokens
  • Cache write
  • Batch discount

Capabilities

reasoning, vision

Notes

"Well-rounded and powerful model across domains" for complex reasoning over text, code, and images. Page notes it has been "succeeded by GPT-5." Knowledge cutoff Jun 1 2024.

Worked cost examples

Estimated monthly cost on three common usage shapes

WorkloadEst. monthly cost

Chatbot

800 input / 200 output tokens per request, 100,000 requests/month

$320.00

RAG pipeline

6,000 input / 700 output tokens per request, 30,000 requests/month

$528.00

Batch summarization

20,000 input / 1,500 output tokens per request, 5,000 requests/month

$260.00

Price history

Changes we have documented since we began tracking

No price changes recorded since we began tracking.

o3 pricing FAQ

How much does o3 cost per 1M tokens?
o3 costs $2.00 per 1M input tokens and $8.00 per 1M output tokens.
How much does o3 cost per request?
A typical request of 800 input and 200 output tokens costs about $0.0032 on o3.
Is o3 cheaper than o1?
o3 is cheaper than o1 on input tokens, by about 87%.
Does o3 have cache pricing?
Yes — cached input reads cost $0.50 per 1M tokens.