OpenAI
OpenAI
Reasoning

o3-mini

Reasoning

Cheaper reasoning model in the o3 generation. Strong on STEM, lower latency than full o3.

o3-mini is a reasoning AI model from OpenAI. It costs $1.100 per million input tokens and $4.400 per million output tokens (blended $2.090/M), with a 200,000-token context window.

INPUT
$1.100/M
per million input tokens
OUTPUT
$4.400/M
per million output tokens
BLENDED 70/30
$2.090/M
unchanged since 3 May
CONTEXT
200,000
tokens
What it is good at
  • Cheap reasoning
  • STEM-tuned
  • Lower latency than o3
Typical use cases
  • Reasoning at scale
  • STEM tutoring
  • Cost-sensitive code review

Benchmarks

vs. best public score
MMLU86%
Multitask academic knowledge across 57 subjects.
Graduate-level science questions, "Google-proof".
MATH95%
High-school competition math problems.
Python function synthesis from docstrings.
Real GitHub issues solved end-to-end.
LMArena Elo1305 Elo
Crowd-sourced head-to-head preference Elo rating.
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.

More from OpenAI

See all 28 →

Frequently asked questions

How much does o3-mini cost?

o3-mini costs $1.100 per million input tokens and $4.400 per million output tokens, for a blended reference rate of $2.090 per million tokens.

What is o3-mini's context window?

o3-mini supports up to 200,000 tokens of context in a single request.

What is o3-mini best for?

o3-mini is well suited to Cheap reasoning, STEM-tuned and Lower latency than o3.

Who makes o3-mini?

o3-mini is developed and served by OpenAI. It was released in Jan 2025.

Terms used on this page