GPT-4o mini
Efficient
The cheap multimodal workhorse of 2024. Largely superseded by GPT-4.1 mini, but still cost-competitive and widely deployed.
GPT-4o mini is a efficient AI model from OpenAI. It costs $0.150 per million input tokens and $0.600 per million output tokens (blended $0.285/M), with a 128,000-token context window.
INPUT
$0.150/M
per million input tokens
OUTPUT
$0.600/M
per million output tokens
CONTEXT
128,000
tokens
What it is good at
- Cheap multimodal
- Fast
- 128K context
Typical use cases
- Cost-sensitive vision
- High-volume chat
- Legacy 4o-mini integrations
Benchmarks
vs. best public score
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.
More from OpenAI
See all 28 →GPT-5.6 Sol
Frontier · 1,050,000 ctx
in $5.000/Mout $30.000/M
GPT-5.6 Terra
Balanced · 1,050,000 ctx
in $2.000/Mout $12.000/M
GPT-5.6 Luna
Efficient · 1,050,000 ctx
in $0.200/Mout $1.200/M
GPT-5.5
Frontier · 1,050,000 ctx
in $5.000/Mout $30.000/M
GPT-5.2
Frontier · 400,000 ctx
in $1.750/Mout $14.000/M
GPT-5.2-Codex
Coding · 400,000 ctx
in $1.750/Mout $14.000/M
Frequently asked questions
How much does GPT-4o mini cost?
GPT-4o mini costs $0.150 per million input tokens and $0.600 per million output tokens, for a blended reference rate of $0.285 per million tokens.
What is GPT-4o mini's context window?
GPT-4o mini supports up to 128,000 tokens of context in a single request.
What is GPT-4o mini best for?
GPT-4o mini is well suited to Cheap multimodal, Fast and 128K context.
Who makes GPT-4o mini?
GPT-4o mini is developed and served by OpenAI. It was released in Jul 2024.