o4 Mini High
text+image+file->text
o4 Mini High is a compact variant — chosen for high throughput and low cost rather than absolute capability. Available via OpenAI. 200,000-token context. Mid-tier pricing.
o4 Mini High is a multimodal AI model from OpenAI. It costs $1.100 per million input tokens and $4.400 per million output tokens (blended $2.090/M), with a 200,000-token context window.
INPUT
$1.100/M
per million input tokens
OUTPUT
$4.400/M
per million output tokens
CONTEXT
200,000
tokens
What it is good at
- Sub-second latency
- Affordable at scale
- Good for routing & classification
- Multimodal: handles images alongside text
Typical use cases
- Routing & classification
- Bulk summarisation
- Realtime chat UX
Benchmarks
No published benchmark scores tracked for this model yet. Frontier and reasoning models from major providers have scores; smaller models and inference-host variants typically inherit the underlying open-weights score.
More from OpenAI
See all 28 →GPT-5.6 Sol
Frontier · 1,050,000 ctx
in $5.000/Mout $30.000/M
GPT-5.6 Terra
Balanced · 1,050,000 ctx
in $2.000/Mout $12.000/M
GPT-5.6 Luna
Efficient · 1,050,000 ctx
in $0.200/Mout $1.200/M
GPT-5.5
Frontier · 1,050,000 ctx
in $5.000/Mout $30.000/M
GPT-5.2
Frontier · 400,000 ctx
in $1.750/Mout $14.000/M
GPT-5.2-Codex
Coding · 400,000 ctx
in $1.750/Mout $14.000/M
Frequently asked questions
How much does o4 Mini High cost?
o4 Mini High costs $1.100 per million input tokens and $4.400 per million output tokens, for a blended reference rate of $2.090 per million tokens.
What is o4 Mini High's context window?
o4 Mini High supports up to 200,000 tokens of context in a single request.
What is o4 Mini High best for?
o4 Mini High is well suited to Sub-second latency, Affordable at scale and Good for routing & classification.
Who makes o4 Mini High?
o4 Mini High is developed and served by OpenAI.