GPT-4o-mini (batch)
The cheap multimodal workhorse of 2024. Largely superseded by GPT-4.1 mini, but still cost-competitive and widely deployed.
GPT-4o-mini (batch) is a multimodal AI model from OpenAI. It costs $0.075 per million input tokens and $0.300 per million output tokens (blended $0.142/M), with a 128,000-token context window.
Profile inherited from upstream GPT-4o mini ↗ — this is a hosted variant of the same open-weights model.
- Cheap multimodal
- Fast
- 128K context
- Cost-sensitive vision
- High-volume chat
- Legacy 4o-mini integrations
Benchmarks
More from OpenAI
See all 28 →Frequently asked questions
How much does GPT-4o-mini (batch) cost?
GPT-4o-mini (batch) costs $0.075 per million input tokens and $0.300 per million output tokens, for a blended reference rate of $0.142 per million tokens.
What is GPT-4o-mini (batch)'s context window?
GPT-4o-mini (batch) supports up to 128,000 tokens of context in a single request.
What is GPT-4o-mini (batch) best for?
GPT-4o-mini (batch) is well suited to Cheap multimodal, Fast and 128K context.
Who makes GPT-4o-mini (batch)?
GPT-4o-mini (batch) is developed and served by OpenAI. It was released in Jul 2024.