GPT-4o-mini (2024-07-18)
The cheap multimodal workhorse of 2024. Largely superseded by GPT-4.1 mini, but still cost-competitive and widely deployed.
GPT-4o-mini (2024-07-18) is a multimodal AI model from OpenAI. It costs $0.150 per million input tokens and $0.600 per million output tokens (blended $0.285/M), with a 128,000-token context window.
Profile inherited from upstream GPT-4o mini ↗ — this is a hosted variant of the same open-weights model.
- Cheap multimodal
- Fast
- 128K context
- Cost-sensitive vision
- High-volume chat
- Legacy 4o-mini integrations
Benchmarks
More from OpenAI
See all 28 →Frequently asked questions
How much does GPT-4o-mini (2024-07-18) cost?
GPT-4o-mini (2024-07-18) costs $0.150 per million input tokens and $0.600 per million output tokens, for a blended reference rate of $0.285 per million tokens.
What is GPT-4o-mini (2024-07-18)'s context window?
GPT-4o-mini (2024-07-18) supports up to 128,000 tokens of context in a single request.
What is GPT-4o-mini (2024-07-18) best for?
GPT-4o-mini (2024-07-18) is well suited to Cheap multimodal, Fast and 128K context.
Who makes GPT-4o-mini (2024-07-18)?
GPT-4o-mini (2024-07-18) is developed and served by OpenAI. It was released in Jul 2024.