GPT-4-32k
Classic
Long-context variant of original GPT-4 (32K). Retired, superseded by GPT-4 Turbo and 4.1.
GPT-4-32k is a frontier AI model from OpenAI. It costs $60.000 per million input tokens and $120.000 per million output tokens (blended $78.000/M), with a 32,000-token context window.
INPUT
$60.000/M
per million input tokens
OUTPUT
$120.000/M
per million output tokens
CONTEXT
32,000
tokens
Benchmarks
No published benchmark scores tracked for this model yet. Frontier and reasoning models from major providers have scores; smaller models and inference-host variants typically inherit the underlying open-weights score.
More from OpenAI
See all 28 →GPT-5.6 Sol
Frontier · 1,050,000 ctx
in $5.000/Mout $30.000/M
GPT-5.6 Terra
Balanced · 1,050,000 ctx
in $2.000/Mout $12.000/M
GPT-5.6 Luna
Efficient · 1,050,000 ctx
in $0.200/Mout $1.200/M
GPT-5.5
Frontier · 1,050,000 ctx
in $5.000/Mout $30.000/M
GPT-5.2
Frontier · 400,000 ctx
in $1.750/Mout $14.000/M
GPT-5.2-Codex
Coding · 400,000 ctx
in $1.750/Mout $14.000/M
Frequently asked questions
How much does GPT-4-32k cost?
GPT-4-32k costs $60.000 per million input tokens and $120.000 per million output tokens, for a blended reference rate of $78.000 per million tokens.
What is GPT-4-32k's context window?
GPT-4-32k supports up to 32,000 tokens of context in a single request.
What is GPT-4-32k best for?
Long-context variant of original GPT-4 (32K). Retired, superseded by GPT-4 Turbo and 4.1.
Who makes GPT-4-32k?
GPT-4-32k is developed and served by OpenAI.