OpenAI
OpenAI
EfficientRETIRED

GPT-3.5 Turbo Instruct

Legacy

Completion-style GPT-3.5 ("instruct"). Retired, use chat completions on a current model.

GPT-3.5 Turbo Instruct is a efficient AI model from OpenAI. It costs $1.500 per million input tokens and $2.000 per million output tokens (blended $1.650/M), with a 4,000-token context window.

INPUT
$1.500/M
per million input tokens
OUTPUT
$2.000/M
per million output tokens
BLENDED 70/30
$1.650/M
unchanged since 3 May
CONTEXT
4,000
tokens

Benchmarks

No published benchmark scores tracked for this model yet. Frontier and reasoning models from major providers have scores; smaller models and inference-host variants typically inherit the underlying open-weights score.

More from OpenAI

See all 28 →

Frequently asked questions

How much does GPT-3.5 Turbo Instruct cost?

GPT-3.5 Turbo Instruct costs $1.500 per million input tokens and $2.000 per million output tokens, for a blended reference rate of $1.650 per million tokens.

What is GPT-3.5 Turbo Instruct's context window?

GPT-3.5 Turbo Instruct supports up to 4,000 tokens of context in a single request.

What is GPT-3.5 Turbo Instruct best for?

Completion-style GPT-3.5 ("instruct"). Retired, use chat completions on a current model.

Who makes GPT-3.5 Turbo Instruct?

GPT-3.5 Turbo Instruct is developed and served by OpenAI.

Terms used on this page