OpenAI
OpenAI
EfficientAUTO PROFILELIVE PRICING

GPT-3.5 Turbo 16k

text->text

GPT-3.5 Turbo 16k is a general-purpose chat-tuned model. Available via OpenAI. 16,385-token context. Mid-tier pricing.

GPT-3.5 Turbo 16k is a efficient AI model from OpenAI. It costs $3.000 per million input tokens and $4.000 per million output tokens (blended $3.300/M), with a 16,385-token context window.

INPUT
$3.000/M
per million input tokens
OUTPUT
$4.000/M
per million output tokens
BLENDED 70/30
$3.300/M
unchanged since 3 May
CONTEXT
16,385
tokens
What it is good at
  • Solid general chat performance
  • Reasonable instruction following
Typical use cases
  • General chat
  • Drafting and rewriting
  • Personal productivity

Benchmarks

No published benchmark scores tracked for this model yet. Frontier and reasoning models from major providers have scores; smaller models and inference-host variants typically inherit the underlying open-weights score.

More from OpenAI

See all 28 →

Frequently asked questions

How much does GPT-3.5 Turbo 16k cost?

GPT-3.5 Turbo 16k costs $3.000 per million input tokens and $4.000 per million output tokens, for a blended reference rate of $3.300 per million tokens.

What is GPT-3.5 Turbo 16k's context window?

GPT-3.5 Turbo 16k supports up to 16,385 tokens of context in a single request.

What is GPT-3.5 Turbo 16k best for?

GPT-3.5 Turbo 16k is well suited to Solid general chat performance, Reasonable instruction following and General chat.

Who makes GPT-3.5 Turbo 16k?

GPT-3.5 Turbo 16k is developed and served by OpenAI.

Terms used on this page