01.AI
01.AI
Efficient

Yi-Large-Turbo

Balanced

Speed-tuned Yi-Large. Cheaper, faster sibling at small quality cost.

Yi-Large-Turbo is a efficient AI model from 01.AI. It costs $0.380 per million input tokens and $0.380 per million output tokens (blended $0.380/M), with a 16,000-token context window.

INPUT
$0.380/M
per million input tokens
OUTPUT
$0.380/M
per million output tokens
BLENDED 70/30
$0.380/M
unchanged since 3 May
CONTEXT
16,000
tokens
What it is good at
  • Fast bilingual chat
  • Cheaper than Yi-Large
Typical use cases
  • Cost-sensitive bilingual chat

Benchmarks

vs. best public score
MMLU75%
Multitask academic knowledge across 57 subjects.
Python function synthesis from docstrings.
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.

More from 01.AI

See all 9 →

Frequently asked questions

How much does Yi-Large-Turbo cost?

Yi-Large-Turbo costs $0.380 per million input tokens and $0.380 per million output tokens, for a blended reference rate of $0.380 per million tokens.

What is Yi-Large-Turbo's context window?

Yi-Large-Turbo supports up to 16,000 tokens of context in a single request.

What is Yi-Large-Turbo best for?

Yi-Large-Turbo is well suited to Fast bilingual chat, Cheaper than Yi-Large and Cost-sensitive bilingual chat.

Who makes Yi-Large-Turbo?

Yi-Large-Turbo is developed and served by 01.AI.

Terms used on this page