Yi-Large-Turbo
Balanced
Speed-tuned Yi-Large. Cheaper, faster sibling at small quality cost.
Yi-Large-Turbo is a efficient AI model from 01.AI. It costs $0.380 per million input tokens and $0.380 per million output tokens (blended $0.380/M), with a 16,000-token context window.
INPUT
$0.380/M
per million input tokens
OUTPUT
$0.380/M
per million output tokens
CONTEXT
16,000
tokens
What it is good at
- Fast bilingual chat
- Cheaper than Yi-Large
Typical use cases
- Cost-sensitive bilingual chat
Benchmarks
vs. best public score
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.
More from 01.AI
See all 9 →Yi-Lightning
Fast · 16,000 ctx
in $0.140/Mout $0.140/M
Yi-Large
Frontier · 32,000 ctx
in $3.000/Mout $3.000/M
Yi-34B-Chat
Open Large · 4,000 ctx
in $0.800/Mout $0.800/M
Yi-34B
Open Base · 4,000 ctx
in $0.800/Mout $0.800/M
Yi-9B
Compact · 4,000 ctx
in $0.300/Mout $0.300/M
Yi-6B
Nano · 4,000 ctx
in $0.150/Mout $0.150/M
Frequently asked questions
How much does Yi-Large-Turbo cost?
Yi-Large-Turbo costs $0.380 per million input tokens and $0.380 per million output tokens, for a blended reference rate of $0.380 per million tokens.
What is Yi-Large-Turbo's context window?
Yi-Large-Turbo supports up to 16,000 tokens of context in a single request.
What is Yi-Large-Turbo best for?
Yi-Large-Turbo is well suited to Fast bilingual chat, Cheaper than Yi-Large and Cost-sensitive bilingual chat.
Who makes Yi-Large-Turbo?
Yi-Large-Turbo is developed and served by 01.AI.