Yi-VL-6B
Vision
Smaller Yi-VL, single-GPU vision model with bilingual capability.
Yi-VL-6B is a multimodal AI model from 01.AI. It costs $0.300 per million input tokens and $0.300 per million output tokens (blended $0.300/M), with a 4,000-token context window.
INPUT
$0.300/M
per million input tokens
OUTPUT
$0.300/M
per million output tokens
CONTEXT
4,000
tokens
What it is good at
- Single-GPU vision
- Open weights
- Bilingual
Typical use cases
- On-prem bilingual vision QA
Benchmarks
No published benchmark scores tracked for this model yet. Frontier and reasoning models from major providers have scores; smaller models and inference-host variants typically inherit the underlying open-weights score.
More from 01.AI
See all 9 →Yi-Lightning
Fast · 16,000 ctx
in $0.140/Mout $0.140/M
Yi-Large
Frontier · 32,000 ctx
in $3.000/Mout $3.000/M
Yi-Large-Turbo
Balanced · 16,000 ctx
in $0.380/Mout $0.380/M
Yi-34B-Chat
Open Large · 4,000 ctx
in $0.800/Mout $0.800/M
Yi-34B
Open Base · 4,000 ctx
in $0.800/Mout $0.800/M
Yi-9B
Compact · 4,000 ctx
in $0.300/Mout $0.300/M
Frequently asked questions
How much does Yi-VL-6B cost?
Yi-VL-6B costs $0.300 per million input tokens and $0.300 per million output tokens, for a blended reference rate of $0.300 per million tokens.
What is Yi-VL-6B's context window?
Yi-VL-6B supports up to 4,000 tokens of context in a single request.
What is Yi-VL-6B best for?
Yi-VL-6B is well suited to Single-GPU vision, Open weights and Bilingual.
Who makes Yi-VL-6B?
Yi-VL-6B is developed and served by 01.AI. It was released in Jan 2024.