01.AI
01.AI
Multimodal

Yi-VL-34B

Vision Large

Vision-language Yi-34B. Open-weights bilingual vision model from 01.AI.

Yi-VL-34B is a multimodal AI model from 01.AI. It costs $3.500 per million input tokens and $3.500 per million output tokens (blended $3.500/M), with a 4,000-token context window.

INPUT
$3.500/M
per million input tokens
OUTPUT
$3.500/M
per million output tokens
BLENDED 70/30
$3.500/M
unchanged since 3 May
CONTEXT
4,000
tokens
What it is good at
  • Open-weights vision
  • Bilingual
  • Document understanding
Typical use cases
  • Self-hosted bilingual vision QA
  • Fine-tune base

Benchmarks

No published benchmark scores tracked for this model yet. Frontier and reasoning models from major providers have scores; smaller models and inference-host variants typically inherit the underlying open-weights score.

More from 01.AI

See all 9 →

Frequently asked questions

How much does Yi-VL-34B cost?

Yi-VL-34B costs $3.500 per million input tokens and $3.500 per million output tokens, for a blended reference rate of $3.500 per million tokens.

What is Yi-VL-34B's context window?

Yi-VL-34B supports up to 4,000 tokens of context in a single request.

What is Yi-VL-34B best for?

Yi-VL-34B is well suited to Open-weights vision, Bilingual and Document understanding.

Who makes Yi-VL-34B?

Yi-VL-34B is developed and served by 01.AI. It was released in Jan 2024.

Terms used on this page