Tencent
Tencent
Multimodal

Hunyuan-Vision

Vision

Vision Hunyuan. Tencent's Chinese-language multimodal model, document and chart QA in Chinese.

Hunyuan-Vision is a multimodal AI model from Tencent. It costs $5.500 per million input tokens and $5.500 per million output tokens (blended $5.500/M), with a 4,000-token context window.

INPUT
$5.500/M
per million input tokens
OUTPUT
$5.500/M
per million output tokens
BLENDED 70/30
$5.500/M
unchanged since 3 May
CONTEXT
4,000
tokens
What it is good at
  • Vision in Chinese
  • Document understanding
  • Tencent Cloud integration
Typical use cases
  • Chinese document parsing
  • Chinese vision QA

Benchmarks

No published benchmark scores tracked for this model yet. Frontier and reasoning models from major providers have scores; smaller models and inference-host variants typically inherit the underlying open-weights score.

More from Tencent

See all 4 →

Frequently asked questions

How much does Hunyuan-Vision cost?

Hunyuan-Vision costs $5.500 per million input tokens and $5.500 per million output tokens, for a blended reference rate of $5.500 per million tokens.

What is Hunyuan-Vision's context window?

Hunyuan-Vision supports up to 4,000 tokens of context in a single request.

What is Hunyuan-Vision best for?

Hunyuan-Vision is well suited to Vision in Chinese, Document understanding and Tencent Cloud integration.

Who makes Hunyuan-Vision?

Hunyuan-Vision is developed and served by Tencent.

Terms used on this page