Hunyuan-Vision
Vision
Vision Hunyuan. Tencent's Chinese-language multimodal model, document and chart QA in Chinese.
Hunyuan-Vision is a multimodal AI model from Tencent. It costs $5.500 per million input tokens and $5.500 per million output tokens (blended $5.500/M), with a 4,000-token context window.
INPUT
$5.500/M
per million input tokens
OUTPUT
$5.500/M
per million output tokens
CONTEXT
4,000
tokens
What it is good at
- Vision in Chinese
- Document understanding
- Tencent Cloud integration
Typical use cases
- Chinese document parsing
- Chinese vision QA
Benchmarks
No published benchmark scores tracked for this model yet. Frontier and reasoning models from major providers have scores; smaller models and inference-host variants typically inherit the underlying open-weights score.
More from Tencent
See all 4 →Frequently asked questions
How much does Hunyuan-Vision cost?
Hunyuan-Vision costs $5.500 per million input tokens and $5.500 per million output tokens, for a blended reference rate of $5.500 per million tokens.
What is Hunyuan-Vision's context window?
Hunyuan-Vision supports up to 4,000 tokens of context in a single request.
What is Hunyuan-Vision best for?
Hunyuan-Vision is well suited to Vision in Chinese, Document understanding and Tencent Cloud integration.
Who makes Hunyuan-Vision?
Hunyuan-Vision is developed and served by Tencent.