MultimodalAUTO PROFILELIVE PRICING

MiMo-V2.6-Pro

text+image+audio+video->text

MiMo-V2.6-Pro sits in the frontier tier, built for high-capability reasoning, tool use, and long-form output where capability outweighs cost. Available via xiaomi. 1,048,576-token context. Efficient pricing.

MiMo-V2.6-Pro is a multimodal AI model from xiaomi. It costs $0.435 per million input tokens and $0.870 per million output tokens (blended $0.566/M), with a 1,048,576-token context window.

INPUT
$0.435/M
per million input tokens
OUTPUT
$0.870/M
per million output tokens
BLENDED 70/30
$0.566/M
unchanged since 3 May
CONTEXT
1,048,576
tokens
What it is good at
  • Strong general reasoning
  • Tool use & function calling
  • Long-form writing
  • Production-grade reliability
  • Multimodal: handles images alongside text
Typical use cases
  • Production assistants
  • Long-document analysis
  • Multi-step agent workflows
  • Engineering copilots

Benchmarks

No published benchmark scores tracked for this model yet. Frontier and reasoning models from major providers have scores; smaller models and inference-host variants typically inherit the underlying open-weights score.

More from xiaomi

See all 4

Frequently asked questions

How much does MiMo-V2.6-Pro cost?

MiMo-V2.6-Pro costs $0.435 per million input tokens and $0.870 per million output tokens, for a blended reference rate of $0.566 per million tokens.

What is MiMo-V2.6-Pro's context window?

MiMo-V2.6-Pro supports up to 1,048,576 tokens of context in a single request.

What is MiMo-V2.6-Pro best for?

MiMo-V2.6-Pro is well suited to Strong general reasoning, Tool use & function calling and Long-form writing.

Who makes MiMo-V2.6-Pro?

MiMo-V2.6-Pro is developed and served by xiaomi.

Terms used on this page