Claude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/MClaude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/M
DeepInfra
DeepInfra
Efficient

Mixtral 8x7B (DI)

MoE

The original Mixtral MoE from late 2023. Kicked off the open-weights MoE wave; now superseded by 8x22B and Llama 3.x.

Mixtral 8x7B (DI) is a efficient AI model from DeepInfra. It costs $0.240 per million input tokens and $0.240 per million output tokens (blended $0.240/M), with a 32,000-token context window.

Profile inherited from upstream Mixtral 8x7B — this is a hosted variant of the same open-weights model.

INPUT
$0.240/M
per million input tokens
OUTPUT
$0.240/M
per million output tokens
BLENDED 70/30
$0.240/M
unchanged since 3 May
CONTEXT
32,000
tokens
What it is good at
  • Open weights
  • 32K context
  • Wide ecosystem
Typical use cases
  • Legacy Mixtral 8x7B deployments
  • Fine-tune base

Benchmarks

vs. best public score
Scores inherited from Mixtral 8x7B — this is a hosted variant of the same open-weights model, so the underlying benchmark scores are identical.
MMLU71%
Multitask academic knowledge across 57 subjects.
MATH28%
High-school competition math problems.
Python function synthesis from docstrings.
LMArena Elo1124 Elo
Crowd-sourced head-to-head preference Elo rating.
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.

More from DeepInfra

See all 11

Frequently asked questions

How much does Mixtral 8x7B (DI) cost?

Mixtral 8x7B (DI) costs $0.240 per million input tokens and $0.240 per million output tokens, for a blended reference rate of $0.240 per million tokens.

What is Mixtral 8x7B (DI)'s context window?

Mixtral 8x7B (DI) supports up to 32,000 tokens of context in a single request.

What is Mixtral 8x7B (DI) best for?

Mixtral 8x7B (DI) is well suited to Open weights, 32K context and Wide ecosystem.

Who makes Mixtral 8x7B (DI)?

Mixtral 8x7B (DI) is developed and served by DeepInfra. It was released in Dec 2023.

Terms used on this page