Claude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/MClaude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/M
DeepInfra
DeepInfra
Frontier

WizardLM-2 8x22B (DI)

Fine-tuned

Largest open-weights Mixtral MoE. Cheap-to-serve frontier-ish quality before Llama 3.1 405B and DeepSeek V3 took the open-weights lead.

WizardLM-2 8x22B (DI) is a frontier AI model from DeepInfra. It costs $0.630 per million input tokens and $0.630 per million output tokens (blended $0.630/M), with a 64,000-token context window.

Profile inherited from upstream Mixtral 8x22B (WizardLM-2 base) — this is a hosted variant of the same open-weights model.

INPUT
$0.630/M
per million input tokens
OUTPUT
$0.630/M
per million output tokens
BLENDED 70/30
$0.630/M
unchanged since 3 May
CONTEXT
64,000
tokens
What it is good at
  • Open-weights MoE
  • Cheap inference per active parameter
  • 64K context
Typical use cases
  • Self-hosted chat at scale
  • Fine-tune base

Benchmarks

vs. best public score
Scores inherited from Mixtral 8x22B (WizardLM-2 base) — this is a hosted variant of the same open-weights model, so the underlying benchmark scores are identical.
MMLU78%
Multitask academic knowledge across 57 subjects.
Graduate-level science questions, "Google-proof".
MATH42%
High-school competition math problems.
Python function synthesis from docstrings.
LMArena Elo1198 Elo
Crowd-sourced head-to-head preference Elo rating.
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.

More from DeepInfra

See all 11

Frequently asked questions

How much does WizardLM-2 8x22B (DI) cost?

WizardLM-2 8x22B (DI) costs $0.630 per million input tokens and $0.630 per million output tokens, for a blended reference rate of $0.630 per million tokens.

What is WizardLM-2 8x22B (DI)'s context window?

WizardLM-2 8x22B (DI) supports up to 64,000 tokens of context in a single request.

What is WizardLM-2 8x22B (DI) best for?

WizardLM-2 8x22B (DI) is well suited to Open-weights MoE, Cheap inference per active parameter and 64K context.

Who makes WizardLM-2 8x22B (DI)?

WizardLM-2 8x22B (DI) is developed and served by DeepInfra. It was released in Apr 2024.

Terms used on this page