Claude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/MClaude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/M
Together AI
Together AI
Frontier

WizardLM-2 8x22B

Fine-tuned

WizardLM-2 8x22B, Microsoft Research's instruction-tuned Mixtral 8x22B variant. Briefly released and pulled in 2024, but widely re-hosted by Together / DeepInfra.

WizardLM-2 8x22B is a frontier AI model from Together AI. It costs $1.200 per million input tokens and $1.200 per million output tokens (blended $1.200/M), with a 64,000-token context window.

INPUT
$1.200/M
per million input tokens
OUTPUT
$1.200/M
per million output tokens
BLENDED 70/30
$1.200/M
unchanged since 3 May
CONTEXT
64,000
tokens
What it is good at
  • Open-weights MoE
  • Strong instruction following
  • Widely hosted
Typical use cases
  • Self-hosted instruction-tuned MoE
  • Cheap quality alternative to closed frontier

Benchmarks

vs. best public score
Scores inherited from Mixtral 8x22B (WizardLM-2 base) — this is a hosted variant of the same open-weights model, so the underlying benchmark scores are identical.
MMLU78%
Multitask academic knowledge across 57 subjects.
Graduate-level science questions, "Google-proof".
MATH42%
High-school competition math problems.
Python function synthesis from docstrings.
LMArena Elo1198 Elo
Crowd-sourced head-to-head preference Elo rating.
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.

More from Together AI

See all 10

Frequently asked questions

How much does WizardLM-2 8x22B cost?

WizardLM-2 8x22B costs $1.200 per million input tokens and $1.200 per million output tokens, for a blended reference rate of $1.200 per million tokens.

What is WizardLM-2 8x22B's context window?

WizardLM-2 8x22B supports up to 64,000 tokens of context in a single request.

What is WizardLM-2 8x22B best for?

WizardLM-2 8x22B is well suited to Open-weights MoE, Strong instruction following and Widely hosted.

Who makes WizardLM-2 8x22B?

WizardLM-2 8x22B is developed and served by Together AI. It was released in Apr 2024.

Terms used on this page