Mistral-Large-2 (NIM)
Enterprise
Mistral Large 2 served via NVIDIA NIM. Identical Mistral weights with NVIDIA enterprise integration.
Mistral-Large-2 (NIM) is a frontier AI model from Nvidia. It costs $2.000 per million input tokens and $6.000 per million output tokens (blended $3.200/M), with a 128,000-token context window.
INPUT
$2.000/M
per million input tokens
OUTPUT
$6.000/M
per million output tokens
CONTEXT
128,000
tokens
What it is good at
- Mistral Large 2 quality
- NIM enterprise packaging
- On-prem NVIDIA hosting
Typical use cases
- NVIDIA-stack frontier inference
- On-prem large-model deployments
Benchmarks
vs. best public score
Scores inherited from Mistral Large 2 — this is a hosted variant of the same open-weights model, so the underlying benchmark scores are identical.
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.
More from Nvidia
See all 6 →Llama-3.1-Nemotron-70B
RLHF-Aligned · 128,000 ctx
in $0.350/Mout $0.400/M
Llama-3.1-Nemotron-Ultra-253B
Ultra · 128,000 ctx
in $1.600/Mout $1.600/M
Nemotron-4-340B
Large · 4,000 ctx
in $4.200/Mout $4.200/M
Mistral-NeMo-12B (NIM)
Efficient · 128,000 ctx
in $0.150/Mout $0.150/M
Phi-3-Mini-4K (NIM)
Nano · 4,000 ctx
in $0.040/Mout $0.040/M
Frequently asked questions
How much does Mistral-Large-2 (NIM) cost?
Mistral-Large-2 (NIM) costs $2.000 per million input tokens and $6.000 per million output tokens, for a blended reference rate of $3.200 per million tokens.
What is Mistral-Large-2 (NIM)'s context window?
Mistral-Large-2 (NIM) supports up to 128,000 tokens of context in a single request.
What is Mistral-Large-2 (NIM) best for?
Mistral-Large-2 (NIM) is well suited to Mistral Large 2 quality, NIM enterprise packaging and On-prem NVIDIA hosting.
Who makes Mistral-Large-2 (NIM)?
Mistral-Large-2 (NIM) is developed and served by Nvidia.