Claude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/MClaude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/M
Meta
Meta
EfficientLIVE INDEX

Llama 3.2 1B Instruct

text->text

Smallest Llama, 1B parameters for ultra-edge inference. Tradeoff: very limited reasoning.

Llama 3.2 1B Instruct is a efficient AI model from Meta. It costs $0.027 per million input tokens and $0.201 per million output tokens (blended $0.079/M), with a 60,000-token context window.

Profile inherited from upstream Llama 3.2 1B — this is a hosted variant of the same open-weights model.

INPUT
$0.027/M
per million input tokens
OUTPUT
$0.201/M
per million output tokens
BLENDED 70/30
$0.079/M

0.0%over 91 days · hover to read

Llama 3.2 1B Instruct — blended price

Reconstructed from 91 days of recorded rate-card changes. Prices hold flat between changes because that is what a posted price does — no value here is interpolated.

CONTEXT
60,000
tokens
What it is good at
  • Smallest Llama
  • Edge / mobile
  • Open weights
Typical use cases
  • On-device routing
  • Tiny-footprint chat

Benchmarks

vs. best public score
Scores inherited from Llama 3.2 1B — this is a hosted variant of the same open-weights model, so the underlying benchmark scores are identical.
MMLU49%
Multitask academic knowledge across 57 subjects.
Graduate-level science questions, "Google-proof".
MATH30%
High-school competition math problems.
Python function synthesis from docstrings.
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.

More from Meta

See all 20

Frequently asked questions

How much does Llama 3.2 1B Instruct cost?

Llama 3.2 1B Instruct costs $0.027 per million input tokens and $0.201 per million output tokens, for a blended reference rate of $0.079 per million tokens.

What is Llama 3.2 1B Instruct's context window?

Llama 3.2 1B Instruct supports up to 60,000 tokens of context in a single request.

What is Llama 3.2 1B Instruct best for?

Llama 3.2 1B Instruct is well suited to Smallest Llama, Edge / mobile and Open weights.

Who makes Llama 3.2 1B Instruct?

Llama 3.2 1B Instruct is developed and served by Meta. It was released in Sep 2024.

Terms used on this page