Granite 4.1 8B
text->text
Granite 4.1 8B is a general-purpose chat-tuned model. Available via IBM. 131,072-token context. Budget pricing.
Granite 4.1 8B is a efficient AI model from IBM. It costs $0.050 per million input tokens and $0.100 per million output tokens (blended $0.065/M), with a 131,072-token context window.
INPUT
$0.050/M
per million input tokens
OUTPUT
$0.100/M
per million output tokens
CONTEXT
131,072
tokens
What it is good at
- Solid general chat performance
- Reasonable instruction following
Typical use cases
- General chat
- Drafting and rewriting
- Personal productivity
Benchmarks
No published benchmark scores tracked for this model yet. Frontier and reasoning models from major providers have scores; smaller models and inference-host variants typically inherit the underlying open-weights score.
More from IBM
See all 5 →Granite 3.1 8B Instruct
Compact · 128,000 ctx
in $0.050/Mout $0.250/M
Granite 3.1 2B Instruct
Nano · 128,000 ctx
in $0.030/Mout $0.100/M
Granite 3.0 8B Dense
Previous Gen · 4,000 ctx
in $0.050/Mout $0.250/M
Granite Code 34B
Code · 8,000 ctx
in $0.350/Mout $1.050/M
Granite 20B Multilingual
Multilingual · 8,000 ctx
in $0.400/Mout $0.400/M
Frequently asked questions
How much does Granite 4.1 8B cost?
Granite 4.1 8B costs $0.050 per million input tokens and $0.100 per million output tokens, for a blended reference rate of $0.065 per million tokens.
What is Granite 4.1 8B's context window?
Granite 4.1 8B supports up to 131,072 tokens of context in a single request.
What is Granite 4.1 8B best for?
Granite 4.1 8B is well suited to Solid general chat performance, Reasonable instruction following and General chat.
Who makes Granite 4.1 8B?
Granite 4.1 8B is developed and served by IBM.