Aya Expanse 8B
Compact ML
Smaller Aya Expanse for single-GPU deployment. Same 23-language coverage at lower cost.
Aya Expanse 8B is a efficient AI model from Cohere. It costs $0.500 per million input tokens and $1.500 per million output tokens (blended $0.800/M), with a 8,000-token context window.
INPUT
$0.500/M
per million input tokens
OUTPUT
$1.500/M
per million output tokens
CONTEXT
8,000
tokens
What it is good at
- Multilingual on a single GPU
- Open weights
- 128K context
Typical use cases
- On-prem multilingual chat
- Multilingual fine-tune base
Benchmarks
vs. best public score
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.
More from Cohere
See all 12 →Command A+
Frontier · 256,000 ctx
in $2.500/Mout $10.000/M
Command A
Previous · 256,000 ctx
in $2.500/Mout $10.000/M
Command R+
Previous · 128,000 ctx
in $2.500/Mout $10.000/M
Command R
Balanced · 128,000 ctx
in $0.150/Mout $0.600/M
Command R7B
Compact · 128,000 ctx
in $0.037/Mout $0.150/M
Aya Expanse 32B
Multilingual · 128,000 ctx
in $0.500/Mout $1.500/M
Frequently asked questions
How much does Aya Expanse 8B cost?
Aya Expanse 8B costs $0.500 per million input tokens and $1.500 per million output tokens, for a blended reference rate of $0.800 per million tokens.
What is Aya Expanse 8B's context window?
Aya Expanse 8B supports up to 8,000 tokens of context in a single request.
What is Aya Expanse 8B best for?
Aya Expanse 8B is well suited to Multilingual on a single GPU, Open weights and 128K context.
Who makes Aya Expanse 8B?
Aya Expanse 8B is developed and served by Cohere. It was released in Oct 2024.