Llama 3.1 70B Instruct
The 2024 70B Llama that defined the open-weights chat baseline before Llama 3.3. Still common where hosts haven't upgraded yet.
Llama 3.1 70B Instruct is a efficient AI model from Meta. It costs $0.400 per million input tokens and $0.400 per million output tokens (blended $0.400/M), with a 131,072-token context window.
Profile inherited from upstream Llama 3.1 70B ↗ — this is a hosted variant of the same open-weights model.
- Open weights
- 128K context
- Wide hosted availability
- Self-hosted chat
- Fine-tune base
- Cost benchmarking
Benchmarks
More from Meta
See all 20 →Frequently asked questions
How much does Llama 3.1 70B Instruct cost?
Llama 3.1 70B Instruct costs $0.400 per million input tokens and $0.400 per million output tokens, for a blended reference rate of $0.400 per million tokens.
What is Llama 3.1 70B Instruct's context window?
Llama 3.1 70B Instruct supports up to 131,072 tokens of context in a single request.
What is Llama 3.1 70B Instruct best for?
Llama 3.1 70B Instruct is well suited to Open weights, 128K context and Wide hosted availability.
Who makes Llama 3.1 70B Instruct?
Llama 3.1 70B Instruct is developed and served by Meta. It was released in Jul 2024.