Meta: Llama 3.1 70B Instruct
by Meta LlamaMeta: Llama 3.1 70B Instruct is a large language model from Meta Llama. It costs $0.400 per million input tokens and $0.400 per million output tokens. Its context window is 131K tokens.
- Input / 1M tokens
- $0.400
- Output / 1M tokens
- $0.400
- Cached input / 1M
- -
- Context window
- 131K
Not supported
tokens
Who serves it cheapest
2 hosts serve Llama 3.1 70B Instruct. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| DeepInfraCheapestfp8 | $0.400 | $0.400 | 131K | - | 99.8% |
| Amazon Bedrock | $0.720 | $0.720 | 131K | - | 100.0% |
The spread between DeepInfra and Amazon Bedrock is 1.8× for identical weights. Quantization and context limits differ, so check both columns before switching.
About Llama 3.1 70B Instruct
Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 70B instruct-tuned version is optimized for high quality dialogue usecases. It has demonstrated strong...
Specifications
| Model ID | meta-llama/llama-3.1-70b-instruct |
|---|---|
| Provider | Meta Llama |
| Context window | 131K tokens |
| Max output | 16K tokens |
| Input modalities | text |
| Output modalities | text |
| Knowledge cutoff | 2023-12-31 |
| Open weights | Yes - meta-llama/Meta-Llama-3.1-70B-Instruct |
| Released | July 23, 2024 |
Cheaper alternatives
Models that cost less than Llama 3.1 70B Instruct while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Meta: Llama 3.1 70B Instruct cost?
$0.400 per million input tokens and $0.400 per million output tokens.
What is the context window of Meta: Llama 3.1 70B Instruct?
131K tokens, with up to 16K tokens of output per request.
Which provider serves Meta: Llama 3.1 70B Instruct cheapest?
DeepInfra at $0.400 per million input tokens - 1.8× cheaper than Amazon Bedrock, the most expensive of the 2 hosts serving it.