Meta: Llama 3.2 1B Instruct
by Meta LlamaMeta: Llama 3.2 1B Instruct is a large language model from Meta Llama. It costs $0.027 per million input tokens and $0.201 per million output tokens. Its context window is 60K tokens.
- Input / 1M tokens
- $0.027
- Output / 1M tokens
- $0.201
- Cached input / 1M
- -
- Context window
- 60K
Not supported
tokens
Who serves it cheapest
1 hosts serve Llama 3.2 1B Instruct. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| CloudflareCheapest | $0.027 | $0.201 | 60K | - | 100.0% |
About Llama 3.2 1B Instruct
Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...
Specifications
| Model ID | meta-llama/llama-3.2-1b-instruct |
|---|---|
| Provider | Meta Llama |
| Context window | 60K tokens |
| Max output | 54K tokens |
| Input modalities | text |
| Output modalities | text |
| Knowledge cutoff | 2023-12-31 |
| Open weights | Yes - meta-llama/Llama-3.2-1B-Instruct |
| Released | September 25, 2024 |
Cheaper alternatives
Models that cost less than Llama 3.2 1B Instruct while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Meta: Llama 3.2 1B Instruct cost?
$0.027 per million input tokens and $0.201 per million output tokens.
What is the context window of Meta: Llama 3.2 1B Instruct?
60K tokens, with up to 54K tokens of output per request.