Meta: Llama Guard 4 12B
by Meta LlamaMeta: Llama Guard 4 12B is a large language model from Meta Llama. It costs $0.180 per million input tokens and $0.180 per million output tokens. Its context window is 164K tokens.
- Input / 1M tokens
- $0.180
- Output / 1M tokens
- $0.180
- Cached input / 1M
- -
- Context window
- 164K
Not supported
tokens
Who serves it cheapest
1 hosts serve Llama Guard 4 12B. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| DeepInfraCheapestbf16 | $0.180 | $0.180 | 164K | - | 100.0% |
About Llama Guard 4 12B
Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...
Specifications
| Model ID | meta-llama/llama-guard-4-12b |
|---|---|
| Provider | Meta Llama |
| Context window | 164K tokens |
| Max output | 16K tokens |
| Input modalities | image, text |
| Output modalities | text |
| Knowledge cutoff | 2024-08-31 |
| Open weights | Yes - meta-llama/Llama-Guard-4-12B |
| Released | April 30, 2025 |
Cheaper alternatives
Models that cost less than Llama Guard 4 12B while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Meta: Llama Guard 4 12B cost?
$0.180 per million input tokens and $0.180 per million output tokens.
What is the context window of Meta: Llama Guard 4 12B?
164K tokens, with up to 16K tokens of output per request.