Meta: Muse Glimmer 30B
by MetaMeta: Muse Glimmer 30B is a large language model from Meta. It costs $0.350 per million input tokens and $1.50 per million output tokens. Its context window is 131K tokens.
- Input / 1M tokens
- $0.350
- Output / 1M tokens
- $1.50
- Cached input / 1M
- $0.040
- Context window
- 131K
On repeated prefixes
tokens
Who serves it cheapest
4 hosts serve Muse Glimmer 30B. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| PhalaCheapest | $0.300 | $1.10 | 131K | - | 95.7% |
| DeepInfrabf16 | $0.300 | $1.20 | 131K | - | 99.8% |
| Fireworks | $0.350 | $1.50 | 131K | - | 99.8% |
| Together | $0.350 | $1.50 | 131K | - | 96.7% |
The spread between Phala and Together is 1.3× for identical weights. Quantization and context limits differ, so check both columns before switching.
About Muse Glimmer 30B
Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...
Specifications
| Model ID | meta/muse-glimmer-30b |
|---|---|
| Provider | Meta |
| Context window | 131K tokens |
| Max output | 118K tokens |
| Input modalities | text, image |
| Output modalities | text |
| Knowledge cutoff | - |
| Open weights | Yes - meta-models/Muse-Glimmer-30B |
| Released | August 9, 2026 |
Cheaper alternatives
Models that cost less than Muse Glimmer 30B while keeping at least half its context window and every input modality it supports.
Qwen3 VL 235B A22B Instruct
$0.210 in · $1.90 out
about the same cheaperERNIE 4.5 VL 424B A47B
$0.420 in · $1.25 out
about the same cheaperQwen3.5 Plus 2026-02-15
$0.260 in · $1.56 out
1.1× cheaperNano Banana 2 Lite (Gemini 3.1 Flash Lite Image)
$0.250 in · $1.50 out
1.1× cheaperGemini 3.1 Flash Lite
$0.250 in · $1.50 out
1.1× cheaperFrequently asked
How much does Meta: Muse Glimmer 30B cost?
$0.350 per million input tokens and $1.50 per million output tokens. Cached input reads cost $0.040 per million tokens.
What is the context window of Meta: Muse Glimmer 30B?
131K tokens, with up to 118K tokens of output per request.
Which provider serves Meta: Muse Glimmer 30B cheapest?
Phala at $0.300 per million input tokens - 1.3× cheaper than Together, the most expensive of the 4 hosts serving it.