Skip to main content
TokenCost logoTokenCost

Muse Glimmer 30B

Last verified August 11, 2026 · Meta pricing

$0.350/1M input · $1.50/1M output · 131K context · Meta

Count Tokens

Tokens
Input Cost
Output Cost

Estimate Monthly Cost

Monthly Cost Estimator

Quick:
<$0.0001/mo

Pick a preset above or enter custom usage

Alternatives to Muse Glimmer 30B

Pricing Details

MetaMeta's first open weights in more than a year, Apache 2.0, ~29.6B dense including a ~1.8B vision encoder (not MoE), logit-distilled from Muse Spark during pre-training (Meta names no version). Text and image in, text out, 100+ languages, knowledge cutoff January 4, 2026. Meta does not serve it on its own API and points developers to Hugging Face. The $0.35/$1.50 rate is Together AI's, and as of August 11, 2026 Together is the only host publishing a per-token price: OpenRouter's endpoints API returns exactly one endpoint (provider Together) at the identical figures, ModelsLab resells at the same card, NVIDIA build.nvidia.com offers a rate-capped free prototyping tier with no per-token rate, and Fireworks was named a launch partner but has shipped no listing. So the cross-host spread is 1.0x, against 11.7x for gpt-oss-120b and 12.4x for Gemma 4 31B. Max output tokens are published nowhere (Together returns null); 16384 here is a catalogue default, not a documented limit. Artificial Analysis' model page lists a 256K context, contradicted by the config, Together and NVIDIA NIM, which all say 131,072; the 32,768 figure in Meta's launch blog sits inside an example agent config. Benchmarks below are Meta-reported. On Meta's own nine-row table it wins five and loses four against Qwen3.6-27B; the 60.7 TerminalBench 2.1 figure widely attributed to Muse Glimmer is Qwen's score, and Meta's card says 51.7. SWE-bench Verified 76.0, one of the rows it loses: Meta's own table puts Qwen3.6-27B in thinking mode at 77.2 on the same suite (both figures read off Meta's model page 2026-08-11). Independent AA figures agree there (52%) but are lower overall: Intelligence Index 35 against Qwen3.6 27B's 38, GDPval-AA v2 Elo 953 against 1141, and an 82% hallucination rate against Qwen's 49%. Openness Index 44. 4-bit GGUF is 15.9 GB and runs on a 24GB card; BF16 shards total 55.7 GB. Estimated tokens (different tokenizer).
Input / 1M tokens
$0.35
Output / 1M tokens
$1.5
Context Window
131K
Max Output
16K

Price History

No price change recorded since 2026-08-11, when this site started tracking this model; it launched on 2026-08-10.

Frequently Asked Questions

Common questions about Muse Glimmer 30B pricing and usage