# Meta releases open-weight Muse Glimmer, its first since Llama

> Meta open-sourced Muse Glimmer, a 30B agent model, on 10 August.

*The 30-billion-parameter model runs local AI agents on a single consumer GPU under the Apache 2.0 licence.*

By Behzad Hosseini · FeaturedDaily
Canonical: https://featureddaily.com/news/meta-releases-open-weight-muse-glimmer-its-first-since-llama

**What happened:** Meta published the weights of Muse Glimmer on Hugging Face on 10 August, under the Apache 2.0 licence. That allows commercial use, modification and redistribution.

**The numbers:** The model has 30 billion parameters. Full BF16 weights run to 59.6 GB; a 4-bit version comes in under 20 GB and runs on a 24 or 32 GB GPU or a Mac, with under 1% accuracy loss. Context length runs to 131,072+ tokens.

**The details:** Muse Glimmer was built by Meta Superintelligence Labs, distilled from Meta's closed flagship Muse Spark 1.2, which launched 5 August. It handles text and images via an 1.8-billion-parameter vision encoder and is trained for multi-step tool use, function calling and recovering when an API call fails. It ships with a small "DFlash" drafter for speculative decoding, giving reported speed-ups of 3.1x on an RTX 5090, 1.5x on an M4 Max and 1.8x on an M5 Max.

**The context:** This is Meta's first fully open model since it retired the Llama line in favour of the proprietary Muse Spark earlier in 2026.

**Why it matters:** Meta is betting that open, locally runnable models will win developers back for agent workloads that don't need a data-centre GPU.

**The rivals:** On the MCP Atlas benchmark, Muse Glimmer scored 75.5, ahead of Gemma 4 31B (54.2) and Qwen 3.6 27B (62.5). Its Artificial Analysis Intelligence Index score is 35.

**In their words:** Mark Zuckerberg said: "Today we're also opening the weights for Muse Glimmer, a great 30B parameter dense model that can run locally."

**What's next:** Zuckerberg said Meta will also open-source Muse Spark 1.2, the closed flagship that Muse Glimmer was distilled from.

Details from [gHacks](https://www.ghacks.net/2026/08/11/meta-releases-muse-glimmer-a-30-billion-parameter-open-weight-ai-model-that-runs-on-a-single-consumer-gpu/).

## Key takeaways

- 30B model runs on one 24-32GB GPU or a Mac
- First fully open Meta model since Llama's retirement
- Beats Gemma 4 and Qwen 3.6 on tool-use benchmark

## Sources

- [Meta Releases Muse Glimmer, a 30-Billion-Parameter Open-Weight AI Model That Runs on a Single Consumer GPU](https://www.ghacks.net/2026/08/11/meta-releases-muse-glimmer-a-30-billion-parameter-open-weight-ai-model-that-runs-on-a-single-consumer-gpu/) — gHacks, 2026-08-11
