Meta releases open source Muse Glimmer, promises weights for Spark

Meta is making open source models central to its AI vision, says Zuckerberg. (Picture: Meta)
Meta claims to have squeezed a 30 billion parameter model onto a «consumer GPU» through clever 4-bit compression, making it weigh just 20GB, needing only 24 GB VRAM.

It not only runs agents, they say, but at a speed and responsiveness that feels natural without long thinking breaks.

At the same time, Mark Zuckerberg is out with a lengthy essay extolling the virtues of open models as central to their strategy, and presenting a «positive vision» for personal superintelligence as a tool for «individual empowerment.»

The selling point for Muse Glimmer is that it can run on a Pro- or Max-level MacBook Pro with sufficient memory, or on the GTX 4090/5090 GPUs from Nvidia on PCs.

These are high-end tools you likely won’t find in a bargain basement workshop in Kampala, but runs significantly below the cost of infrastructure-level Nvidia chips.

As an open model, Meta is only showing comparison benchmarks for Gemma 4-31B and Qwen 3.6-27B, which it seems to beat handily. It’s only barely showing on the LMArena leaderboards, however, at 97th for text and 77th for WebDev, below Gemma 4 and DeepSeek v3.2, but above Qwen 3.5.

You can fetch the model at Hugging Face under an Apache 2.0 license.

Read more: Meta’s presentation, launch post on X. Writeups on CNBC and TechCrunch. Discussions on Hacker News and r/LocalMMaMA.