Meta releases Muse Glimmer, 30B open-weight model for local AI agents

Muse Glimmer ships under Apache 2.0 with a 120K+ context window and runs on a single consumer GPU (24GB VRAM) or Mac/PC. NVIDIA reports 20K tokens/sec on one GPU, and Hugging Face shipped day-0 support in transformers, llama.cpp, and vLLM. Meta also promised Muse Spark 1.2 open weights soon.
How this story unfolded
7 days · 21 reports · 42 community posts · 63 of 66 shown
- Aug 5
- Aug 10
Meta Releases Muse Glimmer AI Model People Can Run on Their Laptopbloomberg.com
Meta is back with Muse Glimmer: local, agentic, multimodal, and open sourcehuggingface.co
meta-models/Muse-Glimmer-30Bhuggingface.co
Meta to open source its most powerful AI model as it takes swipe at OpenAI, Anthropiccnbc.com
unsloth/Muse-Glimmer-30B-GGUFhuggingface.co
Run Local Agentic AI Workflows with Meta’s Muse Glimmer on NVIDIAdeveloper.nvidia.com
meta-models/Muse-Glimmer-30B-GGUFhuggingface.co
Meta Releases AI Model You Can Use at Homebloomberg.com
Meta AI Releases Muse Glimmer: A 30B Open-Weights Agentic Model That Runs on One Consumer GPUmarktechpost.com
Meta’s new Glimmer AI model offers a hint at Zuckerberg’s personal intelligence visiontechcrunch.com
Meta returns to open source with Muse Glimmer, an Apache 2.0 licensed 30B parameter AI model optimized for agents — available nowventurebeat.com
Meta’s Muse Glimmer fits on a laptopthenewstack.io
Meta Brings Powerful AI to the Personal Computerbloomberg.com
Mark Zuckerberg announces Meta's new AI model Muse Glimmer.youtube.com
With new open models, Meta pitches another reboot of its struggling AI strategyarstechnica.com
- Aug 11
- Aug 12
NVIDIA by email
Get an email when NVIDIA has news
No news that day, no email.
More stories today
- MiniMax H3 model integration on Fal to be showcased in live session
- Alibaba releases open weights for 2.4T-parameter Qwen3.8-Max model
- NVIDIA Spectrum-X Ethernet Photonics now in full production
- Agent evals are stuck in the chatbot era, says Raindrop's Ben Hylak
- GitHub blog details strategies for managing AI-generated pull requests