Meta launches Muse Glimmer: What it means for the open-source vs closed AI debate
Meta has launched Muse Glimmer, a new open-weight small language model (SLM) designed to run locally on consumer hardware, marking a shift in the open-source vs closed AI debate. This 30-billion-parameter model is the brainchild of Meta Superintelligence Labs (MSL), a unit led by Alexander Wang, and has been released under a permissive Apache 2.0 license for free download via Hugging Face.
CEO Mark Zuckerberg emphasized Meta's strong support for open-source AI, stating that the company is proud of these releases. He also pointed out that US policy must reduce the additional friction that foreign labs face in order for American open-source models to lead in the future. Zuckerberg added that restricting access to foreign open-source models is not an effective solution.
Muse Glimmer's development involved training the model on the outputs of a larger 'teacher' model through distillation, a common AI/ML technique that has been controversial recently. The training data consisted of synthetic and organic data spanning over 100 languages. To optimize the model for local deployment, Meta applied techniques such as quantization, compression to 4-bit precision, and inference optimization.
The model is composed of two key parts: the main model and a 'drafter model' that processes entire blocks of tokens at once, enabling faster text generation. Muse Glimmer is designed to run on consumer hardware with minimal trade-offs in output quality, making it suitable for tasks such as managing schedules, drafting messages, and organizing files.
Performance benchmarks show that Muse Glimmer excels in agentic task completion, reliable tool use, multi-step reasoning, and multimodal input and reasoning. The model outperformed Google's Gemma and Alibaba's Qwen across various benchmarks, including agentic, coding, multimodal, safety, and reasoning tasks. Muse Glimmer runs efficiently on devices like the Macbook M4 Max chip, M5 Max chip, and Nvidia RTX-5090 GPU, enabling fluid conversation and real-time agent interaction without relying on cloud infrastructure or network access.
Written by urgent.news from The Indian Express's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.
This story
This is one outlet's version. Read the fullest account.