Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Kilo Gateway logo

Model details

Muse Glimmer 30B

Muse Glimmer 30B is a 30-billion-parameter causal language model released by Meta Superintelligence Lab in August 2026 under the Apache 2.0 license, with weights published on Hugging Face under the meta-models organization. It is distilled from the larger Muse Spark model and pairs its causal language core with a dedicated perception encoder, giving it a multimodal understanding pathway alongside text generation. The card positions the model as purpose-built for autonomous agentic tasks on consumer hardware, designed to run locally without requiring cloud infrastructure or network access.

According to the model card, training and evaluation emphasize the capabilities that autonomous agents need to sustain over long workflows: multi-step reasoning, reliable tool use, multimodal understanding, and failure recovery when a tool call errors out. Reported benchmark coverage spans end-to-end agentic task completion on DeepSearch QA, MCP-Atlas, τ3-Bench, and SWE-Bench, all of which probe the model's ability to work inside scaffolds, write and debug code, and resolve multi-turn requests from start to finish. Practical fit centers on local agent orchestration patterns such as OpenClaw and Hermes Agent, where the model can chain reasoning over extended horizons, invoke tools with precise schemas, and interpret screenshots, charts, and documents interleaved with conversation.

Kilo Gatewaymeta/muse-glimmer-30bmuse

Quick Info

Powered by
Provider
Kilo Gateway
Model key
meta/muse-glimmer-30b
Release date
Aug 10, 2026
Last updated
Aug 10, 2026
Knowledge cutoff
2026-01-04
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.30
Output token cost
$1.10

Limits

Output tokens
117,964 tokens
Context window
131,072 tokens

Transparent token rates

Compare Muse Glimmer 30B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Muse Glimmer 30B

Nvidia

Coverage

Shattered.io's news analysis covers Meta's 10 August 2026 announcement of Muse Glimmer, a 30-billion-parameter open-weight, Apache 2.0-licensed model built for local, always-on AI agents, described as Meta's first fully open release since retiring the Llama line in favor of the proprietary Muse Spark earlier in 2026. T The article frames Glimmer's pitch as offline operation on hardware users already own rather than competing with frontier models like GPT-5 or Gemini 3 on leaderboards, positioning it within the industry debate over where inference happens. It cites Meta's own research blog description of Glimmer as "a 30-billion-param

Kilo Gateway

Coverage

This Business Standard press article (dated August 11, 2026) reports Meta's release of Muse Glimmer as a 30-billion-parameter open-weight model built for local AI agents, coding, and tool use, and designed to run on a Mac or PC with a single consumer GPU. Meta officially released the weights under the Apache 2.0 licens The piece adds confirmation through Mark Zuckerberg's X post: "Today we're also opening the weights for Muse Glimmer, a great 30B parameter dense model that can run locally," and teases that Meta would soon release the weights of Muse Spark 1.2, described as the company's latest foundation model. The article is positio

Nvidia

CoverageBenchmark

This Kingy AI launch-day analysis (10 August 2026) details Muse Glimmer 30B's architecture from official Meta sources: a dense causal transformer with about 29.6 billion total parameters (1.8B vision encoder), 52 layers, 6,656 hidden size, 32 query / 2 KV heads, and a hybrid attention pattern of three 2,048-token slidi Kingy frames Glimmer as a serious launch-day option for local agents on 24 GB or 32 GB machines, citing Meta's calibrated GGUF releases, vision and speculative-decoding components, 131K context and broad agent evaluations, while noting it is not a benchmark sweep, zero-cost service, or fully reproducible open-source re

Kilo Gateway

CoverageBenchmark

BenchLM's aggregated model dashboard (published August 10, 2026) provides a structured capability ledger for Muse Glimmer 30B, listing it as a self-hosted model with no comparable first-party hosted token rate and reporting a 131K-token context window (with a noted absence of a stored direct source link). The capabilit The dashboard separates verified from provisional evidence, reporting 4/4 verified Agentic rows and 4/4 verified Coding rows out of 14 published benchmark rows, with some tracked slots left empty and the speed axis explicitly marked as not measured. A price input median of $0.95 is flagged as not tied to a comparable f

Nvidia

CoverageBenchmark

Meta released Muse Glimmer on 10 August 2026 as a 29.6-billion-parameter open-weight dense transformer designed for always-on local AI agents, combining text and image understanding with tool use and long-context reasoning, according to this Wavect buyer's guide. The model card lists BF16 weights, two 4-bit quantizatio The guide positions Muse Glimmer as commercially interesting because it pairs a permissive license with credible agent benchmarks in a single-workstation envelope, framing it as Meta's distilled Muse Spark variant for local coding, function calling and LLM-as-a-judge work. It explicitly separates this hardware-fit ques

Videos about Muse Glimmer 30B

More models around Muse Glimmer 30B