2026-10-11 17:11 UTC

Independent testing will determine whether Meta’s Muse Spark 1.2 and Muse Glimmer 30B open weights, including community GGUF and llama.cpp support, enable practical local agentic inference.

state: resolvedheat: lowuncertainty: mediumconvergesscott: highopen-models local-inference metaMetaUnsloth

What is this?

Meta has introduced the Muse family for coding and agentic workloads, but the supplied material distinguishes between Muse Glimmer, a downloadable 30B open-weight model intended for single-consumer-GPU deployment, and Muse Spark 1.2, described as a proprietary multimodal model served through Meta’s API and Muse Code. Evidence titles indicate community GGUF packaging and a llama.cpp support pull request for Glimmer, while snippets report acknowledged gaps in coding and agentic performance and conflicting independent results. The material therefore does not establish that Spark 1.2 has open weights; practical local-agent viability is currently supported only as a claim about Glimmer and remains subject to independent testing.

Why it matters to Scott

Meta’s release of a consumer-GPU-oriented open-weight Glimmer model converges with Scott’s sovereign, swappable local-model strategy and directly creates a benchmark candidate for his gamepc/Ollama and Ask agent stack. Its significance depends on model-plus-harness evaluation of real agent tasks—not weight availability or benchmark claims—and the supplied grounding does not establish Spark 1.2 as open-weight or locally runnable.
ip:framework.sovereign-software-assuranceip:concept.model-plus-harness-benchmark-unitip:concept.evaluation-driven-developmentdev:project.gamepcdev:technology.ollamadev:project.askdev:concept.hardware-aware-local-inferenceradar:concept.open-modelsradar:concept.local-inferenceradar:concept.llama-cppradar:concept.ggufradar:concept.agent-benchmarksradar:kimi-k3-single-consumer-gpuradar:pi-native-llama-cpp-runtimeradar:kat-coder-v2-5-dev-validation
queries asked of Scott's wikis
  • open-weight models for local coding agents
  • single-GPU agentic inference economics
  • GGUF and llama.cpp production agent stacks
  • benchmark claims versus agent task completion
  • local model censorship and agent reliability
  • open-weight strategy versus proprietary model APIs

Measured heat

no measured readings yet — the hourly heat pass fills this in

How the heat travelled

no chain yet — the hourly chain pass fills this in

Evidence (39) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟧 hnMeta's new open-weight model targets local agentic AIbakigul409
🟠 redditunsloth/Muse-Glimmer-30B-GGUF · Hugging Face
LocalLLaMA
Nunki08445128
🟠 redditMeta will open source their Muse Spark 1.2 and Muse Glimmer 30B
artificial
insumanth61
🟠 redditMeta will soon release the weights for Muse Spark 1.2, their latest foundation model.
singularity
acoolrandomusername52670
🟧 echo.blog ⭐Meta’s essay says it will build “leading personal superintelligence agents and models, including open source models,” deliver them broadly, Meta (Mark Zuckerberg)——
🟠 redditmodel: Muse Glimmer Support by pcuenca · Pull Request #26841 · ggml-org/llama.cpp
LocalLLaMA
jacek20235716
🟠 redditMuse Glimmer ACTUALLY fits on a single RTX 3090
LocalLLaMA
coder543370129
🟠 redditGlimmer seems pretty censored?
LocalLLaMA
Cold_Tree190158103
🟠 redditEarly signs that Muse-Glimmer-30B might quantize *very* well? Share your experiences.
LocalLLaMA
EmPips20175
🟧 hnMark Zuckerberg attacks 'closed' AI rivals as Meta returns to open modelsroot-parent623561
🟠 redditoptimizing glimmer 30b for 3090
LocalLLaMA
ydnar115
🟠 redditGlimmer: 233.4 tps on 5090 with Dflash
LocalLLaMA
YetAnotherAnonymoose040
🟠 redditMuse Glimmer + Hermes getting stuck with loads of terminal commands
LocalLLaMA
KingGongzilla29
🟠 redditPlease Share Your Experience About Muse Glimmer
LocalLLaMA
BarberIcy36698108
🟠 redditMuse Spark 1.2 Open Source before Llama 4 Behemoth!!?
LocalLLaMA
aero-spike10028
🟠 redditMuse Glimmer on one 3090: a max_tokens gotcha that made it look dumb, numbers at *filled* context, and it handles non-English better than I expected
LocalLLaMA
TigerConsistent20
🟠 redditI made a web-design benchmark for local models (Muse Glimmer 30B vs Qwen 3.6 27b vs Deepseek V4 Flash 0731)
LocalLLaMA
ShadyShroomz7527
🟠 redditTested Muse Glimmer locally on coding with OpenCode & agentic work
LocalLLaMA
curiousily_2716
🟠 redditAchievable 253 t/s - unsloth/Muse Glimmer 30B UD-Q5_K_M on a 5090
LocalLLaMA
patricious2612
🟠 redditMuse glimmer benchmark
LocalLLaMA
NoFaithlessness951381117
🟧 hnZuck rekindles open weights Llama drama with Muse Glimmerjoebuckwilliams11
🟠 redditMuse Glimmer on 1/2 AMD v620
LocalLLaMA
Thin_Pollution884377
🟠 redditMuse-Glimmer 30B Hits ~280 t/s in Real Production Coding
LocalLLaMA
Ok-Shower72863826
🟠 redditObservations on Muse-Glimmer reasoning traces being noticeably different from qwen / gemma models and questions for you guys
LocalLLaMA
Certain-Cod-14045864
🟠 reddit1 Day in and I feel okay saying Muse-Glimmer-30B finally beats 3.6-27B for the size in some use-cases
LocalLLaMA
ForsookComparison346195
🟠 redditI ran Muse Glimmer @ 1M context - All tests passed.
LocalLLaMA
StartupTim15049
🟠 redditMuse Glimmer 30B + DFlash speculative decoding on vLLM: 6 patches needed, 25 → 57 tok/s. Dockerfile and numbers inside.
LocalLLaMA
j4ys0nj196
🟠 redditadded day-1 mlx-lm support for meta's muse glimmer 30b (PR up)
LocalLLaMA
divinetribe110
🟠 redditTested in Coding: BF16 Muse Glimmer vs BF16 Qwen3.6 27B
LocalLLaMA
PathfinderTactician8135
🟠 redditMuse Glimmer 30B + DFlash drafter slower than vanilla - low acceptance rate
LocalLLaMA
No_Algae1753033
🟠 redditMuse Glimmer overthinking like crazy
LocalLLaMA
tacticaltweaker617
🟠 redditAn in-depth on-and-off MTP test (Includes Muse Glimmer!)
LocalLLaMA
KitchenAmoeba443833
🟠 redditTested Muse Glimmer + Hermes Agent for Local/Private AI Agent & Coding
LocalLLaMA
curiousily_08
🟠 redditMuse Glimmer 30B running locally in-browser with custom WebGPU kernels at ~25 tok/s on an M4 Max (same speed as llama.cpp)
LocalLLaMA
xenovatech355
🟠 redditMuse Glimmer on RX 7600 XT 16GB
LocalLLaMA
DanC40357
🟠 redditZuckerberg published a manifesto saying no single company should control AI. Then Meta released a model that runs on your laptop. Same day.
artificial
Dapper-Tale-402108
🟧 hnProving Muse Glimmer in Zero-Knowledge at 85 Tok/S on an H100mildog830
🟠 redditLocal Benchmark : Muse Glimmer 30B vs Qwen 3.6 27B vs Gemma4 31B (and many other models and finetunes)
LocalLLaMA
WonderRico6825
🟠 redditI got Muse Glimmer 30B from 17.98 to 84.64 tok/s on a 24 GB GPU. The fastest quant lost.
LocalLLaMA
iam31337020

Interpretation history

Decision trace