2026-10-11 16:37 UTC

Voodoo Quant creator 1ncehost claims the newly MIT-licensed method improves aggressive quantization of smaller Qwen3.5 GGUF models, potentially enabling others to reproduce and extend those local-inference quality gains.

state: seedheat: lowuncertainty: highnovelscott: mediumquantization local-inference gguf1ncehost

What is this?

Voodoo Quant is presented as a quantization method for local models; the case attributes it to creator 1ncehost and reports a new MIT license for improving aggressively compressed smaller Qwen3.5 GGUF models. A Reddit search snippet describes per-tensor optimization rather than Unsloth Dynamic’s tensor-block approach, but the supplied web results do not verify the creator attribution, license change, or reproducibility of the claimed gains. The surrounding Qwen3.5 evaluations establish active work on quantization quality, while Benjamin Marie’s evaluation summary cautions that better KL divergence or perplexity does not guarantee better task benchmarks.

Why it matters to Scott

This offers a new candidate to test for Scott’s gamepc/Ollama bulk classification and generation workloads, where numerical precision and memory pressure are explicit runtime concerns; the hits do not establish that he uses Qwen3.5 or already holds a position on this method. The radar tracks related GGUF optimizations, not this license development, and the unverified license and quality claims warrant evaluation rather than a deployment change.
dev:concept.hardware-aware-local-inferencedev:technology.ollamadev:project.gamepcradar:concept.quantizationradar:concept.ggufradar:bartowski-gguf-tensor-layoutsradar:unsloth-dynamic-3-gguf-validation
queries asked of Scott's wikis
  • local inference memory budgets quality tradeoffs
  • GGUF llama.cpp quantization projects
  • small local models agent workloads
  • quantization evaluation task accuracy versus KLD perplexity
  • permissive licensing reproducible model optimization

Measured heat

now 0 pts/hpeak 0 pts/hcomments 0/hpeers p0momentum: steady1 platformsage 633h
points/hour across evidence · reading as of 2026-10-12 02:59:37.977291+11:00 · deterministic, not a model opinion

How the heat travelled

09-15 06:59⭐ origin directly observedVoodoo Dynamic Quant - Now MIT Licensed
1ncehost on r/LocalLLaMA
—
09-15 06:59amplified on r/LocalLLaMA 👑reddit.post.1wgszma
1ncehost
peak 347 · 40 comments · 100% of case engagement
09-15 07:20our radar first saw it · +0.3hdiscovery anchor: reddit.post.1wgszma—
pace: p81 vs 1032 stories at the 336h mark (now 633h old) — ahead of nvidia-sol-pi-harness-efficiency (1.0x), behind astra-zerobench-human-baseline (1.0x)

Evidence (1) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 reddit ⭐Voodoo Dynamic Quant - Now MIT Licensed
LocalLLaMA
1ncehost34540

Interpretation history

Decision trace