mixture-of-experts
band: warmmomentum: stable
score: 0.523
Episodes (12)
Trajectory notes
- 2026-08-28T20:42:25Z: meta-mobilemoe-on-device-release closed (faded) β Metaβs claimed sub-3GB quality-efficiency frontier converges with Scottβs Model Barbell and Usable Mass positions: cheap, deployable models can be more valuable than larger but impractical capability, especially when hardware co
- 2026-08-27T12:27:34Z: layerstorm-moe-expert-streaming closed (faded) β The radar already tracks this same expert-streaming/offloading question in βAirLLM low-VRAM model streaming,β βExpertCache GPT-OSS 120B M1,β and βllama.cpp hot-expert GPU cache.β LayerStoRm is another unbenchmarked implementation
- 2026-08-25T15:48:01Z: optimal-transport-moe-load-balancing closed (faded) β The case repeats Scottβs Evidence Class Ladder position that seller or paper claims should not be treated as established until independently verified. With TAOT itself and its reported gains uncorroborated in the supplied ma
- 2026-08-12T06:30:53Z: lumabri-peer-to-peer-moe-inference closed (faded) β Lumabri independently applies the volunteer-computing pattern Scott encountered through ChessBrain to his active territory of hardware-aware local inference, extending it from distributed chess search to sharded MoE execution
- 2026-08-07T19:35:00Z: gguf-lora-16gb-moe-training closed (faded) β Scott already treats precision, memory pressure, accelerator placement, and compilation as explicit policy in dev:concept.hardware-aware-local-inference, with dev:project.gamepc providing an active consumer-GPU substrate. The claimed
- 2026-08-03T15:29:58Z: kat-coder-v2-5-dev-validation closed (absorbed) β The need for independent, harness-level validation is already explicit in Scottβs Evaluation-Driven Development position, while Ask and gamepc make a capable 3B-active open coding model a plausible backend candidate. It matters