2026-10-11 17:10 UTC

Independent benchmarks will determine whether the released Qwen3.5-9B triple-loop prototype improves small-model capability through recursive middle-layer computation without disproportionate inference cost.

state: expiredheat: lowuncertainty: highknownscott: mediumopen-models model-architecture local-inferenceQwenNanbeigeLordnyx

What is this?

The case concerns an experimental fine-tune of Qwen/Qwen3.5-9B that reportedly uses “LoopSplit” to execute middle layers recursively, aiming to gain capability without proportionally increasing inference cost. The supplied snippets establish Qwen3.5-9B as an open-weight, dense 9B multimodal model from Alibaba’s Qwen team, with long context and strong reported small-model benchmarks. However, they do not independently document the triple-loop prototype, explain Nanbeige’s or Lordnyx’s roles, or provide measurements of its quality, latency, memory use, or compute cost, so the central claim remains unverified here.

Why it matters to Scott

The radar already tracks essentially the same layer-repetition quality/compute question in `program-of-layers-dynamic-inference` and the closely related small-model claim in `nanbeige-4-2-3b-looped-transformer`. The released 9B artifact is nevertheless a directly testable candidate for Scott’s hardware-aware local-inference stack; measured capability, latency, VRAM, and throughput could affect local model selection, but the supplied evidence contains no such results yet.
ip:concept.inference-time-scalingdev:concept.hardware-aware-local-inferencedev:project.gamepcradar:program-of-layers-dynamic-inferenceradar:nanbeige-4-2-3b-looped-transformerradar:concept.model-architectureradar:concept.inference-efficiencyradar:concept.local-inference
queries asked of Scott's wikis
  • recursive computation and looped transformer layers
  • test-time compute versus inference cost
  • small-model capability per FLOP
  • local inference latency and memory economics
  • independent benchmarking of model architecture claims
  • open-weight architecture experiments

Measured heat

no measured readings yet — the hourly heat pass fills this in

How the heat travelled

no chain yet — the hourly chain pass fills this in

Evidence (2) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 redditQwen3.5-9B Triple-Loop
LocalLLaMA
Important-Farmer-8464717
🟧 echo.other ⭐The Hugging Face model card is the primary artifact: “An experimental fine-tune of Qwen/Qwen3.5-9B using LoopSplit,” with middle layers execLordnyx——

Interpretation history

Decision trace