2026-10-11 18:02 UTC

Independent evaluation and eventual weight access will determine whether the Looping 20B recipe can match or exceed Qwen3 Coder 30B after pretraining on roughly one-tenth as many tokens.

state: expiredheat: lowuncertainty: highknownscott: lowefficient-pretraining open-models coding-models

What is this?

The case describes a reported “Looping 20B” coding model trained on roughly 3.5 trillion tokens—about one-tenth the claimed pretraining volume—and alleges that paper-reported evaluations match or exceed Qwen3 Coder 30B. The supplied web results are unrelated and do not identify the model’s creators, explain the looping recipe, confirm the benchmark results, or establish whether weights are available, so the claim remains ungrounded pending the paper, independent evaluation, and weight access.

Why it matters to Scott

Scott already holds the relevant position in Capability Audit and Evaluation-Driven Development: paper benchmarks are insufficient without repeatable independent testing. Weight access could make the model actionable for his hardware-aware local-inference work, but the supplied material establishes neither the result nor availability, so this is currently an unverified example rather than a consequential update.
ip:concept.capability-auditip:concept.evaluation-driven-developmentdev:concept.hardware-aware-local-inferenceradar:concept.ai-benchmarksradar:concept.benchmark-integrityradar:concept.open-models
queries asked of Scott's wikis
  • compute-efficient pretraining and token economics
  • recursive or looping transformer architectures
  • benchmark claims versus independent model evaluation
  • open-weight access and reproducibility
  • coding-model evaluation and agentic coding performance
  • small-model efficiency versus parameter scaling

Measured heat

no measured readings yet — the hourly heat pass fills this in

How the heat travelled

no chain yet — the hourly chain pass fills this in

Evidence (2) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 reddit20B Looping model (paper) matches or beats Qwen3 Coder 30B at 10% of pre-training tokens
LocalLLaMA
Dany06717
🟧 echo.paper ⭐The paper reports that a 20B Looping model trained on about 3.5 trillion tokens matches or exceeds Qwen3 Coder 30B on reported evaluations.Looping model paper authors——

Interpretation history

Decision trace