long-context
band: warmmomentum: stable
score: 0.463
Episodes (13)
Trajectory notes
- 2026-10-08T00:38:02Z: opus-55-behavior-shift closed (absorbed) — Anthropic has newly arrived in weights at behavior Scott's canon engineers at harness level: the corroborated rebuild/batching break is the Self-Equipping Agent's reuse-don't-rebuild doctrine emerging as default model behavior, while t
- 2026-10-07T01:41:55Z: aleph-alpha-kolibri-1-release closed (window-closed) — Aleph Alpha has independently shipped the exact combination Scott's canon and practice already occupy — a permissively licensed compact sparse-MoE with a 1M-token window aimed at local/agentic inference — which makes Kolibr
- 2026-09-30T11:53:55Z: spark-x25-small-model-release closed (faded) — Scott already distinguishes nominal million-token capacity from usable attention-residence and treats memory pressure, accelerator placement, and runtime support as explicit local-inference policy; those positions are carried by Th
- 2026-08-26T18:40:43Z: sub2bit-disk-context-local-model closed (faded) — Scott already holds the relevant positions in `ip:concept.usable-mass-over-unusable-power` and `ip:framework.context-engineering`: deployable small models can outperform unusable power, while durable histories should be external
- 2026-08-19T01:24:05Z: dots3-note-preview-validation closed (faded) — The evaluation posture is already held in Scott’s “Model-Plus-Harness Benchmark Unit” and “Evaluation-Driven Development”: advertised weights, active parameters, and context length do not establish agent capability without harness-
- 2026-08-11T05:26:55Z: deepseek-v4-flash-1m-rtx5090 closed (window-closed) — The claimed result extends Scott’s hardware-aware local-inference work by proposing a concrete operating point—CPU-offloaded experts plus phase-adaptive speculative decoding—for useful 1M-context inference on one consumer GP
- 2026-08-07T18:36:09Z: poolside-laguna-s-2-1-context-fix closed (faded) — No intersection found in Scott’s wikis, and no radar pages show that this development or its actors are already tracked. The supplied material therefore cannot establish that the weight updates or million-token-context claims b
- 2026-08-07T18:34:00Z: longcat-sparse-24gb-inference closed (faded) — No intersection found: there are no Scott wiki hits tying this architecture or its 24GB long-context inference claim to his positions or projects, and no radar hits showing the development or actor is already tracked.
- 2026-08-07T17:42:39Z: apertus-1-5-open-model-validation closed (faded) — No intersection found in Scott’s wikis or the radar’s accumulated pages. The case raises unresolved questions about open training data, multilingual quality, and practical long-context performance, but the supplied hits do not