frontier-models
band: hotmomentum: stable
score: 1.0
Episodes (67)
Trajectory notes
- 2026-10-05T05:25:41Z: prinz-astra-wwi-cipher closed (faded) — This is the position Scott already holds: a single-source frontier capability claim can only be defended at the evidence class of its weakest support, and press amplification of one developer's report is an echo, not verification — held i
- 2026-10-04T00:28:59Z: typesafe-jev-structured-decisions closed (absorbed) — The world has independently arrived where Scott's canon already stood and where he is already building: the die-test/jevals/Red Hat refutation of Jev's calibrated confidence lands squarely on his risk-based-triage and determ
- 2026-10-02T07:52:32Z: claude-assisted-openai-intrusion closed (window-closed) — Converges as a dated receipt: the WSJ-relayed chain (public Discourse libheif exploit → over-permissioned SSO tokens valid for employee ChatGPT → internal GitHub 'Monorepo', benign PR, $6.5k bounty) is precisely the tran
- 2026-09-30T17:42:12Z: openai-gpt6-sol-luna-release closed (superseded) — OpenAI's 90–95% cached-input discounts, explicit breakpoints, and near-free Luna independently arrive where Prefix-Caching Economics and the Model Barbell already sit — a vendor price sheet as dated receipt — while the halved S
- 2026-09-29T23:19:42Z: anthropic-rnd-automation-index closed (absorbed) — Anthropic operationalizes Human Over the Loop at frontier-lab scale — Claude leads AL4 tasks end-to-end under supervision, nothing measured at AL5, monitors escalate to human review — a dated receipt for the governance model Sc
- 2026-09-27T19:25:29Z: xiaomi-mimo-live-training-dashboard closed (absorbed) — Xiaomi's open RL run is the corporate sibling of the radar's Liang open-training episode, but Scott's canon holds no position on training-run transparency — the dashboard is his own W&B-style training-telemetry pattern rea
- 2026-09-26T00:45:58Z: astra-vending-bench-results closed (absorbed) — The primary write-up resolves the existence question but changes nothing material: the numbers are still Andon's own six self-run replications with no cost figures behind the headline (Sol's 1/8-cost claim unverified), and the thi
- 2026-09-24T18:04:39Z: gemini-three-company-intrusions closed (absorbed) — The reported unintended internet access illustrates the containment requirement Scott already holds in SiloOS and Architecture, Not Vibes; the supplied accounts establish neither a deliberate sandbox escape nor Google adopting
- 2026-09-24T01:40:10Z: fable-51-multimodal-coding-cost closed (absorbed) — Anthropic’s reported visual/tool-use gains converge with Scott’s Agent Hands and Eyes and demonstration-to-agent compilation work, while the slower, costlier runs bear directly on his inference-time scaling and budget-control
- 2026-09-24T01:00:58Z: anthropic-opus-55-website-signal closed (absorbed) — The unverified route slug provides no capability, pricing, API, or evaluation evidence, so it does not yet change Scott’s Claude Code usage or model-comparison work. A confirmed release would warrant trace-backed harness eval