inference-economics
band: hotmomentum: stable
score: 1.0
Episodes (291)
Trajectory notes
- 2026-10-07T22:41:58Z: claude-code-usage-limit-cut closed (absorbed) — Already held: Scott's own wikis carry every position this settled episode re-demonstrates — the cut is Model Perishability (ip:concept.model-perishability) and Vendor Lock-In (ip:concept.vendor-lock-in) landing on a tool he runs d
- 2026-10-07T21:30:01Z: claude-code-headless-metering-3x closed (superseded) — Directly actionable on Scott's live headless scheduling stack: the router project's hourly/weekly Claude Code agents and the cron-ebook's cache-window-tuned cadence are exactly the `claude -p` pattern the claimed ~3x burn t
- 2026-10-07T00:28:41Z: anthropic-context-compaction-cost-reversal closed (absorbed) — Anthropic’s reported cache-sensitive costs converge with Scott’s Prefix-Caching Economics position and give a concrete reason to benchmark his Ask agent’s compaction at different run lengths rather than equating les
- 2026-10-06T14:47:11Z: deepseek-v41-flash-release closed (absorbed) — The expanding runtime and hardware ecosystem converges with Scott’s hardware-aware local-inference work, while the mixed agent results reinforce his position that economics must be measured per successful model-plus-harness task ra
- 2026-10-05T12:51:23Z: qwen38-27b-local-api-substitution closed (absorbed) — Independent builders have now arrived in the field at the posture Scott's own stack was built around — fixed-cost local inference as the working tier with hosted as escalation — and the corroborated reports bear directly on
- 2026-10-04T00:28:59Z: typesafe-jev-structured-decisions closed (absorbed) — The world has independently arrived where Scott's canon already stood and where he is already building: the die-test/jevals/Red Hat refutation of Jev's calibrated confidence lands squarely on his risk-based-triage and determ
- 2026-10-01T17:45:30Z: ukisai-swift-family-release closed (absorbed) — Converges with what dev:project.llmreport's Agent Token Manifesto and his cost-tiered LiteLLM routing already argue — wasted thinking tokens are a real, attackable cost — but the new material converts agreement into a fork his har
- 2026-10-01T07:43:34Z: openai-kimi-distillation-disruption closed (absorbed) — Converges hard on his own dev canon: OpenAI's cross-user replay of *encrypted* reasoning and its warning that portable/replayable reasoning artifacts are an attack surface independently validate the exact rule in dev:conce
- 2026-09-30T17:42:12Z: openai-gpt6-sol-luna-release closed (superseded) — OpenAI's 90–95% cached-input discounts, explicit breakpoints, and near-free Luna independently arrive where Prefix-Caching Economics and the Model Barbell already sit — a vendor price sheet as dated receipt — while the halved S
- 2026-09-29T22:57:57Z: openai-chatgpt-pro-max-500 closed (absorbed) — OpenAI's confirmed move — $200 Pro back on sale at roughly half its usage allowance hours before a possible $500 speed-priced tier — is a dated instance of the nerf-to-upsell pattern his canon already tracks across the Pro-pause, w