model-routing
band: warmmomentum: stable
score: 0.463
Episodes (20)
Trajectory notes
- 2026-09-28T14:10:40Z: vercel-september-open-weight-majority closed (absorbed) — Vercel's dated production index — open weights at 56% of tokens for 14% of spend, third straight monthly price drop — plus the FT's demand-side reporting show the world independently arriving at the task-risk/price barbe
- 2026-09-25T21:40:24Z: claude-code-effort-controls closed (absorbed) — Anthropic’s benchmark discussion and independent Opus 5.5 reports converge directly with Scott’s `High, Not Max` and task-aware routing positions: more inference can worsen accepted coding outcomes while increasing spend, making m
- 2026-09-02T17:58:48Z: dynamic-model-switching-evaluation closed (faded) — The claim independently supports Scott’s evaluation-driven approach while sharpening it for his task-aware model-routing systems: frozen fixtures may mis-rank policies when deployment changes the workload they subsequently rec
- 2026-09-02T13:32:16Z: datadog-ai-usage-cost-optimization closed (faded) — Datadog supplies a consequential enterprise-scale receipt for Scott’s position that production observability, spend attribution, budgets, and model allocation can materially improve AI unit economics. The claimed $1M-plus mont
- 2026-08-30T21:30:34Z: experiential-open-model-gateway closed (faded) — Scott already operates the same core pattern through the self-hosted, OpenAI-compatible LiteLLM gateway and explicit multi-provider routing, while the radar already tracks comparable self-hosted routing layers in NeMo Switchyard,
- 2026-08-20T15:37:30Z: speko-voice-model-router closed (faded) — Speko productises Scott’s existing task-aware, swappable multi-provider routing approach and applies it across the full STT→LLM→TTS voice stack, directly overlapping his LiteLLM routing and Twilio voice-agent experiments. It offers a da
- 2026-08-20T09:37:53Z: specjudge-repository-model-routing closed (faded) — SpecJudge combines Scott’s existing CLAUDE.md-as-repository-spec pattern with his task-aware, cost-tiered model-routing work, while adding a concrete repository-level routing and insufficient-context decision. Independent resu
- 2026-08-13T09:36:37Z: databricks-coding-agent-cost-reduction closed (faded) — Databricks is a consequential enterprise platform vendor reporting the same governed-gateway, task-routing, budget-control, and observability pattern Scott already argues for and implements through task-aware routing and L
- 2026-08-12T14:54:32Z: lupin-claude-code-backend-portability closed (faded) — Scott already holds and implements this position in Ask terminal agent, Markdown OS, and Model-Plus-Harness Benchmark Unit: preserve an inspectable harness while swapping models, then evaluate the model-plus-harness combina
- 2026-08-10T04:30:43Z: claude-code-auto-mode-default closed (disproved) — Anthropic’s default risk-based permission routing independently operationalizes Scott’s risk-based triage, human-over-the-loop supervision, and irreversibility-scaled autonomy in a major coding-agent product, creating a strong