2026-10-11 18:01 UTC

Independent use will determine whether Claude Code’s model and effort-level controls provide predictable quality, latency, and inference-cost tradeoffs for coding-agent workloads.

state: resolvedheat: lowuncertainty: lowconvergesscott: highcoding-agents inference-economics model-routingAnthropicClaude Code
Surfaced 2026-09-20T01:22:22Z — Anthropic introduced and explained model and effort-level controls in Claude Code for adjusting how much inference effort coding tasks recei — Fresh, fast-moving discussion has shifted attention toward a consequential operational claim: changing effort mid-session may invalidate prompt caching and distort the apparent cost tradeoff. The claim remains unverified, but the guide’s rapid uptake and the case’s broad cross-platform spread now warrant near-term attention without advancing evidentiary maturity.

What is this?

Claude Code, Anthropic’s coding-agent product, exposes model and effort-level controls intended to trade inference cost and latency against task quality; Anthropic’s documentation recommends selecting these per request and reports internal agentic-coding benchmarks with important limits, including uneven run counts across configurations. The supplied case records benchmark discussion and several early user reports suggesting Opus 5.5 can perform best around medium effort, with higher settings sometimes adding cost, overengineering, or out-of-scope edits rather than quality. Anthropic has also reportedly confirmed serving experiments that remap numerical effort values, so labels may not be stable across time; the independent evidence remains sparse and does not yet isolate retries, caching, context, delegation, fan-out, or accepted-outcome quality.

Why it matters to Scott

Anthropic’s benchmark discussion and independent Opus 5.5 reports converge directly with Scott’s `High, Not Max` and task-aware routing positions: more inference can worsen accepted coding outcomes while increasing spend, making medium/high—not max—a defensible default. This is a dated-receipts and implementation opportunity for his routing and evaluation work, although unstable effort mappings and uncontrolled workload confounders mean he should validate cost per accepted outcome on trace-backed fixtures before changing policy.
ip:concept.high-not-maxip:concept.token-disciplineip:concept.evaluation-driven-developmentdev:concept.task-aware-model-routingdev:concept.trace-backed-agent-comparisonradar:qwen38-27b-reasoning-effortradar:frontierharness-17x-cost-variationradar:dynamic-model-switching-evaluationradar:claude-subagent-budget-exhaustion
queries asked of Scott's wikis
  • High Not Max reasoning-effort strategy
  • task-aware model routing for coding agents
  • coding-agent evals quality latency cost per accepted outcome
  • adaptive inference budgets and escalation policies
  • prompt caching effects of mid-session model or effort switching
  • subagent fan-out retries and total task economics

Measured heat

no measured readings yet — the hourly heat pass fills this in

How the heat travelled

08-15 10:23 (minted)⭐ origin echo-reconstructedAnthropic introduced and explained model and effort-level controls in Claude Code for adjusting how much inference effort coding tasks recei
Anthropic on blog (echo) · attributed from hn.story.49309309 · published time unknown
—
08-15 10:09first on hacker news · published · lag ?Model v/s Effort Well Explained
ankitg12
—
08-15 10:43first on r/ClaudeAI · published · lag ?How is Opus 5 medium reasoning effort surpassing Fable 5 max and Opus 5 max on FrontierCode v1.1 Main benchmarks?
Lucky_Creme_5208
—
08-22 18:15first on r/artificial · published · lag ?I spent the morning digging into Anthropic so I could write it up properly. The short version
Positive-Ad3618
—
08-25 13:25first on r/singularity · published · lag ?Anthropic’s best AI model struggles to attract users as cheaper tools thrive
yogthos
—
08-25 19:36first on r/OpenAI · published · lag ?ChatGPT (EFFORT) issue
420blzithoe
—
09-05 16:06first on r/LocalLLaMA · published · lag ?Shouldn't the solution to thinking-effort be an adaptive system?
milpster
—
08-15 10:09amplified on hacker newshn.story.49309309
ankitg12
peak 1 · 0 comments · 0% of case engagement
08-15 10:43amplified on r/ClaudeAIreddit.post.1vozlz0
Lucky_Creme_5208
peak 1 · 3 comments · 0% of case engagement
08-16 10:49amplified on r/ClaudeAIreddit.post.1vptwlz
Plenty-Emu3740
peak 191 · 168 comments · 5% of case engagement
08-16 15:29amplified on hacker newshn.story.49320976
thomas_witt
peak 1 · 0 comments · 0% of case engagement
08-18 11:07amplified on r/ClaudeAIreddit.post.1vrm2e0
ThatsNot_True
peak 0 · 26 comments · 0% of case engagement
08-18 23:21amplified on r/ClaudeAIreddit.post.1vs5f9r
FlaTreNeb
peak 438 · 79 comments · 7% of case engagement
47 more amplifiers in ainews.case_chain
08-15 10:21our radar first saw it · lag ?discovery anchor: hn.story.49309309—
08-22 19:40reached heat=high · lag ? · via ledger——

Evidence (54) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟧 hnModel v/s Effort Well Explainedankitg1210
🟧 echo.blog ⭐Anthropic introduced and explained model and effort-level controls in Claude Code for adjusting how much inference effort coding tasks receiAnthropic——
🟠 redditHow is Opus 5 medium reasoning effort surpassing Fable 5 max and Opus 5 max on FrontierCode v1.1 Main benchmarks?
ClaudeAI
Lucky_Creme_520813
🟠 redditThe Absurd Math of $20 AI Coding Subs: Codex vs. Claude Code
ClaudeAI
Plenty-Emu3740184168
🟧 hnI stopped my Claude Code subagents from running on Fable instead of Sonnetthomas_witt10
🟠 redditPlease teach me how to use claude code for building a project. Token usage is getting crazy.
ClaudeAI
ThatsNot_True026
🟠 redditThis is new ... Claude seems to be not in the mood to do some work
ClaudeAI
FlaTreNeb43678
🟠 redditHow is this possible? Claude Code 1M context vs 200k usage almost same
ClaudeAI
Alert_Peak865511
🟧 hnOpus 5.0 drives incoherence into the stratosphereBluestein188170
🟧 hnIs Claude Code a bad harness?wsun1921
🟧 hnSystem reminders – how Claude Code steers itselfBluestein10
🟧 hnAnthropic appears to be A/B testing reduced effort levels in Claude Codematthieu_bl196175
🟠 redditI spent the morning digging into Anthropic so I could write it up properly. The short version
artificial
Positive-Ad361820
🟠 redditAnthropic Stealth Nerfing Effort Levels
ClaudeAI
jcll17550
🟠 redditSonnet 5 Low vs Medium vs High for daily usage
ClaudeAI
Sniffyunt1233127
🟠 redditI don’t understand most of the Opus 5 hate, except…
ClaudeAI
josh_a2934
🟠 redditAnthropic’s best AI model struggles to attract users as cheaper tools thrive
singularity
yogthos752143
🟠 redditChatGPT (EFFORT) issue
OpenAI
420blzithoe78
🟠 redditIs Sonnet actually good enough for Claude Code, or do you mostly stick with Opus?
ClaudeAI
Kiana_Harpel4464
🟠 redditfinding a balance with claude...
ClaudeAI
Lopsided-Start-363125
🟠 redditClaude with low priority mode when 5h limit uses up
ClaudeAI
innovaldragon11
🟠 redditWhen do you actually reach past "high" reasoning effort?
ClaudeAI
Connect-Employer548718
🟠 redditdoes max effort actually write worse code than high
ClaudeAI
andybarriateth03
🟠 redditopus 4.6 randomly "not thinking" in claude code desktop, found out why
ClaudeAI
andybarriateth144
🟧 hnIs it just me, or has Claude Opus gotten worse recently?de6u99er1118
🟠 redditJust Know this about Fable 5.1 Max
ClaudeAI
semiward341122
🟠 redditDoes it happen for you as well?
ClaudeAI
azimovgrbn014
🟠 redditExtended thinking option gone?
ClaudeAI
iamthe0ther0ne2223
🟠 redditDo you really actively change the effort level (i.e. usage) of your Claude?
ClaudeAI
MountainOriginal3663512
🟠 redditShouldn't the solution to thinking-effort be an adaptive system?
LocalLLaMA
milpster19
🟧 hnYou're paying for Claude's thinking and you're not getting itzerotoken20
🟠 redditThe best subagent for Claude?
ClaudeAI
Pichonn2128
🟠 redditClaude Code - question about general experience
ClaudeAI
Ok_Demand_861401
🟠 redditBurnt through Fable usage in one day with MAX20 -- need help
ClaudeAI
Sean_NobleThreads1933
🟠 redditWhere did the latest models and effort levels go?
ClaudeAI
ChickenNBeans03
🟧 hnDive into FrontierHarness Eval: Claude Code cost 5.6× for the same pass rateADD-SP63
🟠 redditWhat is the best and most balanced effort to code with Fable 5 on the $100 plan?
ClaudeAI
JohnyGhost511
🟠 redditWhat is your burned for nothing record?
ClaudeAI
XTemplarxxx817
🟠 redditSurviving Claude Code’s tightened limits: Effort levels, subagent traps, and CLAUDE.md tuning
ClaudeAI
noletovictor011
🟧 hnOpen-Source Skill Makes Claude Code 38% Cheaper and 38% Faster on Small Projectsvmysla10
🟠 redditClaude Code ships a per-model table of what each effort level costs
ClaudeAI
Wsz20207611
🟠 redditModel effort menu missing
ClaudeAI
j0hnqd0303
🟠 redditClaude Code effort level and model selection | Claude
ClaudeAI
theleller45265
🟠 redditIt's time to cancel your subscriptions - Anthropic is silently nerfing Claude's reasoning budget while telling you it's the same model
ClaudeAI
IcyEase2025381
🟠 redditClaude seems both stupider and more expensive than before
ClaudeAI
nickjohnson4022
🟠 redditWhy do the Anthropic models do better with medium effort on agentic coding?
ClaudeAI
ShadowxWarrior14
🟠 redditOpus 5.5 Still Overthinks at Higher Effort
ClaudeAI
Wsz20201720
🟠 redditCan someone explain the Agentic Coding effort drop?
ClaudeAI
keeganwest07
🟠 redditCheap per token, expensive per task: AI model pricing vs. performance [OC]
artificial
Wsz202035
🟠 redditGPT-6 and Opus 5.5's biggest revolution isn't performance, its speed and cost.
singularity
that_90s_guy8124
🟠 redditAnecdata: Opus 5.5 medium cheaper than Sonnet on well specified coding and review
ClaudeAI
andyfsu9930
🟠 redditOpus 5.5 effort level?
ClaudeAI
solishu41122
🟠 redditYour subagents probably aren't running at the effort you think they are.
ClaudeAI
ArgonQQ11
🟧 hnUsing Claude Code: Spending your effortmfiguiere20

Interpretation history

Decision trace