2026-10-11 18:04 UTC

Redditor Brinvik's analysis of Anthropic's own agent cost guide finds that its context-editing and compaction lever cost 74% more on a 20-issue run while saving 32-39% on longer runs, showing the guidance is workload-dependent rather than a universal cost win.

state: resolvedheat: lowuncertainty: lowconvergesscott: mediuminference-economics agent-harnessesAnthropic

What is this?

Anthropic’s Claude Platform cost guide reports workload-dependent results for agent context optimizations: on a 20-issue run, the discussed levers saved nothing and context editing cost 74% more. On a longer run, pruning stale tool results saved 39% and compaction saved 32%, while context editing changed nothing—so the case hypothesis incorrectly groups distinct techniques under the longer-run savings. The guide attributes the need for net-effect measurement to interactions with caching; the supplied web snippets substantiate these measurements but do not independently establish Redditor Brinvik’s analysis or attribution.

Why it matters to Scott

Anthropic’s reported cache-sensitive costs converge with Scott’s Prefix-Caching Economics position and give a concrete reason to benchmark his Ask agent’s compaction at different run lengths rather than equating less context with lower bills; the related radar episodes do not establish prior coverage of these measurements. The result also qualifies blanket transcript-retention economics: longer-run pruning saved 39% and compaction 32%, whereas context editing saved nothing there and cost 74% more on the short run—not the combined optimization win claimed in the hypothesis.
ip:concept.prefix-caching-economicsdev:project.askdev:concept.agent-authored-context-compactionradar:github-tool-output-cost-tradeoffradar:cache-hunter-prompt-cache-debuggingradar:concept.inference-economicsradar:concept.context-compression
queries asked of Scott's wikis
  • agent harness context pruning versus compaction
  • prompt cache invalidation context editing costs
  • workload-dependent inference economics net cost measurement
  • coding agent benchmarks task length cost quality tradeoffs
  • stale tool output reduction task boundary memory

Measured heat

now 0 pts/hpeak 6 pts/hcomments 0/hpeers p30momentum: steady2 platformsage 641h
points/hour across evidence · reading as of 2026-10-07 11:04:53.997080+11:00 · deterministic, not a model opinion

How the heat travelled

09-10 07:24 (minted)⭐ origin echo-reconstructedAnthropic’s official guide states: “On the 20-issue run they saved nothing, and context editing cost 74% more.” It attributes the measuremen
Anthropic on other (echo) · attributed from reddit.post.1wcbkyx · published time unknown
—
09-10 06:58first on r/ClaudeAI · published · lag ?Anthropic's own cost guide contains a lever that cost 74% more
Brinvik
—
09-10 06:58amplified on r/ClaudeAIreddit.post.1wcbkyx
Brinvik
peak 1 · 15 comments · 21% of case engagement
09-15 05:51amplified on r/ClaudeAI 👑reddit.post.1wgrskx
funplayer3s
peak 5 · 41 comments · 60% of case engagement
09-18 07:47amplified on r/ClaudeAIreddit.post.1wjjred
nano-zan
peak 1 · 3 comments · 5% of case engagement
09-18 15:10amplified on r/ClaudeAIreddit.post.1wjsuwr
Muchaszewski
peak 2 · 5 comments · 9% of case engagement
09-18 15:54amplified on r/ClaudeAIreddit.post.1wju17g
Wsz2020
peak 1 · 1 comments · 2% of case engagement
10-06 23:17amplified on r/ClaudeAIreddit.post.1wzhbfc
stichstichstich
peak 1 · 1 comments · 2% of case engagement
09-10 07:20our radar first saw it · lag ?discovery anchor: reddit.post.1wcbkyx—

Evidence (7) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 redditAnthropic's own cost guide contains a lever that cost 74% more
ClaudeAI
Brinvik015
🟧 echo.other ⭐Anthropic’s official guide states: “On the 20-issue run they saved nothing, and context editing cost 74% more.” It attributes the measuremenAnthropic——
🟠 redditDo claude code caches really expire in an hour?
ClaudeAI
funplayer3s341
🟠 redditNew Claude docs seem to use too many tokens
ClaudeAI
nano-zan13
🟠 redditCompact used 2% of my 20x Max Plan. ccusage doesn't show this as usage
ClaudeAI
Muchaszewski05
🟠 redditWhen to Compact
ClaudeAI
Wsz202001
🟠 reddit56% of my Claude Code usage was Claude re-reading the conversation
ClaudeAI
stichstichstich16470

Interpretation history

Decision trace