2026-10-11 17:11 UTC

Reddit user Bitter-Truck1049's controlled six-run /usage measurement claims headless Claude Code (Agent SDK and `claude -p`) consumes roughly 3x more of the 5-hour limit per dollar of API-equivalent tokens than interactive use, and Anthropic documenting, confirming, or correcting that undocumented differential decides who actually bears the usage-limit cut for headless workloads.

state: resolvedheat: lowuncertainty: mediumnovelscott: highagent-harnesses inference-economics claude-codeAnthropic

What is this?

A Reddit user (Bitter-Truck1049, Sep 28, 2026) posted a controlled six-run measurement claiming headless Claude Code entrypoints โ€” the Agent SDK and `claude -p` โ€” burn ~3x more of Claude's 5-hour subscription window per dollar of API-equivalent tokens than interactive use (~5.2% vs ~1.7% per $1), with in-thread corroboration citing transcript entrypoint fields (`sdk-py`/`sdk-cli` vs `cli`/`claude-vscode`). A second independent thread documents subagents defaulting to a 5-minute prompt-cache TTL vs 1 hour for the main session โ€” corroborating a family of undocumented limit-drain mechanics while also offering a mundane mechanism (cold-cache misses inflating real token cost) for part or all of the 3x; no supplied source confirms or corrects the entrypoint-metering claim itself, and Anthropic has said nothing about it. The instrumentation opacity is documented: `/usage` is TUI-only, and the SDK's `rate_limit_event` messages omit utilization fields in normal operation (anthropics/claude-code#50518). The policy backdrop is conflicting in the supplied material: most sources describe the announced June 15, 2026 split of programmatic usage onto a separate API-priced credit as in effect, one guide (checked June 26) reports it paused โ€” and the Sep 28 measurement's premise (headless burn still landing on the 5-hour window) implies the subscription path was still live for that user.

Why it matters to Scott

Directly actionable on Scott's live headless scheduling stack: the router project's hourly/weekly Claude Code agents and the cron-ebook's cache-window-tuned cadence are exactly the `claude -p` pattern the claimed ~3x burn taxes, and if the 5-min subagent cache TTL is the mechanism it hands fresh economic support to prompt-interrupt's warm-kernel design over cold fresh runs (while conditioning the headless applicability of his prefix-caching-economics near-linearity claim). The contested entrypoint-vs-cache attribution is also precisely the controlled, same-fixture comparison his Langfuse / trace-backed-agent-comparison tooling was built to run โ€” a first-to-settle measurement opportunity, not just a topic match.
dev:project.routerip:source.ask-yourself-if-you-re-finished-cron-as-the-poor-man-s-orchestrator-ebookip:framework.prompt-interrupt-architectureip:concept.prefix-caching-economicsdev:technology.claude-codedev:technology.langfusedev:concept.trace-backed-agent-comparisonradar:claude-code-usage-limit-cutradar:claude-code-cache-ttl-analyzerradar:cache-tax-idle-session-warmingradar:replay-prompt-cache-miss-auditradar:frontierharness-17x-cost-variationradar:codepress-subscription-cloud-agentsradar:hermes-claude-directsdkradar:underclass-sticky-subscription-poolradar:claude-phantom-token-billing-bug
queries asked of Scott's wikis
  • claude -p headless cron scheduled orchestration runs
  • subscription vs API arbitrage pricing case
  • prompt cache TTL cache-miss token cost inflation
  • Langfuse trace token attribution comparison tooling
  • metering opacity undocumented rate-limit changes platform trust
  • agent harness token overhead context plumbing

Measured heat

now 4 pts/hpeak 18 pts/hcomments 1/hpeers p91momentum: steady1 platformsage 221h
points/hour across evidence ยท reading as of 2026-10-08 08:06:52.397066+11:00 ยท deterministic, not a model opinion

How the heat travelled

09-28 16:00โญ origin directly observedMeasured it: headless Claude Code (Agent SDK / claude -p) eats ~3x more of the 5-hour limit per token than interactive use (vscode)
Bitter-Truck1049 on r/ClaudeAI
โ€”
10-06 20:58first on r/ClaudeAI ยท published ยท +197.0hSubagents use a 5-min cache by default?!? ๐Ÿคฆ๐Ÿปโ€โ™‚๏ธ
Master-Biscotti-1186
โ€”
09-28 16:00amplified on r/ClaudeAI ๐Ÿ‘‘reddit.post.1wsijnv
Bitter-Truck1049
peak 15 ยท 42 comments ยท 56% of case engagement
10-06 20:58amplified on r/ClaudeAIreddit.post.1wze3rr
Master-Biscotti-1186
peak 3 ยท 11 comments ยท 14% of case engagement
10-07 18:38amplified on r/ClaudeAIreddit.post.1x04j1x
ThePaSch
peak 22 ยท 8 comments ยท 30% of case engagement
09-28 18:20our radar first saw it ยท +2.3hdiscovery anchor: reddit.post.1wsijnvโ€”

Evidence (3) โ€” โญ canonical anchor

sourceobjectauthorscorecomments
๐ŸŸ  reddit โญMeasured it: headless Claude Code (Agent SDK / claude -p) eats ~3x more of the 5-hour limit per token than interactive use (vscode)
ClaudeAI
Bitter-Truck10491446
๐ŸŸ  redditSubagents use a 5-min cache by default?!? ๐Ÿคฆ๐Ÿปโ€โ™‚๏ธ
ClaudeAI
Master-Biscotti-1186714
๐ŸŸ  redditClaude Agent SDK no longer usable with subscription limits; now API-only
ClaudeAI
ThePaSch3010

Interpretation history

Decision trace