2026-10-11 17:11 UTC

Anthropic will confirm and fix a server-side Claude context-injection bug that reports millions of phantom tokens and may overcharge usage-billed accounts.

state: resolvedheat: lowuncertainty: highconvergesscott: lowclaude billing-bug context-window server-side-incidentAnthropic

What is this?

Claude users reported requests failing with context-limit errors claiming roughly 5.5 million excess tokens, even for one-word prompts; a separate Claude Code issue reports runaway input-token consumption that depleted API credits. The supplied search summary says Anthropic confirmed and is fixing a server-side bug, but the underlying snippets do not directly establish that confirmation, explain the alleged token injection, or prove that context-error victims were billed for phantom tokens. Other results describe broader alleged Claude billing discrepancies, though their relationship to this specific incident is unclear.

Why it matters to Scott

If confirmed, the incident strengthens Scott’s position that provider token counts and costs require independent receipts, observability, and enforceable spend controls—especially because his projects route model traffic through LiteLLM and retain Opus rescue paths. The operational implication is meaningful, but provisional: the supplied evidence does not establish Anthropic’s confirmation, phantom-token billing, or a fix.
ip:concept.agent-receiptsdev:technology.langfusedev:technology.litellmdev:project.dev-wiki
queries asked of Scott's wikis
  • LLM usage metering and billing auditability
  • context-window accounting in coding-agent harnesses
  • runaway tool context and prompt injection safeguards
  • provider incident resilience and spend limits
  • token observability for agent workflows
  • verifiable AI API billing and dispute mechanisms

Measured heat

no measured readings yet — the hourly heat pass fills this in

How the heat travelled

no chain yet — the hourly chain pass fills this in

Evidence (2) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 reddit ⭐PSA: Getting "This request exceeds Claude's context limit by about 4-6M tokens" on every message? It's a server-side bug, multiple accounts affected — and usage-billed users are paying for the injected tokens
ClaudeAI
MaleficentSurprise216
🟠 reddit[Resolved / post-mortem] The "context limit exceeded by ~4-6M tokens on every message" bug: what it was, and the memory-backend fix
ClaudeAI
MaleficentSurprise12

Interpretation history

Decision trace