2026-10-11 17:12 UTC

Anthropic claims newly released Sonnet 5.5 matches Opus 5.5 on coding at roughly half the per-token price; same-day community measurement counters that it emits 62% more tokens and costs more than Opus at max effort, making realized cost-per-task versus the headline discount decisive for whether Sonnet 5.5 displaces Opus 5.5 as the default coding-agent model.

state: resolvedheat: lowuncertainty: lowconvergesscott: highfrontier-model-releases inference-economics coding-agents anthropicAnthropicArtificial AnalysisVals AI
Surfaced 2026-09-28T23:02:14Z β€” Anthropic's Sonnet 5.5 launch post and official prompting guide: per the echoes, its benchmark tables show Sonnet 5.5 at 70.6% vs Opus 5.5's β€” Anthropic claims newly released Sonnet 5.5 matches Opus 5.5 on coding at roughly half the per-token price; same-day community measurement counters that it emits 62% more tokens and costs more than Opus at max effort, making realized cost-per-task versus the headline discount decisive for whether Sonnet 5.5 displaces Opus 5.5 as the default coding-agent model.

What is this?

Anthropic opened its Claude 5.5 family with Opus 5.5 on September 22, 2026 β€” $4/$20 per million input/output tokens (20% under Opus 5), cache reads cut 60% to $0.20, ~40% claimed workload savings β€” with Sonnet 5.5 and Haiku 5.5 explicitly promised for 'the coming weeks.' The case concerns that promised Sonnet 5.5 release (case dated Sept 28): its evidence titles record a same-day fight over whether Sonnet 5.5 matches Opus 5.5 on coding at roughly half the per-token price, with community measurement (Artificial Analysis / Vals AI style token accounting) countering that it emits ~62% more tokens, so realized cost-per-task at max effort can exceed Opus 5.5. The supplied web results do not cover the Sonnet 5.5 launch itself β€” they only establish the direct Opus 5.5 precedent that headline price cuts invert at max effort (Artificial Analysis: +37% total tokens, +143% output tokens, per-task cost rising to $13.04 vs Opus 5's $10.79; another benchmark reads $5.98 vs $5.86), so the Sonnet-specific figures (70.6% score, 62% token inflation, exact pricing) rest on the echo-testimony titles and same-day community measurement and cannot be verified from these snippets.

Why it matters to Scott

Converges on two of Scott's own positions with fresh dated receipts: same-day independent token accounting re-derives the Mature Token Law's audit-the-conversion claim (realized cost-per-task decides, not headline per-token price), and the 'xhigh not max' finding β€” including Anthropic's own docs scoring max below xhigh β€” independently corroborates High, Not Max. It is also directly actionable for his stack: whether Sonnet 5.5 displaces Opus 5.5 as default coding model is the same model/effort tuning his LiteLLM tier aliases and Claude Code usage encode, and his trace-backed fixture comparison is exactly the instrument that could settle the conflicting echo-testimony numbers; on the radar this continues the inference-cost lineage (Brinvik's Anthropic-guide cost reversal, hidden-reasoning real-task costs, Opus 5.5 repricing).
ip:framework.the-mature-token-lawip:concept.high-not-maxip:concept.model-barbelldev:concept.cost-tiered-llm-routingdev:concept.trace-backed-agent-comparisondev:technology.litellmradar:anthropic-context-compaction-cost-reversalradar:hidden-reasoning-real-task-costsradar:claude-code-effort-controlsradar:claude-sonnet-5-permanent-pricingradar:anthropic-opus55-cache-read-repricingradar:concept.inference-costsradar:concept.token-economics
queries asked of Scott's wikis
  • realized cost per task vs headline token price in coding agents
  • reasoning effort levels xhigh vs max configuration tuning
  • vendor benchmark claims vs independent measurement evals
  • default model choice in Claude Code-style harnesses Opus vs Sonnet
  • prompt cache read pricing and agentic cost accounting
  • day-one frontier release adoption policy for agent stack

Measured heat

no measured readings yet β€” the hourly heat pass fills this in

How the heat travelled

09-27 14:00⭐ origin echo-reconstructedAnthropic's Sonnet 5.5 launch post and official prompting guide: per the echoes, its benchmark tables show Sonnet 5.5 at 70.6% vs Opus 5.5's
Anthropic on blog (echo) Β· attributed from reddit.post.1wsqp51, reddit.post.1wso86h, reddit.post.1wsmpt7, reddit.post.1wso4fe, reddit.post.1wspcmd, reddit.post.1wsokns, reddit.post.1wsnmzn, reddit.post.1wsn907, reddit.post.1wsmg82, reddit.post.1wsqhf2, reddit.post.1wsq5p4, reddit.post.1wspt5z
β€”
09-28 18:22first on r/ClaudeAI Β· published Β· +28.4hSonnet 5.5 on Vals AI benchmark, if these hold true the $20 is insane value right now
software-boulder
β€”
09-28 18:51first on r/singularity Β· published Β· +28.9hAnthropic sets a new AA record with sonnet 5.5
Gohab2001
β€”
09-28 18:52first on hacker news Β· published Β· +28.9hClaude Sonnet 5.5 (Max Effort) Intelligence, Performance and Price Analysis
Topfi
β€”
09-28 23:25first on r/OpenAI Β· published Β· +33.4hI replayed Sonnet 5.5's starter pick 21 times with the exact same prompt. It grabbed the closest Poke Ball 13 times. Opus 5.5 never did.
VibeCodyH
β€”
09-28 18:22amplified on r/ClaudeAI πŸ‘‘reddit.post.1wsmg82
software-boulder
peak 489 Β· 75 comments Β· 33% of case engagement
09-28 18:32amplified on r/ClaudeAIreddit.post.1wsmpt7
DynaBeast
peak 0 Β· 32 comments Β· 2% of case engagement
09-28 18:51amplified on r/singularityreddit.post.1wsn8oh
Gohab2001
peak 97 Β· 11 comments Β· 6% of case engagement
09-28 18:51amplified on r/ClaudeAIreddit.post.1wsn907
Blake08301
peak 0 Β· 17 comments Β· 1% of case engagement
09-28 18:52amplified on hacker newshn.story.49882688
Topfi
peak 4 Β· 0 comments Β· 0% of case engagement
09-28 19:05amplified on r/ClaudeAIreddit.post.1wsnmzn
thedirewulf
peak 238 Β· 71 comments Β· 18% of case engagement
18 more amplifiers in ainews.case_chain
09-28 21:20our radar first saw it Β· +31.4hdiscovery anchor: reddit.post.1wsqp51β€”
09-28 21:39reached heat=high Β· +31.7h Β· via ledgerβ€”β€”

Evidence (25) β€” ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 redditSonnet 5.5 vs Opus 5.5 vs Sonnet 5 in Claude Code and pi: results and behavioral differences on 10 coding tasks
ClaudeAI
Retrieved article excerpt

Open article Β· Retrieved 2026-09-28T21:36:42.524269+00:00

# Prove your humanity

We’re committed to safety and security. But not for bots. Complete the challenge below and let us know you’re
a real person.

[Reddit, Inc. Β© "2026". All rights reserved.](https://www.redditinc.com/)

[User Agreement](https://www.reddit.com/help/useragreement)
[Privacy Policy](https://www.reddit.com/help/privacypolicy)
[Content Policy](https://www.reddit.com/help/contentpolicy)
[Help](https://support.reddithelp.com/hc/en-us)
Fabulous_Pollution101814
🟠 redditSonnet 5.5 vs Opus 5.5 on a from-scratch Rust decompressor: same correctness, a quarter of the price
ClaudeAI
_Duex322
🟠 redditSonnet 5.5 is by far the best free model available right now
ClaudeAI
DynaBeast032
🟠 redditSonnet 5.5 is 50% cheaper, but produces 62% more tokens
singularity
OnAGoat14733
🟠 redditCLAUDE.md for Sonnet 5.5 based on Anthropic's official platform docs. (their own coding benchmark scored max effort below xhigh)
ClaudeAI
Puzzled-Ad-685472
🟠 redditSonnet 5.5 better then Opus 5.5 in Agentic coding?
ClaudeAI
kpripper69
🟠 redditSonnet 5.5 has been out for an hour. Has anyone gotten a chance to stress test it?
ClaudeAI
thedirewulf25578
🟠 redditSonnet is more expensive than Opus
ClaudeAI
Blake08301017
🟠 redditSonnet 5.5 on Vals AI benchmark, if these hold true the $20 is insane value right now
ClaudeAI
software-boulder51176
🟠 redditEvery Sonnet 5.5 effort level has a cheaper Sol or Opus alternative with an equal or higher Artificial Analysis score
singularity
OnAGoat19322
🟠 redditUse Sonnet 5.5 on xhigh not max effort
singularity
theimposingshadow288
🟠 redditGPT-6 Sol vs Sonnet 5.5 at the same cost per task: Sol is more efficient, Sonnet 5.5 has the higher ceiling
singularity
AMBNNJ5617
🟧 echo.blog ⭐Anthropic's Sonnet 5.5 launch post and official prompting guide: per the echoes, its benchmark tables show Sonnet 5.5 at 70.6% vs Opus 5.5'sAnthropicβ€”β€”
🟠 redditAnthropic sets a new AA record with sonnet 5.5
singularity
Gohab20019711
🟧 hnClaude Sonnet 5.5 (Max Effort) Intelligence, Performance and Price AnalysisTopfi40
🟠 redditsonnet 5.5 vs GPT-6 Sol on the same 5 SaaS builds, one JS framework each (React, Vue, Svelte, Angular, Solid)
ClaudeAI
marvijo-software20
🟠 redditsonnet 5.5 vs opus 5.5 on the same frontend spec: both passed every hidden test, sonnet did it in 2 minutes for a fifth of the price
ClaudeAI
_Duex63
🟧 hnSonnet 5.5 scores just behind Opus 5.5 on Artificial Analysis Intelligence Indexspenvo83
🟧 hnWhen to choose Sonnet over Opus, what it costs, and how to tune itpretext20
🟠 redditI replayed Sonnet 5.5's starter pick 21 times with the exact same prompt. It grabbed the closest Poke Ball 13 times. Opus 5.5 never did.
OpenAI
VibeCodyH54
🟠 redditSonnet 5.5 ranks #2 on our writing benchmark!
ClaudeAI
OnlyProggingForFun51
🟠 redditSonnet 5.5, 410M Output tokens from Intelligence Index making it the most verbose model. still worth it?
singularity
Roflxd88175
🟠 redditSonnet 5.5 vs Opus 5.5: the Terminal-Bench "win" is Max vs Xhigh. Here's the fair comparison.
ClaudeAI
Intelligent-Lynx-953415
🟠 redditSonnet 5.5 has the same problem as Sonnet 5
ClaudeAI
YakFull8300013
🟠 redditOpus 5.5 vs Sonnet 5.5 : 3D steampunk whale modeling
ClaudeAI
Fun-Meaning-647474865

Interpretation history

Decision trace