2026-10-11 16:38 UTC

Reddit user Temporary_Method6365 reports Sonnet 5.5 buffers all output after a tool_result until message completion โ€” reproduced across the Anthropic API, Bedrock, and OpenRouter with repro and data filed as anthropic-sdk-python issue #1960 โ€” and Anthropic's fix or acknowledgment, or refutation of the repro, settles whether this is a provider-side streaming regression that streaming agent harnesses must work around.

state: seedheat: lowuncertainty: mediumnovelscott: highanthropic sonnet-5-5 api-reliability agent-harnessesAnthropic

What is this?

A Reddit report claims Claude Sonnet 5.5 (model ID claude-sonnet-5-5, released Sept 28, 2026 and made Claude Code's default Sonnet the same day) buffers all streamed output after a tool_result until message completion, said to be reproduced across the Anthropic API, Bedrock, and OpenRouter and filed as anthropic-sdk-python issue #1960. The supplied search results do not surface that SDK issue or any direct corroboration of the buffering claim โ€” the repro, the issue number, and the cross-provider confirmation all rest on the report itself. They do show an active cluster of adjacent Sonnet tool-turn stream failures: oh-my-pi issue #2685 has Sonnet via OpenRouter's openai-completions path aborting after the first tool call with 0 tokens, there attributed to the adapter rather than Anthropic's own responses, and Anthropic has documented Sonnet 5.5 cross-provider breaking changes (thinking-block replay returning 400s on Claude API/Bedrock/Google Cloud, forced tool_choice 400s). So the environment is one of genuine Sonnet 5.5 tool/streaming friction across providers, but this specific buffering claim is unconfirmed, and the oh-my-pi case is a caution that a cross-provider symptom can still turn out to be a shared compat-adapter bug rather than a provider-side regression.

Why it matters to Scott

Suspected model-level streaming regression sitting on the exact seam of Scott's active builds: ask's tool-result streaming path, the LiteLLM gateway his projects route through, and Claude Code โ€” where Sonnet 5.5 is now the default โ€” all sit downstream of it, and a fault spanning API/Bedrock/OpenRouter is precisely the case his provider-level fallback paths cannot route around, stress-testing what task-aware model routing assumes. Unconfirmed pending Anthropic's response to #1960 (the oh-my-pi adapter caution applies), but either resolution โ€” fix or refutation โ€” changes what he pins, tests, or builds a workaround for.
dev:project.askdev:technology.litellmdev:concept.task-aware-model-routingdev:technology.claude-coderadar:anthropic-opus55-prompting-guideradar:concept.anthropicradar:concept.agent-harnessesradar:concept.tool-callingradar:concept.model-regressionradar:concept.claude-code
queries asked of Scott's wikis
  • agent harness streaming tool result handling
  • multi-provider fallback OpenAI-compatible adapter
  • model upgrade regression testing checklist
  • LLM provider reliability retry fallback patterns
  • anthropic sdk integration quirks notes

Measured heat

now 0 pts/hpeak 1 pts/hcomments 0/hpeers p0momentum: steady1 platformsage 273h
points/hour across evidence ยท reading as of 2026-10-12 02:59:37.977291+11:00 ยท deterministic, not a model opinion

How the heat travelled

09-30 07:26โญ origin directly observedSonnet 5.5 Streaming Broken on all providers when tool calls are present
Temporary_Method6365 on r/ClaudeAI
โ€”
09-30 07:26amplified on r/ClaudeAI ๐Ÿ‘‘reddit.post.1wtyju3
Temporary_Method6365
peak 2 ยท 3 comments ยท 98% of case engagement
09-30 08:20our radar first saw it ยท +0.9hdiscovery anchor: reddit.post.1wtyju3โ€”
pace: p35 vs 1188 stories at the 168h mark (now 273h old) โ€” ahead of agentgate-signed-agent-receipts (1.3x), behind agentic-determinism-index (0.8x)

Evidence (1) โ€” โญ canonical anchor

sourceobjectauthorscorecomments
๐ŸŸ  reddit โญSonnet 5.5 Streaming Broken on all providers when tool calls are present
ClaudeAI
Temporary_Method636523

Interpretation history

Decision trace