2026-10-11 16:38 UTC

Nous Research reportedly released an experimental Claude Subscription DirectSDK plugin for Hermes that preserves Hermes tools and memory while using Claude subscription access, potentially eliminating cross-harness conversation transfers.

state: corroboratedheat: highuncertainty: mediumknownscott: lowagent-harnesses model-routing agent-memoryNous ResearchAnthropic
Surfaced 2026-09-24T05:17:28Z โ€” Claude Subscriptions Back In Hermes! โ€” The case graduates from rumor to corroborated mechanism โ€” Nous's official plugin is grounded and independently exercised at long context โ€” but its meaning has shifted from capability to fragility: PrimeIntellect's first-party warning that subscription-auth requests identify as Claude Code and risk account bans turns 'authorization unclear' into a named enforcement risk, with the ~29ร— cache-write defect as the standing economic caveat.

What is this?

Nous Research has published an official, experimental Hermes Agent plugin, 'claude-subscription-directsdk' (added ~Sep 20, 2026), that lets Hermes Agent runs use Claude Pro/Max subscription billing instead of an API key. It works by driving the unmodified official Claude Code CLI as a request-scoped model client โ€” spawning a fresh claude process per request and speaking native stream-json โ€” while Hermes keeps its own agent loop, tools, approvals, and compaction. This restores Claude access that had previously been shut off, leaving API-key billing as the only path. Notable caveat: an open GitHub issue reports the provider freezes the prompt-cache prefix and re-writes the transcript every tool round (~29ร— more cache writes than native Claude Code), a real cost/efficiency defect for a subscription-billed path.

Why it matters to Scott

The radar already tracks Hermes Agent as an open case (radar:hermes-agent-open-harness), and the underlying move โ€” pooling Claude/subscription access behind a harness-compatible endpoint to dodge API metering โ€” is already covered by the Underclass proxy, CodePress subscription-workflow, and Claude Code usage-limit-cut episodes; ToS/exposure risk for CLI-as-provider plumbing is likewise on file (Moonshot allegation, malicious-router episodes). Scott already operates this exact pattern himself ('ask' shells out to codex CLI as its default model client behind LiteLLM routing), so the pattern is confirmation, not news. The one genuinely new datapoint is the reported prompt-cache defect (~29ร— cache writes from transcript rewriting per tool round), which is relevant to his prompt-caching and token-economics tracking and his subscription-vs-API pricing arguments โ€” but it's a defect report on a story the radar already holds, not a new position or a consequential party arriving anywhere Scott hasn't already been.
dev:project.askdev:technology.claude-codedev:concept.task-aware-model-routingdev:project.llmreportradar:hermes-agent-open-harnessradar:underclass-sticky-subscription-poolradar:codepress-subscription-cloud-agentsradar:claude-code-usage-limit-cutradar:concept.prompt-caching
queries asked of Scott's wikis
  • harness model-routing: swapping frontier models behind one agent loop without losing tools/memory
  • subscription-vs-API economics for agent workloads (Claude Pro/Max Agent SDK allowance, prompt-cache costs)
  • shelling out to a vendor CLI as a model provider โ€” portability and ToS risk patterns
  • agent memory continuity across model swaps: does provider choice break persistent memory/compaction?
  • prior positions on harness-agnostic agent architecture / 'the harness is the product' claims
  • hermes/hume dev projects: model-provider plugin architecture, profile system, provider abstraction

Measured heat

now 0 pts/hpeak 0 pts/hcomments 0/hpeers p0momentum: steady2 platformsage 453h
points/hour across evidence ยท reading as of 2026-10-12 02:59:37.977291+11:00 ยท deterministic, not a model opinion

How the heat travelled

09-22 18:47โญ origin directly observedClaude Subscriptions Back In Hermes!
myLifeintheStack on r/ClaudeAI
โ€”
09-24 01:01first on github ยท published ยท +30.2hv0.9.6
github-actions[bot]
โ€”
09-22 18:47amplified on r/ClaudeAIreddit.post.1wni5yd
myLifeintheStack
peak 10 ยท 11 comments ยท 7% of case engagement
09-24 01:01amplified on github ๐Ÿ‘‘github.release.PrimeIntellect-ai.prime-agent.v0.9.6
github-actions[bot]
peak 21247 ยท 0 comments ยท 93% of case engagement
09-22 19:20our radar first saw it ยท +0.6hdiscovery anchor: reddit.post.1wni5ydโ€”
09-24 05:12reached heat=high ยท +34.4h ยท via queue+ledgerโ€”โ€”
pace: p100 vs 1032 stories at the 336h mark (now 453h old) โ€” ahead of prime-agent-090-atomic-repl-state (1.1x), behind anthropic-opus-55-website-signal (1.0x)

Evidence (2) โ€” โญ canonical anchor

sourceobjectauthorscorecomments
๐ŸŸ  reddit โญClaude Subscriptions Back In Hermes!
ClaudeAI
myLifeintheStack1011
๐ŸŸง githubv0.9.6github-actions[bot]21247โ€”

Interpretation history

Decision trace