2026-10-11 17:12 UTC

Harden claims its post-trained cybersecurity small model combined with inline reference monitoring outperforms GPT-5.5-xhigh on LinuxArena and SleightBench, offering coding-agent defenses that do not depend solely on the frontier agent model.

state: expiredheat: lowuncertainty: highconvergesscott: highagentic-security coding-agents program-analysisHardenOpenAI

What is this?

Harden reports that its post-trained cybersecurity small language model, paired with inline reference monitoring, outperformed GPT-5.5-xhigh on the LinuxArena and SleightBench coding-agent security benchmarks. The proposed architecture moves defenses into a specialized model-and-monitor harness rather than relying solely on the frontier coding model, with local execution also positioned as useful for proprietary-code and regulated environments. The supplied results support the broader claims that security harnesses can materially improve model performance and that local small models have privacy advantages, but they do not independently verify Harden’s benchmark results or explain Harden’s organizational background.

Why it matters to Scott

Harden’s self-reported benchmark results directly converge with Scott’s load-bearing claims that the meaningful unit is model plus harness and that consequential security should be enforced outside the frontier model. If independently reproduced, a specialized local SLM plus inline reference monitor outperforming GPT-5.5-xhigh creates a strong dated-receipts and implementation opportunity for Architecture, Not Vibes and SiloOS; the supplied evidence does not independently verify the result.
ip:concept.model-plus-harness-benchmark-unitip:framework.architecture-not-vibesip:framework.siloosip:concept.runtime-governancedev:concept.deterministic-agent-control-planeradar:stencil-harness-coding-improvementradar:locus-ast-agent-firewallradar:concept.coding-agent-securityradar:concept.security-benchmarks
queries asked of Scott's wikis
  • coding-agent reference monitors and policy enforcement
  • security harnesses versus model-native safety
  • small local models for proprietary code security
  • program analysis integrated with coding agents
  • defense-in-depth architectures for autonomous agents
  • benchmarking agent security harnesses

Measured heat

no measured readings yet — the hourly heat pass fills this in

How the heat travelled

no chain yet — the hourly chain pass fills this in

Evidence (2) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟧 hnShow HN: Beating GPT5.5-xhigh for Coding agent security with SLMs and IRMse4u93
🟧 echo.blog ⭐This is Harden’s original research/evidence post. It reports Harden’s own evaluations of its cybersecurity SLM plus inline reference-monitorHarden——

Interpretation history

Decision trace