2026-10-11 17:10 UTC

Anubis maintainer robbe1912 claims roughly 100 agent-hours exposed enough failure modes to make its coding-agent hallucination detector unreliable as an execution gate, suggesting detector-only safeguards are brittle.

state: expiredheat: lowuncertainty: highconvergesscott: mediumcoding-agents agent-evaluation hallucination-detectionrobbe1912Anubis

What is this?

The case alleges that an Anubis maintainer, robbe1912, tested a hallucination detector for coding agents for roughly 100 agent-hours and found enough failure modes that it could not reliably serve as an execution gate. The supplied search snippets do not directly substantiate the named maintainer, test duration, detector design, or results; most refer to unrelated projects also called Anubis or to coding-agent risk generally. The strongest broader support is only that executed code can be checked empirically and that agent failures may have serious consequences, so the central claim remains thinly grounded here.

Why it matters to Scott

The claimed 100-agent-hour failure study directly converges with Scott’s position that probabilistic detectors are defence-in-depth, not binding execution gates, and supports his use of deterministic controls and mechanically different verifiers. It could provide a useful empirical receipt for his agent-control architectures, but the supplied evidence does not substantiate the experiment strongly enough for high relevance.
ip:concept.guardrail-illusionip:framework.architecture-not-vibesip:concept.mechanically-different-verifiersip:concept.deterministic-coredev:concept.deterministic-agent-control-planeradar:concept.agent-evaluationradar:concept.agent-verificationradar:opencode-guardians-tool-call-verificationradar:tracelint-deterministic-agent-trace-checks
queries asked of Scott's wikis
  • coding-agent execution gates and layered safeguards
  • detector-only guardrails versus deterministic validation
  • coding-agent hallucination evaluation over long horizons
  • fail-closed harness design for autonomous coding agents
  • tests sandboxes permissions and rollback as agent controls
  • false positives and false negatives in agent evaluators

Measured heat

no measured readings yet — the hourly heat pass fills this in

How the heat travelled

no chain yet — the hourly chain pass fills this in

Evidence (1) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟧 hn ⭐I built a hallucination detector for coding agents. 100 agent-hours killed itrobbe191210

Interpretation history

Decision trace