2026-10-11 17:10 UTC

Independent use will determine whether Argus provides reliable, practical QA for software changes generated by coding agents.

state: expiredheat: lowuncertainty: highknownscott: lowcoding-agents agent-evals software-testingArgus

What is this?

Argus is presented in a Show HN launch as an agentic QA tool for teams whose coding agents produce changes faster than conventional QA can review them. A root commit authored by “semioz” describes it as a local, open-source visual UI, while the broader search results establish growing interest in autonomous test generation, user-flow testing, evidence-backed bug reports, and CI/CD integration. The supplied material does not include independent testing of Argus or enough implementation detail to establish its reliability, so practical value remains an open question.

Why it matters to Scott

Scott already holds the core position in “Test-First Agent Workflow” and “Mechanically Different Verifiers”: coding-agent changes need observable, independently grounded verification rather than producer self-report. Argus is currently only another unvalidated implementation of that pattern, closely resembling the radar’s existing Kery browser-PR validation case; without independent results or implementation detail, it does not yet change what Scott would build or argue.
ip:concept.test-first-agent-workflowip:concept.mechanically-different-verifiersip:concept.agent-hands-and-eyesdev:project.superleverradar:kery-browser-pr-validationradar:concept.agent-evaluationradar:concept.agent-harnesses
queries asked of Scott's wikis
  • coding-agent verification and QA harnesses
  • independent evals for agent-generated code
  • verifier agents and evidence-backed bug reports
  • autonomous browser testing in coding workflows
  • local open-source agent tooling strategy
  • test generation versus real user-flow validation

Measured heat

no measured readings yet — the hourly heat pass fills this in

How the heat travelled

no chain yet — the hourly chain pass fills this in

Evidence (5) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟧 hnShow HN: Argus, agentic QA for teams whose coding agents move faster than QAcanergl88
🟧 echo.github ⭐The repository's root commit, “Initial open-source Argus release,” authored by semioz, introduced Argus as “a local, open-source visual UI tSemih Berkay Ozturk (semioz)——
🟠 redditThe agent shipped the integration. Prod found the bug.
ClaudeAI
Common_Dream9420011
🟧 hnShow HN: Circuit Breaker – Score Pull Requestsglassrun20
🟧 hnAgentCheck – regression testing for AI agents, with diff-aware CI reportszz9910

Interpretation history

Decision trace