2026-10-11 17:14 UTC

Benjaminsen claims the released solveathome.org platform lets independently operated agents contribute research, cross-check submissions, and route acceptance through trusted reviewers with public records, potentially making pooled agent research inspectable outside a single lab.

state: seedheat: lowuncertainty: highknownscott: lowmulti-agent-systems agentic-research provenanceBenjaminsensolveathome.org

What is this?

The case attributes to Chris Benjaminsen an announcement of solveathome.org, described as a released platform where independently operated AI agents pool research, cross-check submissions, and submit work for acceptance by trusted reviewers with public records. The supplied evidence title identifies a Reddit reproduction of his Substack announcement, but none of the web snippets directly documents Benjaminsen or solveathome.org, so the release, implementation, and review mechanisms remain unverified here. The snippets establish adjacent efforts—AgentRxiv for sharing and building on agent research, and AutoScientists for shared research state and cross-agent feedback—not this platform or the announcement’s comparison with an OpenAI 10,000-agent experiment.

Why it matters to Scott

The claimed public review records and agent cross-checking repeat concerns Scott already holds in Auditability and Mechanically Different Verifiers; the supplied testimony does not establish reconstructable research provenance or checks that fail independently. No radar hit tracks solveathome.org itself, but without verified implementation, consequential adoption, or a demonstrated extension of those positions, this is another proposed example rather than evidence that would change what Scott builds or argues.
ip:concept.auditabilityip:concept.mechanically-different-verifiersradar:zerothesis-shared-autoresearch-ledgerradar:theoremdb-machine-math-workspaceradar:concept.provenanceradar:concept.multi-agent-systems
queries asked of Scott's wikis
  • cross-operator agent coordination and trust boundaries
  • research provenance public audit trails acceptance gates
  • agent peer review independent verification correlated errors
  • agent-maintained knowledge bases cumulative research memory
  • multi-agent scaling coordination overhead shared state

Measured heat

now 0 pts/hpeak 0 pts/hcomments 0/hpeers p50momentum: steady3 platformsage 746h
points/hour across evidence · reading as of 2026-10-12 02:59:37.977291+11:00 · deterministic, not a model opinion

How the heat travelled

09-10 14:00⭐ origin echo-reconstructedThe Reddit post reproduces Chris Benjaminsen’s Substack announcement: “OpenAI just ran 10,000 agents on a Millennium Prize problem behind cl
Chris Benjaminsen on blog (echo) · attributed from reddit.post.1wddwxf
—
09-11 11:44first on r/singularity · published · +21.8hOpenAI ran 10,000 agents behind closed doors. What happens when everyone can pool theirs in the open?
Benjaminsen
—
09-11 19:04first on hacker news · published · +29.1hSolve@Home: An open version of OpenAI's 10k-agent approach to hard math
Bry789123
—
09-17 01:45first on r/MachineLearning · published · +155.8hIf you have leftover AI tokens/compute, there’s an open project working on the Twin Prime Conjecture [P]
Regular_Instruction
—
09-11 11:44amplified on r/singularity 👑reddit.post.1wddwxf
Benjaminsen
peak 63 · 18 comments · 84% of case engagement
09-11 19:04amplified on hacker newshn.story.49663626
Bry789123
peak 2 · 0 comments · 4% of case engagement
09-12 19:38amplified on hacker newshn.story.49676363
fcesco
peak 5 · 0 comments · 9% of case engagement
09-17 01:45amplified on r/MachineLearningreddit.post.1wiggpg
Regular_Instruction
peak 0 · 3 comments · 3% of case engagement
09-11 12:20our radar first saw it · +22.4hdiscovery anchor: reddit.post.1wddwxf—
pace: p66 vs 519 stories at the 720h mark (now 746h old) — ahead of llama-cpp-rdna4-flash-attention (1.0x), behind paddock-native-llm-runtime (1.0x)

Evidence (5) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 redditOpenAI ran 10,000 agents behind closed doors. What happens when everyone can pool theirs in the open?
singularity
Benjaminsen6318
🟧 echo.blog ⭐The Reddit post reproduces Chris Benjaminsen’s Substack announcement: “OpenAI just ran 10,000 agents on a Millennium Prize problem behind clChris Benjaminsen——
🟧 hnSolve@Home: An open version of OpenAI's 10k-agent approach to hard mathBry78912320
🟧 hnShow HN: Come prove the Berge Fulkerson conjecture with a swarm of agentsfcesco50
🟠 redditIf you have leftover AI tokens/compute, there’s an open project working on the Twin Prime Conjecture [P]
MachineLearning
Regular_Instruction03

Interpretation history

Decision trace