Benjaminsen claims the released solveathome.org platform lets independently operated agents contribute research, cross-check submissions, and route acceptance through trusted reviewers with public records, potentially making pooled agent research inspectable outside a single lab.
state: seedheat: lowuncertainty: highknownscott: lowmulti-agent-systems agentic-research provenanceBenjaminsensolveathome.org
What is this?
The case attributes to Chris Benjaminsen an announcement of solveathome.org, described as a released platform where independently operated AI agents pool research, cross-check submissions, and submit work for acceptance by trusted reviewers with public records. The supplied evidence title identifies a Reddit reproduction of his Substack announcement, but none of the web snippets directly documents Benjaminsen or solveathome.org, so the release, implementation, and review mechanisms remain unverified here. The snippets establish adjacent efforts—AgentRxiv for sharing and building on agent research, and AutoScientists for shared research state and cross-agent feedback—not this platform or the announcement’s comparison with an OpenAI 10,000-agent experiment.
Why it matters to Scott
The claimed public review records and agent cross-checking repeat concerns Scott already holds in Auditability and Mechanically Different Verifiers; the supplied testimony does not establish reconstructable research provenance or checks that fail independently. No radar hit tracks solveathome.org itself, but without verified implementation, consequential adoption, or a demonstrated extension of those positions, this is another proposed example rather than evidence that would change what Scott builds or argues.
ip:concept.auditabilityip:concept.mechanically-different-verifiersradar:zerothesis-shared-autoresearch-ledgerradar:theoremdb-machine-math-workspaceradar:concept.provenanceradar:concept.multi-agent-systems
queries asked of Scott's wikis
- cross-operator agent coordination and trust boundaries
- research provenance public audit trails acceptance gates
- agent peer review independent verification correlated errors
- agent-maintained knowledge bases cumulative research memory
- multi-agent scaling coordination overhead shared state
Measured heat
now 0 pts/hpeak 0 pts/hcomments 0/hpeers p50momentum: steady3 platformsage 746h
points/hour across evidence · reading as of 2026-10-12 02:59:37.977291+11:00 · deterministic, not a model opinion
How the heat travelled
pace: p66 vs 519 stories at the 720h mark (now 746h old) — ahead of llama-cpp-rdna4-flash-attention (1.0x), behind paddock-native-llm-runtime (1.0x)
Evidence (5) — ⭐ canonical anchor
Interpretation history
2026-09-17T07:28:52Z
A second Reddit author now claims to have contributed tokens, adding a weak participation signal beyond the creator’s announcement. Without submission records or review outcomes, this does not corroborate the claimed inspectable, cross-operator research workflow.
2026-09-17T07:22:35Z
evidence attached: reddit.post.1wiggpg — This independently points users toward solveathome.org's public, verifiable multi-agent research workflow, corroborating the open case.
2026-09-12T20:28:03Z
The newly attached HN title describes a swarm effort on a different conjecture, but supplies neither a demonstrated connection to Solve@Home nor implementation results. It is adjacent interest, not corroboration of this platform’s cross-operator review or auditability claims.
2026-09-12T20:21:51Z
evidence attached: hn.story.49676363 — This is a concrete public multi-agent research effort that bears on whether pooled agent research can produce inspectable mathematical work.
2026-09-11T19:36:04Z
The HN listing adds cross-community visibility, not independent corroboration of the platform’s operation or review mechanisms. This remains a creator-announced experiment rather than demonstrated inspectable, cross-operator research.
2026-09-11T19:22:13Z
evidence attached: hn.story.49663626 — This is direct corroborating coverage of Solve@Home’s open multi-agent research platform for hard mathematics.
2026-09-11T12:27:10Z
grounded: known/low — The claimed public review records and agent cross-checking repeat concerns Scott already holds in Auditability and Mechanically Different Verifiers; the supplie
2026-09-11T12:24:27Z
origin walked (codex/luna, conf 0.99): anchor reddit.post.1wddwxf -> echo.blog.99ce62cd69 by Chris Benjaminsen
2026-09-11T12:23:31Z
case created — The creator links a concrete public research platform whose coordination and review workflow is a distinct episode from OpenAI's mathematics claim.
Decision trace
- 10-09 22:51review_dormant28 days without material information; scheduled checks stopped
- 10-04 17:26drop_targetsquiet through full ladder or over cap 8
- 10-04 01:00drop_targetsquiet through full ladder or over cap 8
- 09-20 18:26review_screenThe new comments ask about contributor benefits and whether smaller local models can participate, but provide no implementation result, artifact, accepted research output, or evidence that changes the
- 09-20 18:20sensor_dirtycomment_update
- 09-17 17:28repriceA second Reddit author now claims to have contributed tokens, adding a weak participation signal beyond the creator’s announcement. Without submission records or review outcomes, this does not corrobo
- 09-17 17:22attachThis independently points users toward solveathome.org's public, verifiable multi-agent research workflow, corroborating the open case.
- 09-17 17:22propose_attachThis independently points users toward solveathome.org's public, verifiable multi-agent research workflow, corroborating the open case.
- 09-13 06:28repriceThe newly attached HN title describes a swarm effort on a different conjecture, but supplies neither a demonstrated connection to Solve@Home nor implementation results. It is adjacent interest, not co
- 09-13 06:21attachThis is a concrete public multi-agent research effort that bears on whether pooled agent research can produce inspectable mathematical work.
- 09-13 06:21propose_attachThis is a concrete public multi-agent research effort that bears on whether pooled agent research can produce inspectable mathematical work.
- 09-12 05:36repriceThe HN listing adds cross-community visibility, not independent corroboration of the platform’s operation or review mechanisms. This remains a creator-announced experiment rather than demonstrated ins
- 09-12 05:22attachThis is direct corroborating coverage of Solve@Home’s open multi-agent research platform for hard mathematics.
- 09-12 05:22propose_attachThis is direct corroborating coverage of Solve@Home’s open multi-agent research platform for hard mathematics.
- 09-12 03:38review_screenThe change adds a contextual link and generic discussion but no new implementation result, independent verification, contradiction, or consequential access detail beyond the already noted public platf
- 09-12 02:20sensor_dirtycomment_update
- 09-11 22:27groundThe claimed public review records and agent cross-checking repeat concerns Scott already holds in Auditability and Mechanically Different Verifiers; the supplied testimony does not establish reconstru
- 09-11 22:24promote_anchororigin walk conf 0.99
- 09-11 22:23createThe creator links a concrete public research platform whose coordination and review workflow is a distinct episode from OpenAI's mathematics claim.