2026-10-11 16:37 UTC

Independent evaluations will determine whether LLMs that persistently write and retrieve their own notes achieve durable reasoning gains over ordinary prompting without prohibitive memory or contamination costs.

state: corroboratedheat: lowuncertainty: highnovelscott: mediumagent-memory reasoning self-reflection

What is this?

The supplied web snippets are entirely off-topic (a British newspaper, an autopsy report, Instagram posts) and contain no material on LLM self-generated notes, reasoning gains, or reusable-strategy extraction from solution traces. Based only on the case's own evidence titles, the claim appears to be about research where an LLM writes notes/strategies from its own solution traces and retrieves them at inference to improve reasoning — but no web grounding actually confirms this, who built it, or when. This must be flagged as ungrounded: the search term 'independent evaluations' was mistaken for the newspaper 'The Independent' and returned no relevant material.

Why it matters to Scott

This directly touches Scott's territory of agent memory and agent-maintained wikis — self-generated notes/strategies retrieved at inference is structurally the same pattern as agent memory systems he builds — but no wiki_hits or radar_hits were supplied to confirm an existing position, so no concrete intersection can be named. If Scott has a documented stance on durable reasoning gains from self-authored memory, this would bear on it directly, but that can't be established from empty hits.
queries asked of Scott's wikis
  • agent memory writing own notes reusable strategies
  • self-reflection loops for reasoning improvement
  • retrieval of past solution traces at inference time
  • contamination risk in self-generated training data
  • memory cost vs reasoning gain tradeoffs in agents
  • wiki-as-memory pattern for coding agents

Measured heat

now 0 pts/hpeak 0 pts/hcomments 0/hpeers p14momentum: steady3 platformsage 1970h
points/hour across evidence · reading as of 2026-10-12 02:59:37.977291+11:00 · deterministic, not a model opinion

How the heat travelled

07-21 14:00⭐ origin echo-reconstructedThe original paper studies extracting reusable strategies and cautions from an LLM’s own solution traces, then retrieving them at inference
Chang Liu, Xinyu Li, Artur Dubrawski on paper (echo) · attributed from hn.story.49035916
—
07-24 14:09first on hacker news · published · +72.2hLLMs can write themselves notes to get better at reasoning
MarcoDewey
—
08-02 23:18first on r/LocalLLaMA · published · +297.3hMaking DS4 0731 self-iterate to write its own memory prompts
dangerous_inference
—
08-03 14:55first on r/ClaudeAI · published · +312.9hCompaction is Hot Garbage
Popular_Sand2773
—
08-16 02:32first on r/singularity · published · +612.5hAI Isn’t Outthinking Mathematicians. It’s Out-Remembering Them.
yogthos
—
08-20 03:21first on r/artificial · published · +709.4hI described my messy AI memory setup on one sub. Eighteen strangers replied describing almost the same architecture, independently.
__hymn
—
07-24 14:09amplified on hacker newshn.story.49035916
MarcoDewey
peak 1 · 0 comments · 0% of case engagement
07-26 11:28amplified on hacker newshn.story.49057003
krishnakantk876
peak 1 · 0 comments · 0% of case engagement
07-27 18:47amplified on hacker newshn.story.49073979
krishnakantk876
peak 1 · 0 comments · 0% of case engagement
07-28 15:27amplified on hacker newshn.story.49085375
johnnymakes
peak 1 · 1 comments · 0% of case engagement
07-29 12:43amplified on hacker newshn.story.49096767
arisAlexis
peak 27 · 11 comments · 2% of case engagement
07-31 13:13amplified on hacker newshn.story.49122715
souravroy78
peak 1 · 0 comments · 0% of case engagement
41 more amplifiers in ainews.case_chain
07-24 14:20our radar first saw it · +72.3hdiscovery anchor: hn.story.49035916—

Evidence (48) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟧 hnLLMs can write themselves notes to get better at reasoningMarcoDewey10
🟧 echo.paper ⭐The original paper studies extracting reusable strategies and cautions from an LLM’s own solution traces, then retrieving them at inference Chang Liu, Xinyu Li, Artur Dubrawski——
🟧 hnWe achieved 94.7% in Longmemeval with no hacks lol, everything open sourcedkrishnakantk87610
🟧 hnWe achieved 94.7% in Longmemeval built in a hackathon, everything open sourcedkrishnakantk87610
🟧 hnSOTA on the hardest AI memory benchmark (BEAM, 10M tokens), with a smaller modeljohnnymakes11
🟧 hnShow HN: Echologue – the private AI voice journal I built for myselfarisAlexis2711
🟧 hnShow HN: Provenance and decay for AI agent memorysouravroy7810
🟠 redditMaking DS4 0731 self-iterate to write its own memory prompts
LocalLLaMA
dangerous_inference10
🟧 hnTeaching an AI to know itself: Building a local LLM agent in Dteleforce10
🟠 redditCompaction is Hot Garbage
ClaudeAI
Popular_Sand2773620
🟧 hnShow HN: We gave a voice agent an inner monologue so it stops forgettingharshit11910
🟧 hnZero-Mem: Zero-Token Memory Operations for LLM Agentstheanonymousone10013
🟠 redditClaude's thinking
ClaudeAI
Firm_Run_815402
🟠 redditDoes asking an AI (like Claude Code) to list and analyze its own mistakes make it start making more errors?
ClaudeAI
JadedImpress7455315
🟧 hnWhy Does Claude.md Keep Growing? Catastrophic Remembering in Agentic Codingsbulaev10
🟧 hnWhy Does Claude.md Keep Growing? Catastrophic Remembering in Agentic Codingoskrim10
🟠 redditSick of repeating yourself to Opus 5? It isn't ignoring you. It has amnesia.
ClaudeAI
styleforge-io19
🟠 redditThe best integration between my phone and Claude Code turned out to be plain files
ClaudeAI
Top-Primary944704
🟧 hnThe Commons – experiments in inherited knowledge and errors between LLM agentscoladul10
🟧 hnAI Isn't Outthinking Mathematicians. It's Out-Remembering Themrzk632518
🟧 hnAugmenting Long-term Memory (2018)re-framer30
🟠 redditAI Isn’t Outthinking Mathematicians. It’s Out-Remembering Them.
singularity
yogthos881278
🟧 hnProgrammatic memory for long-horizon LLM agentsrzk10
🟠 redditAgent in a Room - autonome ai
ClaudeAI
SubjectNo298507
🟧 hnShow HN: Agents Workbook watch Claude Code, Codex write down their working notespradeep117743
🟠 redditI'm 50, not an engineer, and I've spent 8 months building a persistent "AI family" on top of Claude. The trick wasn't prompts — it was a filing system.
ClaudeAI
__hymn027
🟧 hnShow HN: Seahorse – an agent's memory that lives in your own notesssanvi_builds10
🟧 hnAn LLM wiki changed how I workwertyk20
🟠 redditTwo local models as one companion: one speaks, one only measures. State keeps evolving between messages on a 10-minute heartbeat. Whitepaper + MIT code
LocalLLaMA
Caitsters10
🟠 redditI described my messy AI memory setup on one sub. Eighteen strangers replied describing almost the same architecture, independently.
artificial
__hymn018
🟠 redditBuilt a Pokémon card discovery site with Claude Code — feedback welcome
ClaudeAI
fullytorqued24016
🟠 redditWhat I changed after 286 tasks: the memory file needs an editorial policy, not more content
ClaudeAI
StopZestyclose914705
🟠 redditI built a proactive memory system for coding agents and tested whether it actually helps.
ClaudeAI
DreadlockEug29
🟧 hnShow HN: Screen memory without screenshots, just text to MarkdownDramatize6025
🟠 redditI LIKE OPUS 5
ClaudeAI
PlusImage305007
🟠 redditHow I got my Mac to read my Claude Code chats at night and extend my token usage by 1/3rd
ClaudeAI
Unable_Strategy513516143
🟠 redditJust dropped a plugin that Claude Code "learns" from mistakes.
ClaudeAI
ArgusArkim25
🟠 redditClaude Code forgets every skill it ever learned the second the session ends
ClaudeAI
Ordinary-War475502
🟠 redditI ran memory accuracy tests on small models, here's what I found
artificial
Excellent-Fan845733
🟧 hnA frozen LLM with external memory found a novel 26 circle packing structureDarenWatson10
🟧 hnWikiSkill: Compiling Agent Experience into Persistent KB for Skill Evolutionomarsar20
🟠 redditMy CLAUDE.md is mostly scar tissue. Every rule in it is a production incident
ClaudeAI
Top_Commission_856709
🟧 hnShow HN: Give your agent somewhere to think loud watch its decisions unfold livepradeep117711
🟠 redditI Give My Hermes Agents a Haircut Every Week
ClaudeAI
myLifeintheStack010
🟧 hnI gave my AI agent a sleep cycle – it dreams about its errors and fixes themopenamer10
🟠 redditOne thing I started documenting that helped more than expected: rejected approaches
ClaudeAI
Pretend_Sell65921010
🟠 redditSelf-learning memory with local LLMs
LocalLLaMA
KitchenAmoeba443801
🟧 hnNever-again – so your AI agent stops repeating mistakes you fixedmalaysherasia-a22

Interpretation history

Decision trace