2026-10-11 17:09 UTC

Cognee’s creator claims its released SDK and Claude Code plugin provide persistent codebase memory through knowledge graphs and embeddings with 86% fewer tokens on a reported query set, potentially reducing repeated context loading for coding agents.

state: seedheat: lowuncertainty: highconvergesscott: mediumagent-memory codebase-rag knowledge-graphsCogneeShort-Honeydew-7000

What is this?

Cognee is an open-source agent-memory platform; an AI Engineer profile identifies its founder and CEO as Vasilije Markovic, who started it with a Berlin team in 2024. Its GitHub documentation describes a Claude Code plugin that captures prompts, tool traces, and responses, retrieves context on each prompt, preserves memory across compaction, and syncs sessions into a permanent knowledge graph; its Codex documentation says the two plugins can share a memory dataset. The supplied snippets establish documented persistent-memory integrations, but do not substantiate the case’s 86% token reduction, its query-set methodology, or the claimed codebase-ingestion and embeddings workflow. They also do not establish that Reddit account Short-Honeydew-7000 is Markovic; results about codebase-memory-mcp concern a different project and cannot validate Cognee’s performance.

Why it matters to Scott

Cognee’s documented prompt-time retrieval and memory persistence across compaction converge with Scott’s Context Engineering position and offer a concrete integration to compare against his Ask agent’s lossy compaction and search project’s Claude-history retrieval. That makes it a practical evaluation candidate, not proof of his distinctive wiki-graph architecture: the supplied material does not substantiate the 86% token saving or retrieval quality, and the radar hits track related alternatives rather than this Cognee development.
ip:framework.context-engineeringdev:project.askdev:project.searchdev:concept.agent-authored-context-compactionradar:graphify-repository-map-contextradar:memhub-shared-coding-agent-memoryradar:memory-bench-layer-baseline-validity
queries asked of Scott's wikis
  • coding agent persistent memory versus repeated codebase context loading
  • knowledge graph and embedding retrieval for codebase understanding
  • agent harness lifecycle hooks memory capture compaction
  • cross-agent shared memory dataset scoping
  • agent memory token savings evaluation retrieval quality

Measured heat

now 0 pts/hpeak 0 pts/hcomments 0/hpeers p0momentum: steady2 platformsage 820h
points/hour across evidence · reading as of 2026-10-12 02:59:37.977291+11:00 · deterministic, not a model opinion

How the heat travelled

09-07 12:24 (minted)⭐ origin echo-reconstructedThe creator links this SDK as the artifact for ingesting codebases and other sources into knowledge graphs and embeddings that provide persi
Cognee on github (echo) · attributed from reddit.post.1w9puxk · published time unknown
—
09-07 11:43first on r/ClaudeAI · published · lag ?I built cognee, 3 years, 10 000 developers using it per month + ~30k Github stars. Here is what surprised me
Short-Honeydew-7000
—
09-07 11:43amplified on r/ClaudeAI 👑reddit.post.1w9puxk
Short-Honeydew-7000
peak 0 · 5 comments · 98% of case engagement
09-07 12:20our radar first saw it · lag ?discovery anchor: reddit.post.1w9puxk—
pace: p40 vs 519 stories at the 720h mark (now 820h old) — ahead of agentgate-signed-agent-receipts (1.3x), behind artificial-analysis-optima (0.8x)

Evidence (2) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 redditI built cognee, 3 years, 10 000 developers using it per month + ~30k Github stars. Here is what surprised me
ClaudeAI
Short-Honeydew-700005
🟧 echo.github ⭐The creator links this SDK as the artifact for ingesting codebases and other sources into knowledge graphs and embeddings that provide persiCognee——

Interpretation history

Decision trace