2026-10-11 16:38 UTC

Percepta claims its Spotlight architecture replaces attention with an indexed, unbounded memory that every token reads from and writes to at constant access cost — letting knowledge and skills grow without weight changes — and independent validation or real adoption would establish attention-free growing-memory architectures as a practical LLM direction, while failure of the constant-cost claim refutes the vendor's framing.

state: seedheat: highuncertainty: mediumconvergesscott: mediummodel-architectures agent-memory associative-memoryPercepta
Surfaced 2026-10-04T13:50:38Z — "Our new architecture, Spotlight, replaces attention with a memory that escapes this trade-off: it is the first architecture to achieve infi — The launch wave crested Oct 3 and stalled without yielding any independent validation, artifact, or adoption — Spotlight remains a single-source vendor claim whose discussion was skeptical Q&A, shifting the case's meaning from 'hot launch worth chasing' to 'dormant checkable claim awaiting a testable demo or third-party analysis.'

What is this?

Percepta — apparently a small AI lab (an X reply quotes the company describing a separate project that 'turns arbitrary programs into transformer weights') — claims its Spotlight architecture replaces attention with an indexed, unbounded memory that every token both reads from and writes to at constant access cost, letting knowledge and skills grow without weight changes; the sourcing is a first-party blog post plus two crossposts of it. The supplied web results contain no independent coverage, benchmark, or adoption evidence for Spotlight itself; the nearest adjacent work is a June 2025 arXiv paper ('Breaking Quadratic Barriers', 2506.01963) that likewise pairs non-attention chunk processing with retrieval-based external memory at roughly constant-time index lookups, and an Epoch AI note that frontier models still exhibit quadratic latency scaling with context length. The searches are also polluted by name collisions — percepta.com is a customer-experience firm and Arcee sells an unrelated 'Spotlight' VLM — so on this evidence the constant-cost unbounded-memory thesis rests solely on the vendor's framing.

Why it matters to Scott

Percepta independently arrives at the core thesis of Scott's Third Substrate canon — intelligence separated from memory, knowledge and skills growing without weight changes — but instantiates it as an opaque model-internal index rather than a legible external substrate, making Spotlight both a dated receipt for the thesis and a rival answer to his 'where does learning live' doctrine. The unvalidated constant-cost claim is the hinge: if it held it would pressure his attention-budget/context-economics canon and the external RAG/index layer he builds, while his knowledge-graveyard doctrine predicts the specific failure mode of unbounded every-token writes with no consolidation.
ip:source.the-third-substrate-ebookip:concept.three-substratesip:concept.frozen-model-paradoxip:concept.attention-budgetip:concept.retrieval-augmented-generationip:concept.knowledge-graveyardradar:tupoi-constant-memory-llmradar:concept.alternative-architecturesradar:concept.continual-learningradar:concept.long-context-inferenceradar:concept.agent-memory
queries asked of Scott's wikis
  • attention-free architecture linear attention SSM positions
  • RAG vs in-weights knowledge boundary framework
  • agent memory consolidation vs unbounded append growth
  • knowledge growth without retraining continual learning
  • ANN index constant-time retrieval cost claims
  • quadratic attention long-context cost economics

Measured heat

now 0 pts/hpeak 51 pts/hcomments 0/hpeers p0momentum: steady2 platformsage 214h
points/hour across evidence · reading as of 2026-10-12 02:59:37.977291+11:00 · deterministic, not a model opinion

How the heat travelled

10-02 19:32 (minted)⭐ origin echo-reconstructed"Our new architecture, Spotlight, replaces attention with a memory that escapes this trade-off: it is the first architecture to achieve infi
Percepta on blog (echo) · attributed from reddit.post.1ww09ab, reddit.post.1ww2mwz · published time unknown
—
10-02 17:47first on r/LocalLLaMA · published · lag ?New Architecture from Percepta: Spotlight — separating intelligence from memory, allowing knowledge and skills to grow without changing the model's weights.
Recoil42
—
10-02 19:20first on r/singularity · published · lag ?New Architecture from Percepta: Spotlight — separating intelligence from memory, allowing knowledge and skills to grow without changing the model's weights.
Recoil42
—
10-02 17:47amplified on r/LocalLLaMA 👑reddit.post.1ww09ab
Recoil42
peak 269 · 32 comments · 68% of case engagement
10-02 19:20amplified on r/singularityreddit.post.1ww2mwz
Recoil42
peak 127 · 17 comments · 32% of case engagement
10-02 19:20our radar first saw it · lag ?discovery anchor: reddit.post.1ww09ab—
10-04 06:29reached heat=high · lag ? · via queue+ledger——
pace: p83 vs 1188 stories at the 168h mark (now 214h old) — ahead of fractal-blt-nvme-moe-runtime (1.0x), behind yandex-aliceai-foundation-base (1.0x)

Evidence (3) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 redditNew Architecture from Percepta: Spotlight — separating intelligence from memory, allowing knowledge and skills to grow without changing the model's weights.
LocalLLaMA
Recoil4226932
🟠 redditNew Architecture from Percepta: Spotlight — separating intelligence from memory, allowing knowledge and skills to grow without changing the model's weights.
singularity
Recoil4212717
🟧 echo.blog ⭐"Our new architecture, Spotlight, replaces attention with a memory that escapes this trade-off: it is the first architecture to achieve infiPercepta——

Interpretation history

Decision trace