model-security
band: coolmomentum: stable
score: 0.032
Episodes (5)
Trajectory notes
- 2026-10-01T07:43:34Z: openai-kimi-distillation-disruption closed (absorbed) β Converges hard on his own dev canon: OpenAI's cross-user replay of *encrypted* reasoning and its warning that portable/replayable reasoning artifacts are an attack surface independently validate the exact rule in dev:conce
- 2026-09-08T15:36:15Z: proprietary-api-reasoning-trace-extraction closed (faded) β The radar already tracks this development in radar:proprietary-llm-reasoning-trace-extraction; this title-only evidence adds no methods, affected providers, or validated results. It touches Scottβs Verification Boundar
- 2026-09-02T19:31:14Z: trusted-lora-subspace-poisoning-defense closed (faded) β The proposed geometric restriction independently echoes Scottβs βcanβt beats shouldnβtβ approach by constraining what fine-tuning can change rather than relying only on post-training behaviour. It could extend the securit
- 2026-08-23T18:30:54Z: steganeur-llm-covert-channel closed (faded) β If independently validated, Steganeur would expose a concrete way for untrusted models or agents to smuggle data through apparently benign text, extending Scottβs SiloOS and padded-cell threat model to the outbound language channel