Skip to content

Forgetting Is Not a Fix: Path Dependence in Sequential Engram Editing

Jul 2026 · arXiv.org · Vol abs/2607.24805 · 1 citation · 7 references
Computer Science Biology

Abstract

AI Engram (Kwon et al., 2026) formalizes the four engram criteria of neuroscience as a constrained inverse problem in weight space and solves it closed-form: concept-specific memory traces become linear objects that can be extracted once and combined arithmetically. Appendix F states the Compositional Memory States Hypothesis: edited models live on"a commutative manifold where the integration of A and B reaches a consistent equilibrium regardless of the learning sequence."The evidence base is single and paired edits -- in materials terms, single-cycle tests, in which fatigue accumulation is structurally invisible. Whether the hypothesis holds under sequential load is exactly the"temporal dynamics"question the paper defers to future work. We run that test on the authors'own reference implementation, at their reported best edit strength (TOFU alpha=0.6, a choice favoring the linearity hypothesis), with pre-registered predictions, across three model charges (two vendors, two architecture families). Four findings replicate across all three: (1) zero-shot composition and sequential re-calibrated editing diverge by 61-71% of the edit magnitude; (2) cut order is not interchangeable, and the effect scales with concept overlap -- in one charge the order of cutting two Paris landmarks decides whether an uninvolved third concept survives; (3) the survivors'layer-input covariances -- the method's own sufficient statistics, read as strain gauges -- drift monotonically with every further cut, in every surviving concept, in every charge; (4) erased knowledge partially returns under subsequent unrelated cuts. Appendix F's commutative-manifold hypothesis is thereby falsified for sequential editing; the single-edit results of the original paper are untouched. For unlearning-as-compliance: erasure certified today does not certify the artifact after its next edit.

View source

Similar papers

#machine learning Preprint Sep 2026

Pre-carved Niches: The Formation Dynamics of Modular Task Partitions in Early LLM Training

This work tracks formation step by step of a Pythia-410M model from scratch and runs attribution patching at every step, alongside probes for gradient norms, effective updates, weight norms, and first-order loss decomposition across 14 tasks in four cognitive domains, confirming the hypothesis that modularity tracks le...

Guang-Qi Li, Yongxin Li · 0 citations
Jul 2026

The Art of Not Forgetting

We introduce CMP (Cognitive Memory Primitive), an architecture that represents inputs as sparse relational codes, stores them in a two-tier competitive memory, and learns entirely through local, gradient-free updates, with no backpropagation anywhere in the network. We use this architecture to test a specific hypothesi...

Ashmith Atmuri, Akshay Kumar, Yashaswini Rao Bhogarajula · 0 citations
#machine learning Preprint Aug 2026

Kathleen Remembers: Length-Invariant One-Shot Recall Without Attention

This work adds to the Kathleen trunk a second memory layer -- a"notebook": a fixed-key holographic (HRR) associative store with a learned local write gate, a self-gating raw read, and write-triggered forgetting -- 25K parameters that attach to the logits of any trunk.

George Fountzoulas · 0 citations
Book Open access Aug 2026

TTMC: Brain-Inspired Test-Time Memory Calibration with Orthogonal Projection for Online Continual Learning

Test-Time Memory Calibration (TTMC), a novel gradient-free analytic framework that introduces a transductive calibration mechanism that seamlessly fuses the second-order statistics of the unlabelled test stream into the accumulated long-term memory via a closed-form solution, allowing for real-time alignment with the t...

Yuyang Han, Zi-Yu Li, Diwei Su et al. · 0 citations
Jul 2026

The Art of Not Forgetting A Local Learning Architecture for Continual Learning

The results suggest that the combination of sparse representations, local learning, and persistent memory is a promising direction for continual learning, while motivating further investigation into the respective roles of learning rules, representations, and architectural design in mitigating catastrophic forgetting.

Ashmith Atmuri, Yashaswini Rao Bhogarajula · 0 citations
Preprint Aug 2026

When Is a Steerable Concept Representation Real? Measurement Confounds in a Cross-Family Audit of Neuroscience Parallels in LLMs

Large language models (LLMs) are increasingly reported to exhibit human-like neural and cognitive signatures, including concept cells, mental number lines, and cognitive maps. These claims often rely on linear probing and activation steering applied to a single model, yet both methods are highly sensitive to measuremen...

Yuqi Wu, Shengming Zhao, Jie Chen · 3 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.