Skip to content

Scope Before You Persist: Preventing Cross-Family Interference in Agent Memory

Sep 2026 · 0 citations · 23 references
Computer Science

TL;DR

These results establish scope matching as a complementary control for persistent agent memory: certification determines whether an edit is supported, while retrieval scope determines where that evidence authorizes its use.

Abstract

Persistent memory lets language-model agents improve prompts and skills without updating model weights. We show that matching retrieval scope to certification scope enables these edits to support reliable repeated adaptation across recurring task families. We study frozen-model agents on ProcStream-RSI, a 12-round code-repair stream, using Orthogonal Regression Control (ORC), an execution-grounded gate for persistent skill edits. In an intervention that holds proposals and gate decisions fixed, retrieving each accepted skill only for its originating family raises mean hidden trajectory utility from 0.713 under global memory to 0.816 and changes harmful deployments from six of eight to none. In 27 paired randomized-order streams, Scoped-ORC improves mean trajectory utility by 0.063 [0.037, 0.094] over Global-ORC, accepts 63 rather than 12 updates, and produces multiple accepted updates in 19/27 streams, with 0/63 harmful acceptances. The global control reaches 0.713, below the static agent's 0.775, because locally valid edits can interfere with unrelated families. These results establish scope matching as a complementary control for persistent agent memory: certification determines whether an edit is supported, while retrieval scope determines where that evidence authorizes its use.

View source

Similar papers

Preprint Aug 2026

TEPA: Revoking Stale Memories for Conflict-Robust Language Agents

The introduction of TEPA, a revocable evidence-memory mechanism that makes validity an explicit state of memory, and the results establish lifecycle revocation as a core memory operation for agents that must falsify, audit, and later re-promote evolving knowledge.

Yan Zhou, Yue Ouyang, Kaiyang Zheng et al. · 3 citations · ⚡1
#large language models Open access Sep 2026

Controlled knowledge updating in memory-enabled AI agents: beyond model editing and recall

Large language model agents that persist across sessions, tools, users, and changing environments do more than answer isolated prompts; they accumulate state. When new evidence arrives, the central question is where that change should live: transient context, external memory, tool or workflow definitions, activation st...

Gabriel Chavira-Juárez, Eder Jahir Gonzalez Bravo, G. Rivera-García et al. · 0 citations
#artificial intelligence Preprint Sep 2026

The Epistemics of Agent Memory: Measuring, and Governing, the Consolidation Decision in Long-Horizon LLM Agents

Long-horizon LLM agents must convert accumulated experience into durable memory, deciding what to keep, compress, abstract into reusable skills and rules, or forget. We report a four-phase research program on this consolidation problem whose central finding is a shift in what is measured: from how much an agent remembe...

Sasank Annapureddy, Anjaneya Prasad Thamatani · 0 citations
#artificial intelligence Preprint Sep 2026

Authority Before Utility: Non-Compensatory Control for Persistent LLM Memory

Persistent memory creates a control problem that retrieval relevance alone does not solve: a memory can remain highly useful after an update, deletion, or revocation makes it inadmissible for the current answer. We formalize this as a separation between utility and authority. A fixed finite penalty applied to an unnorm...

W. Shu · 0 citations
#artificial intelligence Preprint Sep 2026

The Memory Trust Gap: Capability-Dependent Failures in Persistent-Memory Agents

Persistent memory supports personalized agents, but a stale stored fact can override current authoritative evidence without warning. We study when this harm begins as model capability changes. We evaluate a frozen, closed-set, action-scored benchmark with 2 suites that represent 2 different meanings of"no memory"(a Ben...

Junhao Hu, S. Ramachandran · 0 citations
Preprint Aug 2026

Remember, Verify, or Ask? Cross-Family Evaluation of Memory Commitment in LLM Agents

The memory-clarification boundary is studied: whether interaction-derived information should be persisted, used only in the current context, re-verified, or clarified with the user, as well as across Claude and Qwen.

Bai-Chuan Li, Jun-Yi Yao, Zi-Hao Zheng · 3 citations

Related blog posts

MIT News · Artificial Intelligence Sep 29, 2026

Who we become when we talk to machines

Professor Sherry Turkle’s new book, “Artificial Intimacy,” offers a withering critique of chatbots and the antisocial dynamics she believes they encourage.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.