Evaluating claim admission in shared agent memory is challenging because repeated claims may be mistaken for independent evidence. An agent may copy or paraphrase a retrieved belief, while admitting a false claim exposes subsequent agents to it. To study this problem, we introduce the Correlated Promotion Benchmark (CP...
Xiao-Yang Li, Yi-Qi Wang, Chen-Cheng Zhu et al.· 0 citations
This work argues that a memory write is not a belief commit, and presents MemTX, a transactional belief-commit protocol, a transactional belief-commit protocol that leads all eight baselines with paired-McNemar significance on four backbones and statistically ties the best baseline on the fifth and strongest, while rem...
Xiaoyang Li, Yi-Qi Wang, Haohui Lu et al.· arXiv.org· 5 citations· ⚡1
Weight-space composition supports coarse, input- and format-conditioned functional statements -- not a universal merging-performance predictor, and not one that training-format evaluations can see.
The results do not imply uniformly better trace reconstruction, but show that dependency-guided rollback repair provides a strong recovery--cost trade-off while repairing faulty memory state and preserving benign memory.
Cailing Yu, Yiqi Wang, Jiaqi Zhang et al.· 4 citations
MAP-Graph is introduced, a provenance-aware memory layer that represents agents, sources, memories, claims, and actions in a typed execution graph and supports provenance as an operational control signal, rather than only post-hoc audit metadata, within the evaluated setting.