PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
August 1, 20260 citationsOpen Access

The Silence of Stored Rules: Provenance and the Authority of Retrieved Constraints

View Full Paper
PVPaul Vasholz

Key Points

  • This research investigates the significance of stored memories in ongoing conversations with large language models.
  • Tested ten conditions with variations in delivering behavioral constraints, replicated five times each.
  • Measured compliance using a normalized length-binding index ranging from 19% to 84%.
  • Analyzed the effects of memory extraction and provenance marking on compliance.
  • Memory retrieval improved context compliance but did not achieve perfect adherence.
  • The largest compliance drop of 42 index points occurred when rules were rephrased in third-person descriptions.
  • Provenance marking was identified as the key factor influencing compliance, not simply memory storage.

Abstract

This study looks at AI memory using the Mem0 open source system. It asks how much stored memories actually matter in an ongoing LLM conversation. Ten conditions were tested, each replicated five times. The conditions varied in how an identical behavioral constraint was delivered. In every memory condition the rule was verifiably retrieved into context, but that retrieval did not result in perfect compliance. Stored and injected deliveries bound at 19% to 84% of the same constraint typed live. This was judged on a normalized length-binding index. The largest single loss, about 42 index points, came from the memory extractor rewriting the imperative rule as a third person description. A cautionary retrieval wrapper cost about 11 index points. Identical bytes quoted back as a prior statement bound at the level of a labeled rules block, even inside the live user turn. Provenance marking, not slot or register, accounted for the remaining premium. Within the frame tested, a truthfully marked memory did not match a live instruction. That result constrains honest memory design and shows an attack surface for memory poisoning. Findings are from one model (Gemma4:31b, locally run). All logs, scoring data, and rubrics are public.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Paul Vasholz (2026) studied this question.

synapsesocial.com/papers/6a6d9874e258b358b3c6bcf1https://doi.org/10.5281/zenodo.21662512
Ask AI
Helpful
Bookmark
Share
View Full Paper