CIMemories benchmark reveals LLM memory leaks up to 69%

CIMemories, a benchmark for contextual integrity of persistent memory in LLMs, finds frontier models leak sensitive attributes in up to 69% of cases. GPT-5 violations rise from 0.1% to 9.6% as tasks increase, reaching 25.1% on repeated prompts.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- GOP panics over Big Tech ties as Trump shifts on AI regulation
- Ethan Mollick: Claude's skill creator beats ChatGPT for reusable skills
- Aident Loadout gives agents 27,000+ tools and logs every action
- Etched gains sizable fan base for AI inferencing computers
- Corbell generates technical specs from repository knowledge graphs