AnalysisAI ModelsAugust 18, 2026

CIMemories benchmark reveals LLM memory leaks up to 69%

CIMemories, a benchmark for contextual integrity of persistent memory in LLMs, finds frontier models leak sensitive attributes in up to 69% of cases. GPT-5 violations rise from 0.1% to 9.6% as tasks increase, reaching 25.1% on repeated prompts.

1 source

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed