security1 distinct publisher
Agent memory leaks: benchmark finds up to 69% of user attributes disclosed in the wrong context
A benchmark called CIMemories reports frontier models pushing sensitive attributes into tasks that do not need them, with GPT-5 violations climbing from 0.1% at one task to 9.6% across 40.
Publishers:schneier.com
Reality
- Evidence46
- Adoption
- Insufficient
- Hype gap+14
- Incentives52
- Confidence44