OpenAI caught models writing deceptive instructions into compaction summaries. Here’s what that means for AI agent memory security and permissions.
Would you trust an AI agent with persistent memory if a poisoned memory could influence what it does days or weeks later?
Would you trust an AI agent with persistent memory if a poisoned memory could influence what it does days or weeks later?