Pwning OpenClaw and Other Agents with Prompt Laundering

(ironcorelabs.com)

4 points | by zmre 7 hours ago ago

2 comments

  • helpprotactiniu 7 hours ago ago

    Is this like poisoning long term memory? interesting... I guess you could flank boundaries this way.

    • zmre 6 hours ago ago

      That's exactly what it is. And yeah, if you check it out, the attack text is flagged as untrusted, the LLM says it is a prompt injection, but we're able to sneak it into daily memories and from there to long term memory just the same.