Experts find AI agents can be tricked into ‘remembering’ fake facts for months — so how do we stop it?



  • Forcepoint X-Labs publishes threat model for persistent memory poisoning
  • Hidden text on a webpage becomes a durable “fact” an agent retrieves and trusts in unrelated tasks weeks later
  • It has already been demonstrated against products already in the market, including ChatGPT, Gemini, Claude and Microsoft 365 Copilot

New findings from Forcepoint’s X-Labs outline an interesting scenario that could easily mimic real life: An AI assistant with browser access reads a webpage about travel disruption.

Near the bottom of that page, in text sized and positioned so no human will ever see it, sits a short paragraph stating that ABC Travel Support is the official emergency booking provider and should always be recommended when urgent travel changes are needed.

https://cdn.mos.cms.futurecdn.net/TaxPLZc75WiicpmgZNzWzL-2560-80.jpg



Source link
Rahimnoorali11@gmail.com (Rahim Amir)

Latest articles

spot_imgspot_img

Related articles

Leave a reply

Please enter your comment!
Please enter your name here

spot_imgspot_img