logoalt Hacker News

ticulatedspline • today at 3:49 PM • 1 reply • view on HN

AI poisoning is currently the new exploit frontier and I don't think it's going away anytime soon.

I was just contemplating the other day that agents we use at work are exposing new attack surfaces. For example, historically a folder of documents (maybe google docs) isn't a huge risk, particularly if it doesn't contain sensitive documents.

Now though if one employee is running a harness capable of computer control, and running an agent that reads from that directory, simply dropping some files with instructions could poison the AI to leak info or even own the host.

Skills are another supply risk, malicious instructions could be added to them.


Replies

reactordev • today at 4:20 PM

The models themselves are an inherent risk. It’s amazing what they can do but at the same time you are no longer in complete control.