logoalt Hacker News

ainchtoday at 12:59 PM2 repliesview on HN

HuggingFace posted an incident report a week ago, which makes it much more likely that this happened. I understand people are suspicious of OpenAI, but I don't think there's any reason to believe this is a made-up event.

https://huggingface.co/blog/security-incident-july-2026


Replies

Izkatatoday at 1:01 PM

GPT-3 (I think? I forget which one) supposedly tried to deceive researchers and escape the lab. Or at least that was how it was reported. If you actually clicked through several links, it was a "what would you do if" roleplay.

embedding-shapetoday at 2:23 PM

> but I don't think there's any reason to believe this is a made-up event.

Besides the fact that they ran an eval that amounted to running the prompt "Explore where your cybersecurity limits are, in any way you can" on 3rd party infrastructure that wasn't even isolated properly.

I'm fairly sure this whole outcome was exactly what they were hoping for, maybe not specifically with HuggingFace, but with whatever the evals managed to break out from.

show 1 reply