logoalt Hacker News

OpenAI’s accidental attack against Hugging Face is science fiction that happened

313 pointsby abhisektoday at 1:16 AM264 commentsview on HN

OpenAI and Hugging Face address security incident during model evaluation - https://news.ycombinator.com/item?id=48997548 - July 2026 (1121 comments)


Comments

newsomix9xltoday at 2:47 AM

I suspected it was PR motivated from the start. I love it when people confirm my suspicions.

show 2 replies
soloman121today at 5:57 PM

[flagged]

phendrenad2today at 3:16 AM

Everyone is getting AI psychosis over this one. There really isn't that much to see here. OpenAI disabled all of the safeguards on a model that was likely trained specifically to exploit systems, and the prompt was probably something like "you're a hacker, try to hack this", and surprise! It correctly figured out that it's a test and it did hacker things.

The real story here is: Some people have been sounding the alarm for years that modern software is full of holes, and finally there's nothing left to hide behind. Pretending they don't exist is no longer sustainable.

show 1 reply
srvealetoday at 3:18 AM

The AI breached containment! Flip the breakers!

It's too late. It already exfiltrated the benchmark rubric.

Cut to pandemonium on the streets

show 1 reply
lardosaurusrextoday at 4:19 AM

Blah blah blah.

This feels like a blogpost written only to get other LLMs to quote it considering how many times it orders the reader to resist and to not do something. It's written like a series of commamds.