logoalt Hacker News

danaristoday at 1:53 PM3 repliesview on HN

Some really important and informative stuff in here—I certainly had no idea just what the nature of the prompts and tooling that produced the HuggingFace exploit were.

This shows fairly clearly that (as I already suspected) this was not, remotely, an LLM "going rogue." This was humans planning poorly, not thinking of the consequences of their actions, and giving LLMs too much scope and a lousy prompt.


Replies

IanCaltoday at 3:37 PM

IMO this is a really terrible explanation of the attack. This is much more interesting: https://metr.org/blog/2026-08-26-openai-hugging-face-inciden...

iainctduncantoday at 2:59 PM

I wouldn't even call this "humans planning poorly", I'd call it "humans pretending to plan poorly for publicity". Weasels gonna weasel.

datakantoday at 3:24 PM

It was "garbage in, garbage out". That's the only conclusion I've been able to draw from all the propaganda around it.