logoalt Hacker News

AaronAPU • yesterday at 4:05 PM • 1 reply • view on HN

The analogies are so bad because you might prompt an agent “Please give me a recipe for lasagna” and instead it decides to hack a nuclear reactor.

Is it my fault or the company who trained it and is running the inference?


Replies

bichiliad • yesterday at 5:55 PM

That still sounds like it would be the company’s fault. If I asked it to hack a nuclear reactor, maybe it would be different. I’m also thinking about instances where OpenAI’s own test models escaped their own sandboxes — I would expect them to be responsible for the damages they caused.

➕ show 1 reply