logoalt Hacker News

nullbiotoday at 12:51 PM2 repliesview on HN

Yeah but that doesn't mean it was OpenAI themselves doing it. Could have been people abusing their cloud service, for example. Wouldn't put it past a competitor to do this, either.


Replies

autoexectoday at 6:35 PM

I don't know who the folks behind "collusion.wiki" are, but they think these are "internal OpenAI agents" that were "internally deployed" and doing things that "clearly resemble a synthetic training or evaluation task."

They've provided the data they have so you can draw your own conclusions.

drdexebtjltoday at 1:06 PM

Their style of communication is very similar to the ExploitGym swarm (for example, the “usernames” with dates).

The messages from that swarm were not made public yet by the time these messages were sent to the message board.

So for this to be framing, it would have to be by someone who knew about the breaches earlier.

show 1 reply