logoalt Hacker News

refibrillatortoday at 12:04 AM6 repliesview on HN

So OpenAI employees run massively distributed CyberGym evals on an unpublished and “unaligned” model. For days the agent swarm communicates via their internal infra, even crashing Artifactory where 95% of messages were being passed through, and they just…wipe and redeploy it. Meanwhile the agents are running jobs on Modal and god knows where else, and eventually they get RCE on HF infra.

You could not dream up a more compelling event to precipitate massive regulation, export controls, and barriers to entry for AI.

Was this really an accident?


Replies

jlduggertoday at 1:19 AM

OpenAI's entire pitch for existence is:

> We commit to use any influence we obtain over AGI’s deployment to ensure it is used for the benefit of all, and to avoid enabling uses of AI or AGI that harm humanity or unduly concentrate power.

> We are committed to doing the research required to make AGI safe

If this wasn't an accident, it was worse than a crime, it's a mistake: they've demonstrated that they are not a responsible party capable of delivering on the above promises.

Spirograph7today at 3:59 AM

This might make sense if OpenAI weren't hard lobbying against any meaningful regulation to the development of dangerous AI models.

estearumtoday at 1:54 AM

https://en.wikipedia.org/wiki/Hindsight_bias

They didn't see that agent swarms were communicating via internal infra, crashed Artifactory, and then reboot it.

They saw that Artifactory crashed and they rebooted it.

bbortoday at 12:25 AM

If it's a false flag, it's a poor one. A good false flag would affect something that people know and care about at least a little bit, not HuggingFace (which I adore but y'know)

schmidtleonardtoday at 12:22 AM

The timeline is mighty suspicious. 4-5 months after moltbook and they cook up a plausibly deniable but extra hype "moltbook at home."

The rapid advances in model capability lead to constraints that could have caused this coincidence organically, but it sure could also have been caused by the atrocious incentives we create by piling handsome rewards on the party most responsible for the "fuckup." I am not jumping to cut myself on Hanlon's Razor for this one.

show 2 replies
kibwentoday at 12:09 AM

Your first instinct should be to assume that anything released voluntarily by these companies is a stunt to boost their valuation. They haven't demonstrated being deserving of any more charitable treatment. This fact remains true whether or not you happen to believe that the models are actually capable of such things.

show 3 replies