logoalt Hacker News

burningChromeyesterday at 5:54 PM8 repliesview on HN

The lack of details to me means this was an intentional marketing ploy to try and demonstrate the power of their models to show their technology can compete with the likes of Anthropic and DeepMind.

They created an experiment they knew would generate the outcome they wanted. It would be the similar to what say car companies do to over hype their cars. "This EV can go over 800 miles on a single charge!" And then at the bottom you see all the disclaimers: "Must be on flat ground, with no headwind, with a spare battery in the back seat, with no extra weight added."

Same thing here. Everybody in infosec is calling this out as a marketing stunt and nothing else for a litany of reasons. I'd say look up MG (creator of the OMG cable) on twitter, he has some interesting insights on this one.


Replies

lelanthranyesterday at 6:46 PM

> The lack of details to me means this was an intentional marketing ploy to try and demonstrate the power of their models to show their technology can compete with the likes of Anthropic and DeepMind.

I dunno; Check my posting history, I'm as skeptical of AI companies' claims as anyone, but in this case your theory doesn't explain why:

1. OpenAI guardrails refused to let the target use OpenAI's models to defend against this.

2. Huggingface used GLM (I think) so that they could defend without guardrails.

If this was an intentional marketing ploy, it was marketing for GLM, not for OpenAI nor for Huggingface.

Hence, I don't think it was intentional.

show 2 replies
JoshTriplettyesterday at 7:07 PM

> The lack of details to me means this was an intentional marketing ploy to try and demonstrate the power of their models to show their technology can compete with the likes of Anthropic and DeepMind.

"our model is horribly misaligned and used security exploits to break out of our sandbox and into another company, without being prompted to do so" is not positive marketing.

This is an actual critical problem, not a stunt. We're going to see more of this, and it's going to get much worse.

show 2 replies
rwmjyesterday at 6:17 PM

It's also possible their sandbox was videcoded crap and the AI (which had the guardrails intentionally removed) escaped. This was a oops, but OpenAI turned this into a PR opportunity. They turned lemons into lemonade.

If your AI is really that dangerous you don't need a sandbox at all, you should airgap it from any network.

simonwyesterday at 10:25 PM

> The lack of details to me means this was an intentional marketing ploy

OpenAI said this:

> We are sharing preliminary findings at this stage to help defenders understand what happened and to help calibrate on what models are now capable of. We will continue to conduct a thorough investigation alongside Hugging Face and will share more details on the vulnerabilities, incident, and findings when our investigation is complete.

I suggest giving them a few more days before saying that the lack of detail is proof that this is a "marketing ploy"

hawk_yesterday at 7:05 PM

Concluding this was intentional feels a bit of a stretch. But once it happened, yeah the spin masters got to work and coordinated to turn this into +PR.

jackb4040yesterday at 6:39 PM

> similar to what say car companies do

Another applicable metaphor I've seen floating around is weapons companies testing out a new bomb.

We know the AI labs don't care about negative vs positive public sentiment, and only care that investors see their tech as powerful. The only difference in PR strategy from a weapons company is the latter doesn't care if they get protested.

Zababayesterday at 8:37 PM

>The lack of details to me means this was an intentional marketing ploy to try and demonstrate the power of their models to show their technology can compete with the likes of Anthropic and DeepMind.

DeepMind hasn't been on the frontier for a while, their current best model is behind Anthropic, OpenAI, Moonshot (Kimi k3), xAI (Grok 4.5), Z.AI (GLM 5.2), and even Meta (muse spark). Gemini 3.6 is behind GLM 5.2, released a month earlier, open weights and cheaper.

You can paint the OpenAI story as a way to try to appear as dangerous as Anthropic with all the Mythos stuff.

poloticsyesterday at 6:02 PM

mmh, i think it's "not uphill" (means downhill) "no headwind" (...)