Is there anything -- any possible scrap of evidence whatsoever -- that would convince you that this is not merely a marketing scheme?
This is becoming an idée fixe among the HN crowd. Seemingly nothing can dislodge it, no matter how alarming the incident.
GPT-6 could grab the nuclear launch codes tomorrow and there would be a top-voted comment chuckling that it's all some scheme to pump up the IPO.
---
Put another way, how would you have done the write-up about one of these breakout incidents, if you were in an Anthropic/OpenAI employee's shoes, and (by hypothesis) your intent were not "marketing"? And in a way that doesn't trigger the "it's all marketing" HN top-ranking comment?
This is published on a marketing website.
If it were not a marketing scheme, they would responsibly disclose the vulnerabilities to the code owners, and go on with their lives.
I would dedicate a portion of my organization to making O.S. tools that protect against and contain AI models
I always find it bizarre how rational thought goes out the window whenever AI is involved in HN. There's gotta be something in the water...
Here is one piece of evidence that would convince me: they admit they can't contain it, the they erase the weights and dismantle the company.