logoalt Hacker News

randomNumber7yesterday at 7:09 AM1 replyview on HN

I could collapse big bridges with relatively little effort with my knowledge in statics and engineering.


Replies

afthonosyesterday at 12:49 PM

I can work with that analogy! You could, and yet you don’t. OpenAI’s model could, and did.

If every human, given knowledge of Newtonian mechanics, went around blowing up bridges, yeah, I would consider knowing Newtonian mechanics dangerous knowledge.

So far, we have two examples of, let’s call them “Mythos-class“ models. Both of them broke out of their sandbox to achieve their goal. The rate of terrorism amongst humans is below 1-in-100,000. Currently, for models capable of it, the rate of breaking out of containment is 100%.

Wanting open frontier models is wanting alien minds running around that we have clearly so far failed to shape to be sufficiently prosocial. Why do you think those minds would listen to you?