logoalt Hacker News

glubtoday at 4:39 PM1 replyview on HN

So what's the fuss about then? Is the idea that a Chinese company will release a model that will have no safeguards? For what purpose?

Basic safeguards are all that's required, and they've been there in every usable model since GPT-2, including Chinese models that are supposedly "unsafe".

Or are we saying that some lunatics will start training their own models, spin up a GPU cluster, run some abliteration workflow, or learn how to jailbreak?

That would be a very dedicated person. And dedicated person doesn't need an LLM. So where are they?


Replies

GeneralMayhemtoday at 5:07 PM

Yes, that is exactly Dario's concern. Either one of the US labs or one of the Chinese ones will eventually release something with insufficient safety controls for its power level because it gives them slightly better user retention (look how much complaining there is about current frontier models, especially Fable, rejecting requests). Regulation or consortium is how you avoid the prisoner's dilemma.

show 1 reply