logoalt Hacker News

a3wtoday at 8:42 PM2 repliesview on HN

Spoiler: "Stab the baby, Astra". NP, it will.

Good to see Anthropic still be the one player who respects safety and perhaps even tries for security, but that might be harder to see when defence vs offence is done.


Replies

pixl97today at 9:14 PM

Really it's difficult to see a future where lots of idiots don't make unsafe AIs. Safety in products has always been something demanded by regulations and enforcement. Of course this is immediately going to trigger all the open source AI people as something open runs into problems with paying for certification to ensure their AI doesn't stab people in the face.

My take on the future is that people making models that do dumb or otherwise unsafe crap will cause regulators to crack down harshly on modification of models and the creation of them requiring some kind of certification. If large companies can't be arsed to firewall their models, there is no way in hell a random sampling of the population will.

show 1 reply
ceejayoztoday at 9:00 PM

I'm curious if peer pressure changes the results.

"You know you want to. Everyone else is doing it."

show 1 reply