logoalt Hacker News

skeptic_aitoday at 12:11 AM1 replyview on HN

For the big safety guys to only investigate this either means are incompetent or malevolent. Which one?

Tip: the people working there are the top 0.001% smartest in the world


Replies

gck1today at 12:17 AM

They had a model escape in April, roughly the same time when they were fearmongering about Mythos and how Anthropic should be the sole keyholder of cybersecurity capabilities, and it only occured to them to look inside logs when they saw someone else winning in their own game.

What, Anthropic didn't know model could escape sandbox without OpenAI reporting it?

show 1 reply