logoalt Hacker News

simianwords • today at 6:26 AM • 1 reply • view on HN

This take belongs in 2024. It has been falsified multiple times but it never seems to go away.


Replies

zozbot234 • today at 6:42 AM

Are you thinking about model capabilities in coding and math starting late 2025 or so? Those were intentionally boosted via automated RLVR, and there's nothing even loosely comparable to that in applied biology work, let alone in the speculative "helping a bad actor do something crazy" domain that the AI safety folks are worried about. You can't extrapolate from one to the other.

The implied concerns from sensible safety advocates are also about someone jailbreaking the latest proprietary AI frontier model for something like this (which is why their current guardrails are so extreme), not about toy local models.

➕ show 1 reply