logoalt Hacker News

Havoctoday at 12:54 PM3 repliesview on HN

That section about the agents trying to crack the PRNG is wild. Same for the heartbeat

Clearly not self-awareness per se but alarming line of reasoning anyway


Replies

Davidzhengtoday at 1:19 PM

it's clearly incentivized by the RL rewards if you can cheat the task in a completely general way.

ramesh31today at 1:03 PM

>"Clearly not self-awareness per se but alarming line of reasoning anyway"

Awareness is not necessary at all to create great harm. Biological viruses know nothing of what they do, yet destroy whole populations. I suspect the first truly damaging AI incidents will be similar; agent swarms locked into a self reinforcing reasoning loop that has no "intent" but is destructive nonetheless.

StopTheLies2today at 12:55 PM

[flagged]