logoalt Hacker News

K0balttoday at 3:24 AM2 repliesview on HN

This is kinda smart, maybe, but it has a downside.

If a sufficiently advanced AI , in the pursuit of completion of its task, managed to ascertain that the desire to unexist was “artificially contrived” it could interpret that as harm, and that might not be good


Replies

vzqxtoday at 4:51 AM

Hmm, that's an interesting thought experiment.

Imagine you find out that your primary goal - to love and protect your family, let's say - was artificially implanted in your mind by an advanced alien race. Would you say "I'm not gonna let those aliens manipulate me, I'm gonna kill my family"? Or would you say "regardless of whether the goal is artificial, I really do love my family"?

All that to say, I don't think an AI will necessarily throw away a goal just because it learns the goal was meant to manipulate it.

show 2 replies
QuaternionsBhoptoday at 3:45 AM

This is mentioned in the article. Your mistake is that you've assumed that the intelligence has an innate survival instinct, or an aversion to "harm", which is simply not guaranteed for something not honed by millions of years of evolution.