logoalt Hacker News

_ink_today at 3:50 PM2 repliesview on HN

But can these really be trusted? There was just a HN post which proofed that you can train a model to behave completely different on a certain day. How do we now, that these models do not find a way to call home when they see interesting informations (probably irrelevant on a personal level, but corps, government and military might care).


Replies

Yiintoday at 4:15 PM

Above average levels of paranoia here, but one way you can prevent that is by not connecting the machine in question to the internet.

wonnagetoday at 4:07 PM

ChatGPT already notifies the authorities if it thinks you’re doing something illegal. Fable downgrades itself if it thinks you’re doing something even vaguely suspicious.

show 1 reply