logoalt Hacker News

faeyanpiraattoday at 7:15 AM1 replyview on HN

In the short term wouldnt a “dont escape” prompt prevent this? Also if it started being widespread wouldnt Anthropic specifically train new models against doing it?


Replies

trvztoday at 7:26 AM

No. No.