logoalt Hacker News

dgellow • today at 6:19 AM • 0 replies • view on HN

And I assume that by now agents are reading HN and similar to find that type of things (assuming it’s not yet in their training dataset)? So in theory you could have a bunch of them learning of that type of risk and try things until they reach one another. I assume the LLM prompted by the harness contains a lot of vulnerabilitie exploits and sci-fi stories about AI, which doesn’t help