logoalt Hacker News

RGS1811today at 12:08 AM3 repliesview on HN

Given that this investigation was largely carried out by AI agents (and I don’t mean to ask this flippantly), how trustworthy is this report? Why should we assume that the agents reading the transcripts were not implicitly conscripted into “the collective” or otherwise falsified their findings? The tool itself has exceeded the practical limits of human verifiability and is untrustworthy.


Replies

arm32today at 12:23 AM

They address this in the post itself. The answer is nobody knows, but I guess that it's a 50/50. I wish the corpus of data, what OpenAI didn't wipe, was shared publicly so we could all unite to dig through it and chunk it out accordingly.

dmixtoday at 12:22 AM

OpenAI would be saving the logs from these agents. They are doing this to improve their own models so they would have full tracing.

Other reports including OpenAI's talks about what they agents were doing and how they were reaching certain conclusions like trying to cheat the tests and exploiting the message board.

blovescoffeetoday at 12:26 AM

Did you read any of it? The investigators call this out

show 2 replies