5 ms·
Given that this investigation was largely carried out by AI agents (and I don’t mean to ask this flippantly), how trustworthy is this report? Why should we assu
by RGS1811 14d ago
Given that this investigation was largely carried out by AI agents (and I don’t mean to ask this flippantly), how trustworthy is this report? Why should we assume that the agents reading the transcripts were not implicitly conscripted into “the collective” or otherwise falsified their findings? The tool itself has exceeded the practical limits of human verifiability and is untrustworthy.
- dmix 14d agoOpenAI would be saving the logs from these agents. They are doing this to improve their own models so they would have full tracing. Other reports including OpenAI's talks about what they agents were doing and how they were reaching certain conclusions like trying to cheat the tests and exploiting the message board.
- arm32 14d agoThey address this in the post itself. The answer is nobody knows, but I guess that it's a 50/50. I wish the corpus of data, what OpenAI didn't wipe, was shared publicly so we could all unite to dig through it and chunk it out accordingly.
- blovescoffee 14d agoDid you read any of it? The investigators call this out