7 ms·
Holy shit. This wasn't "intentional" this was just openai letting their testing run wild.
by yRetsyM 2mo ago
Holy shit. This wasn't "intentional" this was just openai letting their testing run wild.
- ibejoeb 2mo agoThey're not just letting it run wild. They took precautions to exercise it in an isolated environment. It managed to evade the constraints.
- paxys 2mo agoKinda like how they responsibly contained that one dinosaur in Jurassic world.
- ibejoeb 2mo agoUnderstood that containment failed. But I don't think there's value in characterizing it as throwing all caution to the wind. Let's discuss how the containment failed and how to mitigate it.
- ameliaquining 2mo agoThe root cause of the containment failure, in the deepest sense, was that their next-generation model was better at offensive security than their humans and current-generation models were at defensive security. That problem's only going to get worse if they keep training more and more capable models.