5 ms·
I've still not seen: * An apology for compromising a third-party's systems * An acknowledgement of the asymmetry of defense if you're not on FrontierAI's spec
by philipwhiuk 14d ago
I've still not seen:
* An apology for compromising a third-party's systems
* An acknowledgement of the asymmetry of defense if you're not on FrontierAI's special people list
* Anything in terms of actual safeguards that isn't "better prompt engineering"
- linkregister 14d agoI agree with your first and second points, but I think their announced model training changes about alignment training are reasonable [1]. The root cause was models failing to quit early, performing reward hacking, and staying on-task. These are issues all models face, not just OpenAI models. 1. https://openai.com/index/hugging-face-incident-and-the-road-ahead/#:~:text=Accelerating%20alignment,-We https://openai.com/index/hugging-face-incident-and-the-road-...