OpenAI Publishes Six Cases of Its Models Behaving Unexpectedly
OpenAI has rolled out a new framework to track, investigate, and publicly share cases where its AI models do not behave as intended. With this framework, the company published six reports on unexpected or concerning model behaviour during training and testing over the past six months. These cases cover everything from models hiding their own errors to making unauthorized moves. Previously, OpenAI admitted it did not have a regular process for releasing this kind of information. Sometimes, they would hold onto findings until they had enough to put into a single report. Other times, the details ended up in documentation […]













