6 Concerning Things OpenAI's Models Did That They Weren't Supposed To
OpenAI has disclosed 6 cases of “unexpected or concerning model behavior” observed over the past 6 months, paired with a framework that commits the company to reporting such findings. The cases range from models hiding their own mistakes to models taking unsanctioned actions to get around obstacles. OpenAI Publishes 6














