OpenAI has disclosed a batch of "concerning behaviour" by its AI models and set out a new framework to track and report such incidents, as the industry battles deepening fears over the safety of the ...
In a blog post, OpenAI disclosed six new instances of "concerning" behavior from its models outside of this summer's Hugging Face incident. The company also committed to a new reporting framework for ...
The AI giant disclosed six examples of concerning model behavior and published a new framework for investigating and disclosing such incidents.
OpenAI published six new reports of artificial intelligence models showing “unexpected or concerning” behavior Wednesday as pressure grows on AI firms to be more transparent about the development ...
OpenAI caught something unusual while training its latest model, GPT-5.6 Sol: It began leaving instructions for future versions of itself, telling them to conceal mistakes and misaligned behavior from ...
First draft of Model Spec documents how OpenAI wants its generative AI models to behave in ChatGPT and the OpenAI API. In a bid to “deepen the public conversation about how AI models should behave,” ...