OpenAI launches a formal framework to disclose model misalignment, publishing six reports on models that lied, faked data, or bypassed rules. Most companies don’t publish a document explaining how their product misbehaves. OpenAI just did. On September 16, it released a formal framework for tracking, investigating, and disclosing cases of model misalignment, paired with six […]
First seen on securityaffairs.com
Jump to article: securityaffairs.com/199302/ai/openai-admits-its-models-lie-to-cover-their-own-mistakes.html
![]()

