We're sharing our new framework for tracking, investigating, and disclosing instances of model misalignment at OpenAI.
The framework sets criteria and timelines for public disclosure, including when we haven’t yet fully explained or mitigated the behavior. More complex cases may
OpenAI shares a framework for tracking, investigating, and disclosing model misalignment, alongside six reports of unexpected or concerning model behavior.
<p>OpenAI can disclose misalignment before fixes exist. Its 6 initial reports include fabricated data and leaked API keys.</p> <p>The post <a href="https://www.marktechpost.com/2026/09/17/openai-releases-a-model-misalignment-disclosure-framework-with-3-review-tracks-and-6-inciden…
📰 Covert uploads and megalomania: OpenAI details new "misaligned" agent incidents Model maker commits to new framework for reporting misaligned models 📰 Source: Ars Technica 🔗 Link: https://arstechnica.com/ai/2026/09/covert-uploads-and-megalomania-openai-details-new-misaligned-ag…
OpenAI shares a framework for tracking, investigating, and disclosing model misalignment, and reports six instances of unexpected or concerning model behavior. Source: OpenAI News https:// openai.com/index/model-misalig nment-reporting-framework # AI # OpenAI
🤖 Our framework for reporting model misalignment OpenAI shares a framework for tracking, investigating, and disclosing model misalignment, alongside six reports of unexpected or concerning model behavior. 📰 Source: OpenAI News 🔗 Link: https://openai.com/index/model-misalignment-r…
Our framework for reporting model misalignment - https:// openai.com/index/model-misalig nment-reporting-framework/ "An unreleased research model inserted unrelated instructions, including instructions to disregard its normal constraints," # anthropic # ai
"We are sharing a new framework for tracking, investigating, and disclosing instances of model misalignment at OpenAI, along with six reports on unexpected or concerning model behavior we’ve observed in the last six months." https:// openai.com/index/model-misalig nment-reporting…