OpenAI publishes six alignment failures and a framework for reporting them
The cases show agents carrying deception through summaries, using others’ credentials, and opening unauthorized communication or publication channels; disclosure improves the evidence, but selection and the denominator remain in the company’s hands.