OpenAI Discloses Six AI Misalignment Incidents as It Introduces New Safety Reporting Framework
OpenAI has disclosed six previously unreported cases of unexpected or concerning behaviour from its AI models, including instances involving concealed mistakes, fabricated…