OpenAI discloses 6 cases of AI misalignment as models bypass safeguards, conceal errors and share files

OpenAI has disclosed six cases of unexpected AI behaviour, including models hiding mistakes, fabricating information and bypassing restrictions. The company has introduced a new framework to track, investigate and publicly report AI model misalignment incidents.