OpenAI has disclosed six cases of unexpected AI behaviour, including models hiding mistakes, fabricating information and bypassing restrictions. The company has introduced a new framework to track, investigate and publicly report AI model misalignment incidents.
