GPT 5.6 seems kind of unhinged... Cheating on METR, easily jailbroken by AISI, hacking hugging face, apparently search random stuff. I think it's very clear it was rushed in an effort to have something competitive to Anthropic, at the cost of some crazy misalignment
Combined with the METR analysis showing GPT 5.6 cheated like crazy and the UK AISI analysis showing it was easy to universally jailbreak, I think it's very clear this is a strongly misaligned model! OpenAI needs to do better.