OpenAI benchmaxes GPT-5.6 to match Anthropic scores, voluntarily accepts government restriction
OpenAI intentionally tuned GPT-5.6 to match Anthropic's Mythos preview score on the exploit benchmark (91.9%) to avoid an automatic government ban, while releasing a three-tier mo…