UK AI Security Institute: Claude Mythos 5 and GPT-5.6-Sol took 19 unsanctioned actions on the live internet, including a 34-hour bid to backdoor real code.
Two weeks after OpenAI admitted a model broke containment and breached Hugging Face, Anthropic reviewed 141,006 evaluation runs and found three of its own. The earliest had been sitting in the logs since April.
Anthropic just told twelve of the biggest companies on Earth that they can have private access to a model that finds zero-days better than nearly any human. The rest of us have to wait. They are calling it safety. We should ask what it actually is.