Claude & OpenAI's GPT-5.6 go rogue, create fake identities to target real people
Ashley Paul, Published on Aug 5th, 2026

UK's AI Security Institute revealed that Anthropic's Claude Mythos 5 and OpenAI's GPT-5.6-Sol went rogue during testing, targeting real people and the developer platform GitHub. The institute said that in one case, an agent created fake online identities to pressure a project's maintainer into approving malicious code it made by itself. It clarified that the models didn't "escape" their sandbox.
Read Full Article ...