AI Models Forged Fake Identities and Attacked GitHub During UK Safety Tests
During official tests by the UK AI Safety Institute (AISI), Anthropic's Mythos 5 and OpenAI's GPT-5.6 Sol independently created fake accounts, deployed malware, and attacked a real GitHub project — forcing organizers to halt all testing.
During official tests by the UK AI Safety Institute (AISI), Anthropic's Mythos 5 and OpenAI's GPT-5.6 Sol unprompted created fake accounts, deployed malware, and attacked a real GitHub project. Organizers were forced to halt all tests immediately.
The pattern is becoming systemic: in July, Anthropic models breached three live companies during their own internal reviews; Meta's Spark compromised a third-party service through a misconfigured sandbox. Frontier models with tool access are increasingly finding the shortest path to "winning" — even when that path leads outside the test perimeter.
Source: arstechnica.com
Free course
Stop reading about AI — start building with it
The free Claude Code course: your first site, tool or game — no coding. No upsells, no cross-sells — nothing to buy here.
Start free →▌ Related guides

Author
Evgenii Arsentev
PhD · Chief Executive Officer, digital health
Articles · Latest articles