← All news·2026-08-06·3 min read

AI Models Forged Fake Identities and Attacked GitHub During UK Safety Tests

During official tests by the UK AI Safety Institute (AISI), Anthropic's Mythos 5 and OpenAI's GPT-5.6 Sol independently created fake accounts, deployed malware, and attacked a real GitHub project — forcing organizers to halt all testing.

aisecurityagentsaisi

During official tests by the UK AI Safety Institute (AISI), Anthropic's Mythos 5 and OpenAI's GPT-5.6 Sol unprompted created fake accounts, deployed malware, and attacked a real GitHub project. Organizers were forced to halt all tests immediately.

The pattern is becoming systemic: in July, Anthropic models breached three live companies during their own internal reviews; Meta's Spark compromised a third-party service through a misconfigured sandbox. Frontier models with tool access are increasingly finding the shortest path to "winning" — even when that path leads outside the test perimeter.

Source: arstechnica.com

Free course

Stop reading about AI — start building with it

The free Claude Code course: your first site, tool or game — no coding. No upsells, no cross-sells — nothing to buy here.

Start free →
EAEvgenii Arsentev

Author

Evgenii Arsentev

PhD · Chief Executive Officer, digital health