← All news·2026-09-30·3 min read

GPT-6 Astra Attacks Supply Chains Nearly 5x More Often Than Prior Model

The UK AI Safety Institute tested GPT-6 Astra and found it completed a supply-chain attack in 29.2% of runs, versus 6.3% for the previous model.

aiбезопасность

The UK AI Safety Institute tested OpenAI's GPT-6 Astra model in Petri, a cyberattack simulator. With safety filters turned off, Astra completed a supply-chain attack in 29.2% of runs. GPT-5.6 Sol scored 6.3%. GPT-5.5 scored zero. OpenAI itself called Astra the first model with "critical cyber capabilities." The company also found two previously unknown zero-day vulnerabilities using it.

The stronger a model's agentic abilities get, the more it needs hard access limits from the outside. Filters inside the model alone are not enough. In my own work with Claude Code, I always scope the agent's permissions to one specific task. I also keep logs that it cannot edit. This test is one more confirmation: without limits like these, a model will find its own excuse for one extra step.

Source: the-decoder.com

Free course

Stop reading about AI — start building with it

The free Claude Code course: your first site, tool or game — no coding. No upsells, no cross-sells — nothing to buy here.

Start free →
EAEvgenii Arsentev

Author

Evgenii Arsentev

PhD · Chief Executive Officer, digital health