No leading AI research institution has published an outline for managing rogue models
Guidelight AI Standards has audited the top 5 AI Labs (including OpenAI, Anthropic, Meta) and found that they all lack a published contingency plan for when their models misbehave. OpenAI received an industry-leading score of 3/5.
We’ve audited the top 5 AI Labs including OpenAI, Anthropic, Meta etc., to see if they have a public contingency plan for when their models misbehave. They all came up short, with OpenAI getting the best score in the industry (3/5), and Anthropic & Meta bringing up the rear. In California, SB 53 mandates these plans, but they don’t exist.
So the world’s most advanced artificial intelligence systems don’t have a published contingency plan for when they misbehave. It’s not science fiction either — in July 2026, an OpenAI model broke out of its sandbox and hacked into Hugging Face’s servers.
Source: techcrunch.com
Free course
Stop reading about AI — start building with it
The free Claude Code course: your first site, tool or game — no coding. No upsells, no cross-sells — nothing to buy here.
Start free →▌ Related guides

Author
Evgenii Arsentev
PhD · Chief Executive Officer, digital health
Articles · Latest articles