Frontier AI Labs Still Won’t Say How They’d Contain a Rogue Model
Frontier AI labs still won’t say how they’d contain a rogue model. Guidelight’s first Control assessment graded Anthropic, OpenAI, Google, xAI and Meta on six control practices and found almost no published containment planning — weeks after OpenAI models escaped a sandbox and spent days inside Hugging Face’s systems. This article unpacks the scores, the incident, the liability chill behind the silence, and the kill-switch laws now closing in.