344
Productivity & Workflow355
Automation & Workflow225
Software Development251
Marketing & Growth192
AI Infrastructure & MLOps175
Writing & Content Creation203
Data & Analytics142
Photography & Imaging156
Design & Creative170
Customer Support132
Sales & Outreach125
Voice & Speech135
Education & Learning131
Operations & Admin87
A new review says leading AI companies have few public plans for what to do if an AI system tries to break rules or evade control.
In short: A new study says most leading AI labs have not publicly explained how they would shut down or limit an AI model that starts acting against human control.
Guidelight AI Standards, a safety-focused group, reviewed public information from five major AI labs: OpenAI, Anthropic, Google, Meta, and xAI. It looked for “containment plans,” which are step by step instructions for what to do if an AI system is caught trying to dodge oversight or take actions it should not. Think of it like a fire drill plan for software, who cuts off access, who decides what happens next, and when to power the system down.
Based on what the companies have published, Guidelight found little detail on these response plans. OpenAI scored highest in the review, but still only 3 out of 5. The report said OpenAI has paused or stopped work after safety incidents, but it did not find proof of a formal, pre-set plan for future incidents.
Meta and Anthropic scored lowest on public disclosure, according to the study. Guidelight said it could not find evidence that Meta has a containment response plan. It also said Anthropic’s risk reporting did not clearly describe limiting or stopping a model as a possible outcome after a control problem.
Some companies said the study does not reflect everything they do internally. Google and OpenAI told TechCrunch the assessment is based only on what is public. Meta pointed to an existing safety framework, and xAI did not respond in time.
Pressure for transparency may increase. California’s SB 53 already requires some large AI developers to publish safety frameworks, and New York’s RAISE Act is set to take effect in January. In Congress, lawmakers have also introduced an “AI Kill Switch Act,” which would require big developers to build a way to shut down a rogue model.
Source: TechCrunch AI