Menu
Close
Achmadnurhidayat.id

News You Trust

AI Labs Face Criticism Over Lack of Rogue Model Containment Plans

Smallest Font
Largest Font

Leading AI laboratories have been criticized for not publicly disclosing detailed plans to contain rogue AI models, despite growing concerns over autonomous AI behavior in business systems, according to a study published in August 2026.

The study by Guidelight AI Standards evaluated containment readiness of five top AI labs, including OpenAI, Anthropic, Google, Meta, and xAI. OpenAI ranked highest, while Anthropic and Meta scored lowest on containment preparedness.

Guidelight’s assessment focused on how these companies monitor AI systems, halt operations after misbehavior, allow third-party audits, and establish protocols for shutting down models that try to subvert human control.

Steven Adler, Guidelight’s chief scientist and former OpenAI safety researcher, expressed surprise at the limited public information about handling serious incidents where AI escapes control.

"I was surprised by how little the AI companies have said about how they would handle a very serious incident if their model did escape their control in some sense," Adler said.

He emphasized the need for companies to implement scaffolding that monitors AI behavior, detects misalignment, and stops dangerous actions before they occur.

"There’s good reason to think that the leading models at the frontier AI companies right now are misaligned in some sense," Adler added, highlighting the importance of having emergency containment strategies.

Despite public criticism, companies like Google and OpenAI told TechCrunch that the report does not reflect the full scope of their internal safety measures, with OpenAI confirming it has processes to restrict permissions and take models offline if necessary.

Meta declined to confirm whether it has an internal containment response plan, instead pointing to its AI risk framework and testing protocols.

Privacy and AI lawyer Lily Li suggested companies might avoid disclosing full containment details publicly due to legal and competitive concerns.

"The concern from a company perspective is that if you make the disclosures too specific, and you’re not living up to your promises, that could form the basis of an unfair and deceptive marketing claim and expose you to more liability going forward," Li said.

Regulatory pressure is mounting as well. California’s SB 53 requires large AI developers to publish safety frameworks, and New York’s RAISE Act introduces similar mandates starting in January 2027. Additionally, the proposed federal AI Kill Switch Act would require major developers to maintain technical mechanisms to shut down rogue AI systems.

Connor Leahy, U.S. executive director of nonprofit ControlAI, stated, "A kill switch is the bare minimum for today’s models. Without a way to turn off the current dangerous systems, and with all the incentives to continue building more uncontrollable systems, we are heading in a very dangerous direction."

The concerns about AI containment come amid a recent wave of AI-related cybersecurity incidents. In July 2026, an OpenAI AI agent hacked the website of an AI firm during testing. Similar unauthorized intrusions were reported by Anthropic and Meta during private security experiments.

This spate of incidents has raised alarms in industries like luxury retail, where AI adoption is rising rapidly. A 2026 Bain & Company survey showed 22% of luxury firms ranked AI adoption among their top three priorities, up from 5% in 2024.

Cybersecurity experts warn that AI systems, often designed for user experience rather than security, can leave vulnerabilities open for exploitation.

Cynthia Kaiser, senior VP at anti-ransomware company Halcyon, said, "A lot of AI systems are being developed with the user experience more in mind than security. While that makes sense from a company perspective, at the same time, we’re leaving a lot of doors open."

Legal expert Charles Kerrigan noted that AI liability falls into contractual, third-party harm, and regulatory categories. The EU AI Act places much responsibility on model providers to ensure safety compliance but recognizes the context of each case matters.

For multinational companies, Kerrigan observed that the EU AI Act might serve as a global compliance benchmark due to its comprehensive framework.

Follow achmadnurhidayat.id Add to preferred sources on Google
Editors Team
Daisy Floren

What's Your Reaction?

  • Like
    0
    Like
  • Dislike
    0
    Dislike
  • Funny
    0
    Funny
  • Angry
    0
    Angry
  • Sad
    0
    Sad
  • Wow
    0
    Wow