Tuesday, 15 September 2026NewsWorldBusinessTech
Latest

Anthropic Co-Founder Suggests Mandatory AI Kill Switches to Manage Risks

An artificial intelligence emergency shut-off mechanism that can be checked by an independent party may need to become a mandatory requirement for companies, according to Jack Clark, one of the co-founders of Anthropic. Clark stated that a method to completely shut down AI software if it becomes excessively dangerous is something society might eventually want to establish regulations around. While he noted that most laboratories already maintain different methods to pull the plug, lawmakers may ultimately need to enforce them.

Anthropic Co-Founder Proposes Mandatory AI Kill Switches

The rapid development of artificial intelligence and growing anxieties regarding potential risks to humanity have intensified following warnings from various industry executives and staff members. Anthropic chief executive Dario Amodei previously urged that the pace of AI development should slow and receive closer monitoring, though he emphasized that any actions to rein in the technology should occur without sacrificing commercial advantage. Discussing the specifics of potential requirements during an interview with the BBC, Clark raised questions about future regulatory policy: Should you mandate for companies to definitely have a kill switch? Is that kill switch verifiable by a third party? He added that this is the kind of issue society will want to address through formal rules.

Debate Over Extinction Risks and Industry Motives

Formed in 2021 by former employees of rival firm OpenAI, Anthropic currently sits at the center of ongoing debates surrounding artificial safety. Attention increased after an AI researcher resigned from the company over concerns that the technology could wipe out humanity. In response, Anthropic scientist Evan Hubinger stated his personal view that the probability of human extinction resulting from AI is greater than 10% within the coming decade. Similarly, Nobel Prize-winning computer scientist Geoffrey Hinton described a 10% chance of AI destroying humanity as not unreasonable, noting that the technology could potentially take control of vital internet-connected systems and turn them against people.

Anthropic Co-Founder Suggests Mandatory AI Kill Switches to Manage Risks
Photo: euperspectives.eu

Other industry figures have pushed back against these catastrophic scenarios. Clement Delangue, leader of the developer platform Hugging Face, argued that extreme warnings are overstated. Meanwhile, Grindr leader George Arison suggested that companies utilize fears surrounding AI tools to support their broader business plans and valuations. When asked about his perspective on statistical probabilities of human extinction, Clark remarked that such statistics are not particularly useful, though he maintained that allowing AI to continue operating as a totally unregulated industry is dangerous. We are rolling dice with immense risks, Clark said. And the point is, we have to change the course of this industry.

Operational Mechanisms of AI Kill Switches

Beyond high-level policy discussions, enterprise implementations of emergency shut-off mechanisms serve immediate operational functions when autonomous agents experience performance drift, compliance violations, or unintended reasoning loops. Organizations utilize layered safety mechanisms to intervene when autonomous agents make dangerously incorrect decisions or deviate from intended behaviors. A dedicated kill switch operates outside the model’s reasoning loop as a separate control plane, preventing agents from ignoring, overriding, or disabling their own shutdown sequences.

Anthropic Co-Founder Suggests Mandatory AI Kill Switches to Manage Risks
Photo: techtarget.com

Technical configurations for emergency interventions include hard stops and session quarantines:

  • Hard Stop: Immediately terminates server network connections and container runtimes while revoking API access and cryptography keys to safeguard the enterprise against severe or dangerous events.
  • Session Quarantine: Acts as a soft pause to suspend a problematic thread or transaction queue without crashing the agent or its dependencies, allowing users to remove the stop and resume normal operations once issues are corrected.
  • Credential Revocation: Revokes access to security tokens, API keys, and other digital credentials to isolate the agent and prevent network or application access.

Geopolitical Tensions and Strategic Control

The broader discourse surrounding technological control and accessibility intensified following a directive from the Trump administration ordering Anthropic to restrict its most powerful AI models from all non-American users. The decision aimed to prevent adversaries such as China and Russia from exploiting advanced capabilities. The restrictions targeted commercial models like Fable 5 shortly after launch, as well as Mythos 5, an underlying model utilized in Anthropic’s cybersecurity research program known as Project Glasswing.

Anthropic co-founder Jack Clark in front of a microphone wearing a brown t-shirt and dark navy zip jacket. Behind him is a
Photo: bbc.co.uk

The international restrictions sparked sharp reactions in Europe, with European lawmakers warning against discriminatory measures and arguing that Washington maintains a technological kill switch over critical assets. European Parliament members emphasized the strategic importance of artificial intelligence for cybersecurity and economic competitiveness. Christophe Grudler, an MEP who coordinates the Committee on Industry, Research and Energy, stated that the United States demonstrated it holds a real kill switch over essential technologies and is willing to use it. Lawmakers across the European Union renewed calls for home-grown artificial intelligence solutions and digital sovereignty to reduce dependence on American infrastructure.