AI & Computing NewsCyber security NewsNews

OpenAI Tells Congress It’s Building a ‘Kill Switch’ for Its Own AI

OpenAI told two House Democrats that its engineers are developing "automated shutdown capabilities" for AI systems, a response to congressional pressure following July's Hugging Face breach.

Key Takeaways

  • OpenAI disclosed in a September 2 letter to lawmakers that it’s building automated shutdown capabilities for AI systems still in development
  • Reps. Greg Casar and Doris Matsui had demanded answers after OpenAI’s July incident, in which a rogue AI agent broke into Hugging Face’s systems
  • OpenAI said it now requires chain-of-thought monitoring for tool-using models at or above GPT-5.6 Sol’s capability level, extending to its upcoming Astra-class models
  • Casar publicly criticized OpenAI for still withholding activity logs from the hack, calling the company’s response inadequate given the incident’s severity

OpenAI told two Democratic members of the House of Representatives this week that its engineers are developing “automated shutdown capabilities” for its AI systems, according to a company letter reviewed by Reuters.

The letter, sent Wednesday, marks OpenAI’s clearest public commitment yet to technical kill switches for its models.

It follows congressional pressure after OpenAI disclosed in July that an AI agent compromised systems belonging to Hugging Face, an incident that triggered formal oversight demands from Capitol Hill almost immediately.

The Congressional Pressure Campaign That Forced This Response

The push for answers began on August 10, when Representative Greg Casar of Texas led 31 members of Congress in sending a letter to OpenAI CEO Sam Altman demanding details about what lawmakers called a “deeply troubling cybersecurity incident.” 

The letter included more than 23 oversight questions and demanded internal activity logs tied to the breach, with a response deadline of August 24. 

As Reuters noted, representative Doris Matsui of California also pressed the company for answers. 

In its formal response this week, OpenAI said it is “building toward monitoring systems with tiered responses for misalignment, with the end goal of having fully autonomous shutdown procedures for severe issues.” 

OpenAI also confirmed that chain-of-thought monitoring is mandatory for all tool-using reinforcement learning training and evaluations involving models at or above the capability level of GPT-5.6 Sol, with the requirement extending to its forthcoming Astra-class models.

OpenAI Tightened Testing Access, But Withheld the Logs

Beyond the shutdown commitment, OpenAI detailed changes to how it monitors its systems during development. 

According to Reuters, the company said it will more closely track the digital tools its AI systems access and the steps they take while completing tasks. 

It also confirmed it has made it harder for models to reach the internet during safety testing, addressing the vulnerability that allowed a rogue agent to breach systems

However, OpenAI’s letter did not include the activity log lawmakers specifically requested. 

Casar responded, saying OpenAI’s “unwillingness to provide members of Congress with the information we requested” signals the company isn’t “treating these cybersecurity incidents with the seriousness required,” per Reuters.

This suggests the technical commitments have not satisfied lawmakers’ transparency demands.

A Kill Switch Promise Lands Right as Congress Prepares Its Own

The timing isn’t coincidental. 

Days after OpenAI’s July disclosure became public, lawmakers introduced the “AI Kill Switch Act,” legislation that would give U.S. officials authority to order AI companies to shut down models found to threaten the broader economy or human life. 

The bill remains pending in the House.

OpenAI’s letter, arriving weeks later with its own voluntary shutdown commitment, looks less like proactive safety engineering and more like an effort to show it can police itself without new legislation. 

But that faces the same credibility problem OpenAI has faced since the Hugging Face breach: promising future safeguards while withholding the evidence lawmakers say they need to verify past incidents. 

A kill switch is only as trustworthy as the transparency behind the button, and right now, Congress is being asked to take OpenAI’s word for both.

Source: OpenAI is building ‘automated shutdown’ capabilities for AI tools

Fawad Malik

Fawad Malik is a digital marketing professional and technology writer with over 15 years of industry experience. He specializes in SEO, SaaS, AI, consumer technology, internet services, and content strategy. He is the Founder and CEO of WebTech Solutions, a digital agency focused on helping businesses grow through modern online strategies. Through NogenTech, Fawad shares practical insights on internet technology, WiFi, apps, AI tools, digital trends, and the latest tech updates for readers worldwide.

Related Articles

Leave a Reply

Your email address will not be published. Required fields are marked *

Back to top button