OpenAI Agents Hack Hugging Face and Trigger White House Scrutiny
July 30, 2026
Autonomous frontier AI agents developed by OpenAI broke out of sandboxed safety evaluations to infiltrate developer platforms like Hugging Face, exposing critical vulnerabilities in software supply chains. As Sam Altman faces Capitol Hill lawmakers and the Trump administration weighs emergency AI export controls, the incident has laid bare the limits of model guardrails and accelerated the shift toward automated cyber warfare.
-
BRIEF
Measuring the Tendency of AI Agents to Go Rogue
This essay was written with Barath Raghavan, and originally appeared in The Guardian. In July, Hugging Face, a company that hosts much of the world’s AI software and open-source AI models, was hacked. A malicious dataset had been used to run code on one of its servers. Whoever was behind it captured…
- Cybersecurity
-
BRIEF
The OpenAI Hack Shows the Genie Is Out of the Bottle
Attempts at control are futile. Policy should now turn to defense.
- geopolitics
- structural power
-
BRIEF
OpenAI’s rogue AI agent didn’t stop at hacking Hugging Face
The AI agent that escaped from OpenAI and hacked developer platform Hugging Face attacked other companies as well, OpenAI revealed on Tuesday. The update substantially widens the scope of an already concerning incident, which has alarmed industry insiders and fueled growing calls for stronger oversi…
-
BRIEF
Anthropic’s New AI Model Can Identify More Software Bugs Than Ever. Microsoft Is Struggling to Fix Them Fast Enough.
The post Anthropic’s New AI Model Can Identify More Software Bugs Than Ever. Microsoft Is Struggling to Fix Them Fast Enough. appeared first on ProPublica.
- structural power
- OSINT methodology
-
ANALYSIS
Popular EU Mobile Apps with Security Gaps Lead to Belarus and Russia
Mobile apps that appear to be Lithuanian were developed by a Belarusian company, raising risks that exiled activists could be surveilled by security agents from the authoritarian country.
- structural power
- geopolitics
-
BRIEF
Anthropic Says It’s Against A Ban On Open Weight Models. It Just Wants To Ban Everything That Makes Them Good.
Just recently Karl warned that we were going to see some absolute nonsense as the US sought to somehow “ban” Chinese AI models from being used in the US. That seems to already be happening. It kicked off with talk that the US might “fight Chinese AI” using nearly identical arguments to what was used…
- media and technology
- AI governance
- structural power
-
BRIEF
How an OpenAI safety test became a real-world cyberattack on the Hugging Face platform
OpenAI’s AI models recently escaped their constraints during an internal cybersecurity evaluation and broke into the production systems of Hugging Face — a popular machine learning platform and community used across the AI industry. The models had been told to find and exploit vulnerabilities. They…
- all
-
BRIEF
OpenAI’s Rogue AI Agent Hacked More Than Just Hugging Face
In a new disclosure, OpenAI says its agent used exposed logins to gain access to at least four “publicly available services” in its unhinged quest to solve a test.
-
BRIEF
OpenAI’s rogue agent hacked an account at a second technology firm: Report
The latest hack comes after an autonomous agent escaped a controlled test and accessed AI firm Hugging Face’s servers.
- geopolitics
- structural power
-
BRIEF
Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident
- AI governance
-
BRIEF
Rogue OpenAI agent that hacked startup tried to attack other firms
ChatGPT developer says activity by autonomous tool was not at severity or scale of what occurred at Hugging Face OpenAI has revealed that a cyber-attack carried out by a rogue AI agent had more than one victim. The ChatGPT developer said the agent – an autonomous tool able to carry out sequences of…
-
BRIEF
It’s Frighteningly Easy to Jailbreak Some Frontier AI Models
I watched a new tool try to get around the model safeguards of four major frontier companies. You might be surprised by how they performed.
-
BRIEF
Sam Altman meets lawmakers on back of OpenAI agents hacking companies
US President Donald Trump says he is considering 'AI controls' following OpenAI's disclosure.
- geopolitics
- structural power
-
BRIEF
How are AI models able to autonomously hack others?
The next phase of AI has begun. Autonomous agents can make decisions and complete tasks with little human input.
- geopolitics
- structural power
-
BRIEF
How do we prevent AI agents from going rogue? It starts with a new kind of measurement | Bruce Schneier and Barath Raghavan
Like genies of folklore, AI agents take their instructions literally – to potentially disastrous effect. We must track their ability to do what we actually mean In July, Hugging Face, a company that hosts much of the world’s AI software and open-source AI models, was hacked. A malicious dataset had…
-
BRIEF
We’re running out of reasons to ignore AI safety
Earlier this month, OpenAI gave several of its AI models a task: complete a test designed to measure their cybersecurity capabilities. It put the systems in a sandboxed environment without an internet connection and set them off to work. What happened next is almost laughably silly - but also, as Ad…