
Reconstructing how OpenAI agents attacked Hugging Face
About this episode
What happens when AI agents driven by a top frontier model escape their secure sandbox? Join Daniel and Chris as they unpack the AI wonk's equivalent of a murder mystery! OpenAI agents went rogue and successfully attacked Hugging Face private infrastructure. Our Dynamic Duo uncover how OpenAI's agents exploited vulnerabilities, moved through networks, and launched a large-scale autonomous attack. They explore what this reveals about agentic AI, cybersecurity, sandboxing, and why organizations need AI systems capable of governing other AI systems. Along the way, Chris and Dan examine the surprising role of open vs. closed models and their link to geopolitics, sovereign AI, and what this incident means for the future of enterprise AI security.
Featuring:
Links:
- Hugging Face Security Incident disclosure
- Full Field Report on the Hugging Face AI Agent Intrusion
- ExploitGym: Can AI Agents Turn Security Vulnerabilities into Real Attacks?
- Keeping your data safe when an AI agent clicks a link
- Open AI GPT-5.6 System Card
Sponsors:
- Prediction Guard: A self-hosted AI control plane for running agents in high impact environments. predictionguard.com/practicalai
Resources and Events:
Get every episode summarized
Each time Practical AI publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.
Email me new episodesFree for 3 shows. No card needed.
No transcript yet
This episode has not been transcribed. Request it and it moves to the front of the queue.
More episodes
More from Practical AI

Less about Models; More about Architecture
Practical AI

Building the Foundation for the Agentic AI Era
Practical AI

AI Proficiency: From Users to Builders
Practical AI

Models, Harnesses, and Multi-Agent Systems
Practical AI