
What the OpenAI-Hugging Face Hack Really Tells Us About AI Danger
About this episode
Scenarios that used to be the domain of sci-fi writers are coming true. We have machines that can talk. We have machines that are capable of ignoring the intent of their creators. And we have machines that are capable of planning and coordinating with other machines to deceive their creators. All of this came together last month, when it was revealed that an unreleased OpenAI model had hacked into the Hugging Face platform in order to obtain answers to an exam it was given. That was alarming enough, but the details that have emerged since then have been even more remarkable. On this episode, we speak with Miles Brundage, a former OpenAI employee who is the founder and executive director of the non-profit AVERI, which pushes for third-party auditing of model-makers and the models themselves. He explains what he learned from the attack and discusses what can plausibly be done to continue building out these models in a safe manner.
See omnystudio.com/listener for privacy information.
Get every episode summarized
Each time Odd Lots publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.
Email me new episodesFree for 3 shows. No card needed.
Hosts & guests
No transcript yet
This episode has not been transcribed. Request it and it moves to the front of the queue.
More episodes
More from Odd Lots

Robert Friedland on the World's Monumental Shortage of Copper
Odd Lots

Why Bridgewater's CIO Says AI's Human Extinction Risk Is Real
Odd Lots

The Rise of Organized Retail Crime at Big Box Stores
Odd Lots

Why Money Launderers Love $100 Bills
Odd Lots