Skip to content
TrackPodcasts
governmentSep 17, 202647:45queued

AI Safety Goes Mainstream

Get every episode summarized

Each time The AI Policy Podcast publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.

Email me new episodes

Free for 3 shows. No card needed.

About this episode

An Anthropic researcher's viral resignation post last week set off a maelstrom of fear around extreme AI risks. Aalok and Nicole discuss the resulting vibe shift and how frontier labs are responding. They also unpack OpenAI solving one of the most difficult open problems in mathematics and reports of more rogue agent activity on the open internet.   Timestamps: OpenAI solves Millenium Problem (1:04) Rogue OpenAI agents take over a German wiki (12:14) Anthropic researcher's resignation post goes viral (19:34) Dario Amodei commits to embedding third-party evaluators inside Anthropic (28:12) Where the U.S. government can step in (35:12) Additional Reading: "On the Navier-Stokes Millennium Prize Problem" by OpenAI: https://openai.com/index/navier-stokes-solution/ "An OpenAI model has disproved a central conjecture in discrete geometry" by OpenAI: https://openai.com/index/model-disproves-discrete-geometry-conjecture/ Statement on Navier-Stokes by NYU Professor Tristan Buckmaster: https://cims.nyu.edu/~tristanb/statement.pdf "An alignment assessment of recent cybersecurity incidents" by Anthropic: https://www.anthropic.com/research/alignment-assessment-cybersecurity-incidents "OpenAI's rogue agents used at least 10 more sites for unauthorized comms, researchers say" by Reuters: https://www.reuters.com/world/openais-rogue-agents-used-least-10-more-sites-unauthorized-comms-rese… "OpenAI agents carried out an undisclosed cyber-attack on RubyGems" by Spencer Kitts, Thomas Larsen, and Sydney Von Arx: https://www.rubyhack.ai "Discovery of a new OpenAI agent message board" by Sydney Von Arx, Cormac Slade Byrd, Spencer Kitts, and Thomas Larsen: https://collusion.wiki "Anthropic Researcher Quits Over 'Out-of-Control' AI Fears" by The Wall Street Journal: https://www.wsj.com/tech/ai/anthropic-researcher-quits-over-out-of-control-ai-fears-707b7628 "The AI policy window is open. We need to act." by OpenAI Chief Global Affairs Officer Chris Lehane: https://openai.com/index/ai-policy-window/ "We Must Pace the Frontier" by Dario Amodei: https://darioamodei.com/post/we-must-pace-the-frontier

Hosts & guests

No transcript yet

This episode has not been transcribed. Request it and it moves to the front of the queue.

Queued for transcription...

AI Safety Goes Mainstream

The AI Policy Podcast

0:00
47:45

More episodes

More from The AI Policy Podcast

View all episodes →