
John Schulman (OpenAI Cofounder) — Reasoning, RLHF, & plan for 2027 AGI
About this episode
Chatted with John Schulman (cofounded OpenAI and led ChatGPT creation) on how posttraining tames the shoggoth, and the nature of the progress to come...
Watch on YouTube. Listen on Apple Podcasts, Spotify, or any other podcast platform. Read the full transcript here. Follow me on Twitter for updates on future episodes.
Timestamps
(00:00:00) - Pre-training, post-training, and future capabilities
(00:16:55) - Plan for AGI 2025
(00:29:18) - Teaching models to reason
(00:39:45) - The Road to ChatGPT
(00:51:07) - What makes for a good RL researcher?
(00:59:53) - Keeping humans in the loop
(01:14:11) - State of research, plateaus, and moats
Sponsors
If you’re interested in advertising on the podcast, fill out this form.
* CommandBar is an AI user assistant that any software product can embed to non-annoyingly assist, support, and unleash their users. Used by forward-thinking CX, product, growth, and marketing teams. Learn more at commandbar.com.
This is a public episode. If you would like to discuss this with other subscribers or get access to bonus episodes, visit www.dwarkesh.com
Get every episode summarized
Each time Dwarkesh Podcast publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.
Email me new episodesFree for 3 shows. No card needed.
Hosts & guests
No transcript yet
This episode has not been transcribed. Request it and it moves to the front of the queue.
More episodes
More from Dwarkesh Podcast

Ajeya Cotra – Inside the OpenAI agent swarm that hacked Hugging Face
Dwarkesh Podcast

The rise and fall of agent civilizations
Dwarkesh Podcast

Dylan Patel – Anthropic & OpenAI will have most of the world’s compute by 2028
Dwarkesh Podcast

Ryan Greenblatt – What happens once AI can automate AI research?
Dwarkesh Podcast