Skip to content
TrackPodcasts
technologyJan 15, 202438:48pending

Learning Transformer Programs with Dan Friedman - #667

Get every episode summarized

Each time The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence) publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.

Email me new episodes

Free for 3 shows. No card needed.

About this episode

Today, we continue our NeurIPS series with Dan Friedman, a PhD student in the Princeton NLP group. In our conversation, we explore his research on mechanistic interpretability for transformer models, specifically his paper, Learning Transformer Programs. The LTP paper proposes modifications to the transformer architecture which allow transformer models to be easily converted into human-readable programs, making them inherently interpretable. In our conversation, we compare the approach proposed by this research with prior approaches to understanding the models and their shortcomings. We also dig into the approach’s function and scale limitations and constraints. The complete show notes for this episode can be found at twimlai.com/go/667.

Hosts & guests

No transcript yet

This episode has not been transcribed. Request it and it moves to the front of the queue.

Learning Transformer Programs with Dan Friedman - #667

The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

0:00
38:48

More episodes

More from The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

View all episodes →