Skip to content
TrackPodcasts
technologyAug 19, 20251:09:33pending

915: How to Jailbreak LLMs (and How to Prevent It), with Michelle Yi

Get every episode summarized

Each time Super Data Science: ML & AI Podcast with Jon Krohn publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.

Email me new episodes

Free for 3 shows. No card needed.

About this episode

Tech leader, investor, and Generationship cofounder Michelle Yi talks to Jon Krohn about finding ways to trust and secure AI systems, the methods that hackers use to jailbreak code, and what users can do to build their own trustworthy AI systems. Learn all about “red teaming” and how tech teams can handle other key technical terms like data poisoning, prompt stealing, jailbreaking and slop squatting.  This episode is brought to you by ⁠Trainium2, the latest AI chip from AWS⁠ and by the ⁠Dell AI Factory with NVIDIA⁠. Additional materials: ⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠⁠www.superdatascience.com/915⁠⁠⁠⁠⁠ Interested in sponsoring a SuperDataScience Podcast episode? Email [email protected] for sponsorship information. In this episode you will learn: (03:31) What “trustworthy AI” means      (31:15) How to build trustworthy AI systems  (46:55) About Michelle’s “sorry bench”   (48:13) How LLMs help construct causal graphs   (51:45) About Generationship

Hosts & guests

No transcript yet

This episode has not been transcribed. Request it and it moves to the front of the queue.

915: How to Jailbreak LLMs (and How to Prevent It), with Michelle Yi

Super Data Science: ML & AI Podcast with Jon Krohn

0:00
1:09:33

More episodes

More from Super Data Science: ML & AI Podcast with Jon Krohn

View all episodes →