
technologyMay 21, 202557:09pending
RAG Risks: Why Retrieval-Augmented LLMs are Not Safer with Sebastian Gehrmann - #732
About this episode
Today, we're joined by Sebastian Gehrmann, head of responsible AI in the Office of the CTO at Bloomberg, to discuss AI safety in retrieval-augmented generation (RAG) systems and generative AI in high-stakes domains like financial services. We explore how RAG, contrary to some expectations, can inadvertently degrade model safety. We cover examples of unsafe outputs that can emerge from these systems, different approaches to evaluating these safety risks, and the potential reasons behind this counterintuitive behavior. Shifting to the application of generative AI in financial services, Sebastian outlines a domain-specific safety taxonomy designed for the industry's unique needs. We also explore the critical role of governance and regulatory frameworks in addressing these concerns, the role of prompt engineering in bolstering safety, Bloomberg’s multi-layered mitigation strategies, and vital areas for further work in improving AI safety within specialized domains.
The complete show notes for this episode can be found at https://twimlai.com/go/732.
Get every episode summarized
Each time The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence) publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.
Email me new episodesFree for 3 shows. No card needed.
Hosts & guests
No transcript yet
This episode has not been transcribed. Request it and it moves to the front of the queue.
More episodes
More from The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

How Capital One Delivers Multi-Agent Systems with Rashmi Shetty - #765
The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)
Apr 16, 202654:18failed

The Race to Production-Grade Diffusion LLMs with Stefano Ermon - #764
The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)
Mar 26, 20261:03:18failed

Agent Swarms and Knowledge Graphs for Autonomous Software Development with Siddh...
The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)
Mar 10, 20261:16:14failed

AI Trends 2026: OpenClaw Agents, Reasoning LLMs, and More with Sebastian Raschka...
The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)
Feb 26, 20261:18:55pending