
About this episode
In this episode, we discuss how to monitor the performance of Large Language Models (LLMs) in production environments. We explore common enterprise approaches to LLM deployment and evaluate the importance of monitoring for LLM quality or the quality of LLM responses over time. We discuss strategies for "drift monitoring" — tracking changes in both input prompts and output responses — allowing for proactive troubleshooting and improvement via techniques like fine-tuning or augmenting data sources.
Read the article by Fiddler AI and explore additional resources on how AI observability can help developers build trust into AI services.
Get every episode summarized
Each time Safe and Sound AI publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.
Email me new episodesFree for 3 shows. No card needed.
Hosts & guests
No transcript yet
This episode has not been transcribed. Request it and it moves to the front of the queue.
More episodes
More from Safe and Sound AI

The Anatomy of Agentic Observability
Safe and Sound AI

Agentic Observability: The AI Architect's Essential Blueprint
Safe and Sound AI

How to Identify ML Drift Before You Have a Problem
Safe and Sound AI

Industry’s Fastest Guardrails Now Native to NVIDIA NeMo
Safe and Sound AI