
About this episode
Production ML systems include more than just the model. In these complicated systems, how do you ensure quality over time, especially when you are constantly updating your infrastructure, data and models? Tania Allard joins us to discuss the ins and outs of testing ML systems. Among other things, she presents a simple formula that helps you score your progress towards a robust system and identify problem areas.
Featuring:
- Tania Allard – Website, GitHub, X
- Chris Benson – Website, GitHub, LinkedIn, X
- Daniel Whitenack – Website, GitHub, X
Show Notes:
- “What’s your ML score” talk
- “Jupyter Notebooks: Friends or Foes?” talk
- Joel Grus’s episode: “AI code that facilitates good science”
- Papermill
- nbdev
- nbval
Books
Upcoming Events:
- Register for upcoming webinars here!
Get every episode summarized
Each time Practical AI publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.
Email me new episodesFree for 3 shows. No card needed.
No transcript yet
This episode has not been transcribed. Request it and it moves to the front of the queue.
More episodes
More from Practical AI

Less about Models; More about Architecture
Practical AI

Building the Foundation for the Agentic AI Era
Practical AI

AI Proficiency: From Users to Builders
Practical AI

Models, Harnesses, and Multi-Agent Systems
Practical AI