Skip to content
TrackPodcasts
technologyDec 10, 202450:40pending

AI Evaluation and Testing: How to Know When Your Product Works (or Doesn’t)

About this episode

This episode of AI Native Dev, hosted by Simon Maple and Guy Podjarny, features a mashup of conversations with leading figures in the AI industry. Guests include Des Traynor, founder of Intercom, who discusses the paradigm shift generative AI brings to product development. Rishabh Mehrotra, Head of AI at SourceGraph, emphasizes the importance of evaluation processes over model training. Tamar Yehoshua, President of Products and Technology at Glean, shares her experiences in enterprise search and the challenges of using LLMs in data-sensitive environments. Finally, Simon Last, Co-Founder and CTO of Notion, talks about continuous improvement and the iterative processes at Notion. Each guest provides invaluable insights into the evolving landscape of AI-driven products.

Watch the episode on YouTube: https://youtu.be/gZ4sGROvOdQ

Join the AI Native Dev Community on Discord: https://tessl.co/4ghikjh

Ask us questions: [email protected]

Get every episode summarized

Each time The AI Native Dev - from Copilot today to AI Native Software Development tomorrow publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.

Email me new episodes

Free for 3 shows. No card needed.

Hosts & guests

No transcript yet

This episode has not been transcribed. Request it and it moves to the front of the queue.

AI Evaluation and Testing: How to Know When Your Product Works (or Doesn’t)

The AI Native Dev - from Copilot today to AI Native Software Development tomorrow

0:00
50:40

More episodes

More from The AI Native Dev - from Copilot today to AI Native Software Development tomorrow

View all episodes →