Skip to content
TrackPodcasts
scienceOct 15, 202443:28pending

Google's NotebookLM and the Future of AI-Generated Audio

Deep Papers

About this episode

This week, Aman Khan and Harrison Chu explore NotebookLM’s unique features, including its ability to generate realistic-sounding podcast episodes from text (but this podcast is very real!). They dive into some technical underpinnings of the product, specifically the SoundStorm model used for generating high-quality audio, and how it leverages a hierarchical vector quantization approach (RVQ) to maintain consistency in speaker voice and tone throughout long audio durations. 

The discussion also touches on ethical implications of such technology, particularly the potential for hallucinations and the need to balance creative freedom with factual accuracy. We close out with a few hot takes, and speculate on the future of AI-generated audio. 


Learn more about AI observability and evaluation, join the Arize AI Slack community or get the latest on LinkedIn and X.

Get every episode summarized

Each time Deep Papers publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.

Email me new episodes

Free for 3 shows. No card needed.

Hosts & guests

No transcript yet

This episode has not been transcribed. Request it and it moves to the front of the queue.

Google's NotebookLM and the Future of AI-Generated Audio

Deep Papers

0:00
43:28

More episodes

More from Deep Papers

View all episodes →