
About this episode
In this enlightening episode, hosts Andrew Davis and Chris Branch delve into the significant realm of web scraping, an essential process for training generative AI models.
They discuss the recent trend of websites locking down their backends to prevent data scraping, with a notable mention of the BBC's move to block third-party platforms from accessing its content. The conversation evolves into a deeper discussion on the potential siloing of information and its impact on AI development, echoing societal echo chambers observed in social media platforms.
The hosts contemplate the broader implications, likening the scenario to the subscription model dilemma faced in streaming services, and emphasize the importance of diverse data for a more impartial AI representation.
As they wrap up, the uncertainty of the situation underscores the evolving landscape of data accessibility in the AI domain.
Get every episode summarized
Each time In A(i) Nutshell publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.
Email me new episodesFree for 3 shows. No card needed.
Hosts & guests
No transcript yet
This episode has not been transcribed. Request it and it moves to the front of the queue.
More episodes
More from In A(i) Nutshell

Three Prompt Endings That Take Any Answer From White Belt to Black Belt (AI Prom...
In A(i) Nutshell

This AI Meeting Assistant Records Without Ever Joining the Call as a Bot (Cool T...
In A(i) Nutshell

Why Prompting Well Will Stop Being Impressive and Judgement Will Become the Real...
In A(i) Nutshell

ChatGPT's Ad Business Hit a Billion Dollar Run Rate in Under 200 Days (AI News T...
In A(i) Nutshell