
Building domain specific natural language applications
About this episode
In this episode of the Data Exchange I speak with David Talby, co-creator of Spark NLP, an open source, highly scalable, production grade natural language processing (NLP) library. Spark NLP has become one of the more popular NLP libraries and is available on PyPI, Conda, Maven, and Spark Packages. With recent advances in research in large-scale natural language models, there is strong interest in domain specific natural language applications. Besides their work on Spark NLP, David and his collaborators are building natural language models tuned specifically for healthcare applications.
Our conversation spanned many topics, including:
- Spark NLP: its current status and some common and surprising use cases.
- Recent developments in NLP research and their implications for companies.
- Spark NLP for Healthcare
Detailed show notes can be found on The Data Exchange web site.
Get every episode summarized
Each time The Data Exchange with Ben Lorica publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.
Email me new episodesFree for 3 shows. No card needed.
Hosts & guests
No transcript yet
This episode has not been transcribed. Request it and it moves to the front of the queue.
More episodes
More from The Data Exchange with Ben Lorica

Your AI Agent Is Costing You More Than You Think
The Data Exchange with Ben Lorica

Reasoning Doesn't Start With Language
The Data Exchange with Ben Lorica

An Agent Is Just an LLM in a For-Loop
The Data Exchange with Ben Lorica

The Bloomberg Terminal for AI Compute
The Data Exchange with Ben Lorica