
Unleashing the power of large language models
About this episode
Maarten Grootendorst, is a data scientist at IKNL, and more importantly, he’s the author of two open source libraries that I’ve come to love: BERTopic (topic modeling with transformers and c-TF-IDF) and PolyFuzz (fuzzy string matching). Both these projects bring the power of transformers and other leading edge models, and package them with simple APIs, clear documentation, and visualization tools.
Download a FREE copy of our recent NLP Industry Survey Results: https://gradientflow.com/2021nlpsurvey/
Subscribe: Apple • Android • Spotify • Stitcher • Google • AntennaPod • RSS.
Detailed show notes can be found on The Data Exchange web site.
Get every episode summarized
Each time The Data Exchange with Ben Lorica publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.
Email me new episodesFree for 3 shows. No card needed.
Hosts & guests
No transcript yet
This episode has not been transcribed. Request it and it moves to the front of the queue.
More episodes
More from The Data Exchange with Ben Lorica

Your AI Agent Is Costing You More Than You Think
The Data Exchange with Ben Lorica

Reasoning Doesn't Start With Language
The Data Exchange with Ben Lorica

An Agent Is Just an LLM in a For-Loop
The Data Exchange with Ben Lorica

The Bloomberg Terminal for AI Compute
The Data Exchange with Ben Lorica