Skip to content
TrackPodcasts
technologyMar 13, 202545:45pending

The Evolution of Reinforcement Fine-Tuning in AI

About this episode

Travis Addair is Co-Founder & CTO at Predibase. In this episode, the discussion centers on transforming pre-trained foundation models into domain-specific assets through advanced customization techniques.

Subscribe to the Gradient Flow Newsletter 📩  https://gradientflow.substack.com/

Support our work by leaving a small tip 💰 https://buymeacoffee.com/gradientflow

Subscribe: Apple · Spotify · Overcast · Pocket Casts · AntennaPod · Podcast Addict · Amazon ·  RSS.

Detailed show notes - with links to many references - can be found on The Data Exchange web site.

Get every episode summarized

Each time The Data Exchange with Ben Lorica publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.

Email me new episodes

Free for 3 shows. No card needed.

Hosts & guests

No transcript yet

This episode has not been transcribed. Request it and it moves to the front of the queue.

The Evolution of Reinforcement Fine-Tuning in AI

The Data Exchange with Ben Lorica

0:00
45:45

More episodes

More from The Data Exchange with Ben Lorica

View all episodes →