Skip to content
TrackPodcasts
technologyJun 13, 202435:32pending

Fine-tuning and Preference Alignment in a Single Streamlined Process

About this episode

Jiwoo Hong and  Noah Lee of KAIST AI are co-authors of ORPO: Monolithic Preference Optimization without Reference Model

Subscribe to the Gradient Flow Newsletterhttps://gradientflow.substack.com/

Subscribe: AppleSpotify OvercastPocket CastsAntennaPodPodcast AddictAmazon •  RSS.

Detailed show notes can be found on The Data Exchange web site.

Get every episode summarized

Each time The Data Exchange with Ben Lorica publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.

Email me new episodes

Free for 3 shows. No card needed.

Hosts & guests

No transcript yet

This episode has not been transcribed. Request it and it moves to the front of the queue.

Fine-tuning and Preference Alignment in a Single Streamlined Process

The Data Exchange with Ben Lorica

0:00
35:32

More episodes

More from The Data Exchange with Ben Lorica

View all episodes →