使用Player FM应用程序离线!
Fine-tuning and Preference Alignment in a Single Streamlined Process
Manage episode 423374192 series 2570898
Jiwoo Hong and Noah Lee of KAIST AI are co-authors of ORPO: Monolithic Preference Optimization without Reference Model.
Subscribe to the Gradient Flow Newsletter: https://gradientflow.substack.com/
Subscribe: Apple • Spotify • Overcast • Pocket Casts • AntennaPod • Podcast Addict • Amazon • RSS.
Detailed show notes can be found on The Data Exchange web site.
278集单集
Manage episode 423374192 series 2570898
Jiwoo Hong and Noah Lee of KAIST AI are co-authors of ORPO: Monolithic Preference Optimization without Reference Model.
Subscribe to the Gradient Flow Newsletter: https://gradientflow.substack.com/
Subscribe: Apple • Spotify • Overcast • Pocket Casts • AntennaPod • Podcast Addict • Amazon • RSS.
Detailed show notes can be found on The Data Exchange web site.
278集单集
Tutti gli episodi
×欢迎使用Player FM
Player FM正在网上搜索高质量的播客,以便您现在享受。它是最好的播客应用程序,适用于安卓、iPhone和网络。注册以跨设备同步订阅。