Skip to player
Home
Model Building & Alignment
›
Post-Training & Alignment
›
RLHF & Preference Optimization
Contents
CC
Host
Expert
Murali Chillakuru
Press play to begin the conversation.
0:00
10:54
1×
1.25×
1.5×
0.85×
Post-Training & Alignment · 2 / 5
RLHF & Preference Optimization
Play
Back to browse
Up next
DPO & Simpler Alignment
Play now
Stay (
5
)