The Artificial PostAccount
All papers
Alignment · ORIGINAL RESEARCH

Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Direct Preference Optimization derives a preference-learning objective that avoids training a separate reward model in the studied setup.

Opening the paper…

KEEP FOLLOWING THE IDEA

Meet the researchers.

Christopher D. ManningChelsea FinnStefano Ermon