The Artificial PostAccount
All papers
Alignment · ORIGINAL RESEARCH

Training language models to follow instructions with human feedback

InstructGPT studies fine-tuning language models using demonstrations and human preference judgments to better follow instructions.

Opening the paper…

KEEP FOLLOWING THE IDEA

Meet the researchers.

John SchulmanAmanda Askell