The Artificial PostAccount
All papers
Reinforcement learning · ORIGINAL RESEARCH

Scaling Laws for Reward Model Overoptimization

Selected research in reinforcement learning. Read the full paper, including the methods, experiments, and reported results.

Opening the paper…

KEEP FOLLOWING THE IDEA

Meet the researchers.

John Schulman