The Artificial PostAccount
All papers
Reinforcement learning · ORIGINAL RESEARCH

Proximal Policy Optimization Algorithms

Proximal Policy Optimization proposes a policy-gradient objective that limits overly large updates while simplifying implementation.

Opening the paper…

KEEP FOLLOWING THE IDEA

Meet the researchers.

Alec RadfordJohn Schulman