The Artificial PostAccount
All papers
Robotics · ORIGINAL RESEARCH

Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor

Soft Actor-Critic combines off-policy learning with an objective that rewards both successful actions and policy entropy.

Opening the paper…

KEEP FOLLOWING THE IDEA

Meet the researchers.

Pieter AbbeelSergey Levine