Playing Atari with Deep Reinforcement Learning
Selected research in reinforcement learning. Read the full paper, including the methods, experiments, and reported results.
Paper & contextLearning decisions through interaction and search.
Silver studies how agents learn to act through experience, rewards, and planning. These papers follow deep reinforcement learning from Atari to game-playing systems that combine learned models with search, alongside work on improving training and evaluation.
10 papers
Selected research in reinforcement learning. Read the full paper, including the methods, experiments, and reported results.
Paper & contextDouble Q-learning uses two value estimates to reduce the tendency of standard Q-learning to overestimate action values, and combines this with deep networks for learning from high-dimensional observations.
Paper & contextSelected research in reinforcement learning. Read the full paper, including the methods, experiments, and reported results.
Paper & contextSelected research in ai for science, reinforcement learning. Read the full paper, including the methods, experiments, and reported results.
Paper & contextSelected research in ai for science, reinforcement learning. Read the full paper, including the methods, experiments, and reported results.
Paper & contextSelected research in ai for science, reinforcement learning. Read the full paper, including the methods, experiments, and reported results.
Paper & contextSelected research in ai for science, reinforcement learning. Read the full paper, including the methods, experiments, and reported results.
Paper & contextSelected research in reinforcement learning, efficient ai. Read the full paper, including the methods, experiments, and reported results.
Paper & contextSelected research in ai for science, reinforcement learning. Read the full paper, including the methods, experiments, and reported results.
Paper & contextSelected research in ai for science, reinforcement learning. Read the full paper, including the methods, experiments, and reported results.
Paper & contextAn independent editorial profile. Inclusion does not imply Council membership or endorsement. Research is collaborative; coauthorship does not imply sole credit.
Selection & sources