Skip to main content
BookintermediatePaid

Deep Reinforcement Learning Hands-On

by Maxim Lapan · Maxim Lapan

A PyTorch-based walkthrough of deep RL methods, from deep Q-networks and value iteration through policy gradients, TRPO, and AlphaGo Zero, applied to Atari, stock trading, and chatbots. You finish able to implement and debug these agents.

Visit resource

This link may earn us a small commission at no extra cost to you. Affiliate disclosure

More resources on Reinforcement Learning

See all Reinforcement Learning resources →