Skip to main content
WebsitebeginnerFree

Sutton & Barto Book (Free PDF)

Unknown

Richard Sutton's page hosting the complete second-edition draft of the standard RL textbook, plus code and errata. It covers bandits, Markov decision processes, temporal-difference learning, function approximation, and policy gradients, giving you the field's theoretical foundation.

Visit resource

More resources on Reinforcement Learning

See all Reinforcement Learning resources →