Scinovex
articleTop 1% cited

Reinforcement Learning: An Introduction

IEEE Transactions on Neural Networks · 2005 · Vol. 16(1) · pp. 285–286

Abstract

Reinforcement learning, one of the most active research areas in artificial intelligence, is a computational approach to learning whereby an agent tries to maximize the total amount of reward it receives when interacting with a complex, uncertain environment. In Reinforcement Learning, Richard Sutton and Andrew Barto provide a clear and simple account of the key ideas and algorithms of reinforcement learning. Their discussion ranges from the history of the field's intellectual foundations to the most recent developments and applications. The only necessary mathematical background is familiarity with elementary concepts of probability. The book is divided into three parts. Part I defines the reinforcement learning problem in terms of Markov decision processes. Part II provides basic solution methods: dynamic programming, Monte Carlo methods, and temporal-difference learning. Part III presents a unified view of the solution methods and incorporates artificial neural networks, eligibility traces, and planning; the two final chapters present case studies and consider the future of reinforcement learning.

Reinforcement Learning in RoboticsData Stream Mining TechniquesEvolutionary Algorithms and ApplicationsReinforcement learningComputer scienceArtificial intelligenceReinforcementMachine learningEngineering
Citations
25,702
FWCI
1306.58
field-weighted impact
References
82
Percentile
100%
vs. same field & year
Citations per year
Cited by
Near-Optimal Reinforcement Learning in Polynomial Time
Machine Learning · 2002 · 851 citations
Artificial Intelligence in Surgery: Promises and Perils
Annals of Surgery · 2018 · 1,261 citations
Semisupervised Deep Reinforcement Learning in Support of IoT and Smart City Services
IEEE Internet of Things Journal · 2017 · 429 citations
Multi-Agent Deep Reinforcement Learning for Large-Scale Traffic Signal Control
IEEE Transactions on Intelligent Transportation Systems · 2019 · 940 citations
Online learning: A comprehensive survey
Neurocomputing · 2021 · 551 citations
Habits, Rituals, and the Evaluative Brain
Annual Review of Neuroscience · 2008 · 1,657 citations
References
Dynamic Programming and Markov Processes.
Journal of the American Statistical Association · 1961 · 3,447 citations
Purposive Behavior in Animals and Men
The American Journal of Psychology · 1950 · 2,470 citations
Learning to Predict by the Methods of Temporal Differences
Machine Learning · 1988 · 3,908 citations
Dynamic Programming
Science · 1966 · 13,052 citations
Genetic algorithms in search, optimization, and machine learning
Choice Reviews Online · 1989 · 49,283 citations
Citation Network

How this paper connects to the literature. Drag to explore, click any node to open that paper.