article Open AccessTop 10% cited
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Machine Learning · 1992 · Vol. 8(3-4) · pp. 229–256
Ronald J. Williams✉(Northeastern University)
Evolutionary Algorithms and Applicationsstochastic dynamics and bifurcationVLSI and FPGA Design TechniquesReinforcement learningConnectionismComputer scienceSimple (philosophy)ReinforcementAssociative propertyArtificial intelligenceAlgorithmArtificial neural networkBackpropagation
Funding
- National Science Foundation
Citations
7,410
FWCI
14.43
field-weighted impact
References
41
Percentile
99%
vs. same field & year
Citations per year
Cited by
Multi-Agent Deep Reinforcement Learning for Large-Scale Traffic Signal Control
IEEE Transactions on Intelligent Transportation Systems · 2019 · 940 citations
The Helmholtz Machine
Neural Computation · 1995 · 1,207 citations
DeepMimic
ACM Transactions on Graphics · 2018 · 784 citations
Space/Aerial-Assisted Computing Offloading for IoT Applications: A Learning-Based Approach
IEEE Journal on Selected Areas in Communications · 2019 · 846 citations
Deep reinforcement learning for de novo drug design
Science Advances · 2018 · 1,110 citations
Multi-Agent Reinforcement Learning: A Review of Challenges and Applications
Applied Sciences · 2021 · 333 citations
A review on the attention mechanism of deep learning
Neurocomputing · 2021 · 3,096 citations
Reinforcement learning of motor skills with policy gradients
Neural Networks · 2008 · 854 citations
References
Adaptive filtering, prediction and control
Automatica · 1985 · 4,488 citations
Learning to Predict by the Methods of Temporal Differences
Machine Learning · 1988 · 3,908 citations
Citation Network
How this paper connects to the literature. Drag to explore, click any node to open that paper.
