article Open AccessTop 1% cited
Learning to predict by the methods of temporal differences
Machine Learning · 1988 · Vol. 3(1) · pp. 9–44
Richard S. Sutton✉(TE Laboratories (Ireland))
Evolutionary Algorithms and ApplicationsArtificial Intelligence in GamesSports Analytics and PerformanceTemporal difference learningComputer scienceMachine learningHeuristicArtificial intelligenceConvergence (economics)Supervised learningsortComputationReinforcement learning
Citations
2,774
FWCI
36.49
field-weighted impact
References
30
Percentile
100%
vs. same field & year
Citations per year
Cited by
Near-Optimal Reinforcement Learning in Polynomial Time
Machine Learning · 2002 · 851 citations
Predictive Reward Signal of Dopamine Neurons
Journal of Neurophysiology · 1998 · 4,543 citations
TD-Gammon, a Self-Teaching Backgammon Program, Achieves Master-Level Play
Neural Computation · 1994 · 789 citations
Making Working Memory Work: A Computational Model of Learning in the Prefrontal Cortex and Basal Ganglia
Neural Computation · 2005 · 1,096 citations
Reinforcement Learning in Continuous Time and Space
Neural Computation · 2000 · 979 citations
Multi-Agent Reinforcement Learning: A Review of Challenges and Applications
Applied Sciences · 2021 · 333 citations
Dynamic Dopamine Modulation in the Basal Ganglia: A Neurocomputational Account of Cognitive Deficits in Medicated and Nonmedicated Parkinsonism
Journal of Cognitive Neuroscience · 2005 · 947 citations
The Linear Programming Approach to Approximate Dynamic Programming
Operations Research · 2003 · 701 citations
References
Matrix Iterative Analysis
Mathematics of Computation · 1963 · 4,153 citations
Citation Network
How this paper connects to the literature. Drag to explore, click any node to open that paper.
