article Open AccessTop 1% cited
Learning to Predict by the Methods of Temporal Differences
Machine Learning · 1988 · Vol. 3(1) · pp. 9–44
Evolutionary Algorithms and ApplicationsMetaheuristic Optimization Algorithms ResearchArtificial Intelligence in GamesTemporal difference learningComputer scienceArtificial intelligenceMachine learningHeuristicsortSupervised learningConvergence (economics)Reinforcement learningArtificial neural network
Funding
- Air Force Office of Scientific Research
Citations
3,908
FWCI
60.24
field-weighted impact
References
50
Percentile
100%
vs. same field & year
Citations per year
Cited by
A framework for mesencephalic dopamine systems based on predictive Hebbian learning
Journal of Neuroscience · 1996 · 2,093 citations
Near-Optimal Reinforcement Learning in Polynomial Time
Machine Learning · 2002 · 851 citations
Practical issues in temporal difference learning
Machine Learning · 1992 · 795 citations
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Machine Learning · 1992 · 7,410 citations
Technical Note: Q-Learning
Machine Learning · 1992 · 3,640 citations
Self-improving reactive agents based on reinforcement learning, planning and teaching
Machine Learning · 1992 · 1,641 citations
TD-Gammon, a Self-Teaching Backgammon Program, Achieves Master-Level Play
Neural Computation · 1994 · 789 citations
Temporal difference learning and TD-Gammon
Communications of the ACM · 1995 · 1,472 citations
References
Dynamic Programming
Science · 1966 · 13,052 citations
Matrix Iterative Analysis
Mathematics of Computation · 1963 · 4,153 citations
Citation Network
How this paper connects to the literature. Drag to explore, click any node to open that paper.
