article Open AccessTop 1% cited
Q-learning
Machine Learning · 1992 · Vol. 8(3-4) · pp. 279–292
Christopher J. Watkins✉(The FRAM Centre)Peter Dayan(University of Edinburgh)
Reinforcement Learning in RoboticsSketchAction (physics)Markov processMarkov decision processConvergence (economics)Computer scienceQ-learningSimple (philosophy)Mathematical optimizationDynamic programming
Citations
8,916
FWCI
19.43
field-weighted impact
References
14
Percentile
99%
vs. same field & year
Citations per year
Cited by
Near-Optimal Reinforcement Learning in Polynomial Time
Machine Learning · 2002 · 851 citations
Multi-Agent Deep Reinforcement Learning for Large-Scale Traffic Signal Control
IEEE Transactions on Intelligent Transportation Systems · 2019 · 940 citations
Path planning and obstacle avoidance for AUV: A review
Ocean Engineering · 2021 · 331 citations
Reinforcement learning for demand response: A review of algorithms and modeling techniques
Applied Energy · 2018 · 770 citations
Optimized Computation Offloading Performance in Virtual Edge Computing Systems Via Deep Reinforcement Learning
IEEE Internet of Things Journal · 2018 · 681 citations
On the Convergence of Stochastic Iterative Dynamic Programming Algorithms
Neural Computation · 1994 · 796 citations
Multi-Agent Reinforcement Learning: A Review of Challenges and Applications
Applied Sciences · 2021 · 333 citations
Multiagent Reinforcement Learning for Integrated Network of Adaptive Traffic Signal Controllers (MARLIN-ATSC): Methodology and Large-Scale Application on Downtown Toronto
IEEE Transactions on Intelligent Transportation Systems · 2013 · 509 citations
References
Self-improving reactive agents based on reinforcement learning, planning and teaching
Machine Learning · 1992 · 1,641 citations
Citation Network
How this paper connects to the literature. Drag to explore, click any node to open that paper.
