articleTop 1% cited
Technical Note: Q-Learning
Machine Learning · 1992 · Vol. 8(3-4) · pp. 279–292
Christopher J. Watkins✉(Highbury Hospital)Peter Dayan(University of Edinburgh)
Reinforcement Learning in RoboticsMachine Learning and AlgorithmsData Stream Mining TechniquesSketchAction (physics)Q-learningConvergence (economics)Markov processSimple (philosophy)MathematicsMarkov decision processDynamic programmingDiscrete mathematics
Citations
3,640
FWCI
27.75
field-weighted impact
References
17
Percentile
100%
vs. same field & year
Citations per year
Cited by
Near-Optimal Reinforcement Learning in Polynomial Time
Machine Learning · 2002 · 851 citations
Path planning and obstacle avoidance for AUV: A review
Ocean Engineering · 2021 · 331 citations
Reinforcement learning for demand response: A review of algorithms and modeling techniques
Applied Energy · 2018 · 770 citations
Ten simple rules for the computational modeling of behavioral data
eLife · 2019 · 695 citations
On-Line Building Energy Optimization Using Deep Reinforcement Learning
IEEE Transactions on Smart Grid · 2018 · 609 citations
Deep learning in neural networks: An overview
Neural Networks · 2014 · 17,774 citations
Human-level control through deep reinforcement learning
Nature · 2015 · 29,167 citations
From model-based control to data-driven control: Survey, classification and perspective
Information Sciences · 2012 · 1,321 citations
References
Learning to Predict by the Methods of Temporal Differences
Machine Learning · 1988 · 3,908 citations
Self-improving reactive agents based on reinforcement learning, planning and teaching
Machine Learning · 1992 · 1,641 citations
Citation Network
How this paper connects to the literature. Drag to explore, click any node to open that paper.
