Who Cited It

Reinforcement learning in robotics: A survey

2013 · The International Journal of Robotics Research · 3,154 citations · 7 from inside this corpus

Jens Kober, J. Andrew Bagnell, Jan Peters

Reinforcement learning offers to robotics a framework and set of tools for the design of sophisticated and hard-to-engineer behaviors. Conversely, the challenges of robotic problems provide both inspiration, impact, and validation for developments in reinforcement learning. The relationship between disciplines has sufficient promise to be likened to that between physics and mathematics. In this article, we attempt to strengthen the links between the two research communities by providing a survey of work in reinforcement learning for behavior generation in robots. We highlight both key challenges in robot reinforcement learning as well as notable successes. We discuss how contributions tamed the complexity of the domain and study the role of algorithms, representations, and prior knowledge in achieving these successes. As a result, a particular focus of our paper lies on the choice between model-based and model-free as well as between value-function-based and policy-search methods. By analyzing a simple problem in some detail we demonstrate how reinforcement learning approaches may be profitably applied, and we note throughout open questions and the tremendous potential for future research.

Reinforcement learning in robotics: A survey (2013)Reinforcement learning in rob…Genetic Algorithms (2002)Genetic AlgorithmsGaussian Processes for Machine Learning (2005)Gaussian Processes for Machin…Reinforcement Learning: A Survey (1996)Reinforcement Learning: A Sur…Markov Decision Processes: Discrete Stochastic Dynamic Programming. (1995)Markov Decision Processes: Di…Pattern Recognition and Machine Learning (Information Science and Statistics) (2006)Pattern Recognition and Machi…Simple statistical gradient-following algorithms for connectionist reinforcement learning (1992)Simple statistical gradient-f…Introduction to Reinforcement Learning (1998)Introduction to Reinforcement…Markov Decision Processes (1994)Markov Decision ProcessesPolicy Gradient Methods for Reinforcement Learning with Function Approximation (1999)Policy Gradient Methods for R…A survey of robot learning from demonstration (2008)A survey of robot learning fr…Apprenticeship learning via inverse reinforcement learning (2004)Apprenticeship learning via i…International Conference on Machine Learning (ICML)-2005 (2006)International Conference on M…Advances in Neural Information Processing Systems 27 (NIPS 2014) (2014)Advances in Neural Informatio…Maximum Entropy Inverse Reinforcement Learning (2018)Maximum Entropy Inverse Reinf…Introduction to Stochastic Search and Optimization (2003)Introduction to Stochastic Se…Gaussian Processes for Machine Learning (Adaptive Computation and Machine Learning) (2005)Gaussian Processes for Machin…Advances in Neural Information Processing Systems 19 (NIPS 2006) (2006)Advances in Neural Informatio…Adaptive Computation and Machine Learning (2012)Adaptive Computation and Mach…Integrated Architectures for Learning, Planning, and Reacting Based on Approximating Dyna… (1990)Mastering the game of Go without human knowledge (2017)Mastering the game of Go with…How artificial intelligence will change the future of marketing (2019)How artificial intelligence w…End-to-end training of deep visuomotor policies (2016)End-to-end training of deep v…A State-of-the-Art Survey on Deep Learning Theory and Architectures (2019)A State-of-the-Art Survey on …Target-driven visual navigation in indoor scenes using deep reinforcement learning (2017)Target-driven visual navigati…Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates (2017)Deep reinforcement learning f…End-to-End Training of Deep Visuomotor Policies (2015)End-to-End Training of Deep V…
26 of 26 neighbouring works in this corpus. Blue is what this paper cites; orange is what cites it, and a dashed line is one neighbour citing another. Only the largest labels are drawn — every node carries its full title on hover.
this paper works it cites works citing it node size = global citations · hover for the full title

What this paper cites, inside the corpus

What cites it, inside the corpus

Topics

Reinforcement Learning in RoboticsComputer Science
Evolutionary Algorithms and ApplicationsComputer Science
Advanced Multi-Objective Optimization AlgorithmsComputer Science

Is this record sound?

complete

Nothing in this record contradicts itself and no field we check is missing.

  • supports3 author record(s) attached.
  • supports262 reference(s) recorded.
  • neutralThe DOI carries no year to check against.
  • supportsA title is present.

Provenance

Everything above was read from one stored OpenAlex payload, fetched 2026-09-04T03:58:47+00:00.

sha256 a08467ae9504f237…