Who Cited It

Simple statistical gradient-following algorithms for connectionist reinforcement learning

1992 · Machine Learning · 7,519 citations · 24 from inside this corpus

Ronald J. Williams

No abstract in the source record.

Simple statistical gradient-following algorithms for connectionist reinforcement learning (1992)Simple statistical gradient-f…Learning Internal Representations by Error Propagation (1985)Learning Internal Representat…Parallel Distributed Processing (1986)Parallel Distributed Processi…Learning from Delayed Rewards (1989)Learning from Delayed RewardsLearning to Predict by the Methods of Temporal Differences (1988)Learning to Predict by the Me…Parallel distributed processing: explorations in the microstructure of cognition, vol. 1:… (1986)Parallel distributed processi…Learning Automata: An Introduction (1989)Learning Automata: An Introdu…Forward Models: Supervised Learning with a Distal Teacher (1992)Forward Models: Supervised Le…Deep learning in neural networks: An overview (2014)Deep learning in neural netwo…Mastering the game of Go with deep neural networks and tree search (2016)Mastering the game of Go with…Reinforcement Learning: A Survey (1996)Reinforcement Learning: A Sur…Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks (2017)Model-Agnostic Meta-Learning …Policy Gradient Methods for Reinforcement Learning with Function Approximation (1999)Policy Gradient Methods for R…Deep Reinforcement Learning: A Brief Survey (2017)Deep Reinforcement Learning: …Neural Architecture Search with Reinforcement Learning (2016)Neural Architecture Search wi…Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochast… (2018)Soft Actor-Critic: Off-Policy…Reinforcement learning in robotics: A survey (2013)Reinforcement learning in rob…Recent Trends in Deep Learning Based Natural Language Processing [Review Article] (2018)Recent Trends in Deep Learnin…Automatic differentiation in machine learning: a survey (2017)Automatic differentiation in …Soft Actor-Critic Algorithms and Applications (2018)Soft Actor-Critic Algorithms …Actor-critic algorithms (2002)Actor-critic algorithmsHigh-Dimensional Continuous Control Using Generalized Advantage Estimation (2015)High-Dimensional Continuous C…Deterministic Policy Gradient Algorithms (2014)Deterministic Policy Gradient…Counterfactual Multi-Agent Policy Gradients (2018)Counterfactual Multi-Agent Po…End-to-end training of deep visuomotor policies (2016)End-to-end training of deep v…AutoML: A survey of the state-of-the-art (2020)AutoML: A survey of the state…Asynchronous Methods for Deep Reinforcement Learning (2016)Asynchronous Methods for Deep…Transfer Learning for Reinforcement Learning Domains: A Survey (2009)Transfer Learning for Reinfor…Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates (2017)Deep reinforcement learning f…Hybrid computing using a neural network with dynamic external memory (2016)Hybrid computing using a neur…Neural Architecture Search: A Survey (2018)Neural Architecture Search: A…End-to-End Training of Deep Visuomotor Policies (2015)End-to-End Training of Deep V…
31 of 31 neighbouring works in this corpus. Blue is what this paper cites; orange is what cites it, and a dashed line is one neighbour citing another. Only the largest labels are drawn — every node carries its full title on hover.
this paper works it cites works citing it node size = global citations · hover for the full title

What this paper cites, inside the corpus

What cites it, inside the corpus

PaperYearCited
Deep learning in neural networks: An overview201418,236
Mastering the game of Go with deep neural networks and tree search201616,014
Reinforcement Learning: A Survey19968,953
Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks20175,794
Policy Gradient Methods for Reinforcement Learning with Function Approximation19994,966
Deep Reinforcement Learning: A Brief Survey20174,434
Neural Architecture Search with Reinforcement Learning20163,881
Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochast…20183,500
Reinforcement learning in robotics: A survey20133,154
Recent Trends in Deep Learning Based Natural Language Processing [Review Article]20182,907
Automatic differentiation in machine learning: a survey20172,094
Soft Actor-Critic Algorithms and Applications20181,994
Actor-critic algorithms20021,818
High-Dimensional Continuous Control Using Generalized Advantage Estimation20151,746
Deterministic Policy Gradient Algorithms20141,736
Counterfactual Multi-Agent Policy Gradients20181,727
End-to-end training of deep visuomotor policies20161,706
AutoML: A survey of the state-of-the-art20201,705
Asynchronous Methods for Deep Reinforcement Learning20161,689
Transfer Learning for Reinforcement Learning Domains: A Survey20091,564
Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates20171,468
Hybrid computing using a neural network with dynamic external memory20161,448
Neural Architecture Search: A Survey20181,406
End-to-End Training of Deep Visuomotor Policies20151,400

Topics

Evolutionary Algorithms and ApplicationsComputer Science
stochastic dynamics and bifurcationPhysics and Astronomy
VLSI and FPGA Design TechniquesEngineering

Is this record sound?

complete

Nothing in this record contradicts itself and no field we check is missing.

  • supports1 author record(s) attached.
  • supports41 reference(s) recorded.
  • neutralThe DOI carries no year to check against.
  • supportsA title is present.

Provenance

Everything above was read from one stored OpenAlex payload, fetched 2026-09-04T03:58:43+00:00.

sha256 5cad55ac5d4d41e1…