Continuous control with deep reinforcement learning
Explore this paper's citation graph
- Type
- article
- Published
- 2016-07-22
- Cited by
- 6,787
- References
- 31
- Access
- Open access
- OpenAlex
- https://openalex.org/W2963864421
Keywords
Reinforcement learning, Computer science, Domain (mathematical analysis), Artificial intelligence, Action (physics)
References
- Playing Atari with Deep Reinforcement Learning
- Reinforcement learning in feedback control
- Actor–Critic-Based Optimal Tracking for Partially Unknown Nonlinear Discrete-Time Systems
- Autonomous reinforcement learning with experience replay.
- A Survey on Policy Search for Robotics
- On the Theory of the Brownian Motion
- Active perception and recognition learning system based on Actor-Q architecture
- Synthesis and stabilization of complex behaviors through online trajectory optimization
- Real-time reinforcement learning by sequential Actor-Critics and experience replay
- PILCO: A Model-Based and Data-Efficient Approach to Policy Search
- Human-level control through deep reinforcement learning
- States versus Rewards: Dissociable neural prediction error signals underlying model-based and model-free reinforcement learning
- Deep Sparse Rectifier Neural Networks
- MuJoCo: A physics engine for model-based control
- ImageNet classification with deep convolutional neural networks
- Deterministic Policy Gradient Algorithms
- Evolving deep unsupervised convolutional networks for vision-based reinforcement learning
- A generalized iterative LQG method for locally-optimal feedback control of constrained nonlinear stochastic systems
- Actor-Critic Reinforcement Learning with Energy-Based Policies
- End-to-end training of deep visuomotor policies
Cited by
- PLATO: Policy learning using adaptive trajectory optimization
- Learning hand-eye coordination for robotic grasping with deep learning and large-scale data collection
- Towards Deep Developmental Learning
- Learning in constrained stochastic dynamic potential games
- Looking at Humans in the Age of Self-Driving and Highly Automated Vehicles
- Vision-based behavior for UAV reactive avoidance by using a reinforcement learning method
- Deep reinforcement learning for tensegrity robot locomotion
- Reset-free guided policy search: Efficient deep reinforcement learning with stochastic initial states
- Path integral guided policy search
- Collective robot reinforcement learning with distributed asynchronous guided policy search
- Improved deep reinforcement learning for robotics through distribution-based experience retention
- A general adaptive dynamic programming approach with experience replay
- Convolutional fitted Q iteration for vision-based control problems
- Bayesian Optimization with Robust Bayesian Neural Networks
- Nonparametric General Reinforcement Learning
- Learning control policy for parcel singulation
- Approximate Newton Methods for Policy Search in Markov Decision Processes
- The Predictron: End-To-End Learning and Planning
- Stable reinforcement learning with autoencoders for tactile and visual data
- Adaptive Neural Network Control of AUVs With Control Input Nonlinearities Using Reinforcement Learning
Related papers
- End-to-end training of deep visuomotor policies
- Addressing Function Approximation Error in Actor-Critic Methods
- Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor
- Mastering the game of Go without human knowledge
- Proximal Policy Optimization Algorithms
- Mastering the game of Go with deep neural networks and tree search