Introduction to Reinforcement Learning
Explore this paper's citation graph
Summary
In Reinforcement Learning, Richard Sutton and Andrew Barto provide a clear and simple account of the key ideas and algorithms of reinforcement learning.
- Type
- book
- Published
- 1998-03-01
- Cited by
- 3,299
- References
- 0
- OpenAlex
- https://openalex.org/W1515851193
- Semantic Scholar
- https://api.semanticscholar.org/CorpusID:261579713
Keywords
Reinforcement learning, Reinforcement, Key (lock), Simple (philosophy), Field (mathematics)
References
No references recorded for this paper.
Cited by
- Machine Learning for Adaptive Image Interpretation
- Rational and Convergent Learning in Stochastic Games
- Efficient No-Regret Multiagent Learning
- Reinforcement Learning for Context Aware Segmentation
- Toward Off-Policy Learning Control with Function Approximation
- Explanation-Based Acquisition of Planning Operators
- Design of multi agent adaptive neuro-fuzzy based intelligent controllers for multi-objective nonlinear system
- Chained learning architectures in a simple closed-loop behavioural context
- Agent Learning using Action-Dependent Learning Rates in Computer Role-Playing Games
- Posterior Weighted Reinforcement Learning with State Uncertainty
- Graph partitioning techniques for Markov Decision Processes decomposition
- Toward Nonlinear Local Reinforcement Learning Rules Through Neuroevolution
- Frontal Theta Oscillatory Activity Is a Common Mechanism for the Computation of Unexpected Outcomes and Learning Rate
- Adaptive Modeling and Planning for Reactive Agents
- Utile Distinctions for Relational Reinforcement Learning
- Stochastic Enforced Hill-Climbing
- Examining Extended Dynamic Scripting in a Tactical Game Framework
- Using Homomorphisms to Transfer Options across Continuous Reinforcement Learning Domains
- Exploration strategies in n-Person general-sum multiagent reinforcement learning with sequential action selection
- On Policy Learning in Restricted Policy Spaces
Related papers
- Technical Note: Q-Learning
- Learning from delayed rewards
- Reinforcement learning
- Dynamic Programming
- Markov Decision Processes: Discrete Stochastic Dynamic Programming
- Mastering the game of Go with deep neural networks and tree search
- Policy Gradient Methods for Reinforcement Learning with Function Approximation
- Human-level control through deep reinforcement learning
- Least-Squares Policy Iteration
- Artificial intelligence: a modern approach