Transcription of Bonus Lecture: Introduction to Reinforcement Learning
{{id}} {{{paragraph}}}
Bonus lecture : Introduction to Reinforcement Learning Garima Lalwani, Karan Ganju and Unnat Jain Credits: These slides and images are borrowed from slides by David Silver and Peter Abbeel Outline 1 RL Problem Formulation 2 Model-based Prediction and Control 3 Model-free Prediction 4 Model-free Control 5 Summary Part 1: RL Problem Formulation Characteristics of Reinforcement Learning What makes Reinforcement Learning different from other machine Learning paradigms? There is no supervisor, only a reward signal Feedback is delayed, not instantaneous Time really matters (correlated, non data). Agent's actions affect the subsequent data it receives Agent and Environment Agent Observed state action St At reward Rt Environment Rewards A reward Rt is a scalar feedback signal Indicates how well agent is doing at step t The agent's job is to maximise cumulative reward Rod Balancing Demo Learn to swing up and
Bonus Lecture: Introduction to Reinforcement Learning Garima Lalwani, Karan Ganju and Unnat Jain Credits: These slides and images are borrowed from slides by David Silver and Peter Abbeel
Domain:
Source:
Link to this page:
Please notify us if you found a problem with this document:
{{id}} {{{paragraph}}}
190: Reinforcement Learning: An, 190: Reinforcement Learning: An Introduction, Reinforcement learning, Brief Introduction to Reinforcement Learning, REINFORCEMENT LEARNING: AN INTRODUCTION, Introduction, Reinforcement Learning and Control, Learning, Introduction to reinforcement learning, Reinforcement Learning. Richard S. Sutton