Transcription of Reinforcement Learning - Lecture 1: Introduction
{{id}} {{{paragraph}}}
Reinforcement LearningLecture 1: IntroductionAlexandre Proutiere, Sadegh Talebi, Jungseul OkKTH, The Royal Institute of TechnologyLecture 1: Outline1. Generic models for sequential decision making2. Overview and schedule of the course2 Lecture 1: models for sequential decision making2. Overview and schedule of the course3 Sequential Decision a sequential action selection / control policymaximising rewards4 Sequential Decision MakingProblem definition1. System dynamics2. Set of available policies available information or feedback to thedecision maker3. Reward structure5 Applications6 Sequential Decision few examples: Linear:st+1=Ast+Bat Deterministic and stationary:st+1=F(st,at) Markovian:P(st+1=s |ht,st=s,at=a) =pt(s |s,a)where s pt(s |s,a) = 1; homogenous ifpt(s |s,a) =p(s |s,a)7 Sequential Decision MakingInformation - Set of few examples: Markov Decision Process (MDP)- Fully observable state and reward- Known reward distribution
Reinforcement Learning Lecture 1: Introduction Alexandre Proutiere, Sadegh Talebi, Jungseul Ok KTH, The Royal Institute of Technology
Domain:
Source:
Link to this page:
Please notify us if you found a problem with this document:
{{id}} {{{paragraph}}}
190: Reinforcement Learning: An, 190: Reinforcement Learning: An Introduction, Reinforcement learning, Brief Introduction to Reinforcement Learning, REINFORCEMENT LEARNING: AN INTRODUCTION, Introduction, Reinforcement Learning and Control, Learning, Introduction to reinforcement learning, Reinforcement Learning. Richard S. Sutton