PDF4PRO ⚡AMP

Modern search engine that looking for books and documents around the web

Example: barber

Reinforcement Learning - Lecture 1: Introduction

Reinforcement LearningLecture 1: IntroductionAlexandre Proutiere, Sadegh Talebi, Jungseul OkKTH, The Royal Institute of TechnologyLecture 1: Outline1. Generic models for sequential decision making2. Overview and schedule of the course2 Lecture 1: models for sequential decision making2. Overview and schedule of the course3 Sequential Decision a sequential action selection / control policymaximising rewards4 Sequential Decision MakingProblem definition1. System dynamics2. Set of available policies available information or feedback to thedecision maker3. Reward structure5 Applications6 Sequential Decision few examples: Linear:st+1=Ast+Bat Deterministic and stationary:st+1=F(st,at) Markovian:P(st+1=s |ht,st=s,at=a) =pt(s |s,a)where s pt(s |s,a) = 1; homogenous ifpt(s |s,a) =p(s |s,a)7 Sequential Decision MakingInformation - Set of few examples: Markov Decision Process (MDP)- Fully observable state and reward- Known reward distribution

Reinforcement Learning Lecture 1: Introduction Alexandre Proutiere, Sadegh Talebi, Jungseul Ok KTH, The Royal Institute of Technology

Loading..

Tags:

  Lecture, Introduction, Learning, Reinforcement, Reinforcement learning lecture 1

Information

Domain:

Source:

Link to this page:

Please notify us if you found a problem with this document:

Spam in document Broken preview Other abuse

Transcription of Reinforcement Learning - Lecture 1: Introduction

Related search queries