Example: barber
Deep Reinforcement Learning with Double Q-learning

Deep Reinforcement Learning with Double Q-learning

Back to document page

using Q-learning (Watkins, 1989), a form of temporal dif-ference learning (Sutton, 1988). Most interesting problems are too large to learn all action values in all states sepa-rately. Instead, we can learn a parameterized value function Q(s;a; t). The standard Q-learning update for the param-eters after taking action At in state St and ...

  Learning, Q learning

Download Deep Reinforcement Learning with Double Q-learning


Information

Domain:

Source:

Link to this page:

Please notify us if you found a problem with this document:

Other abuse

Advertisement

Related search queries