Example: air traffic controller
Deep Reinforcement Learning with Double Q-learning
tion. This is the idea behind Double Q-learning (van Hasselt, 2010). In the original Double Q-learning algorithm, two value functions are learned by assigning each experience ran-domly to update one of the two value functions, such that there are two sets of …
Download Deep Reinforcement Learning with Double Q-learning
Information
Domain:
Source:
Link to this page: