Example: tourism industry
Maximum Entropy Inverse Reinforcement Learning

Maximum Entropy Inverse Reinforcement Learning

Back to document page

sition distribution, T. Paths in these MDPs (Figure 1d) are now determined by the action choices of the agent and the random outcomes of the MDP. Our distribution over paths must take this randomness into account. We use the maximum entropy distribution of paths con-ditioned on the transition distribution, T, and constrained to

  Distribution, Learning, Maximum, Reinforcement, Inverse, Entropy, Maximum entropy inverse reinforcement learning

Download Maximum Entropy Inverse Reinforcement Learning


Information

Domain:

Source:

Link to this page:

Please notify us if you found a problem with this document:

Other abuse

Advertisement

Related search queries