Analysis of Inverse Reinforcement Learning with Perturbed Demonstrations

doi:https://doi.org/10.3233/978-1-60750-606-5-349

Open AccessBook Chapter10.3233/978-1-60750-606-5-349

Analysis of Inverse Reinforcement Learning with Perturbed Demonstrations

Francisco S. Melo,Manuel Lopes,Ricardo Ferreira-2010-01-01-Frontiers in artificial intelligence and applications

12

TL;DRAbstract

Inverse reinforcement learning (IRL) addresses the problem of recovering the unknown reward function for a given Markov decision problem (MDP) given the corresponding optimal policy or a perturbed version thereof. This paper studies the space of possible solutions to the general IRL problem, when the agent is provided with incomplete/imperfect information regarding the optimal policy for the MDP whose reward must be estimated. We focus on scenarios with finite state-action spaces and discuss the constraints imposed on the set of possible solutions when the agent is provided with (i) perturbed policies; (ii) optimal policies; and (iii) incomplete policies. We discuss previous works on IRL in light of our analysis and show that, with our characterization of the solution space, it is possible to determine non-trivial closed-form solutions for the IRL problem. We also discuss several other interesting aspects of the IRL problem that stem from our analysis.

Chat with Paper

AI Agents for this Paper

Inverse reinforcement learning (IRL) addresses the problem of recovering the unknown reward function for a given Markov decision problem (MDP) given the corresponding optimal policy or a perturbed version thereof. This paper studies the space of possible solutions to the general IRL problem, when the agent is provided with incomplete/imperfect information regarding the optimal policy for the MDP whose reward must be estimated. We focus on scenarios with finite state-action spaces and discuss the constraints imposed on the set of possible solutions when the agent is provided with (i) perturbed policies; (ii) optimal policies; and (iii) incomplete policies. We discuss previous works on IRL in light of our analysis and show that, with our characterization of the solution space, it is possible to determine non-trivial closed-form solutions for the IRL problem. We also discuss several other interesting aspects of the IRL problem that stem from our analysis.

Keywords

InverseReinforcementReinforcement learningComputer scienceArtificial intelligencePsychologyMathematicsSocial psychology

Chat

Click to start Chat