Premium Content

Imitation and Inverse RL

Learning from demonstrations and inferring rewards

This chapter requires a subscription to access.

What you'll unlock:

  • 1. Behavior Cloning
  • 2. Distribution Shift and the BC Failure Mode
  • 3. DAgger: Dataset Aggregation
  • 4. Maximum-Entropy Inverse RL
  • 5. GAIL and AIRL
Subscribe to Unlock

Already have an account? Sign in