Premium ContentImitation and Inverse RL
Learning from demonstrations and inferring rewards
This chapter requires a subscription to access.
What you'll unlock:
- 1. Behavior Cloning
- 2. Distribution Shift and the BC Failure Mode
- 3. DAgger: Dataset Aggregation
- 4. Maximum-Entropy Inverse RL
- 5. GAIL and AIRL
Subscribe to UnlockAlready have an account? Sign in