Learning Causally Invariant Reward Functions from Diverse Demonstrations
- Type
- Preprint / manuscript
- Venue / status
- Independent preprint
Status: Preprint.
Summary: Studies reward learning under environment shifts and proposes learning reward structure that is stable across diverse demonstrations.