arxivcs.LGcs.AI2026-07-09
Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning
Ali Larian, Qian Lin, Chang Zong Wu, Daniel S. Brown
As autonomous agents are increasingly deployed across diverse operational contexts, aligning their behavior with human intent demands reward functions that remain robust to such changes rather than overfitting to any single environment. Inverse reinforcement learning (IRL) provid…