arxivcs.LGcs.AI2026-07-21
From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning
Garvit Singla, Uma Maheswari Natarajan, Raghuram Bharadwaj Diddigi
Model-Agnostic Meta-Learning (MAML) is a widely used framework for reinforcement learning (RL) that enables efficient transfer by learning global policy parameters that can be rapidly adapted to new tasks. MAML training proceeds in two loops: an inner loop where the global parame…