arxivcs.AIcs.LG2026-06-29
Exploration and Online Transfer with Behavioral Foundation Models
Louis Bagot, Mathieu Lefort, Laëtitia Matignon
Zero-shot Transfer in Reinforcement Learning (RL) aims to train an agent that can generate optimal policies for any reward function, without additional learning at transfer time, while training only on reward-free trajectories. For their generality over tasks, such models are som…