arxivcs.LGcs.AI2026-07-23
Relative Value Learning
Marc Höftmann, Jan Robine, Stefan Harmeling
In reinforcement learning, critics typically estimate absolute state values $V(s)$, estimating how good a particular situation is in isolation. However, it turns out that only differences in value are relevant for control. Motivated by this, we propose Relative Value Learning (RV…