CORTEXA
← Browse

Seungjun Oh

1 paper indexed

arxivcs.LGcs.AI2026-07-01

Learning Generalizable Skill Policy with Data-Efficient Unsupervised RL

Jongchan Park, Seungjun Oh, Seungho Baek, Yusung Kim

Unsupervised Reinforcement Learning (URL) aims to pre-train scalable, skill-conditioned policies without extrinsic rewards, serving as a foundation for downstream control tasks. Despite recent progress, we argue that current off-policy URL methods are limited by two critical, ove…

View free PDFSource page