CORTEXA
← Browse

Hexian Ni

1 paper indexed

arxivcs.RO2026-07-02

CoRe: Combined Rewards with Vision-Language Model Feedback for Preference-Aligned Reinforcement Learning

Hexian Ni, Tao Lu, Yinghao Cai

Reward design remains a central challenge in reinforcement learning (RL). Hand-crafted rewards are often difficult to specify and may lead to suboptimal policies, while learned rewards from preferences can suffer from inefficiency and unstable training. Inspired by the dual natur…

View free PDFSource page