CORTEXA
← Browse

Zesheng Shi

1 paper indexed

arxivcs.AI2026-07-04

Agent Reinforcement Learning via Pivotal-Aware Self-Feedback Retry

Weiyang Guo, Zesheng Shi, Longhui Zhang, Zeen Zhu, Min Zhang, Jing Li

Large language model (LLM) agents have shown strong decision-making capabilities in long-horizon interactive tasks, yet they still struggle to effectively leverage failed trajectories: full retries incur high interaction costs, while experience retrieval tends to dilute critical…

View free PDFSource page