CORTEXA
← Browse

Qitai Tan

3 papers indexed

arxivcs.AI2026-07-23

PATS: Policy-Aware Training Scaffolding for Agentic Reinforcement Learning

Yipeng Shi, Zhipeng Ma, Yue Wang, Qitai Tan, Yang Li, Peng Chen, et al.

In long-horizon LLM agent reinforcement learning, weak policies often repeat similar failures, producing uninformative rollout trajectories and limiting effective policy optimization. Existing skill-centric methods improve exploration by optimizing, filtering, or internalizing re…

View free PDFSource page
arxivcs.LG2026-07-10

GatedLinear: Adaptive Routing of Complementary Linear Bases for Time Series Forecasting

Qitai Tan, Ruiwen Gu, Yilin Su, Mo Li, Xu Lin, Xiao-Ping Zhang

Time series forecasting requires models to capture diverse, often mutually exclusive, temporal dynamics, from smooth trend continuation to nonstationary drift and strict phase-aligned recurrence. While recent deep learning models have improved accuracy, they typically force these…

View free PDFSource page
arxivcs.AI2026-06-26

ATOD: Annealed Turn-aware On-policy Distillation for Multi-turn Autonomous Agents

Qitai Tan, Zefang Zong, Yang Li, Peng Chen

Training small language-model agents for long-horizon interactive tasks requires both fast imitation and reward-driven improvement. On-policy distillation (OPD) provides dense teacher guidance and typically improves rapidly in the early stage, but its gains saturate once the stud…

View free PDFSource page