CORTEXA
← Browse

Zhixuan Li

4 papers indexed

arxivcs.LGcs.CL2026-07-13

SCOPE-RL: Optimizing Reasoning Paths Before and After Success

Xiaojian Liu, Han Xu, Jianqiang Xia, Zhixuan Li, Ke Xu, Yiwei Dai, et al.

Reinforcement learning with verifiable rewards (RLVR) optimizes LLMs using sparse verifiable final-answer rewards. This sparse anchor reliably verifies whether a trajectory succeeds but provides no direct feedback on the reasoning path that produced it. Before success, prerequisi…

View free PDFSource page
arxivcs.LGcs.AI2026-06-26

Let the Data Decide: Supervision Analysis, Capability Trade-offs, and Adaptive Objective Routing in Continued Pre-Training via Off-Policy Distillation

Jiangan Yuan, Zhixuan Li, Han Xu

Off-policy distillation is now central to large language model pre-training, yet how training data, objective parameterization, and model capabilities interact remains poorly characterized. We studies top-$k$-truncated, temperature-scaled off-policy distillation by decomposing th…

View free PDFSource page
crossrefRemote Sensing2023-07-23Cited by 47

Urban Flood Risk Assessment through the Integration of Natural and Human Resilience Based on Machine Learning Models

Wenting Zhang, Bin Hu, Yongzhi Liu, Xingnan Zhang, Zhixuan Li

Flood risk assessment and mapping are considered essential tools for the improvement of flood management. This research aims to construct a more comprehensive flood assessment framework by emphasizing factors related to human resilience and integrating them with meteorological an…

View free PDFSource page