CORTEXA
← Browse

Jiangan Yuan

2 papers indexed

arxivcs.LGcs.AI2026-06-26

Let the Data Decide: Supervision Analysis, Capability Trade-offs, and Adaptive Objective Routing in Continued Pre-Training via Off-Policy Distillation

Jiangan Yuan, Zhixuan Li, Han Xu

Off-policy distillation is now central to large language model pre-training, yet how training data, objective parameterization, and model capabilities interact remains poorly characterized. We studies top-$k$-truncated, temperature-scaled off-policy distillation by decomposing th…

View free PDFSource page