CORTEXA
← Browse

Jian Zhao

6 papers indexed

arxivcs.AI2026-07-16

Step-Level Preference Learning for Generative Agents in Social Simulations

Wenchang Gao, Pingyue Sheng, Lanlan Qiu, Yunfei Ma, Jian Zhao, Baicheng Chen, et al.

Large language model (LLM)-based generative agents simulate human behavior through long-horizon decision-making processes that comprise intermediate steps such as planning, memory retrieval, reflection, and action selection. However, fine-grained human annotations of these interm…

View free PDFSource page
arxivcs.LG2026-07-05

RL Forgets! Towards Continual Policy Optimization

Mao-Lin Luo, Zhe-Xu Wang, Zi-Hao Zhou, Bo Ye, Jian Zhao, Min-Ling Zhang, et al.

Continual post-training is becoming a central paradigm for adapting vision-language models to evolving tasks. Recent work has increasingly favored reinforcement learning over supervised fine-tuning, driven by the belief that reinforcement learning is inherently less prone to forg…

View free PDFSource page
crossrefPhotonics2021-03-15Cited by 10

Optical Machine Learning Using Time-Lens Deep Neural NetWorks

Luhe Zhang, Caiyun Li, Jiangyong He, Yange Liu, Jian Zhao, Huiyi Guo, et al.

As a high-throughput data analysis technique, photon time stretching (PTS) is widely used in the monitoring of rare events such as cancer cells, rough waves, and the study of electronic and optical transient dynamics. The PTS technology relies on high-speed data collection, and t…

View free PDFSource page