CORTEXA
← Browse

Peng Jiang

2 papers indexed

arxivcs.LGcs.AI2026-07-03

ACPO: Adaptive Credit Policy Optimization via Fine-Grained Surrogate Entropy

Zijun Xie, Yuyang You, Yongzhi Li, Enlei Gong, Zeyu Chen, Quan Chen, et al.

Reinforcement Learning (RL) has substantially improved the reasoning ability of large language models (LLMs), but sparse outcome rewards still make token-level credit assignment difficult. Existing scalable RL methods typically assign trajectory-level rewards uniformly across tok…

View free PDFSource page
arxiveess.SP2026-07-03

Asymptotically Optimal Local Receiver in Uplink CF-mMIMO: A Functional-Variational Analysis

Jiafei Fu, Peng Jiang, Dongming Wang, Pengcheng Zhu

In cell-free massive multiple-input multiple-output (CF-mMIMO) systems, the canonical uplink local receiver is the local minimum mean square error (LMMSE) receiver with large-scale fading decoding (LSFD) at the central processing unit (CPU). The LSFD coefficients are derived unde…

View free PDFSource page