This paper proposes a human-in-the-loop distributed consensus control approach for demand-side management across multiple buildings. Specifically, a novel framework is introduced in which a human acts as the non-autonomous leader in consensus control of cooperative buildings part…
We develop the Continuous Distributed Coupled Policy Gradient (CDCPG) algorithm for cooperative reinforcement learning in networked Markov decision processes with continuous state and action spaces. Each agent maintains a local actor over a bounded graph neighborhood, and a local…
In cell-free massive multiple-input multiple-output (CF-mMIMO) systems, the canonical uplink local receiver is the local minimum mean square error (LMMSE) receiver with large-scale fading decoding (LSFD) at the central processing unit (CPU). The LSFD coefficients are derived unde…