CORTEXA
← Browse
arxiveess.SY2026-07-10Cited by 0

An Improved Deep Reinforcement Learning Control Strategy for Traction Dual Rectifiers in EMUs

Zhigang Liu, Mingwei Tang, Xiangyu Meng, Hui Wang, Qiao Zhang, Haoyu Wang, Mengru Li

Due to the use of PI-based d q current decoupling in the pulse rectifier of CRH5 high-speed trains, the PI parameters directly affect the traction system's control performance. Linearized control may have issues with reference trajectory changes or model mismatches, leading to a decrease in system performance, while nonlinear control may have problems with jitter and poor steady-state accuracy. This paper proposes a new control strategy that replaces all PI in the d q current decoupling control with a single intelligent agent. This method based on Deep Reinforcement Learning (DRL) can avoid various drawbacks of linearization and nonlinear control and ensure the stability of intermediate DC voltage. However, when EMUs are in different working conditions and switching, the Twin Delayed Deep Deterministic Policy Gradient (TD3) algorithm used in traction dual rectifiers does not have a good control effect. Focusing on the issue, Reward Shaping (RS) is added to re-design a nonlinear reward function, which can be combined with Prioritized Experience Replay (PER) to increase the convergence speed of the episode reward. The simulation results show that the improved control strategy can be effectively applied to EMUs working in multiple conditions. Finally, the stability analysis is carried out using Lyapunov's second method and the verification results of the hardware-in-the-loop (HIL) simulation platform show that the DRL control has a good effect.

View free PDFSource page

Related papers

arxiveess.SY2026-07-24

Fast Frequency Services from HVDC-Connected Offshore Wind Power Plants: A Review in the European Context

Zhenghua Xu, George Alin Raducu, Behnam Nouri, Oscar Saborío-Romano, Nicolaos A. Cutululis

The rapid expansion of offshore wind energy is central to the European Union's climate-neutrality targets, with High Voltage Direct Current-connected offshore wind power plants (HVDC-OWPPs) becoming increasingly important for integrating gigawatt-scale renewable generation over l…

View free PDFSource page
arxivcs.ROeess.SY2026-07-24

Conformal Constraint Tightening for Chance-Constrained Motion Planning with Unknown Dynamics

Shubham Natraj, Bruno Sinopoli, Yiannis Kantaros

Motion planning algorithms compute control sequences that drive autonomous robots to goal regions while avoiding unsafe states. Existing methods, from sampling-based planning to deep reinforcement learning, typically provide task-completion guarantees only with respect to a nomin…

View free PDFSource page
arxivcs.NIcs.MAeess.SY2026-07-24

Predictive Lightweight MARL for Resilient Coverage in Sparse-Signaling Aerial Networks

Chuan-Chi Lai, Ang-Hsun Tsai

This letter proposes the Predictive Lightweight Multi-Agent Reinforcement Learning (PL-MARL) framework to ensure resilient coverage in bandwidth-constrained UAV swarms. To counter coordination collapse caused by sparse signaling and information aging, we introduce a Kinematic-Awa…

View free PDFSource page
arxiveess.SY2026-07-24

Physics-Informed Neural Network for Modeling the Dynamic Behavior of Grid-Forming Converters

Hussein Jaffal, Arianna Fois, Sarra Bouchkati, Amirali Mahjoob, Andreas Ulbig

This paper investigates physics-informed neural networks for modeling the full dynamic behavior of droop-controlled grid-forming converters. The approach is trained on synthetic data generated via numerical solvers and benchmarked against both traditional integration methods and…

View free PDFSource page
arxiveess.SY2026-07-24

StateFormer: A Multivariate Transformer for Learning History-Dependent Battery State Dynamics and Long-Horizon Health Forecasting

Zhe Bai, Stephen Harris

This paper introduces a novel multivariate Transformer \emph{StateFormer} that forecasts degradation dynamics of large-scale battery systems. The model learns across time scales, from short-term thermal fluctuations to long-term aging trajectories, enabling accurate prediction of…

View free PDFSource page