CORTEXA
← Browse
crossrefRobotics2025-05-31Cited by 2

Guided Reinforcement Learning with Twin Delayed Deep Deterministic Policy Gradient for a Rotary Flexible-Link System

Carlos Saldaña Enderica, José Ramon Llata, Carlos Torre-Ferrero

This study proposes a robust methodology for vibration suppression and trajectory tracking in rotary flexible-link systems by leveraging guided reinforcement learning (GRL). The approach integrates the twin delayed deep deterministic policy gradient (TD3) algorithm with a linear quadratic regulator (LQR) acting as a guiding controller during training. Flexible-link mechanisms common in advanced robotics and aerospace systems exhibit oscillatory behavior that complicates precise control. To address this, the system is first identified using experimental input-output data from a Quanser® virtual plant, generating an accurate state-space representation suitable for simulation-based policy learning. The hybrid control strategy enhances sample efficiency and accelerates convergence by incorporating LQR-generated trajectories during TD3 training. Internally, the TD3 agent benefits from architectural features such as twin critics, delayed policy updates, and target action smoothing, which collectively improve learning stability and reduce overestimation bias. Comparative results show that the guided TD3 controller achieves superior performance in terms of vibration damping, transient response, and robustness, when compared to conventional LQR, fuzzy logic, neural networks, and GA-LQR approaches. Although the controller was validated using a high-fidelity digital twin, it has not yet been deployed on the physical plant. Future work will focus on real-time implementation and structural robustness testing under parameter uncertainty. Overall, this research demonstrates that guided reinforcement learning can yield stable and interpretable policies that comply with classical control criteria, offering a scalable and generalizable framework for intelligent control of flexible mechanical systems.

View free PDFSource page

Related papers

crossrefRobotics2022-09-09Cited by 13

Deep Reinforcement Learning for Autonomous Dynamic Skid Steer Vehicle Trajectory Tracking

Sandeep Srikonda, William Robert Norris, Dustin Nottage, Ahmet Soylemezoglu

Designing controllers for skid-steered wheeled robots is complex due to the interaction of the tires with the ground and wheel slip due to the skid-steer driving mechanism, leading to nonlinear dynamics. Due to the recent success of reinforcement learning algorithms for mobile ro…

View free PDFSource page
crossrefRobotics2024-11-17Cited by 2

Trajectory Aware Deep Reinforcement Learning Navigation Using Multichannel Cost Maps

Tareq A. Fahmy, Omar M. Shehata, Shady A. Maged

Deep reinforcement learning (DRL)-based navigation in an environment with dynamic obstacles is a challenging task due to the partially observable nature of the problem. While DRL algorithms are built around the Markov property (assumption that all the necessary information for ma…

View free PDFSource page
crossrefRobotics2025-08-18

Autonomous Grasping of Deformable Objects with Deep Reinforcement Learning: A Study on Spaghetti Manipulation

Prem Gamolped, Nattapat Koomklang, Abbe Mowshowitz, Eiji Hayashi

Packing food into lunch boxes requires the correct portion to be selected. Food items such as fried chicken, eggs, and sausages are straightforward to manipulate when packing. In contrast, deformable objects like spaghetti can give challenges to lunch box packing due to their fra…

View free PDFSource page
crossrefRobotics2023-09-28

An Advisor-Based Architecture for a Sample-Efficient Training of Autonomous Navigation Agents with Reinforcement Learning

Rukshan Darshana Wijesinghe, Dumindu Tissera, Mihira Kasun Vithanage, Alex Xavier, Subha Fernando, Jayathu Samarawickrama

Recent advancements in artificial intelligence have enabled reinforcement learning (RL) agents to exceed human-level performance in various gaming tasks. However, despite the state-of-the-art performance demonstrated by model-free RL algorithms, they suffer from high sample compl…

View free PDFSource page
crossrefRobotics2026-02-02

Visual and Visual–Inertial SLAM for UGV Navigation in Unstructured Natural Environments: A Survey of Challenges and Deep Learning Advances

Tiago Pereira, Carlos Viegas, Salviano Soares, Nuno Ferreira

Localization and mapping remain critical challenges for Unmanned Ground Vehicles (UGVs) operating in unstructured natural environments, such as forests and agricultural fields. While Visual SLAM (VSLAM) and Visual–Inertial SLAM (VI-SLAM) have matured significantly in structured a…

View free PDFSource page
crossrefRobotics2026-05-11Cited by 1

Attention-Guided Path Planning: Learning Efficient Heuristics for Mobile Robot Navigation via Deep Neural Networks

Abderrahim Waga, Said Benhlima, Ali Bekri, Fatima Zahrae Saber, Jawad Abdouni, Toufik Mzili, et al.

Path planning in cluttered environments constitutes a critical challenge for mobile robotics. Although optimal solutions can be obtained by classical methods such as A*, they have the disadvantage of being computationally expensive in complex environments. In this paper, we propo…

View free PDFSource page