CORTEXA
← Browse
crossrefRobotics2023-09-28Cited by 0

An Advisor-Based Architecture for a Sample-Efficient Training of Autonomous Navigation Agents with Reinforcement Learning

Rukshan Darshana Wijesinghe, Dumindu Tissera, Mihira Kasun Vithanage, Alex Xavier, Subha Fernando, Jayathu Samarawickrama

Recent advancements in artificial intelligence have enabled reinforcement learning (RL) agents to exceed human-level performance in various gaming tasks. However, despite the state-of-the-art performance demonstrated by model-free RL algorithms, they suffer from high sample complexity. Hence, it is uncommon to find their applications in robotics, autonomous navigation, and self-driving, as gathering many samples is impractical in real-world hardware systems. Therefore, developing sample-efficient learning algorithms for RL agents is crucial in deploying them in real-world tasks without sacrificing performance. This paper presents an advisor-based learning algorithm, incorporating prior knowledge into the training by modifying the deep deterministic policy gradient algorithm to reduce the sample complexity. Also, we propose an effective method of employing an advisor in data collection to train autonomous navigation agents to maneuver physical platforms, minimizing the risk of collision. We analyze the performance of our methods with the support of simulation and physical experimental setups. Experiments reveal that incorporating an advisor into the training phase significantly reduces the sample complexity without compromising the agent’s performance compared to various benchmark approaches. Also, they show that the advisor’s constant involvement in the data collection process diminishes the agent’s performance, while the limited involvement makes training more effective.

View free PDFSource page

Related papers

crossrefRobotics2024-11-17Cited by 2

Trajectory Aware Deep Reinforcement Learning Navigation Using Multichannel Cost Maps

Tareq A. Fahmy, Omar M. Shehata, Shady A. Maged

Deep reinforcement learning (DRL)-based navigation in an environment with dynamic obstacles is a challenging task due to the partially observable nature of the problem. While DRL algorithms are built around the Markov property (assumption that all the necessary information for ma…

View free PDFSource page
crossrefRobotics2022-09-09Cited by 13

Deep Reinforcement Learning for Autonomous Dynamic Skid Steer Vehicle Trajectory Tracking

Sandeep Srikonda, William Robert Norris, Dustin Nottage, Ahmet Soylemezoglu

Designing controllers for skid-steered wheeled robots is complex due to the interaction of the tires with the ground and wheel slip due to the skid-steer driving mechanism, leading to nonlinear dynamics. Due to the recent success of reinforcement learning algorithms for mobile ro…

View free PDFSource page
crossrefRobotics2026-05-11Cited by 1

Attention-Guided Path Planning: Learning Efficient Heuristics for Mobile Robot Navigation via Deep Neural Networks

Abderrahim Waga, Said Benhlima, Ali Bekri, Fatima Zahrae Saber, Jawad Abdouni, Toufik Mzili, et al.

Path planning in cluttered environments constitutes a critical challenge for mobile robotics. Although optimal solutions can be obtained by classical methods such as A*, they have the disadvantage of being computationally expensive in complex environments. In this paper, we propo…

View free PDFSource page
crossrefRobotics2024-01-15Cited by 3

Genetic Algorithm-Based Data Optimization for Efficient Transfer Learning in Convolutional Neural Networks: A Brain–Machine Interface Implementation

Goragod Pongthanisorn, Genci Capi

In brain–machine interface (BMI) systems, the performance of trained Convolutional Neural Networks (CNNs) is significantly influenced by the quality of the training data. Another issue is the training time of CNNs. This paper introduces a novel approach by combining transfer lear…

View free PDFSource page
crossrefRobotics2025-05-31Cited by 2

Guided Reinforcement Learning with Twin Delayed Deep Deterministic Policy Gradient for a Rotary Flexible-Link System

Carlos Saldaña Enderica, José Ramon Llata, Carlos Torre-Ferrero

This study proposes a robust methodology for vibration suppression and trajectory tracking in rotary flexible-link systems by leveraging guided reinforcement learning (GRL). The approach integrates the twin delayed deep deterministic policy gradient (TD3) algorithm with a linear…

View free PDFSource page
crossrefRobotics2025-08-18

Autonomous Grasping of Deformable Objects with Deep Reinforcement Learning: A Study on Spaghetti Manipulation

Prem Gamolped, Nattapat Koomklang, Abbe Mowshowitz, Eiji Hayashi

Packing food into lunch boxes requires the correct portion to be selected. Food items such as fried chicken, eggs, and sausages are straightforward to manipulate when packing. In contrast, deformable objects like spaghetti can give challenges to lunch box packing due to their fra…

View free PDFSource page