CORTEXA
← Browse
crossrefMachine Learning and Knowledge Extraction2023-12-01Cited by 0

Solving Partially Observable 3D-Visual Tasks with Visual Radial Basis Function Network and Proximal Policy Optimization

Julien Hautot, Céline Teulière, Nourddine Azzaoui

Visual Reinforcement Learning (RL) has been largely investigated in recent decades. Existing approaches are often composed of multiple networks requiring massive computational power to solve partially observable tasks from high-dimensional data such as images. Using State Representation Learning (SRL) has been shown to improve the performance of visual RL by reducing the high-dimensional data into compact representation, but still often relies on deep networks and on the environment. In contrast, we propose a lighter, more generic method to extract sparse and localized features from raw images without training. We achieve this using a Visual Radial Basis Function Network (VRBFN), which offers significant practical advantages, including efficient and accurate training with minimal complexity due to its two linear layers. For real-world applications, its scalability and resilience to noise are essential, as real sensors are subject to change and noise. Unlike CNNs, which may require extensive retraining, this network might only need minor fine-tuning. We test the efficiency of the VRBFN representation to solve different RL tasks using Proximal Policy Optimization (PPO). We present a large study and comparison of our extraction methods with five classical visual RL and SRL approaches on five different first-person partially observable scenarios. We show that this approach presents appealing features such as sparsity and robustness to noise and that the obtained results when training RL agents are better than other tested methods on four of the five proposed scenarios.

View free PDFSource page

Related papers

crossrefMachine Learning and Knowledge Extraction2021-07-15Cited by 69

Recent Advances in Deep Reinforcement Learning Applications for Solving Partially Observable Markov Decision Processes (POMDP) Problems: Part 1—Fundamentals and Applications in Games, Robotics and Natural Language Processing

Xuanchen Xiang, Simon Foo

The first part of a two-part series of papers provides a survey on recent advances in Deep Reinforcement Learning (DRL) applications for solving partially observable Markov decision processes (POMDP) problems. Reinforcement Learning (RL) is an approach to simulate the human’s nat…

View free PDFSource page
crossrefMachine Learning and Knowledge Extraction2024-03-11Cited by 2

Enhancing Docking Accuracy with PECAN2, a 3D Atomic Neural Network Trained without Co-Complex Crystal Structures

Heesung Shim, Jonathan E. Allen, W. F. Drew Bennett

Decades of drug development research have explored a vast chemical space for highly active compounds. The exponential growth of virtual libraries enables easy access to billions of synthesizable molecules. Computational modeling, particularly molecular docking, utilizes physics-b…

View free PDFSource page
crossrefMachine Learning and Knowledge Extraction2024-05-05Cited by 39

Multilayer Perceptron Neural Network with Arithmetic Optimization Algorithm-Based Feature Selection for Cardiovascular Disease Prediction

Fahad A. Alghamdi, Haitham Almanaseer, Ghaith Jaradat, Ashraf Jaradat, Mutasem K. Alsmadi, Sana Jawarneh, et al.

In the healthcare field, diagnosing disease is the most concerning issue. Various diseases including cardiovascular diseases (CVDs) significantly influence illness or death. On the other hand, early and precise diagnosis of CVDs can decrease chances of death, resulting in a bette…

View free PDFSource page
crossrefMachine Learning and Knowledge Extraction2025-07-14Cited by 1

Application of Dueling Double Deep Q-Network for Dynamic Traffic Signal Optimization: A Case Study in Danang City, Vietnam

Tho Cao Phan, Viet Dinh Le, Teron Nguyen

This study investigates the application of the Dueling Double Deep Q-Network (3DQN) algorithm to optimize traffic signal control at a major urban intersection in Danang City, Vietnam. The objective is to enhance signal timing efficiency in response to mixed traffic flow and real-…

View free PDFSource page
crossrefMachine Learning and Knowledge Extraction2025-05-31Cited by 5

Simulated Annealing-Based Hyperparameter Optimization of a Convolutional Neural Network for MRI Brain Tumor Classification

Sofia El Amoury, Youssef Smili, Youssef Fakhri

Brain tumor classification poses significant challenges in medical imaging, largely due to the heterogeneity and structural complexity of tumors. With Magnetic Resonance Imaging (MRI) serving as a cornerstone for diagnosis, manual interpretation by radiologists is time-consuming…

View free PDFSource page
crossrefMachine Learning and Knowledge Extraction2025-10-13Cited by 2

Learning to Partition: Dynamic Deep Neural Network Model Partitioning for Edge-Assisted Low-Latency Video Analytics

Yan Lyu, Likai Liu, Xuezhi Wang, Zhiyu Fan, Jinchen Wang, Guanyu Gao

In edge-assisted low-latency video analytics, a critical challenge is balancing on-device inference latency against the high bandwidth costs and network delays of offloading. Ineffectively managing this trade-off degrades performance and hinders critical applications like autonom…

View free PDFSource page