CORTEXA
← Browse
crossrefFrontiers in Robotics and AI2026-05-13Cited by 0

Morphological symmetry-aware generalized policy network for deep reinforcement learning

Ryo Hakoda, Yubin Liu, Matthew Hwang, Yoshihiro Sato, Jun Takamatsu, Katsushi Ikeuchi, Takeshi Oishi

Exploiting the morphological symmetry of robotic systems, such as humanoid and quadruped robots, is a promising direction for improving robot learning. In deep reinforcement learning (DRL) for robot control, prior studies have leveraged such symmetry to improve learning efficiency through data augmentation, equivariant multilayer perceptrons (EMLPs), and multi-agent reinforcement learning (MARL) formulations. However, DRL training is inherently unstable, as the data distribution strongly depends on exploration, which is driven by stochasticity in the environment. To address this issue, we propose a symmetry-assisted, general-purpose DRL framework for morphologically symmetric robots that enables stable and robust learning. The framework models the environment as a symmetric Markov decision process (MDP) and constructs a full-body policy from a single-sided base policy using symmetry operators. We further propose a symmetric PPO objective with a coupled importance-sampling ratio. This objective aligns the policy optimization process with the imposed symmetry and serves as a principled alternative to MAPPO-style multi-agent formulations. Experimental results demonstrate that the proposed method outperforms existing approaches on most symmetric tasks, while still maintaining performance comparable to or better than standard PPO on asymmetric tasks, where symmetry is less directly exploitable.

View free PDFSource page

Related papers

crossrefFrontiers in Robotics and AI2026-07-01

Exploring deep reinforcement learning acceleration by superscaling data augmentation via branched fractal symmetries

Ryan Vander Stelt, Cleiver Ruiz-Martinez, Caeden Rosen, Blake Hull, Juan Rojas

Learning deep reinforcement learning (DRL) policies directly in physical robots remains bottlenecked by slow wall-clock training times. We present preliminary research on Branched Euclidean Group Fractal Symmetries , a trajectory-level augmentation framework that super-scales gro…

View free PDFSource page
crossrefFrontiers in Robotics and AI2026-07-16

Structural predictors and latent maturity regimes of robotic readiness in global health systems: evidence from machine learning-based latent clustering and class prediction

Moumita Mukherjee, Raja Hashim Ali

Background The systematic integration of robotics into health service delivery systems requires periodic assessment of robotic readiness in terms of digital-health maturity regimes across countries. The current study aims to cluster 169 countries into maturity regimes and classif…

View free PDFSource page