Development of a PPO-Reinforcement Learned Walking Tripedal Soft-Legged Robot using SOFA

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Mokhtar, Yomna, Shohdy, Tarek, Hassan, Abdallah A., Eshra, Mostafa, Elmenawy, Omar, Khalil, Osama, El-Hussieny, Haitham
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866912324553539584
author Mokhtar, Yomna
Shohdy, Tarek
Hassan, Abdallah A.
Eshra, Mostafa
Elmenawy, Omar
Khalil, Osama
El-Hussieny, Haitham
author_facet Mokhtar, Yomna
Shohdy, Tarek
Hassan, Abdallah A.
Eshra, Mostafa
Elmenawy, Omar
Khalil, Osama
El-Hussieny, Haitham
contents Rigid robots were extensively researched, whereas soft robotics remains an underexplored field. Utilizing soft-legged robots in performing tasks as a replacement for human beings is an important stride to take, especially under harsh and hazardous conditions over rough terrain environments. For the demand to teach any robot how to behave in different scenarios, a real-time physical and visual simulation is essential. When it comes to soft robots specifically, a simulation framework is still an arduous problem that needs to be disclosed. Using the simulation open framework architecture (SOFA) is an advantageous step. However, neither SOFA's manual nor prior public SOFA projects show its maximum capabilities the users can reach. So, we resolved this by establishing customized settings and handling the framework components appropriately. Settling on perfect, fine-tuned SOFA parameters has stimulated our motivation towards implementing the state-of-the-art (SOTA) reinforcement learning (RL) method of proximal policy optimization (PPO). The final representation is a well-defined, ready-to-deploy walking, tripedal, soft-legged robot based on PPO-RL in a SOFA environment. Robot navigation performance is a key metric to be considered for measuring the success resolution. Although in the simulated soft robots case, an 82\% success rate in reaching a single goal is a groundbreaking output, we pushed the boundaries to further steps by evaluating the progress under assigning a sequence of goals. While trailing the platform steps, outperforming discovery has been observed with an accumulative squared error deviation of 19 mm. The full code is publicly available at \href{https://github.com/tarekshohdy/PPO_SOFA_Soft_Legged_Robot.git}{github.com/tarekshohdy/PPO$\textunderscore$SOFA$\textunderscore$Soft$\textunderscore$Legged$\textunderscore$ Robot.git}
format Preprint
id arxiv_https___arxiv_org_abs_2504_09242
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Development of a PPO-Reinforcement Learned Walking Tripedal Soft-Legged Robot using SOFA
Mokhtar, Yomna
Shohdy, Tarek
Hassan, Abdallah A.
Eshra, Mostafa
Elmenawy, Omar
Khalil, Osama
El-Hussieny, Haitham
Robotics
Artificial Intelligence
Rigid robots were extensively researched, whereas soft robotics remains an underexplored field. Utilizing soft-legged robots in performing tasks as a replacement for human beings is an important stride to take, especially under harsh and hazardous conditions over rough terrain environments. For the demand to teach any robot how to behave in different scenarios, a real-time physical and visual simulation is essential. When it comes to soft robots specifically, a simulation framework is still an arduous problem that needs to be disclosed. Using the simulation open framework architecture (SOFA) is an advantageous step. However, neither SOFA's manual nor prior public SOFA projects show its maximum capabilities the users can reach. So, we resolved this by establishing customized settings and handling the framework components appropriately. Settling on perfect, fine-tuned SOFA parameters has stimulated our motivation towards implementing the state-of-the-art (SOTA) reinforcement learning (RL) method of proximal policy optimization (PPO). The final representation is a well-defined, ready-to-deploy walking, tripedal, soft-legged robot based on PPO-RL in a SOFA environment. Robot navigation performance is a key metric to be considered for measuring the success resolution. Although in the simulated soft robots case, an 82\% success rate in reaching a single goal is a groundbreaking output, we pushed the boundaries to further steps by evaluating the progress under assigning a sequence of goals. While trailing the platform steps, outperforming discovery has been observed with an accumulative squared error deviation of 19 mm. The full code is publicly available at \href{https://github.com/tarekshohdy/PPO_SOFA_Soft_Legged_Robot.git}{github.com/tarekshohdy/PPO$\textunderscore$SOFA$\textunderscore$Soft$\textunderscore$Legged$\textunderscore$ Robot.git}
title Development of a PPO-Reinforcement Learned Walking Tripedal Soft-Legged Robot using SOFA
topic Robotics
Artificial Intelligence
url https://arxiv.org/abs/2504.09242