Continual Domain Randomization

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Josifovski, Josip, Auddy, Sayantan, Malmir, Mohammadhossein, Piater, Justus, Knoll, Alois, Navarro-Guerrero, Nicolás
Format: Preprint
Veröffentlicht: 2024
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866917440867270656
author Josifovski, Josip
Auddy, Sayantan
Malmir, Mohammadhossein
Piater, Justus
Knoll, Alois
Navarro-Guerrero, Nicolás
author_facet Josifovski, Josip
Auddy, Sayantan
Malmir, Mohammadhossein
Piater, Justus
Knoll, Alois
Navarro-Guerrero, Nicolás
contents Domain Randomization (DR) is commonly used for sim2real transfer of reinforcement learning (RL) policies in robotics. Most DR approaches require a simulator with a fixed set of tunable parameters from the start of the training, from which the parameters are randomized simultaneously to train a robust model for use in the real world. However, the combined randomization of many parameters increases the task difficulty and might result in sub-optimal policies. To address this problem and to provide a more flexible training process, we propose Continual Domain Randomization (CDR) for RL that combines domain randomization with continual learning to enable sequential training in simulation on a subset of randomization parameters at a time. Starting from a model trained in a non-randomized simulation where the task is easier to solve, the model is trained on a sequence of randomizations, and continual learning is employed to remember the effects of previous randomizations. Our robotic reaching and grasping tasks experiments show that the model trained in this fashion learns effectively in simulation and performs robustly on the real robot while matching or outperforming baselines that employ combined randomization or sequential randomization without continual learning. Our code and videos are available at https://continual-dr.github.io/.
format Preprint
id arxiv_https___arxiv_org_abs_2403_12193
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Continual Domain Randomization
Josifovski, Josip
Auddy, Sayantan
Malmir, Mohammadhossein
Piater, Justus
Knoll, Alois
Navarro-Guerrero, Nicolás
Robotics
Domain Randomization (DR) is commonly used for sim2real transfer of reinforcement learning (RL) policies in robotics. Most DR approaches require a simulator with a fixed set of tunable parameters from the start of the training, from which the parameters are randomized simultaneously to train a robust model for use in the real world. However, the combined randomization of many parameters increases the task difficulty and might result in sub-optimal policies. To address this problem and to provide a more flexible training process, we propose Continual Domain Randomization (CDR) for RL that combines domain randomization with continual learning to enable sequential training in simulation on a subset of randomization parameters at a time. Starting from a model trained in a non-randomized simulation where the task is easier to solve, the model is trained on a sequence of randomizations, and continual learning is employed to remember the effects of previous randomizations. Our robotic reaching and grasping tasks experiments show that the model trained in this fashion learns effectively in simulation and performs robustly on the real robot while matching or outperforming baselines that employ combined randomization or sequential randomization without continual learning. Our code and videos are available at https://continual-dr.github.io/.
title Continual Domain Randomization
topic Robotics
url https://arxiv.org/abs/2403.12193