Saved in:
Bibliographic Details
Main Authors: Chung, Hojun, Lee, Junseo, Kim, Minsoo, Kim, Dohyeong, Oh, Songhwai
Format: Preprint
Published: 2024
Subjects:
Online Access:https://arxiv.org/abs/2410.19715
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866917838538670080
author Chung, Hojun
Lee, Junseo
Kim, Minsoo
Kim, Dohyeong
Oh, Songhwai
author_facet Chung, Hojun
Lee, Junseo
Kim, Minsoo
Kim, Dohyeong
Oh, Songhwai
contents Training agents that are robust to environmental changes remains a significant challenge in deep reinforcement learning (RL). Unsupervised environment design (UED) has recently emerged to address this issue by generating a set of training environments tailored to the agent's capabilities. While prior works demonstrate that UED has the potential to learn a robust policy, their performance is constrained by the capabilities of the environment generation. To this end, we propose a novel UED algorithm, adversarial environment design via regret-guided diffusion models (ADD). The proposed method guides the diffusion-based environment generator with the regret of the agent to produce environments that the agent finds challenging but conducive to further improvement. By exploiting the representation power of diffusion models, ADD can directly generate adversarial environments while maintaining the diversity of training environments, enabling the agent to effectively learn a robust policy. Our experimental results demonstrate that the proposed method successfully generates an instructive curriculum of environments, outperforming UED baselines in zero-shot generalization across novel, out-of-distribution environments. Project page: https://rllab-snu.github.io/projects/ADD
format Preprint
id arxiv_https___arxiv_org_abs_2410_19715
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Adversarial Environment Design via Regret-Guided Diffusion Models
Chung, Hojun
Lee, Junseo
Kim, Minsoo
Kim, Dohyeong
Oh, Songhwai
Machine Learning
Artificial Intelligence
Training agents that are robust to environmental changes remains a significant challenge in deep reinforcement learning (RL). Unsupervised environment design (UED) has recently emerged to address this issue by generating a set of training environments tailored to the agent's capabilities. While prior works demonstrate that UED has the potential to learn a robust policy, their performance is constrained by the capabilities of the environment generation. To this end, we propose a novel UED algorithm, adversarial environment design via regret-guided diffusion models (ADD). The proposed method guides the diffusion-based environment generator with the regret of the agent to produce environments that the agent finds challenging but conducive to further improvement. By exploiting the representation power of diffusion models, ADD can directly generate adversarial environments while maintaining the diversity of training environments, enabling the agent to effectively learn a robust policy. Our experimental results demonstrate that the proposed method successfully generates an instructive curriculum of environments, outperforming UED baselines in zero-shot generalization across novel, out-of-distribution environments. Project page: https://rllab-snu.github.io/projects/ADD
title Adversarial Environment Design via Regret-Guided Diffusion Models
topic Machine Learning
Artificial Intelligence
url https://arxiv.org/abs/2410.19715