CURIOUS: Intrinsically Motivated Modular Multi-Goal Reinforcement Learning

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Colas, Cédric, Fournier, Pierre, Sigaud, Olivier, Chetouani, Mohamed, Oudeyer, Pierre-Yves
Format: Preprint
Published: 2018
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866917228536922112
author Colas, Cédric
Fournier, Pierre
Sigaud, Olivier
Chetouani, Mohamed
Oudeyer, Pierre-Yves
author_facet Colas, Cédric
Fournier, Pierre
Sigaud, Olivier
Chetouani, Mohamed
Oudeyer, Pierre-Yves
contents In open-ended environments, autonomous learning agents must set their own goals and build their own curriculum through an intrinsically motivated exploration. They may consider a large diversity of goals, aiming to discover what is controllable in their environments, and what is not. Because some goals might prove easy and some impossible, agents must actively select which goal to practice at any moment, to maximize their overall mastery on the set of learnable goals. This paper proposes CURIOUS, an algorithm that leverages 1) a modular Universal Value Function Approximator with hindsight learning to achieve a diversity of goals of different kinds within a unique policy and 2) an automated curriculum learning mechanism that biases the attention of the agent towards goals maximizing the absolute learning progress. Agents focus sequentially on goals of increasing complexity, and focus back on goals that are being forgotten. Experiments conducted in a new modular-goal robotic environment show the resulting developmental self-organization of a learning curriculum, and demonstrate properties of robustness to distracting goals, forgetting and changes in body properties.
format Preprint
id arxiv_https___arxiv_org_abs_1810_06284
institution arXiv
publishDate 2018
record_format arxiv
spellingShingle CURIOUS: Intrinsically Motivated Modular Multi-Goal Reinforcement Learning
Colas, Cédric
Fournier, Pierre
Sigaud, Olivier
Chetouani, Mohamed
Oudeyer, Pierre-Yves
Artificial Intelligence
In open-ended environments, autonomous learning agents must set their own goals and build their own curriculum through an intrinsically motivated exploration. They may consider a large diversity of goals, aiming to discover what is controllable in their environments, and what is not. Because some goals might prove easy and some impossible, agents must actively select which goal to practice at any moment, to maximize their overall mastery on the set of learnable goals. This paper proposes CURIOUS, an algorithm that leverages 1) a modular Universal Value Function Approximator with hindsight learning to achieve a diversity of goals of different kinds within a unique policy and 2) an automated curriculum learning mechanism that biases the attention of the agent towards goals maximizing the absolute learning progress. Agents focus sequentially on goals of increasing complexity, and focus back on goals that are being forgotten. Experiments conducted in a new modular-goal robotic environment show the resulting developmental self-organization of a learning curriculum, and demonstrate properties of robustness to distracting goals, forgetting and changes in body properties.
title CURIOUS: Intrinsically Motivated Modular Multi-Goal Reinforcement Learning
topic Artificial Intelligence
url https://arxiv.org/abs/1810.06284