ReLIC: A Recipe for 64k Steps of In-Context Reinforcement Learning for Embodied AI

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Elawady, Ahmad, Chhablani, Gunjan, Ramrakhya, Ram, Yadav, Karmesh, Batra, Dhruv, Kira, Zsolt, Szot, Andrew
Natura: Preprint
Pubblicazione: 2024
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866912056931778560
author Elawady, Ahmad
Chhablani, Gunjan
Ramrakhya, Ram
Yadav, Karmesh
Batra, Dhruv
Kira, Zsolt
Szot, Andrew
author_facet Elawady, Ahmad
Chhablani, Gunjan
Ramrakhya, Ram
Yadav, Karmesh
Batra, Dhruv
Kira, Zsolt
Szot, Andrew
contents Intelligent embodied agents need to quickly adapt to new scenarios by integrating long histories of experience into decision-making. For instance, a robot in an unfamiliar house initially wouldn't know the locations of objects needed for tasks and might perform inefficiently. However, as it gathers more experience, it should learn the layout of its environment and remember where objects are, allowing it to complete new tasks more efficiently. To enable such rapid adaptation to new tasks, we present ReLIC, a new approach for in-context reinforcement learning (RL) for embodied agents. With ReLIC, agents are capable of adapting to new environments using 64,000 steps of in-context experience with full attention while being trained through self-generated experience via RL. We achieve this by proposing a novel policy update scheme for on-policy RL called "partial updates'' as well as a Sink-KV mechanism that enables effective utilization of a long observation history for embodied agents. Our method outperforms a variety of meta-RL baselines in adapting to unseen houses in an embodied multi-object navigation task. In addition, we find that ReLIC is capable of few-shot imitation learning despite never being trained with expert demonstrations. We also provide a comprehensive analysis of ReLIC, highlighting that the combination of large-scale RL training, the proposed partial updates scheme, and the Sink-KV are essential for effective in-context learning. The code for ReLIC and all our experiments is at https://github.com/aielawady/relic
format Preprint
id arxiv_https___arxiv_org_abs_2410_02751
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle ReLIC: A Recipe for 64k Steps of In-Context Reinforcement Learning for Embodied AI
Elawady, Ahmad
Chhablani, Gunjan
Ramrakhya, Ram
Yadav, Karmesh
Batra, Dhruv
Kira, Zsolt
Szot, Andrew
Machine Learning
Intelligent embodied agents need to quickly adapt to new scenarios by integrating long histories of experience into decision-making. For instance, a robot in an unfamiliar house initially wouldn't know the locations of objects needed for tasks and might perform inefficiently. However, as it gathers more experience, it should learn the layout of its environment and remember where objects are, allowing it to complete new tasks more efficiently. To enable such rapid adaptation to new tasks, we present ReLIC, a new approach for in-context reinforcement learning (RL) for embodied agents. With ReLIC, agents are capable of adapting to new environments using 64,000 steps of in-context experience with full attention while being trained through self-generated experience via RL. We achieve this by proposing a novel policy update scheme for on-policy RL called "partial updates'' as well as a Sink-KV mechanism that enables effective utilization of a long observation history for embodied agents. Our method outperforms a variety of meta-RL baselines in adapting to unseen houses in an embodied multi-object navigation task. In addition, we find that ReLIC is capable of few-shot imitation learning despite never being trained with expert demonstrations. We also provide a comprehensive analysis of ReLIC, highlighting that the combination of large-scale RL training, the proposed partial updates scheme, and the Sink-KV are essential for effective in-context learning. The code for ReLIC and all our experiments is at https://github.com/aielawady/relic
title ReLIC: A Recipe for 64k Steps of In-Context Reinforcement Learning for Embodied AI
topic Machine Learning
url https://arxiv.org/abs/2410.02751