Incremental Learning of Humanoid Robot Behavior from Natural Interaction and Large Language Models

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Bärmann, Leonard, Kartmann, Rainer, Peller-Konrad, Fabian, Niehues, Jan, Waibel, Alex, Asfour, Tamim
Format: Preprint
Published: 2023
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866909345011204096
author Bärmann, Leonard
Kartmann, Rainer
Peller-Konrad, Fabian
Niehues, Jan
Waibel, Alex
Asfour, Tamim
author_facet Bärmann, Leonard
Kartmann, Rainer
Peller-Konrad, Fabian
Niehues, Jan
Waibel, Alex
Asfour, Tamim
contents Natural-language dialog is key for intuitive human-robot interaction. It can be used not only to express humans' intents, but also to communicate instructions for improvement if a robot does not understand a command correctly. Of great importance is to endow robots with the ability to learn from such interaction experience in an incremental way to allow them to improve their behaviors or avoid mistakes in the future. In this paper, we propose a system to achieve incremental learning of complex behavior from natural interaction, and demonstrate its implementation on a humanoid robot. Building on recent advances, we present a system that deploys Large Language Models (LLMs) for high-level orchestration of the robot's behavior, based on the idea of enabling the LLM to generate Python statements in an interactive console to invoke both robot perception and action. The interaction loop is closed by feeding back human instructions, environment observations, and execution results to the LLM, thus informing the generation of the next statement. Specifically, we introduce incremental prompt learning, which enables the system to interactively learn from its mistakes. For that purpose, the LLM can call another LLM responsible for code-level improvements of the current interaction based on human feedback. The improved interaction is then saved in the robot's memory, and thus retrieved on similar requests. We integrate the system in the robot cognitive architecture of the humanoid robot ARMAR-6 and evaluate our methods both quantitatively (in simulation) and qualitatively (in simulation and real-world) by demonstrating generalized incrementally-learned knowledge.
format Preprint
id arxiv_https___arxiv_org_abs_2309_04316
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle Incremental Learning of Humanoid Robot Behavior from Natural Interaction and Large Language Models
Bärmann, Leonard
Kartmann, Rainer
Peller-Konrad, Fabian
Niehues, Jan
Waibel, Alex
Asfour, Tamim
Robotics
Artificial Intelligence
Natural-language dialog is key for intuitive human-robot interaction. It can be used not only to express humans' intents, but also to communicate instructions for improvement if a robot does not understand a command correctly. Of great importance is to endow robots with the ability to learn from such interaction experience in an incremental way to allow them to improve their behaviors or avoid mistakes in the future. In this paper, we propose a system to achieve incremental learning of complex behavior from natural interaction, and demonstrate its implementation on a humanoid robot. Building on recent advances, we present a system that deploys Large Language Models (LLMs) for high-level orchestration of the robot's behavior, based on the idea of enabling the LLM to generate Python statements in an interactive console to invoke both robot perception and action. The interaction loop is closed by feeding back human instructions, environment observations, and execution results to the LLM, thus informing the generation of the next statement. Specifically, we introduce incremental prompt learning, which enables the system to interactively learn from its mistakes. For that purpose, the LLM can call another LLM responsible for code-level improvements of the current interaction based on human feedback. The improved interaction is then saved in the robot's memory, and thus retrieved on similar requests. We integrate the system in the robot cognitive architecture of the humanoid robot ARMAR-6 and evaluate our methods both quantitatively (in simulation) and qualitatively (in simulation and real-world) by demonstrating generalized incrementally-learned knowledge.
title Incremental Learning of Humanoid Robot Behavior from Natural Interaction and Large Language Models
topic Robotics
Artificial Intelligence
url https://arxiv.org/abs/2309.04316