Data-Efficient Multitask DAgger

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Fu, Haotian, Gong, Ran, Zhang, Xiaohan, Minniti, Maria Vittoria, Patel, Jigarkumar, Schmeckpeper, Karl
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866912615831175168
author Fu, Haotian
Gong, Ran
Zhang, Xiaohan
Minniti, Maria Vittoria
Patel, Jigarkumar
Schmeckpeper, Karl
author_facet Fu, Haotian
Gong, Ran
Zhang, Xiaohan
Minniti, Maria Vittoria
Patel, Jigarkumar
Schmeckpeper, Karl
contents Generalist robot policies that can perform many tasks typically require extensive expert data or simulations for training. In this work, we propose a novel Data-Efficient multitask DAgger framework that distills a single multitask policy from multiple task-specific expert policies. Our approach significantly increases the overall task success rate by actively focusing on tasks where the multitask policy underperforms. The core of our method is a performance-aware scheduling strategy that tracks how much each task's learning process benefits from the amount of data, using a Kalman filter-based estimator to robustly decide how to allocate additional demonstrations across tasks. We validate our approach on MetaWorld, as well as a suite of diverse drawer-opening tasks in IsaacLab. The resulting policy attains high performance across all tasks while using substantially fewer expert demonstrations, and the visual policy learned with our method in simulation shows better performance than naive DAgger and Behavior Cloning when transferring zero-shot to a real robot without using real data.
format Preprint
id arxiv_https___arxiv_org_abs_2509_25466
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Data-Efficient Multitask DAgger
Fu, Haotian
Gong, Ran
Zhang, Xiaohan
Minniti, Maria Vittoria
Patel, Jigarkumar
Schmeckpeper, Karl
Machine Learning
Artificial Intelligence
Generalist robot policies that can perform many tasks typically require extensive expert data or simulations for training. In this work, we propose a novel Data-Efficient multitask DAgger framework that distills a single multitask policy from multiple task-specific expert policies. Our approach significantly increases the overall task success rate by actively focusing on tasks where the multitask policy underperforms. The core of our method is a performance-aware scheduling strategy that tracks how much each task's learning process benefits from the amount of data, using a Kalman filter-based estimator to robustly decide how to allocate additional demonstrations across tasks. We validate our approach on MetaWorld, as well as a suite of diverse drawer-opening tasks in IsaacLab. The resulting policy attains high performance across all tasks while using substantially fewer expert demonstrations, and the visual policy learned with our method in simulation shows better performance than naive DAgger and Behavior Cloning when transferring zero-shot to a real robot without using real data.
title Data-Efficient Multitask DAgger
topic Machine Learning
Artificial Intelligence
url https://arxiv.org/abs/2509.25466