Zero-Shot Adaptation of Behavioral Foundation Models to Unseen Dynamics

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Bobrin, Maksim, Zisman, Ilya, Nikulin, Alexander, Kurenkov, Vladislav, Dylov, Dmitry
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866910188916703232
author Bobrin, Maksim
Zisman, Ilya
Nikulin, Alexander
Kurenkov, Vladislav
Dylov, Dmitry
author_facet Bobrin, Maksim
Zisman, Ilya
Nikulin, Alexander
Kurenkov, Vladislav
Dylov, Dmitry
contents Behavioral Foundation Models (BFMs) proved successful in producing policies for arbitrary tasks in a zero-shot manner, requiring no test-time training or task-specific fine-tuning. Among the most promising BFMs are the ones that estimate the successor measure learned in an unsupervised way from task-agnostic offline data. However, these methods fail to react to changes in the dynamics, making them inefficient under partial observability or when the transition function changes. This hinders the applicability of BFMs in a real-world setting, e.g., in robotics, where the dynamics can unexpectedly change at test time. In this work, we demonstrate that Forward-Backward (FB) representation, one of the methods from the BFM family, cannot distinguish between distinct dynamics, leading to an interference among the latent directions, which parametrize different policies. To address this, we propose a FB model with a transformer-based belief estimator, which greatly facilitates zero-shot adaptation. We also show that partitioning the policy encoding space into dynamics-specific clusters, aligned with the context-embedding directions, yields additional gain in performance. These traits allow our method to respond to the dynamics observed during training and to generalize to unseen ones. Empirically, in the changing dynamics setting, our approach achieves up to a 2x higher zero-shot returns compared to the baselines for both discrete and continuous tasks.
format Preprint
id arxiv_https___arxiv_org_abs_2505_13150
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Zero-Shot Adaptation of Behavioral Foundation Models to Unseen Dynamics
Bobrin, Maksim
Zisman, Ilya
Nikulin, Alexander
Kurenkov, Vladislav
Dylov, Dmitry
Machine Learning
Behavioral Foundation Models (BFMs) proved successful in producing policies for arbitrary tasks in a zero-shot manner, requiring no test-time training or task-specific fine-tuning. Among the most promising BFMs are the ones that estimate the successor measure learned in an unsupervised way from task-agnostic offline data. However, these methods fail to react to changes in the dynamics, making them inefficient under partial observability or when the transition function changes. This hinders the applicability of BFMs in a real-world setting, e.g., in robotics, where the dynamics can unexpectedly change at test time. In this work, we demonstrate that Forward-Backward (FB) representation, one of the methods from the BFM family, cannot distinguish between distinct dynamics, leading to an interference among the latent directions, which parametrize different policies. To address this, we propose a FB model with a transformer-based belief estimator, which greatly facilitates zero-shot adaptation. We also show that partitioning the policy encoding space into dynamics-specific clusters, aligned with the context-embedding directions, yields additional gain in performance. These traits allow our method to respond to the dynamics observed during training and to generalize to unseen ones. Empirically, in the changing dynamics setting, our approach achieves up to a 2x higher zero-shot returns compared to the baselines for both discrete and continuous tasks.
title Zero-Shot Adaptation of Behavioral Foundation Models to Unseen Dynamics
topic Machine Learning
url https://arxiv.org/abs/2505.13150