Goal Discovery with Causal Capacity for Efficient Reinforcement Learning

Fuente: arXiv
Enregistré dans:
Détails bibliographiques
Auteurs principaux: Yu, Yan, Yang, Yaodong, Lu, Zhengbo, Ma, Chengdong, Zhou, Wengang, Li, Houqiang
Format: Preprint
Publié: 2025
Sujets:
Accès en ligne:
Tags: Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
_version_ 1866915443949699072
author Yu, Yan
Yang, Yaodong
Lu, Zhengbo
Ma, Chengdong
Zhou, Wengang
Li, Houqiang
author_facet Yu, Yan
Yang, Yaodong
Lu, Zhengbo
Ma, Chengdong
Zhou, Wengang
Li, Houqiang
contents Causal inference is crucial for humans to explore the world, which can be modeled to enable an agent to efficiently explore the environment in reinforcement learning. Existing research indicates that establishing the causality between action and state transition will enhance an agent to reason how a policy affects its future trajectory, thereby promoting directed exploration. However, it is challenging to measure the causality due to its intractability in the vast state-action space of complex scenarios. In this paper, we propose a novel Goal Discovery with Causal Capacity (GDCC) framework for efficient environment exploration. Specifically, we first derive a measurement of causality in state space, \emph{i.e.,} causal capacity, which represents the highest influence of an agent's behavior on future trajectories. After that, we present a Monte Carlo based method to identify critical points in discrete state space and further optimize this method for continuous high-dimensional environments. Those critical points are used to uncover where the agent makes important decisions in the environment, which are then regarded as our subgoals to guide the agent to make exploration more purposefully and efficiently. Empirical results from multi-objective tasks demonstrate that states with high causal capacity align with our expected subgoals, and our GDCC achieves significant success rate improvements compared to baselines.
format Preprint
id arxiv_https___arxiv_org_abs_2508_09624
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Goal Discovery with Causal Capacity for Efficient Reinforcement Learning
Yu, Yan
Yang, Yaodong
Lu, Zhengbo
Ma, Chengdong
Zhou, Wengang
Li, Houqiang
Machine Learning
Artificial Intelligence
Causal inference is crucial for humans to explore the world, which can be modeled to enable an agent to efficiently explore the environment in reinforcement learning. Existing research indicates that establishing the causality between action and state transition will enhance an agent to reason how a policy affects its future trajectory, thereby promoting directed exploration. However, it is challenging to measure the causality due to its intractability in the vast state-action space of complex scenarios. In this paper, we propose a novel Goal Discovery with Causal Capacity (GDCC) framework for efficient environment exploration. Specifically, we first derive a measurement of causality in state space, \emph{i.e.,} causal capacity, which represents the highest influence of an agent's behavior on future trajectories. After that, we present a Monte Carlo based method to identify critical points in discrete state space and further optimize this method for continuous high-dimensional environments. Those critical points are used to uncover where the agent makes important decisions in the environment, which are then regarded as our subgoals to guide the agent to make exploration more purposefully and efficiently. Empirical results from multi-objective tasks demonstrate that states with high causal capacity align with our expected subgoals, and our GDCC achieves significant success rate improvements compared to baselines.
title Goal Discovery with Causal Capacity for Efficient Reinforcement Learning
topic Machine Learning
Artificial Intelligence
url https://arxiv.org/abs/2508.09624