SPRIG: Stackelberg Perception-Reinforcement Learning with Internal Game Dynamics

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Martinez-Lopez, Fernando, Chen, Juntao, Lu, Yingdong
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866916622490402816
author Martinez-Lopez, Fernando
Chen, Juntao
Lu, Yingdong
author_facet Martinez-Lopez, Fernando
Chen, Juntao
Lu, Yingdong
contents Deep reinforcement learning agents often face challenges to effectively coordinate perception and decision-making components, particularly in environments with high-dimensional sensory inputs where feature relevance varies. This work introduces SPRIG (Stackelberg Perception-Reinforcement learning with Internal Game dynamics), a framework that models the internal perception-policy interaction within a single agent as a cooperative Stackelberg game. In SPRIG, the perception module acts as a leader, strategically processing raw sensory states, while the policy module follows, making decisions based on extracted features. SPRIG provides theoretical guarantees through a modified Bellman operator while preserving the benefits of modern policy optimization. Experimental results on the Atari BeamRider environment demonstrate SPRIG's effectiveness, achieving around 30% higher returns than standard PPO through its game-theoretical balance of feature extraction and decision-making.
format Preprint
id arxiv_https___arxiv_org_abs_2502_14264
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle SPRIG: Stackelberg Perception-Reinforcement Learning with Internal Game Dynamics
Martinez-Lopez, Fernando
Chen, Juntao
Lu, Yingdong
Artificial Intelligence
Deep reinforcement learning agents often face challenges to effectively coordinate perception and decision-making components, particularly in environments with high-dimensional sensory inputs where feature relevance varies. This work introduces SPRIG (Stackelberg Perception-Reinforcement learning with Internal Game dynamics), a framework that models the internal perception-policy interaction within a single agent as a cooperative Stackelberg game. In SPRIG, the perception module acts as a leader, strategically processing raw sensory states, while the policy module follows, making decisions based on extracted features. SPRIG provides theoretical guarantees through a modified Bellman operator while preserving the benefits of modern policy optimization. Experimental results on the Atari BeamRider environment demonstrate SPRIG's effectiveness, achieving around 30% higher returns than standard PPO through its game-theoretical balance of feature extraction and decision-making.
title SPRIG: Stackelberg Perception-Reinforcement Learning with Internal Game Dynamics
topic Artificial Intelligence
url https://arxiv.org/abs/2502.14264