Geometric Active Exploration in Markov Decision Processes: the Benefit of Abstraction
Fuente:
arXiv
Saved in:
| Main Authors: | De Santi, Riccardo, Joseph, Federico Arangath, Liniger, Noah, Mutti, Mirco, Krause, Andreas |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HiPPO-Prophecy: State-Space Models can Provably Learn Dynamical Systems in Context
by: Joseph, Federico Arangath, et al.
Published: (2024)
by: Joseph, Federico Arangath, et al.
Published: (2024)
Enhancing Diversity in Parallel Agents: A Maximum State Entropy Exploration Story
by: De Paola, Vincenzo, et al.
Published: (2025)
by: De Paola, Vincenzo, et al.
Published: (2025)
Exploiting Causal Graph Priors with Posterior Sampling for Reinforcement Learning
by: Mutti, Mirco, et al.
Published: (2023)
by: Mutti, Mirco, et al.
Published: (2023)
The Limits of Pure Exploration in POMDPs: When the Observation Entropy is Enough
by: Zamboni, Riccardo, et al.
Published: (2024)
by: Zamboni, Riccardo, et al.
Published: (2024)
Global Reinforcement Learning: Beyond Linear and Convex Rewards via Submodular Semi-gradient Methods
by: De Santi, Riccardo, et al.
Published: (2024)
by: De Santi, Riccardo, et al.
Published: (2024)
Towards Principled Unsupervised Multi-Agent Reinforcement Learning
by: Zamboni, Riccardo, et al.
Published: (2025)
by: Zamboni, Riccardo, et al.
Published: (2025)
K-Myriad: Jump-starting reinforcement learning with unsupervised parallel agents
by: De Paola, Vincenzo, et al.
Published: (2026)
by: De Paola, Vincenzo, et al.
Published: (2026)
Provable Maximum Entropy Manifold Exploration via Diffusion Models
by: De Santi, Riccardo, et al.
Published: (2025)
by: De Santi, Riccardo, et al.
Published: (2025)
Test-Time Regret Minimization in Meta Reinforcement Learning
by: Mutti, Mirco, et al.
Published: (2024)
by: Mutti, Mirco, et al.
Published: (2024)
Topology-Aware State Abstraction with Tangle Cores for Markov Decision Processes
by: Shihab, Ibne Farabi, et al.
Published: (2026)
by: Shihab, Ibne Farabi, et al.
Published: (2026)
How to Explore with Belief: State Entropy Maximization in POMDPs
by: Zamboni, Riccardo, et al.
Published: (2024)
by: Zamboni, Riccardo, et al.
Published: (2024)
From Parameters to Behaviors: Unsupervised Compression of the Policy Space
by: Tenedini, Davide, et al.
Published: (2025)
by: Tenedini, Davide, et al.
Published: (2025)
Verifier-Constrained Flow Expansion for Discovery Beyond the Data
by: De Santi, Riccardo, et al.
Published: (2026)
by: De Santi, Riccardo, et al.
Published: (2026)
A Unified Density Operator View of Flow Control and Merging
by: De Santi, Riccardo, et al.
Published: (2026)
by: De Santi, Riccardo, et al.
Published: (2026)
Constrained Flow Optimization via Sequential Fine Tuning for Molecular Design
by: Gutjahr, Sven, et al.
Published: (2026)
by: Gutjahr, Sven, et al.
Published: (2026)
Model-Based Exploration in Monitored Markov Decision Processes
by: Kazemipour, Alireza, et al.
Published: (2025)
by: Kazemipour, Alireza, et al.
Published: (2025)
Reward Compatibility: A Framework for Inverse RL
by: Lazzati, Filippo, et al.
Published: (2025)
by: Lazzati, Filippo, et al.
Published: (2025)
Lambda-Skip Connections: the architectural component that prevents Rank Collapse
by: Joseph, Federico Arangath, et al.
Published: (2024)
by: Joseph, Federico Arangath, et al.
Published: (2024)
Unsupervised Behavioral Compression: Learning Low-Dimensional Policy Manifolds through State-Occupancy Matching
by: Fraschini, Andrea, et al.
Published: (2026)
by: Fraschini, Andrea, et al.
Published: (2026)
Logarithmic Regret of Exploration in Average Reward Markov Decision Processes
by: Boone, Victor, et al.
Published: (2025)
by: Boone, Victor, et al.
Published: (2025)
Efficient Personalization of Generative Models via Optimal Experimental Design
by: Schacht, Guy, et al.
Published: (2025)
by: Schacht, Guy, et al.
Published: (2025)
Transition Constrained Bayesian Optimization via Markov Decision Processes
by: Folch, Jose Pablo, et al.
Published: (2024)
by: Folch, Jose Pablo, et al.
Published: (2024)
Offline Inverse RL: New Solution Concepts and Provably Efficient Algorithms
by: Lazzati, Filippo, et al.
Published: (2024)
by: Lazzati, Filippo, et al.
Published: (2024)
How does Inverse RL Scale to Large State Spaces? A Provably Efficient Approach
by: Lazzati, Filippo, et al.
Published: (2024)
by: Lazzati, Filippo, et al.
Published: (2024)
Quantum Logic Gate Synthesis as a Markov Decision Process
by: Alam, M. Sohaib, et al.
Published: (2019)
by: Alam, M. Sohaib, et al.
Published: (2019)
A Classification View on Meta Learning Bandits
by: Mutti, Mirco, et al.
Published: (2025)
by: Mutti, Mirco, et al.
Published: (2025)
Learn A Flexible Exploration Model for Parameterized Action Markov Decision Processes
by: Wang, Zijian, et al.
Published: (2025)
by: Wang, Zijian, et al.
Published: (2025)
Flow Density Control: Generative Optimization Beyond Entropy-Regularized Fine-Tuning
by: De Santi, Riccardo, et al.
Published: (2025)
by: De Santi, Riccardo, et al.
Published: (2025)
ActSafe: Active Exploration with Safety Constraints for Reinforcement Learning
by: As, Yarden, et al.
Published: (2024)
by: As, Yarden, et al.
Published: (2024)
A Theoretical Framework for Partially Observed Reward-States in RLHF
by: Kausik, Chinmaya, et al.
Published: (2024)
by: Kausik, Chinmaya, et al.
Published: (2024)
Monitored Markov Decision Processes
by: Parisi, Simone, et al.
Published: (2024)
by: Parisi, Simone, et al.
Published: (2024)
Temporal Difference Calibration in Sequential Tasks: Application to Vision-Language-Action Models
by: Francis-Meretzki, Shelly, et al.
Published: (2026)
by: Francis-Meretzki, Shelly, et al.
Published: (2026)
Efficient Tail-Aware Generative Optimization via Flow Model Fine-Tuning
by: Wang, Zifan, et al.
Published: (2026)
by: Wang, Zifan, et al.
Published: (2026)
Federated Control in Markov Decision Processes
by: Jin, Hao, et al.
Published: (2024)
by: Jin, Hao, et al.
Published: (2024)
Generalized Linear Markov Decision Process
by: Zhang, Sinian, et al.
Published: (2025)
by: Zhang, Sinian, et al.
Published: (2025)
Towards Robust Knowledge Removal in Federated Learning with High Data Heterogeneity
by: Santi, Riccardo, et al.
Published: (2025)
by: Santi, Riccardo, et al.
Published: (2025)
Learning in Markov Decision Processes with Exogenous Dynamics
by: Maran, Davide, et al.
Published: (2026)
by: Maran, Davide, et al.
Published: (2026)
Directed Exploration in Reinforcement Learning from Linear Temporal Logic
by: Bagatella, Marco, et al.
Published: (2024)
by: Bagatella, Marco, et al.
Published: (2024)
Optimal Decision Tree Policies for Markov Decision Processes
by: Vos, Daniël, et al.
Published: (2023)
by: Vos, Daniël, et al.
Published: (2023)
Policy Testing in Markov Decision Processes
by: Ariu, Kaito, et al.
Published: (2025)
by: Ariu, Kaito, et al.
Published: (2025)
Similar Items
-
HiPPO-Prophecy: State-Space Models can Provably Learn Dynamical Systems in Context
by: Joseph, Federico Arangath, et al.
Published: (2024) -
Enhancing Diversity in Parallel Agents: A Maximum State Entropy Exploration Story
by: De Paola, Vincenzo, et al.
Published: (2025) -
Exploiting Causal Graph Priors with Posterior Sampling for Reinforcement Learning
by: Mutti, Mirco, et al.
Published: (2023) -
The Limits of Pure Exploration in POMDPs: When the Observation Entropy is Enough
by: Zamboni, Riccardo, et al.
Published: (2024) -
Global Reinforcement Learning: Beyond Linear and Convex Rewards via Submodular Semi-gradient Methods
by: De Santi, Riccardo, et al.
Published: (2024)