How to Explore with Belief: State Entropy Maximization in POMDPs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zamboni, Riccardo, Cirino, Duilio, Restelli, Marcello, Mutti, Mirco |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Limits of Pure Exploration in POMDPs: When the Observation Entropy is Enough
von: Zamboni, Riccardo, et al.
Veröffentlicht: (2024)
von: Zamboni, Riccardo, et al.
Veröffentlicht: (2024)
Towards Principled Unsupervised Multi-Agent Reinforcement Learning
von: Zamboni, Riccardo, et al.
Veröffentlicht: (2025)
von: Zamboni, Riccardo, et al.
Veröffentlicht: (2025)
From Parameters to Behaviors: Unsupervised Compression of the Policy Space
von: Tenedini, Davide, et al.
Veröffentlicht: (2025)
von: Tenedini, Davide, et al.
Veröffentlicht: (2025)
Unsupervised Behavioral Compression: Learning Low-Dimensional Policy Manifolds through State-Occupancy Matching
von: Fraschini, Andrea, et al.
Veröffentlicht: (2026)
von: Fraschini, Andrea, et al.
Veröffentlicht: (2026)
Enhancing Diversity in Parallel Agents: A Maximum State Entropy Exploration Story
von: De Paola, Vincenzo, et al.
Veröffentlicht: (2025)
von: De Paola, Vincenzo, et al.
Veröffentlicht: (2025)
K-Myriad: Jump-starting reinforcement learning with unsupervised parallel agents
von: De Paola, Vincenzo, et al.
Veröffentlicht: (2026)
von: De Paola, Vincenzo, et al.
Veröffentlicht: (2026)
How Log-Barrier Helps Exploration in Policy Optimization
von: Cesani, Leonardo, et al.
Veröffentlicht: (2026)
von: Cesani, Leonardo, et al.
Veröffentlicht: (2026)
A Theoretical Framework for Partially Observed Reward-States in RLHF
von: Kausik, Chinmaya, et al.
Veröffentlicht: (2024)
von: Kausik, Chinmaya, et al.
Veröffentlicht: (2024)
Scalable Multi-Agent Offline Reinforcement Learning and the Role of Information
von: Zamboni, Riccardo, et al.
Veröffentlicht: (2025)
von: Zamboni, Riccardo, et al.
Veröffentlicht: (2025)
Inverse Reinforcement Learning with Sub-optimal Experts
von: Poiani, Riccardo, et al.
Veröffentlicht: (2024)
von: Poiani, Riccardo, et al.
Veröffentlicht: (2024)
Exploiting Causal Graph Priors with Posterior Sampling for Reinforcement Learning
von: Mutti, Mirco, et al.
Veröffentlicht: (2023)
von: Mutti, Mirco, et al.
Veröffentlicht: (2023)
Building surrogate models using trajectories of agents trained by Reinforcement Learning
von: Cestero, Julen, et al.
Veröffentlicht: (2025)
von: Cestero, Julen, et al.
Veröffentlicht: (2025)
Online Planning in POMDPs with State-Requests
von: Avalos, Raphael, et al.
Veröffentlicht: (2024)
von: Avalos, Raphael, et al.
Veröffentlicht: (2024)
ESCORT: Efficient Stein-variational and Sliced Consistency-Optimized Temporal Belief Representation for POMDPs
von: Zhang, Yunuo, et al.
Veröffentlicht: (2025)
von: Zhang, Yunuo, et al.
Veröffentlicht: (2025)
Statistical Analysis of Policy Space Compression Problem
von: Molaei, Majid, et al.
Veröffentlicht: (2024)
von: Molaei, Majid, et al.
Veröffentlicht: (2024)
Projection by Convolution: Optimal Sample Complexity for Reinforcement Learning in Continuous-Space MDPs
von: Maran, Davide, et al.
Veröffentlicht: (2024)
von: Maran, Davide, et al.
Veröffentlicht: (2024)
Efficient Learning of POMDPs with Known Observation Model in Average-Reward Setting
von: Russo, Alessio, et al.
Veröffentlicht: (2024)
von: Russo, Alessio, et al.
Veröffentlicht: (2024)
Rethinking Transformers in Solving POMDPs
von: Lu, Chenhao, et al.
Veröffentlicht: (2024)
von: Lu, Chenhao, et al.
Veröffentlicht: (2024)
Limitations of Physics-Informed Neural Networks: a Study on Smart Grid Surrogation
von: Cestero, Julen, et al.
Veröffentlicht: (2025)
von: Cestero, Julen, et al.
Veröffentlicht: (2025)
Optimizing Energy Management of Smart Grid using Reinforcement Learning aided by Surrogate models built using Physics-informed Neural Networks
von: Cestero, Julen, et al.
Veröffentlicht: (2025)
von: Cestero, Julen, et al.
Veröffentlicht: (2025)
Achieving $\widetilde{\mathcal{O}}(\sqrt{T})$ Regret in Average-Reward POMDPs with Known Observation Models
von: Russo, Alessio, et al.
Veröffentlicht: (2025)
von: Russo, Alessio, et al.
Veröffentlicht: (2025)
Toward Learning POMDPs Beyond Full-Rank Actions and State Observability
von: Shaw, Seiji, et al.
Veröffentlicht: (2026)
von: Shaw, Seiji, et al.
Veröffentlicht: (2026)
Pessimistic Iterative Planning with RNNs for Robust POMDPs
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2024)
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2024)
Scalable Policy-Based RL Algorithms for POMDPs
von: Anjarlekar, Ameya, et al.
Veröffentlicht: (2025)
von: Anjarlekar, Ameya, et al.
Veröffentlicht: (2025)
Reinforcement Learning for Generative AI: State of the Art, Opportunities and Open Research Challenges
von: Franceschelli, Giorgio, et al.
Veröffentlicht: (2023)
von: Franceschelli, Giorgio, et al.
Veröffentlicht: (2023)
Value of Information and Reward Specification in Active Inference and POMDPs
von: Wei, Ran
Veröffentlicht: (2024)
von: Wei, Ran
Veröffentlicht: (2024)
Missingness-MDPs: Bridging the Theory of Missing Data and POMDPs
von: Wendland, Joshua, et al.
Veröffentlicht: (2026)
von: Wendland, Joshua, et al.
Veröffentlicht: (2026)
Sequential Monte Carlo for Policy Optimization in Continuous POMDPs
von: Abdulsamad, Hany, et al.
Veröffentlicht: (2025)
von: Abdulsamad, Hany, et al.
Veröffentlicht: (2025)
Parameterized Projected Bellman Operator
von: Vincent, Théo, et al.
Veröffentlicht: (2023)
von: Vincent, Théo, et al.
Veröffentlicht: (2023)
Manifold Sampling via Entropy Maximization
von: Braun, Cornelius V., et al.
Veröffentlicht: (2026)
von: Braun, Cornelius V., et al.
Veröffentlicht: (2026)
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2025)
von: Galesloot, Maris F. L., et al.
Veröffentlicht: (2025)
Selecting Belief-State Approximations in Simulators with Latent States
von: Jiang, Nan
Veröffentlicht: (2025)
von: Jiang, Nan
Veröffentlicht: (2025)
Statistical Tractability of Off-policy Evaluation of History-dependent Policies in POMDPs
von: Zhang, Yuheng, et al.
Veröffentlicht: (2025)
von: Zhang, Yuheng, et al.
Veröffentlicht: (2025)
How does Inverse RL Scale to Large State Spaces? A Provably Efficient Approach
von: Lazzati, Filippo, et al.
Veröffentlicht: (2024)
von: Lazzati, Filippo, et al.
Veröffentlicht: (2024)
Generalizing Multi-Step Inverse Models for Representation Learning to Finite-Memory POMDPs
von: Wu, Lili, et al.
Veröffentlicht: (2024)
von: Wu, Lili, et al.
Veröffentlicht: (2024)
The Belief State Transformer
von: Hu, Edward S., et al.
Veröffentlicht: (2024)
von: Hu, Edward S., et al.
Veröffentlicht: (2024)
Learning Interpretable Policies in Hindsight-Observable POMDPs through Partially Supervised Reinforcement Learning
von: Lanier, Michael, et al.
Veröffentlicht: (2024)
von: Lanier, Michael, et al.
Veröffentlicht: (2024)
Belief-State Query Policies for User-Aligned POMDPs
von: Bramblett, Daniel, et al.
Veröffentlicht: (2024)
von: Bramblett, Daniel, et al.
Veröffentlicht: (2024)
Exploring Token-Space Manipulation in Latent Audio Tokenizers
von: Paissan, Francesco, et al.
Veröffentlicht: (2026)
von: Paissan, Francesco, et al.
Veröffentlicht: (2026)
Theoretical Insights in Model Inversion Robustness and Conditional Entropy Maximization for Collaborative Inference Systems
von: Xia, Song, et al.
Veröffentlicht: (2025)
von: Xia, Song, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
The Limits of Pure Exploration in POMDPs: When the Observation Entropy is Enough
von: Zamboni, Riccardo, et al.
Veröffentlicht: (2024) -
Towards Principled Unsupervised Multi-Agent Reinforcement Learning
von: Zamboni, Riccardo, et al.
Veröffentlicht: (2025) -
From Parameters to Behaviors: Unsupervised Compression of the Policy Space
von: Tenedini, Davide, et al.
Veröffentlicht: (2025) -
Unsupervised Behavioral Compression: Learning Low-Dimensional Policy Manifolds through State-Occupancy Matching
von: Fraschini, Andrea, et al.
Veröffentlicht: (2026) -
Enhancing Diversity in Parallel Agents: A Maximum State Entropy Exploration Story
von: De Paola, Vincenzo, et al.
Veröffentlicht: (2025)