Scalable Solution Methods for Dec-POMDPs with Deterministic Dynamics
Fuente:
arXiv
Saved in:
| Main Authors: | You, Yang, Schutz, Alex, Li, Zhikun, Lacerda, Bruno, Skilton, Robert, Hawes, Nick |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Finite-State Controller Based Offline Solver for Deterministic POMDPs
by: Schutz, Alex, et al.
Published: (2025)
by: Schutz, Alex, et al.
Published: (2025)
Partially Observable Monte-Carlo Graph Search
by: You, Yang, et al.
Published: (2025)
by: You, Yang, et al.
Published: (2025)
Neural Value Iteration
by: You, Yang, et al.
Published: (2025)
by: You, Yang, et al.
Published: (2025)
Tackling GNARLy Problems: Graph Neural Algorithmic Reasoning Reimagined through Reinforcement Learning
by: Schutz, Alex, et al.
Published: (2025)
by: Schutz, Alex, et al.
Published: (2025)
Return Capping: Sample-Efficient CVaR Policy Gradient Optimisation
by: Mead, Harry, et al.
Published: (2025)
by: Mead, Harry, et al.
Published: (2025)
Monte Carlo Tree Search with Boltzmann Exploration
by: Painter, Michael, et al.
Published: (2024)
by: Painter, Michael, et al.
Published: (2024)
Gaussian Process Aggregation for Root-Parallel Monte Carlo Tree Search with Continuous Actions
by: Xiao, Junlin, et al.
Published: (2025)
by: Xiao, Junlin, et al.
Published: (2025)
JaxWildfire: A GPU-Accelerated Wildfire Simulator for Reinforcement Learning
by: Çakır, Ufuk, et al.
Published: (2025)
by: Çakır, Ufuk, et al.
Published: (2025)
Solving Collaborative Dec-POMDPs with Deep Reinforcement Learning Heuristics
by: Soffair, Nitsan
Published: (2022)
by: Soffair, Nitsan
Published: (2022)
Computing the Reachability Value of Posterior-Deterministic POMDPs
by: Fijalkow, Nathanaël, et al.
Published: (2026)
by: Fijalkow, Nathanaël, et al.
Published: (2026)
No Regrets: Investigating and Improving Regret Approximations for Curriculum Discovery
by: Rutherford, Alexander, et al.
Published: (2024)
by: Rutherford, Alexander, et al.
Published: (2024)
Scalable Policy-Based RL Algorithms for POMDPs
by: Anjarlekar, Ameya, et al.
Published: (2025)
by: Anjarlekar, Ameya, et al.
Published: (2025)
Towards Optimal Performance and Action Consistency Guarantees in Dec-POMDPs with Inconsistent Beliefs and Limited Communication
by: Shimron, Moshe Rafaeli, et al.
Published: (2025)
by: Shimron, Moshe Rafaeli, et al.
Published: (2025)
Tighter Value-Function Approximations for POMDPs
by: Krale, Merlijn, et al.
Published: (2025)
by: Krale, Merlijn, et al.
Published: (2025)
Synthetic POMDPs to Challenge Memory-Augmented RL: Memory Demand Structure Modeling
by: Wang, Yongyi, et al.
Published: (2025)
by: Wang, Yongyi, et al.
Published: (2025)
No Compromise in Solution Quality: Speeding Up Belief-dependent Continuous POMDPs via Adaptive Multilevel Simplification
by: Zhitnikov, Andrey, et al.
Published: (2023)
by: Zhitnikov, Andrey, et al.
Published: (2023)
Rethinking Transformers in Solving POMDPs
by: Lu, Chenhao, et al.
Published: (2024)
by: Lu, Chenhao, et al.
Published: (2024)
Multi-Environment POMDPs with Finite-Horizon Objectives
by: Brice, Léonard, et al.
Published: (2026)
by: Brice, Léonard, et al.
Published: (2026)
Simplification of Risk Averse POMDPs with Performance Guarantees
by: Pariente, Yaacov, et al.
Published: (2024)
by: Pariente, Yaacov, et al.
Published: (2024)
What should be observed for optimal reward in POMDPs?
by: Konsta, Alyzia-Maria, et al.
Published: (2024)
by: Konsta, Alyzia-Maria, et al.
Published: (2024)
Planning under Distribution Shifts with Causal POMDPs
by: Ceriscioli, Matteo, et al.
Published: (2026)
by: Ceriscioli, Matteo, et al.
Published: (2026)
BetaZero: Belief-State Planning for Long-Horizon POMDPs using Learned Approximations
by: Moss, Robert J., et al.
Published: (2023)
by: Moss, Robert J., et al.
Published: (2023)
Generalizing Multi-Step Inverse Models for Representation Learning to Finite-Memory POMDPs
by: Wu, Lili, et al.
Published: (2024)
by: Wu, Lili, et al.
Published: (2024)
Belief-State Query Policies for User-Aligned POMDPs
by: Bramblett, Daniel, et al.
Published: (2024)
by: Bramblett, Daniel, et al.
Published: (2024)
Online Planning in POMDPs with State-Requests
by: Avalos, Raphael, et al.
Published: (2024)
by: Avalos, Raphael, et al.
Published: (2024)
Inducing Individual Students' Learning Strategies through Homomorphic POMDPs
by: Gao, Huifan, et al.
Published: (2024)
by: Gao, Huifan, et al.
Published: (2024)
Factored Online Planning in Many-Agent POMDPs
by: Galesloot, Maris F. L., et al.
Published: (2023)
by: Galesloot, Maris F. L., et al.
Published: (2023)
Pessimistic Iterative Planning with RNNs for Robust POMDPs
by: Galesloot, Maris F. L., et al.
Published: (2024)
by: Galesloot, Maris F. L., et al.
Published: (2024)
Multi-Environment POMDPs: Discrete Model Uncertainty Under Partial Observability
by: Bovy, Eline M., et al.
Published: (2025)
by: Bovy, Eline M., et al.
Published: (2025)
Improving Regret Approximation for Unsupervised Dynamic Environment Generation
by: Mead, Harry, et al.
Published: (2026)
by: Mead, Harry, et al.
Published: (2026)
DITTO: Offline Imitation Learning with World Models
by: DeMoss, Branton, et al.
Published: (2023)
by: DeMoss, Branton, et al.
Published: (2023)
Sequential Monte Carlo for Policy Optimization in Continuous POMDPs
by: Abdulsamad, Hany, et al.
Published: (2025)
by: Abdulsamad, Hany, et al.
Published: (2025)
Missingness-MDPs: Bridging the Theory of Missing Data and POMDPs
by: Wendland, Joshua, et al.
Published: (2026)
by: Wendland, Joshua, et al.
Published: (2026)
Optimizing Task Completion Time Updates Using POMDPs
by: Eddy, Duncan, et al.
Published: (2026)
by: Eddy, Duncan, et al.
Published: (2026)
Value of Information and Reward Specification in Active Inference and POMDPs
by: Wei, Ran
Published: (2024)
by: Wei, Ran
Published: (2024)
How to Explore with Belief: State Entropy Maximization in POMDPs
by: Zamboni, Riccardo, et al.
Published: (2024)
by: Zamboni, Riccardo, et al.
Published: (2024)
Finite-State Controllers for (Hidden-Model) POMDPs using Deep Reinforcement Learning
by: Hudák, David, et al.
Published: (2026)
by: Hudák, David, et al.
Published: (2026)
Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs
by: Galesloot, Maris F. L., et al.
Published: (2025)
by: Galesloot, Maris F. L., et al.
Published: (2025)
Point-Based Value Iteration for POMDPs with Neural Perception Mechanisms
by: Yan, Rui, et al.
Published: (2023)
by: Yan, Rui, et al.
Published: (2023)
Online Risk-Averse Planning in POMDPs Using Iterated CVaR Value Function
by: Pariente, Yaacov, et al.
Published: (2026)
by: Pariente, Yaacov, et al.
Published: (2026)
Similar Items
-
A Finite-State Controller Based Offline Solver for Deterministic POMDPs
by: Schutz, Alex, et al.
Published: (2025) -
Partially Observable Monte-Carlo Graph Search
by: You, Yang, et al.
Published: (2025) -
Neural Value Iteration
by: You, Yang, et al.
Published: (2025) -
Tackling GNARLy Problems: Graph Neural Algorithmic Reasoning Reimagined through Reinforcement Learning
by: Schutz, Alex, et al.
Published: (2025) -
Return Capping: Sample-Efficient CVaR Policy Gradient Optimisation
by: Mead, Harry, et al.
Published: (2025)