BetaZero: Belief-State Planning for Long-Horizon POMDPs using Learned Approximations
Fuente:
arXiv
Saved in:
| Main Authors: | Moss, Robert J., Corso, Anthony, Caers, Jef, Kochenderfer, Mykel J. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Conditional Deep Generative Models for Belief State Planning
by: Bigeard, Antoine, et al.
Published: (2025)
by: Bigeard, Antoine, et al.
Published: (2025)
ConstrainedZero: Chance-Constrained POMDP Planning using Learned Probabilistic Failure Surrogates and Adaptive Safety Constraints
by: Moss, Robert J., et al.
Published: (2024)
by: Moss, Robert J., et al.
Published: (2024)
Constrained Hierarchical Monte Carlo Belief-State Planning
by: Jamgochian, Arec, et al.
Published: (2023)
by: Jamgochian, Arec, et al.
Published: (2023)
The future of AI in critical mineral exploration
by: Caers, Jef
Published: (2025)
by: Caers, Jef
Published: (2025)
Managing Geological Uncertainty in Critical Mineral Supply Chains: A POMDP Approach with Application to U.S. Lithium Resources
by: Arief, Mansur, et al.
Published: (2025)
by: Arief, Mansur, et al.
Published: (2025)
Intelligent prospector v2.0: exploration drill planning under epistemic model uncertainty
by: Mern, John, et al.
Published: (2024)
by: Mern, John, et al.
Published: (2024)
On Technique Identification and Threat-Actor Attribution using LLMs and Embedding Models
by: Guru, Kyla, et al.
Published: (2025)
by: Guru, Kyla, et al.
Published: (2025)
Scaling Recurrent Neural Networks to a Billion Parameters with Zero-Order Optimization
by: Chaubard, Francois, et al.
Published: (2025)
by: Chaubard, Francois, et al.
Published: (2025)
Adaptive Science Operations in Deep Space Missions Using Offline Belief State Planning
by: Kim, Grace Ra, et al.
Published: (2025)
by: Kim, Grace Ra, et al.
Published: (2025)
Optimizing Task Completion Time Updates Using POMDPs
by: Eddy, Duncan, et al.
Published: (2026)
by: Eddy, Duncan, et al.
Published: (2026)
Robust Planning for Autonomous Vehicles with Diffusion-Based Failure Samplers
by: Wang, Juanran, et al.
Published: (2025)
by: Wang, Juanran, et al.
Published: (2025)
Addressing Myopic Constrained POMDP Planning with Recursive Dual Ascent
by: Stocco, Paula, et al.
Published: (2024)
by: Stocco, Paula, et al.
Published: (2024)
Semi-Markovian Planning to Coordinate Aerial and Maritime Medical Evacuation Platforms
by: Al-Husseini, Mahdi, et al.
Published: (2024)
by: Al-Husseini, Mahdi, et al.
Published: (2024)
Failure Probability Estimation for Black-Box Autonomous Systems using State-Dependent Importance Sampling Proposals
by: Delecki, Harrison, et al.
Published: (2024)
by: Delecki, Harrison, et al.
Published: (2024)
Large-Scale Multi-Robot Assembly Planning for Autonomous Manufacturing
by: Brown, Kyle, et al.
Published: (2023)
by: Brown, Kyle, et al.
Published: (2023)
Belief-State Query Policies for User-Aligned POMDPs
by: Bramblett, Daniel, et al.
Published: (2024)
by: Bramblett, Daniel, et al.
Published: (2024)
Diffusion-Based Failure Sampling for Evaluating Safety-Critical Autonomous Systems
by: Delecki, Harrison, et al.
Published: (2024)
by: Delecki, Harrison, et al.
Published: (2024)
Adaptive mine planning under geological uncertainty: A POMDP framework for sequential decision-making
by: Khalifi, Hamza, et al.
Published: (2026)
by: Khalifi, Hamza, et al.
Published: (2026)
How to Explore with Belief: State Entropy Maximization in POMDPs
by: Zamboni, Riccardo, et al.
Published: (2024)
by: Zamboni, Riccardo, et al.
Published: (2024)
Beyond Gradient Averaging in Parallel Optimization: Improved Robustness through Gradient Agreement Filtering
by: Chaubard, Francois, et al.
Published: (2024)
by: Chaubard, Francois, et al.
Published: (2024)
Online Planning in POMDPs with State-Requests
by: Avalos, Raphael, et al.
Published: (2024)
by: Avalos, Raphael, et al.
Published: (2024)
Graph Q-Learning for Combinatorial Optimization
by: Dax, Victoria M., et al.
Published: (2024)
by: Dax, Victoria M., et al.
Published: (2024)
A Semi-Decentralized Approach to Multiagent Control
by: Al-Husseini, Mahdi, et al.
Published: (2026)
by: Al-Husseini, Mahdi, et al.
Published: (2026)
Scene Informer: Anchor-based Occlusion Inference and Trajectory Prediction in Partially Observable Environments
by: Lange, Bernard, et al.
Published: (2023)
by: Lange, Bernard, et al.
Published: (2023)
Optimal Ground Station Selection for Low-Earth Orbiting Satellites
by: Eddy, Duncan, et al.
Published: (2024)
by: Eddy, Duncan, et al.
Published: (2024)
Multi-Environment POMDPs with Finite-Horizon Objectives
by: Brice, Léonard, et al.
Published: (2026)
by: Brice, Léonard, et al.
Published: (2026)
Improving the Resilience of Quadrotors in Underground Environments by Combining Learning-based and Safety Controllers
by: Ward, Isaac Ronald, et al.
Published: (2025)
by: Ward, Isaac Ronald, et al.
Published: (2025)
Aircraft Collision Avoidance Systems: Technological Challenges and Solutions on the Path to Regulatory Acceptance
by: Katz, Sydney M., et al.
Published: (2025)
by: Katz, Sydney M., et al.
Published: (2025)
Tighter Value-Function Approximations for POMDPs
by: Krale, Merlijn, et al.
Published: (2025)
by: Krale, Merlijn, et al.
Published: (2025)
AI-Driven Optimization under Uncertainty for Mineral Processing Operations
by: Xu, William, et al.
Published: (2025)
by: Xu, William, et al.
Published: (2025)
Diffusion Models for Safety Validation of Autonomous Driving Systems
by: Wang, Juanran, et al.
Published: (2025)
by: Wang, Juanran, et al.
Published: (2025)
Reward Bias Substitution: Single-Axis Bias Mitigations Redirect Optimization Pressure
by: Lamparth, Max, et al.
Published: (2026)
by: Lamparth, Max, et al.
Published: (2026)
A New Strategy for Verifying Reach-Avoid Specifications in Neural Feedback Systems
by: Akinwande, Samuel I., et al.
Published: (2026)
by: Akinwande, Samuel I., et al.
Published: (2026)
Zono-Conformal Prediction: Zonotope-Based Uncertainty Quantification for Regression and Classification Tasks
by: Lützow, Laura, et al.
Published: (2025)
by: Lützow, Laura, et al.
Published: (2025)
Importance Sampling-Guided Meta-Training for Intelligent Agents in Highly Interactive Environments
by: Arief, Mansur, et al.
Published: (2024)
by: Arief, Mansur, et al.
Published: (2024)
Finite-State Controllers for (Hidden-Model) POMDPs using Deep Reinforcement Learning
by: Hudák, David, et al.
Published: (2026)
by: Hudák, David, et al.
Published: (2026)
Pessimistic Iterative Planning with RNNs for Robust POMDPs
by: Galesloot, Maris F. L., et al.
Published: (2024)
by: Galesloot, Maris F. L., et al.
Published: (2024)
Planning Transformer: Long-Horizon Offline Reinforcement Learning with Planning Tokens
by: Clinton, Joseph, et al.
Published: (2024)
by: Clinton, Joseph, et al.
Published: (2024)
One Bias After Another: Mechanistic Reward Shaping and Persistent Biases in Language Reward Models
by: Fein, Daniel, et al.
Published: (2026)
by: Fein, Daniel, et al.
Published: (2026)
The FABRIC Strategy for Verifying Neural Feedback Systems
by: Akinwande, Samuel I., et al.
Published: (2026)
by: Akinwande, Samuel I., et al.
Published: (2026)
Similar Items
-
Conditional Deep Generative Models for Belief State Planning
by: Bigeard, Antoine, et al.
Published: (2025) -
ConstrainedZero: Chance-Constrained POMDP Planning using Learned Probabilistic Failure Surrogates and Adaptive Safety Constraints
by: Moss, Robert J., et al.
Published: (2024) -
Constrained Hierarchical Monte Carlo Belief-State Planning
by: Jamgochian, Arec, et al.
Published: (2023) -
The future of AI in critical mineral exploration
by: Caers, Jef
Published: (2025) -
Managing Geological Uncertainty in Critical Mineral Supply Chains: A POMDP Approach with Application to U.S. Lithium Resources
by: Arief, Mansur, et al.
Published: (2025)