Provably Safe Reinforcement Learning for Stochastic Reach-Avoid Problems with Entropy Regularization
Fuente:
arXiv
Saved in:
| Main Authors: | Mazumdar, Abhijit, Wisniewski, Rafal, Bujorianu, Manuela L. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Safe Reinforcement Learning for Constrained Markov Decision Processes with Stochastic Stopping Time
by: Mazumdar, Abhijit, et al.
Published: (2024)
by: Mazumdar, Abhijit, et al.
Published: (2024)
Data-Driven Robust Safety Verification for Markov Decision Processes
by: Mazumdar, Abhijit, et al.
Published: (2025)
by: Mazumdar, Abhijit, et al.
Published: (2025)
Robust Correlated Equilibrium: Definition and Computation
by: Misra, Rahul, et al.
Published: (2023)
by: Misra, Rahul, et al.
Published: (2023)
Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning
by: Pan, Jingduo, et al.
Published: (2026)
by: Pan, Jingduo, et al.
Published: (2026)
An Online Multiobjective Policy Gradient for Long-run Average-reward Markov Decision Process
by: Misra, Rahul, et al.
Published: (2025)
by: Misra, Rahul, et al.
Published: (2025)
Large Deviations in Safety-Critical Systems with Probabilistic Initial Conditions
by: Gomez, Aitor R., et al.
Published: (2024)
by: Gomez, Aitor R., et al.
Published: (2024)
Solving Minimum-Cost Reach Avoid using Reinforcement Learning
by: So, Oswin, et al.
Published: (2024)
by: So, Oswin, et al.
Published: (2024)
Implicit Safe Set Algorithm for Provably Safe Reinforcement Learning
by: Zhao, Weiye, et al.
Published: (2024)
by: Zhao, Weiye, et al.
Published: (2024)
SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning
by: Anisimov, Maksim, et al.
Published: (2026)
by: Anisimov, Maksim, et al.
Published: (2026)
Distributionally Robust Safety Verification for Markov Decision Processes
by: Mazumdar, Abhijit, et al.
Published: (2024)
by: Mazumdar, Abhijit, et al.
Published: (2024)
Leveraging Analytic Gradients in Provably Safe Reinforcement Learning
by: Walter, Tim, et al.
Published: (2025)
by: Walter, Tim, et al.
Published: (2025)
Dynamic Model Predictive Shielding for Provably Safe Reinforcement Learning
by: Banerjee, Arko, et al.
Published: (2024)
by: Banerjee, Arko, et al.
Published: (2024)
Imitate the Good and Avoid the Bad: An Incremental Approach to Safe Reinforcement Learning
by: Hoang, Huy, et al.
Published: (2023)
by: Hoang, Huy, et al.
Published: (2023)
State Entropy Regularization for Robust Reinforcement Learning
by: Ashlag, Yonatan, et al.
Published: (2025)
by: Ashlag, Yonatan, et al.
Published: (2025)
Provable Traffic Rule Compliance in Safe Reinforcement Learning on the Open Sea
by: Krasowski, Hanna, et al.
Published: (2024)
by: Krasowski, Hanna, et al.
Published: (2024)
Solving Reach-Avoid-Stay Problems Using Deep Deterministic Policy Gradients
by: Chenevert, Gabriel, et al.
Published: (2024)
by: Chenevert, Gabriel, et al.
Published: (2024)
Avoiding Premature Collapse: Adaptive Annealing for Entropy-Regularized Structural Inference
by: Liu, Yizhi
Published: (2026)
by: Liu, Yizhi
Published: (2026)
Infinite-Horizon Reach-Avoid Zero-Sum Games via Deep Reinforcement Learning
by: Li, Jingqi, et al.
Published: (2022)
by: Li, Jingqi, et al.
Published: (2022)
Convergence Theorems for Entropy-Regularized and Distributional Reinforcement Learning
by: Jhaveri, Yash, et al.
Published: (2025)
by: Jhaveri, Yash, et al.
Published: (2025)
A Provable Approach for End-to-End Safe Reinforcement Learning
by: Wachi, Akifumi, et al.
Published: (2025)
by: Wachi, Akifumi, et al.
Published: (2025)
The Edge-of-Reach Problem in Offline Model-Based Reinforcement Learning
by: Sims, Anya, et al.
Published: (2024)
by: Sims, Anya, et al.
Published: (2024)
Reinforcement Learning with Adaptive Regularization for Safe Control of Critical Systems
by: Tian, Haozhe, et al.
Published: (2024)
by: Tian, Haozhe, et al.
Published: (2024)
Provably Safe Model Updates
by: Elmecker-Plakolm, Leo, et al.
Published: (2025)
by: Elmecker-Plakolm, Leo, et al.
Published: (2025)
Entropy Regularized Task Representation Learning for Offline Meta-Reinforcement Learning
by: Nakhaei, Mohammadreza, et al.
Published: (2024)
by: Nakhaei, Mohammadreza, et al.
Published: (2024)
Entropy-Regularized Adjoint Matching for Offline Reinforcement Learning
by: Ghanem, Abdelghani, et al.
Published: (2026)
by: Ghanem, Abdelghani, et al.
Published: (2026)
Global Convergence of Wasserstein Policy Gradient for Entropy-Regularized Reinforcement Learning
by: Zhu, Zhaoyu, et al.
Published: (2026)
by: Zhu, Zhaoyu, et al.
Published: (2026)
Provably Convergent Actor-Critic for MARL through Risk-aversion
by: Zhang, Yizhou, et al.
Published: (2026)
by: Zhang, Yizhou, et al.
Published: (2026)
Solving Parameter-Robust Avoid Problems with Unknown Feasibility using Reinforcement Learning
by: So, Oswin, et al.
Published: (2026)
by: So, Oswin, et al.
Published: (2026)
Viability of Future Actions: Robust Safety in Reinforcement Learning via Entropy Regularization
by: Massiani, Pierre-François, et al.
Published: (2025)
by: Massiani, Pierre-François, et al.
Published: (2025)
SafeOR-Gym: A Benchmark Suite for Safe Reinforcement Learning Algorithms on Practical Operations Research Problems
by: Ramanujam, Asha, et al.
Published: (2025)
by: Ramanujam, Asha, et al.
Published: (2025)
Towards Provable Emergence of In-Context Reinforcement Learning
by: Wang, Jiuqi, et al.
Published: (2025)
by: Wang, Jiuqi, et al.
Published: (2025)
Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies
by: Li, Kevin, et al.
Published: (2025)
by: Li, Kevin, et al.
Published: (2025)
Relative Entropy Regularized Reinforcement Learning for Efficient Encrypted Policy Synthesis
by: Suh, Jihoon, et al.
Published: (2025)
by: Suh, Jihoon, et al.
Published: (2025)
Convergent Reinforcement Learning Algorithms for Stochastic Shortest Path Problem
by: Guin, Soumyajit, et al.
Published: (2025)
by: Guin, Soumyajit, et al.
Published: (2025)
Row-Stochastic Matrices Can Provably Outperform Doubly Stochastic Matrices in Decentralized Learning
by: Liu, Bing, et al.
Published: (2025)
by: Liu, Bing, et al.
Published: (2025)
Model-Free Robust $ϕ$-Divergence Reinforcement Learning Using Both Offline and Online Data
by: Panaganti, Kishan, et al.
Published: (2024)
by: Panaganti, Kishan, et al.
Published: (2024)
Safely Learning Controlled Stochastic Dynamics
by: Brogat-Motte, Luc, et al.
Published: (2025)
by: Brogat-Motte, Luc, et al.
Published: (2025)
Provable Partially Observable Reinforcement Learning with Privileged Information
by: Cai, Yang, et al.
Published: (2024)
by: Cai, Yang, et al.
Published: (2024)
Invertible ResNets for Inverse Imaging Problems: Competitive Performance with Provable Regularization Properties
by: Arndt, Clemens, et al.
Published: (2024)
by: Arndt, Clemens, et al.
Published: (2024)
Stabilizing Information Flow Entropy: Regularization for Safe and Interpretable Autonomous Driving Perception
by: Yang, Haobo, et al.
Published: (2025)
by: Yang, Haobo, et al.
Published: (2025)
Similar Items
-
Safe Reinforcement Learning for Constrained Markov Decision Processes with Stochastic Stopping Time
by: Mazumdar, Abhijit, et al.
Published: (2024) -
Data-Driven Robust Safety Verification for Markov Decision Processes
by: Mazumdar, Abhijit, et al.
Published: (2025) -
Robust Correlated Equilibrium: Definition and Computation
by: Misra, Rahul, et al.
Published: (2023) -
Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning
by: Pan, Jingduo, et al.
Published: (2026) -
An Online Multiobjective Policy Gradient for Long-run Average-reward Markov Decision Process
by: Misra, Rahul, et al.
Published: (2025)