Safe Reinforcement Learning for Constrained Markov Decision Processes with Stochastic Stopping Time
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mazumdar, Abhijit, Wisniewski, Rafal, Bujorianu, Manuela L. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Provably Safe Reinforcement Learning for Stochastic Reach-Avoid Problems with Entropy Regularization
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2026)
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2026)
Data-Driven Robust Safety Verification for Markov Decision Processes
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2025)
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2025)
An Online Multiobjective Policy Gradient for Long-run Average-reward Markov Decision Process
von: Misra, Rahul, et al.
Veröffentlicht: (2025)
von: Misra, Rahul, et al.
Veröffentlicht: (2025)
Large Deviations in Safety-Critical Systems with Probabilistic Initial Conditions
von: Gomez, Aitor R., et al.
Veröffentlicht: (2024)
von: Gomez, Aitor R., et al.
Veröffentlicht: (2024)
Distributionally Robust Safety Verification for Markov Decision Processes
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2024)
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2024)
Flipping-based Policy for Chance-Constrained Markov Decision Processes
von: Shen, Xun, et al.
Veröffentlicht: (2024)
von: Shen, Xun, et al.
Veröffentlicht: (2024)
Non-stationary and Varying-discounting Markov Decision Processes for Reinforcement Learning
von: Chen, Zhizuo, et al.
Veröffentlicht: (2025)
von: Chen, Zhizuo, et al.
Veröffentlicht: (2025)
Online Reinforcement Learning in Markov Decision Process Using Linear Programming
von: Leon, Vincent, et al.
Veröffentlicht: (2023)
von: Leon, Vincent, et al.
Veröffentlicht: (2023)
Achieving Instance-dependent Sample Complexity for Constrained Markov Decision Process
von: Jiang, Jiashuo, et al.
Veröffentlicht: (2024)
von: Jiang, Jiashuo, et al.
Veröffentlicht: (2024)
Weakly Time-Coupled Approximation of Markov Decision Processes
von: Soheili, Negar, et al.
Veröffentlicht: (2026)
von: Soheili, Negar, et al.
Veröffentlicht: (2026)
Finite-Time Complexity of Online Primal-Dual Natural Actor-Critic Algorithm for Constrained Markov Decision Processes
von: Zeng, Sihan, et al.
Veröffentlicht: (2021)
von: Zeng, Sihan, et al.
Veröffentlicht: (2021)
Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form
von: Kitamura, Toshinori, et al.
Veröffentlicht: (2024)
von: Kitamura, Toshinori, et al.
Veröffentlicht: (2024)
Robust Correlated Equilibrium: Definition and Computation
von: Misra, Rahul, et al.
Veröffentlicht: (2023)
von: Misra, Rahul, et al.
Veröffentlicht: (2023)
Reinforcement Learning with Function Approximation for Non-Markov Processes
von: Kara, Ali Devran
Veröffentlicht: (2026)
von: Kara, Ali Devran
Veröffentlicht: (2026)
On Convergence of Average-Reward Q-Learning in Weakly Communicating Markov Decision Processes
von: Wan, Yi, et al.
Veröffentlicht: (2024)
von: Wan, Yi, et al.
Veröffentlicht: (2024)
Risk-sensitive Markov Decision Process and Learning under General Utility Functions
von: Wu, Zhengqi, et al.
Veröffentlicht: (2023)
von: Wu, Zhengqi, et al.
Veröffentlicht: (2023)
Online Markov Decision Processes with Terminal Law Constraints
von: Moreno, Bianca Marin, et al.
Veröffentlicht: (2026)
von: Moreno, Bianca Marin, et al.
Veröffentlicht: (2026)
Risk-Aware Safe Reinforcement Learning for Control of Stochastic Linear Systems
von: Esmaeili, Babak, et al.
Veröffentlicht: (2025)
von: Esmaeili, Babak, et al.
Veröffentlicht: (2025)
Optimal Sample Complexity for Average Reward Markov Decision Processes
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
Learning Sequential Decisions from Multiple Sources via Group-Robust Markov Decision Processes
von: Xu, Mingyuan, et al.
Veröffentlicht: (2026)
von: Xu, Mingyuan, et al.
Veröffentlicht: (2026)
Efficient Algorithms for Robust Markov Decision Processes with $s$-Rectangular Ambiguity Sets
von: Ho, Chin Pang, et al.
Veröffentlicht: (2026)
von: Ho, Chin Pang, et al.
Veröffentlicht: (2026)
Learning to Stop: Deep Learning for Mean Field Optimal Stopping
von: Magnino, Lorenzo, et al.
Veröffentlicht: (2024)
von: Magnino, Lorenzo, et al.
Veröffentlicht: (2024)
Resilient Constrained Reinforcement Learning
von: Ding, Dongsheng, et al.
Veröffentlicht: (2023)
von: Ding, Dongsheng, et al.
Veröffentlicht: (2023)
Bellman Optimality of Average-Reward Robust Markov Decision Processes with a Constant Gain
von: Wang, Shengbo, et al.
Veröffentlicht: (2025)
von: Wang, Shengbo, et al.
Veröffentlicht: (2025)
Fast Two-Time-Scale Stochastic Gradient Method with Applications in Reinforcement Learning
von: Zeng, Sihan, et al.
Veröffentlicht: (2024)
von: Zeng, Sihan, et al.
Veröffentlicht: (2024)
Stochastic-Constrained Stochastic Optimization with Markovian Data
von: Kim, Yeongjong, et al.
Veröffentlicht: (2023)
von: Kim, Yeongjong, et al.
Veröffentlicht: (2023)
Bayesian Ambiguity Contraction-based Adaptive Robust Markov Decision Processes for Adversarial Surveillance Missions
von: Choi, Jimin, et al.
Veröffentlicht: (2025)
von: Choi, Jimin, et al.
Veröffentlicht: (2025)
Accelerating Optimization via Differentiable Stopping Time
von: Xie, Zhonglin, et al.
Veröffentlicht: (2025)
von: Xie, Zhonglin, et al.
Veröffentlicht: (2025)
A Two-Time-Scale Stochastic Optimization Framework with Applications in Control and Reinforcement Learning
von: Zeng, Sihan, et al.
Veröffentlicht: (2021)
von: Zeng, Sihan, et al.
Veröffentlicht: (2021)
Provably Efficient Representation Selection in Low-rank Markov Decision Processes: From Online to Offline RL
von: Zhang, Weitong, et al.
Veröffentlicht: (2021)
von: Zhang, Weitong, et al.
Veröffentlicht: (2021)
Sampling-based Safe Reinforcement Learning for Nonlinear Dynamical Systems
von: Suttle, Wesley A., et al.
Veröffentlicht: (2024)
von: Suttle, Wesley A., et al.
Veröffentlicht: (2024)
Rate-Optimal Regret for the Safe Learning-based Control of the Constrained Linear Quadratic Regulator
von: Hutchinson, Spencer, et al.
Veröffentlicht: (2026)
von: Hutchinson, Spencer, et al.
Veröffentlicht: (2026)
Safe-EF: Error Feedback for Nonsmooth Constrained Optimization
von: Islamov, Rustem, et al.
Veröffentlicht: (2025)
von: Islamov, Rustem, et al.
Veröffentlicht: (2025)
Structured Reinforcement Learning for Combinatorial Decision-Making
von: Hoppe, Heiko, et al.
Veröffentlicht: (2025)
von: Hoppe, Heiko, et al.
Veröffentlicht: (2025)
Safe Time-Varying Optimization based on Gaussian Processes with Spatio-Temporal Kernel
von: Li, Jialin, et al.
Veröffentlicht: (2024)
von: Li, Jialin, et al.
Veröffentlicht: (2024)
Deep Learning for the Multiple Optimal Stopping Problem
von: Laurière, Mathieu, et al.
Veröffentlicht: (2025)
von: Laurière, Mathieu, et al.
Veröffentlicht: (2025)
humancompatible.train: Implementing Optimization Algorithms for Stochastically-Constrained Stochastic Optimization Problems
von: Kliachkin, Andrii, et al.
Veröffentlicht: (2025)
von: Kliachkin, Andrii, et al.
Veröffentlicht: (2025)
Unified Precision-Guaranteed Stopping Rules for Contextual Learning
von: Ding, Mingrui, et al.
Veröffentlicht: (2026)
von: Ding, Mingrui, et al.
Veröffentlicht: (2026)
Constrained Stochastic Spectral Preconditioning Converges for Nonconvex Objectives
von: Oikonomidis, Konstantinos, et al.
Veröffentlicht: (2026)
von: Oikonomidis, Konstantinos, et al.
Veröffentlicht: (2026)
Hierarchical Reinforcement Learning Framework for Stochastic Spaceflight Campaign Design
von: Takubo, Yuji, et al.
Veröffentlicht: (2021)
von: Takubo, Yuji, et al.
Veröffentlicht: (2021)
Ähnliche Einträge
-
Provably Safe Reinforcement Learning for Stochastic Reach-Avoid Problems with Entropy Regularization
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2026) -
Data-Driven Robust Safety Verification for Markov Decision Processes
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2025) -
An Online Multiobjective Policy Gradient for Long-run Average-reward Markov Decision Process
von: Misra, Rahul, et al.
Veröffentlicht: (2025) -
Large Deviations in Safety-Critical Systems with Probabilistic Initial Conditions
von: Gomez, Aitor R., et al.
Veröffentlicht: (2024) -
Distributionally Robust Safety Verification for Markov Decision Processes
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2024)