Controlling Underestimation Bias in Constrained Reinforcement Learning for Safe Exploration
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gao, Shiqing, Ding, Jiaxin, Fu, Luoyi, Wang, Xinbing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Extreme Value Policy Optimization for Safe Reinforcement Learning
von: Gao, Shiqing, et al.
Veröffentlicht: (2026)
von: Gao, Shiqing, et al.
Veröffentlicht: (2026)
Exterior Penalty Policy Optimization with Penalty Metric Network under Constraints
von: Gao, Shiqing, et al.
Veröffentlicht: (2024)
von: Gao, Shiqing, et al.
Veröffentlicht: (2024)
Characterizing the Influence of Topology on Graph Learning Tasks
von: Wu, Kailong, et al.
Veröffentlicht: (2024)
von: Wu, Kailong, et al.
Veröffentlicht: (2024)
Off-Policy Safe Reinforcement Learning with Constrained Optimistic Exploration
von: Li, Guopeng, et al.
Veröffentlicht: (2026)
von: Li, Guopeng, et al.
Veröffentlicht: (2026)
CHAINSFORMER: Numerical Reasoning on Knowledge Graphs from a Chain Perspective
von: Zhao, Ze, et al.
Veröffentlicht: (2025)
von: Zhao, Ze, et al.
Veröffentlicht: (2025)
Constrained Reinforcement Learning for Safe Heat Pump Control
von: Zhang, Baohe, et al.
Veröffentlicht: (2024)
von: Zhang, Baohe, et al.
Veröffentlicht: (2024)
Reinforcement Learning by Guided Safe Exploration
von: Yang, Qisong, et al.
Veröffentlicht: (2023)
von: Yang, Qisong, et al.
Veröffentlicht: (2023)
Safe and Balanced: A Framework for Constrained Multi-Objective Reinforcement Learning
von: Gu, Shangding, et al.
Veröffentlicht: (2024)
von: Gu, Shangding, et al.
Veröffentlicht: (2024)
Vulnerability Analysis of Safe Reinforcement Learning via Inverse Constrained Reinforcement Learning
von: Fan, Jialiang, et al.
Veröffentlicht: (2026)
von: Fan, Jialiang, et al.
Veröffentlicht: (2026)
Kernel-Based Safe Exploration in Deep Reinforcement Learning
von: Majumdar, Rupak, et al.
Veröffentlicht: (2026)
von: Majumdar, Rupak, et al.
Veröffentlicht: (2026)
Revisiting Safe Exploration in Safe Reinforcement learning
von: Eckel, David, et al.
Veröffentlicht: (2024)
von: Eckel, David, et al.
Veröffentlicht: (2024)
Enhance Exploration in Safe Reinforcement Learning with Contrastive Representation Learning
von: Doan, Duc Kien, et al.
Veröffentlicht: (2025)
von: Doan, Duc Kien, et al.
Veröffentlicht: (2025)
ActSafe: Active Exploration with Safety Constraints for Reinforcement Learning
von: As, Yarden, et al.
Veröffentlicht: (2024)
von: As, Yarden, et al.
Veröffentlicht: (2024)
Latent Safety-Constrained Policy Approach for Safe Offline Reinforcement Learning
von: Koirala, Prajwal, et al.
Veröffentlicht: (2024)
von: Koirala, Prajwal, et al.
Veröffentlicht: (2024)
RADAR: Reasoning as Discrimination with Aligned Representations for LLM-based Knowledge Graph Reasoning
von: Xue, Bo, et al.
Veröffentlicht: (2026)
von: Xue, Bo, et al.
Veröffentlicht: (2026)
Provably Efficient Exploration in Inverse Constrained Reinforcement Learning
von: Yue, Bo, et al.
Veröffentlicht: (2024)
von: Yue, Bo, et al.
Veröffentlicht: (2024)
Optimistic Exploration for Risk-Averse Constrained Reinforcement Learning
von: McCarthy, James, et al.
Veröffentlicht: (2025)
von: McCarthy, James, et al.
Veröffentlicht: (2025)
Feasibility Consistent Representation Learning for Safe Reinforcement Learning
von: Cen, Zhepeng, et al.
Veröffentlicht: (2024)
von: Cen, Zhepeng, et al.
Veröffentlicht: (2024)
OXYGENERATOR: Reconstructing Global Ocean Deoxygenation Over a Century with Deep Learning
von: Lu, Bin, et al.
Veröffentlicht: (2024)
von: Lu, Bin, et al.
Veröffentlicht: (2024)
Towards Safe Reinforcement Learning via Constraining Conditional Value-at-Risk
von: Ying, Chengyang, et al.
Veröffentlicht: (2022)
von: Ying, Chengyang, et al.
Veröffentlicht: (2022)
OASIS: Conditional Distribution Shaping for Offline Safe Reinforcement Learning
von: Yao, Yihang, et al.
Veröffentlicht: (2024)
von: Yao, Yihang, et al.
Veröffentlicht: (2024)
Certificated Actor-Critic: Hierarchical Reinforcement Learning with Control Barrier Functions for Safe Navigation
von: Xie, Junjun, et al.
Veröffentlicht: (2025)
von: Xie, Junjun, et al.
Veröffentlicht: (2025)
Reinforcement Learning with Adaptive Regularization for Safe Control of Critical Systems
von: Tian, Haozhe, et al.
Veröffentlicht: (2024)
von: Tian, Haozhe, et al.
Veröffentlicht: (2024)
Safe Reinforcement Learning for Constrained Markov Decision Processes with Stochastic Stopping Time
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2024)
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2024)
Offline Inverse Constrained Reinforcement Learning for Safe-Critical Decision Making in Healthcare
von: Fang, Nan, et al.
Veröffentlicht: (2024)
von: Fang, Nan, et al.
Veröffentlicht: (2024)
Resilient Constrained Reinforcement Learning
von: Ding, Dongsheng, et al.
Veröffentlicht: (2023)
von: Ding, Dongsheng, et al.
Veröffentlicht: (2023)
Mitigating System Bias in Resource Constrained Asynchronous Federated Learning Systems
von: Gao, Jikun, et al.
Veröffentlicht: (2024)
von: Gao, Jikun, et al.
Veröffentlicht: (2024)
Safe Reinforcement Learning for Real-World Engine Control
von: Bedei, Julian, et al.
Veröffentlicht: (2025)
von: Bedei, Julian, et al.
Veröffentlicht: (2025)
Counterfactually Safe Reinforcement Learning
von: Li, Jingyi, et al.
Veröffentlicht: (2026)
von: Li, Jingyi, et al.
Veröffentlicht: (2026)
Safe In-Context Reinforcement Learning
von: Moeini, Amir, et al.
Veröffentlicht: (2025)
von: Moeini, Amir, et al.
Veröffentlicht: (2025)
Enhancing Efficiency of Safe Reinforcement Learning via Sample Manipulation
von: Gu, Shangding, et al.
Veröffentlicht: (2024)
von: Gu, Shangding, et al.
Veröffentlicht: (2024)
A CMDP-within-online framework for Meta-Safe Reinforcement Learning
von: Khattar, Vanshaj, et al.
Veröffentlicht: (2024)
von: Khattar, Vanshaj, et al.
Veröffentlicht: (2024)
Secure Resource Allocation via Constrained Deep Reinforcement Learning
von: Sun, Jianfei, et al.
Veröffentlicht: (2025)
von: Sun, Jianfei, et al.
Veröffentlicht: (2025)
Efficient Exploration in Average-Reward Constrained Reinforcement Learning: Achieving Near-Optimal Regret With Posterior Sampling
von: Provodin, Danil, et al.
Veröffentlicht: (2024)
von: Provodin, Danil, et al.
Veröffentlicht: (2024)
Hierarchical Reinforcement Learning for Safe Mapless Navigation with Congestion Estimation
von: Gao, Jianqi, et al.
Veröffentlicht: (2025)
von: Gao, Jianqi, et al.
Veröffentlicht: (2025)
Off-Policy Primal-Dual Safe Reinforcement Learning
von: Wu, Zifan, et al.
Veröffentlicht: (2024)
von: Wu, Zifan, et al.
Veröffentlicht: (2024)
Exploration and Evaluation of Bias in Cyberbullying Detection with Machine Learning
von: Root, Andrew, et al.
Veröffentlicht: (2024)
von: Root, Andrew, et al.
Veröffentlicht: (2024)
Policy Bifurcation in Safe Reinforcement Learning
von: Zou, Wenjun, et al.
Veröffentlicht: (2024)
von: Zou, Wenjun, et al.
Veröffentlicht: (2024)
Learning-Driven Exploration for Reinforcement Learning
von: Usama, Muhammad, et al.
Veröffentlicht: (2019)
von: Usama, Muhammad, et al.
Veröffentlicht: (2019)
Safety Constrained Multi-Agent Reinforcement Learning for Active Voltage Control
von: Qu, Yang, et al.
Veröffentlicht: (2024)
von: Qu, Yang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Extreme Value Policy Optimization for Safe Reinforcement Learning
von: Gao, Shiqing, et al.
Veröffentlicht: (2026) -
Exterior Penalty Policy Optimization with Penalty Metric Network under Constraints
von: Gao, Shiqing, et al.
Veröffentlicht: (2024) -
Characterizing the Influence of Topology on Graph Learning Tasks
von: Wu, Kailong, et al.
Veröffentlicht: (2024) -
Off-Policy Safe Reinforcement Learning with Constrained Optimistic Exploration
von: Li, Guopeng, et al.
Veröffentlicht: (2026) -
CHAINSFORMER: Numerical Reasoning on Knowledge Graphs from a Chain Perspective
von: Zhao, Ze, et al.
Veröffentlicht: (2025)