Efficient Policy Evaluation with Safety Constraint for Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Claire, Liu, Shuze Daniel, Zhang, Shangtong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Efficient Multi-Policy Evaluation for Reinforcement Learning
von: Liu, Shuze Daniel, et al.
Veröffentlicht: (2024)
von: Liu, Shuze Daniel, et al.
Veröffentlicht: (2024)
Doubly Optimal Policy Evaluation for Reinforcement Learning
von: Liu, Shuze Daniel, et al.
Veröffentlicht: (2024)
von: Liu, Shuze Daniel, et al.
Veröffentlicht: (2024)
Efficient Policy Evaluation with Offline Data Informed Behavior Policy Design
von: Liu, Shuze, et al.
Veröffentlicht: (2023)
von: Liu, Shuze, et al.
Veröffentlicht: (2023)
Convergence of Two-Timescale Markovian Stochastic Approximations with Applications in Reinforcement Learning
von: Mahadevan, Vagul, et al.
Veröffentlicht: (2026)
von: Mahadevan, Vagul, et al.
Veröffentlicht: (2026)
The ODE Method for Stochastic Approximation and Reinforcement Learning with Markovian Noise
von: Liu, Shuze Daniel, et al.
Veröffentlicht: (2024)
von: Liu, Shuze Daniel, et al.
Veröffentlicht: (2024)
Beyond Linear Attention: Softmax Transformers Implement In-Context Reinforcement Learning
von: Xie, Zixuan, et al.
Veröffentlicht: (2026)
von: Xie, Zixuan, et al.
Veröffentlicht: (2026)
Predicting Plasticity in Deep Continual Learning: A Theoretical Perspective
von: Wang, Jiuqi, et al.
Veröffentlicht: (2026)
von: Wang, Jiuqi, et al.
Veröffentlicht: (2026)
Towards Formalizing Reinforcement Learning Theory
von: Zhang, Shangtong
Veröffentlicht: (2025)
von: Zhang, Shangtong
Veröffentlicht: (2025)
CRASH: Challenging Reinforcement-Learning Based Adversarial Scenarios For Safety Hardening
von: Kulkarni, Amar, et al.
Veröffentlicht: (2024)
von: Kulkarni, Amar, et al.
Veröffentlicht: (2024)
MathlibLemma: Folklore Lemma Generation and Benchmark for Formal Mathematics
von: Liu, Xinyu, et al.
Veröffentlicht: (2026)
von: Liu, Xinyu, et al.
Veröffentlicht: (2026)
Extensions of Robbins-Siegmund Theorem with Applications in Reinforcement Learning
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
Towards Provable Emergence of In-Context Reinforcement Learning
von: Wang, Jiuqi, et al.
Veröffentlicht: (2025)
von: Wang, Jiuqi, et al.
Veröffentlicht: (2025)
Convergence and Emergence of In-Context Reinforcement Learning with Chain of Thought
von: Xie, Zixuan, et al.
Veröffentlicht: (2026)
von: Xie, Zixuan, et al.
Veröffentlicht: (2026)
Counterfactual Explanations for Continuous Action Reinforcement Learning
von: Dong, Shuyang, et al.
Veröffentlicht: (2025)
von: Dong, Shuyang, et al.
Veröffentlicht: (2025)
Almost Sure Convergence Rates of Stochastic Approximation and Reinforcement Learning via a Poisson-Moreau Drift
von: Liu, Xinyu, et al.
Veröffentlicht: (2026)
von: Liu, Xinyu, et al.
Veröffentlicht: (2026)
Concurrent Learning of Policy and Unknown Safety Constraints in Reinforcement Learning
von: Yifru, Lunet, et al.
Veröffentlicht: (2024)
von: Yifru, Lunet, et al.
Veröffentlicht: (2024)
Sample-Efficient Policy Constraint Offline Deep Reinforcement Learning based on Sample Filtering
von: Chen, Yuanhao, et al.
Veröffentlicht: (2025)
von: Chen, Yuanhao, et al.
Veröffentlicht: (2025)
Transformers Can Learn Temporal Difference Methods for In-Context Reinforcement Learning
von: Wang, Jiuqi, et al.
Veröffentlicht: (2024)
von: Wang, Jiuqi, et al.
Veröffentlicht: (2024)
Adaptive Policy Selection and Fine-Tuning under Interaction Budgets for Offline-to-Online Reinforcement Learning
von: Bozkurt, Alper Kamil, et al.
Veröffentlicht: (2026)
von: Bozkurt, Alper Kamil, et al.
Veröffentlicht: (2026)
Almost Sure Convergence Rates and Concentration of Stochastic Approximation and Reinforcement Learning with Markovian Noise
von: Qian, Xiaochi, et al.
Veröffentlicht: (2024)
von: Qian, Xiaochi, et al.
Veröffentlicht: (2024)
Revisiting a Design Choice in Gradient Temporal Difference Learning
von: Qian, Xiaochi, et al.
Veröffentlicht: (2023)
von: Qian, Xiaochi, et al.
Veröffentlicht: (2023)
On the Divergence of Differential Temporal Difference Learning without Local Clocks
von: Antrobius, David, et al.
Veröffentlicht: (2026)
von: Antrobius, David, et al.
Veröffentlicht: (2026)
Group Fairness in Multi-Task Reinforcement Learning
von: Song, Kefan, et al.
Veröffentlicht: (2025)
von: Song, Kefan, et al.
Veröffentlicht: (2025)
Probabilistic Constraint for Safety-Critical Reinforcement Learning
von: Chen, Weiqin, et al.
Veröffentlicht: (2023)
von: Chen, Weiqin, et al.
Veröffentlicht: (2023)
Global Optimality and Finite Sample Analysis of Softmax Off-Policy Actor Critic under State Distribution Mismatch
von: Zhang, Shangtong, et al.
Veröffentlicht: (2021)
von: Zhang, Shangtong, et al.
Veröffentlicht: (2021)
Conservative Distributional Reinforcement Learning with Safety Constraints
von: Zhang, Hengrui, et al.
Veröffentlicht: (2022)
von: Zhang, Hengrui, et al.
Veröffentlicht: (2022)
Online Estimation and Inference for Robust Policy Evaluation in Reinforcement Learning
von: Liu, Weidong, et al.
Veröffentlicht: (2023)
von: Liu, Weidong, et al.
Veröffentlicht: (2023)
Almost Sure Convergence of Linear Temporal Difference Learning with Arbitrary Features
von: Wang, Jiuqi, et al.
Veröffentlicht: (2024)
von: Wang, Jiuqi, et al.
Veröffentlicht: (2024)
Policy Constraint by Only Support Constraint for Offline Reinforcement Learning
von: Gao, Yunkai, et al.
Veröffentlicht: (2025)
von: Gao, Yunkai, et al.
Veröffentlicht: (2025)
Multi-agent Markov Entanglement
von: Chen, Shuze, et al.
Veröffentlicht: (2025)
von: Chen, Shuze, et al.
Veröffentlicht: (2025)
Offline Two-Player Zero-Sum Markov Games with KL Regularization
von: Chen, Claire, et al.
Veröffentlicht: (2026)
von: Chen, Claire, et al.
Veröffentlicht: (2026)
A Survey of In-Context Reinforcement Learning
von: Moeini, Amir, et al.
Veröffentlicht: (2025)
von: Moeini, Amir, et al.
Veröffentlicht: (2025)
Safe In-Context Reinforcement Learning
von: Moeini, Amir, et al.
Veröffentlicht: (2025)
von: Moeini, Amir, et al.
Veröffentlicht: (2025)
Linear $Q$-Learning Does Not Diverge in $L^2$: Convergence Rates to a Bounded Set
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
Constraint-Conditioned Policy Optimization for Versatile Safe Reinforcement Learning
von: Yao, Yihang, et al.
Veröffentlicht: (2023)
von: Yao, Yihang, et al.
Veröffentlicht: (2023)
Adaptive Scaling of Policy Constraints for Offline Reinforcement Learning
von: Jing, Tan, et al.
Veröffentlicht: (2025)
von: Jing, Tan, et al.
Veröffentlicht: (2025)
Efficient Online Reinforcement Learning for Diffusion Policy
von: Ma, Haitong, et al.
Veröffentlicht: (2025)
von: Ma, Haitong, et al.
Veröffentlicht: (2025)
Safety-Oriented Pruning and Interpretation of Reinforcement Learning Policies
von: Gross, Dennis, et al.
Veröffentlicht: (2024)
von: Gross, Dennis, et al.
Veröffentlicht: (2024)
Safe Reinforcement Learning with Learned Non-Markovian Safety Constraints
von: Low, Siow Meng, et al.
Veröffentlicht: (2024)
von: Low, Siow Meng, et al.
Veröffentlicht: (2024)
ActSafe: Active Exploration with Safety Constraints for Reinforcement Learning
von: As, Yarden, et al.
Veröffentlicht: (2024)
von: As, Yarden, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Efficient Multi-Policy Evaluation for Reinforcement Learning
von: Liu, Shuze Daniel, et al.
Veröffentlicht: (2024) -
Doubly Optimal Policy Evaluation for Reinforcement Learning
von: Liu, Shuze Daniel, et al.
Veröffentlicht: (2024) -
Efficient Policy Evaluation with Offline Data Informed Behavior Policy Design
von: Liu, Shuze, et al.
Veröffentlicht: (2023) -
Convergence of Two-Timescale Markovian Stochastic Approximations with Applications in Reinforcement Learning
von: Mahadevan, Vagul, et al.
Veröffentlicht: (2026) -
The ODE Method for Stochastic Approximation and Reinforcement Learning with Markovian Noise
von: Liu, Shuze Daniel, et al.
Veröffentlicht: (2024)