Safety-Aware Reinforcement Learning for Control via Risk-Sensitive Action-Value Iteration and Quantile Regression
Fuente:
arXiv
Guardado en:
| Autores principales: | Enwerem, Clinton, Puranic, Aniruddh G., Baras, John S., Belta, Calin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Learning from Imperfect Demonstrations via Temporal Behavior Tree-Guided Trajectory Repair
por: Puranic, Aniruddh G., et al.
Publicado: (2026)
por: Puranic, Aniruddh G., et al.
Publicado: (2026)
Risk-Constrained Belief-Space Optimization for Safe Control under Latent Uncertainty
por: Enwerem, Clinton, et al.
Publicado: (2026)
por: Enwerem, Clinton, et al.
Publicado: (2026)
Variational Neural Belief Parameterizations for Robust Dexterous Grasping under Multimodal Uncertainty
por: Enwerem, Clinton, et al.
Publicado: (2026)
por: Enwerem, Clinton, et al.
Publicado: (2026)
Safe Collective Control under Noisy Inputs and Competing Constraints via Non-Smooth Barrier Functions
por: Enwerem, Clinton, et al.
Publicado: (2023)
por: Enwerem, Clinton, et al.
Publicado: (2023)
On the Stability and Realizability of Recurrent Polynomial Surrogate Ternary Logic Gate Networks
por: Damera, Sai Sandeep, et al.
Publicado: (2026)
por: Damera, Sai Sandeep, et al.
Publicado: (2026)
Robust Stochastic Shortest-Path Planning via Risk-Sensitive Incremental Sampling
por: Enwerem, Clinton, et al.
Publicado: (2024)
por: Enwerem, Clinton, et al.
Publicado: (2024)
Ternary Logic Encodings of Temporal Behavior Trees with Application to Control Synthesis
por: Matheu, Ryan, et al.
Publicado: (2026)
por: Matheu, Ryan, et al.
Publicado: (2026)
Accelerated Learning with Linear Temporal Logic using Differentiable Simulation
por: Bozkurt, Alper Kamil, et al.
Publicado: (2025)
por: Bozkurt, Alper Kamil, et al.
Publicado: (2025)
Polynomial Surrogate Training for Differentiable Ternary Logic Gate Networks
por: Damera, Sai Sandeep, et al.
Publicado: (2026)
por: Damera, Sai Sandeep, et al.
Publicado: (2026)
Learning Safety for Obstacle Avoidance via Control Barrier Functions
por: Liu, Shuo, et al.
Publicado: (2025)
por: Liu, Shuo, et al.
Publicado: (2025)
On Learning the Tail Quantiles of Driving Behavior Distributions via Quantile Regression and Flows
por: Tee, Jia Yu, et al.
Publicado: (2023)
por: Tee, Jia Yu, et al.
Publicado: (2023)
SPLASH! Sample-efficient Preference-based inverse reinforcement learning for Long-horizon Adversarial tasks from Suboptimal Hierarchical demonstrations
por: Crowley, Peter, et al.
Publicado: (2025)
por: Crowley, Peter, et al.
Publicado: (2025)
Iterative Convex Optimization with Control Barrier Functions for Obstacle Avoidance among Polytopes
por: Liu, Shuo, et al.
Publicado: (2026)
por: Liu, Shuo, et al.
Publicado: (2026)
Safety-Critical Planning and Control for Dynamic Obstacle Avoidance Using Control Barrier Functions
por: Liu, Shuo, et al.
Publicado: (2024)
por: Liu, Shuo, et al.
Publicado: (2024)
Quantile-Coupled Flow Matching for Distributional Reinforcement Learning
por: Groom, Michael, et al.
Publicado: (2026)
por: Groom, Michael, et al.
Publicado: (2026)
Risk-Sensitive Reinforcement Learning with Exponential Criteria
por: Noorani, Erfaun, et al.
Publicado: (2022)
por: Noorani, Erfaun, et al.
Publicado: (2022)
Constraint-Aware Reinforcement Learning via Adaptive Action Scaling
por: Dawood, Murad, et al.
Publicado: (2025)
por: Dawood, Murad, et al.
Publicado: (2025)
Learning Risk-Aware Quadrupedal Locomotion using Distributional Reinforcement Learning
por: Schneider, Lukas, et al.
Publicado: (2023)
por: Schneider, Lukas, et al.
Publicado: (2023)
Value Iteration for Learning Concurrently Executable Robotic Control Tasks
por: Tahmid, Sheikh A., et al.
Publicado: (2025)
por: Tahmid, Sheikh A., et al.
Publicado: (2025)
Auxiliary-Variable Adaptive Control Barrier Functions for Safety Critical Systems
por: Liu, Shuo, et al.
Publicado: (2023)
por: Liu, Shuo, et al.
Publicado: (2023)
Enhanced Quantile Regression with Spiking Neural Networks for Long-Term System Health Prognostics
por: Poland, David J
Publicado: (2025)
por: Poland, David J
Publicado: (2025)
Robustness Evaluation of Offline Reinforcement Learning for Robot Control Against Action Perturbations
por: Ayabe, Shingo, et al.
Publicado: (2024)
por: Ayabe, Shingo, et al.
Publicado: (2024)
Communication-Aware Reinforcement Learning for Cooperative Adaptive Cruise Control
por: Jiang, Sicong, et al.
Publicado: (2024)
por: Jiang, Sicong, et al.
Publicado: (2024)
Safe Hybrid-Action Reinforcement Learning-Based Decision and Control for Discretionary Lane Change
por: Xu, Ruichen, et al.
Publicado: (2024)
por: Xu, Ruichen, et al.
Publicado: (2024)
Ethics-Aware Safe Reinforcement Learning for Rare-Event Risk Control in Interactive Urban Driving
por: Li, Dianzhao, et al.
Publicado: (2025)
por: Li, Dianzhao, et al.
Publicado: (2025)
Risk-Aware Safe Reinforcement Learning for Control of Stochastic Linear Systems
por: Esmaeili, Babak, et al.
Publicado: (2025)
por: Esmaeili, Babak, et al.
Publicado: (2025)
Incorporating System-level Safety Requirements in Perception Models via Reinforcement Learning
por: Fan, Weisi, et al.
Publicado: (2024)
por: Fan, Weisi, et al.
Publicado: (2024)
Multiagent Reinforcement Learning with Neighbor Action Estimation
por: Luo, Zhenglong, et al.
Publicado: (2026)
por: Luo, Zhenglong, et al.
Publicado: (2026)
Vintix: Action Model via In-Context Reinforcement Learning
por: Polubarov, Andrey, et al.
Publicado: (2025)
por: Polubarov, Andrey, et al.
Publicado: (2025)
RobustVLA: Robustness-Aware Reinforcement Post-Training for Vision-Language-Action Models
por: Zhang, Hongyin, et al.
Publicado: (2025)
por: Zhang, Hongyin, et al.
Publicado: (2025)
Reinforcement Learning with Ensemble Model Predictive Safety Certification
por: Gronauer, Sven, et al.
Publicado: (2024)
por: Gronauer, Sven, et al.
Publicado: (2024)
Imitation Learning with Limited Actions via Diffusion Planners and Deep Koopman Controllers
por: Bi, Jianxin, et al.
Publicado: (2024)
por: Bi, Jianxin, et al.
Publicado: (2024)
Adaptive Outer-Loop Control of Quadrotors via Reinforcement Learning
por: Saj, Vishnu, et al.
Publicado: (2026)
por: Saj, Vishnu, et al.
Publicado: (2026)
STL-based Optimization of Biomolecular Neural Networks for Regression and Control
por: Palanques-Tost, Eric, et al.
Publicado: (2025)
por: Palanques-Tost, Eric, et al.
Publicado: (2025)
Reinforcement Learning with Action Chunking
por: Li, Qiyang, et al.
Publicado: (2025)
por: Li, Qiyang, et al.
Publicado: (2025)
Hysteresis-Aware Neural Network Modeling and Whole-Body Reinforcement Learning Control of Soft Robots
por: Chen, Zongyuan, et al.
Publicado: (2025)
por: Chen, Zongyuan, et al.
Publicado: (2025)
Learning from Suboptimal Data in Continuous Control via Auto-Regressive Soft Q-Network
por: Liu, Jijia, et al.
Publicado: (2025)
por: Liu, Jijia, et al.
Publicado: (2025)
ActSafe: Active Exploration with Safety Constraints for Reinforcement Learning
por: As, Yarden, et al.
Publicado: (2024)
por: As, Yarden, et al.
Publicado: (2024)
Handling Long-Term Safety and Uncertainty in Safe Reinforcement Learning
por: Günster, Jonas, et al.
Publicado: (2024)
por: Günster, Jonas, et al.
Publicado: (2024)
Stability Enhancement in Reinforcement Learning via Adaptive Control Lyapunov Function
por: Chen, Donghe, et al.
Publicado: (2025)
por: Chen, Donghe, et al.
Publicado: (2025)
Ejemplares similares
-
Learning from Imperfect Demonstrations via Temporal Behavior Tree-Guided Trajectory Repair
por: Puranic, Aniruddh G., et al.
Publicado: (2026) -
Risk-Constrained Belief-Space Optimization for Safe Control under Latent Uncertainty
por: Enwerem, Clinton, et al.
Publicado: (2026) -
Variational Neural Belief Parameterizations for Robust Dexterous Grasping under Multimodal Uncertainty
por: Enwerem, Clinton, et al.
Publicado: (2026) -
Safe Collective Control under Noisy Inputs and Competing Constraints via Non-Smooth Barrier Functions
por: Enwerem, Clinton, et al.
Publicado: (2023) -
On the Stability and Realizability of Recurrent Polynomial Surrogate Ternary Logic Gate Networks
por: Damera, Sai Sandeep, et al.
Publicado: (2026)