RE-SAC: Disentangling aleatoric and epistemic risks in bus fleet control: A stable and robust ensemble DRL approach
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yifan, Zheng, Liang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adaptive Autopilot: Constrained DRL for Diverse Driving Behaviors
von: Selvaraj, Dinesh Cyril, et al.
Veröffentlicht: (2024)
von: Selvaraj, Dinesh Cyril, et al.
Veröffentlicht: (2024)
Interpretable DRL-based Maneuver Decision of UCAV Dogfight
von: Han, Haoran, et al.
Veröffentlicht: (2024)
von: Han, Haoran, et al.
Veröffentlicht: (2024)
Evaluating deep learning models for fault diagnosis of a rotating machinery with epistemic and aleatoric uncertainty
von: Jalayer, Reza, et al.
Veröffentlicht: (2024)
von: Jalayer, Reza, et al.
Veröffentlicht: (2024)
Vision-based DRL Autonomous Driving Agent with Sim2Real Transfer
von: Li, Dianzhao, et al.
Veröffentlicht: (2023)
von: Li, Dianzhao, et al.
Veröffentlicht: (2023)
BAPR: Bayesian amnesic piecewise-robust reinforcement learning for non-stationary continuous control
von: Zhang, Yifan, et al.
Veröffentlicht: (2026)
von: Zhang, Yifan, et al.
Veröffentlicht: (2026)
Probabilistic unifying relations for modelling epistemic and aleatoric uncertainty: semantics and automated reasoning with theorem proving
von: Ye, Kangfeng, et al.
Veröffentlicht: (2023)
von: Ye, Kangfeng, et al.
Veröffentlicht: (2023)
SAC Flow: Sample-Efficient Reinforcement Learning of Flow-Based Policies via Velocity-Reparameterized Sequential Modeling
von: Zhang, Yixian, et al.
Veröffentlicht: (2025)
von: Zhang, Yixian, et al.
Veröffentlicht: (2025)
HAIM-DRL: Enhanced Human-in-the-loop Reinforcement Learning for Safe and Efficient Autonomous Driving
von: Huang, Zilin, et al.
Veröffentlicht: (2024)
von: Huang, Zilin, et al.
Veröffentlicht: (2024)
AirPilot: Interpretable PPO-based DRL Auto-Tuned Nonlinear PID Drone Controller for Robust Autonomous Flights
von: Zhang, Junyang, et al.
Veröffentlicht: (2024)
von: Zhang, Junyang, et al.
Veröffentlicht: (2024)
FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control
von: Kim, Donghu, et al.
Veröffentlicht: (2026)
von: Kim, Donghu, et al.
Veröffentlicht: (2026)
Disentangled Unsupervised Skill Discovery for Efficient Hierarchical Reinforcement Learning
von: Hu, Jiaheng, et al.
Veröffentlicht: (2024)
von: Hu, Jiaheng, et al.
Veröffentlicht: (2024)
Multi-UAV Speed Control with Collision Avoidance and Handover-aware Cell Association: DRL with Action Branching
von: Yan, Zijiang, et al.
Veröffentlicht: (2023)
von: Yan, Zijiang, et al.
Veröffentlicht: (2023)
Uncertainty-Aware DRL for Autonomous Vehicle Crowd Navigation in Shared Space
von: Golchoubian, Mahsa, et al.
Veröffentlicht: (2024)
von: Golchoubian, Mahsa, et al.
Veröffentlicht: (2024)
stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation
von: Maes, Lucas, et al.
Veröffentlicht: (2026)
von: Maes, Lucas, et al.
Veröffentlicht: (2026)
Learning-based social coordination to improve safety and robustness of cooperative autonomous vehicles in mixed traffic
von: Valiente, Rodolfo, et al.
Veröffentlicht: (2022)
von: Valiente, Rodolfo, et al.
Veröffentlicht: (2022)
Agile and versatile bipedal robot tracking control through reinforcement learning
von: Li, Jiayi, et al.
Veröffentlicht: (2024)
von: Li, Jiayi, et al.
Veröffentlicht: (2024)
Aerobatic maneuvers in insect-scale flapping-wing aerial robots via deep-learned robust tube model predictive control
von: Hsiao, Yi-Hsuan, et al.
Veröffentlicht: (2025)
von: Hsiao, Yi-Hsuan, et al.
Veröffentlicht: (2025)
Reinforcement Learning Position Control of a Quadrotor Using Soft Actor-Critic (SAC)
von: Mahran, Youssef, et al.
Veröffentlicht: (2025)
von: Mahran, Youssef, et al.
Veröffentlicht: (2025)
SOMTP: Self-Supervised Learning-Based Optimizer for MPC-Based Safe Trajectory Planning Problems in Robotics
von: Liu, Yifan, et al.
Veröffentlicht: (2024)
von: Liu, Yifan, et al.
Veröffentlicht: (2024)
cc-DRL: a Convex Combined Deep Reinforcement Learning Flight Control Design for a Morphing Quadrotor
von: Yang, Tao, et al.
Veröffentlicht: (2024)
von: Yang, Tao, et al.
Veröffentlicht: (2024)
Balancing Signal and Variance: Adaptive Offline RL Post-Training for VLA Flow Models
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
von: Zhang, Hongyin, et al.
Veröffentlicht: (2025)
Learning collision risk proactively from naturalistic driving data at scale
von: Jiao, Yiru, et al.
Veröffentlicht: (2025)
von: Jiao, Yiru, et al.
Veröffentlicht: (2025)
Neural Motion Simulator: Pushing the Limit of World Models in Reinforcement Learning
von: Hao, Chenjie, et al.
Veröffentlicht: (2025)
von: Hao, Chenjie, et al.
Veröffentlicht: (2025)
Multi-Agent Generative Adversarial Interactive Self-Imitation Learning for AUV Formation Control and Obstacle Avoidance
von: Fang, Zheng, et al.
Veröffentlicht: (2024)
von: Fang, Zheng, et al.
Veröffentlicht: (2024)
RobocupGym: A challenging continuous control benchmark in Robocup
von: Beukman, Michael, et al.
Veröffentlicht: (2024)
von: Beukman, Michael, et al.
Veröffentlicht: (2024)
CubeDAgger: Interactive Imitation Learning for Dynamic Systems with Efficient yet Low-risk Interaction
von: Kobayashi, Taisuke
Veröffentlicht: (2025)
von: Kobayashi, Taisuke
Veröffentlicht: (2025)
Compliant Residual DAgger: Improving Real-World Contact-Rich Manipulation with Human Corrections
von: Xu, Xiaomeng, et al.
Veröffentlicht: (2025)
von: Xu, Xiaomeng, et al.
Veröffentlicht: (2025)
State-wise Constrained Policy Optimization
von: Zhao, Weiye, et al.
Veröffentlicht: (2023)
von: Zhao, Weiye, et al.
Veröffentlicht: (2023)
Off-dynamics Conditional Diffusion Planners
von: Ng, Wen Zheng Terence, et al.
Veröffentlicht: (2024)
von: Ng, Wen Zheng Terence, et al.
Veröffentlicht: (2024)
Natural Humanoid Robot Locomotion with Generative Motion Prior
von: Zhang, Haodong, et al.
Veröffentlicht: (2025)
von: Zhang, Haodong, et al.
Veröffentlicht: (2025)
BaTCAVe: Trustworthy Explanations for Robot Behaviors
von: Sagar, Som, et al.
Veröffentlicht: (2024)
von: Sagar, Som, et al.
Veröffentlicht: (2024)
Dichotomous Diffusion Policy Optimization
von: Liang, Ruiming, et al.
Veröffentlicht: (2025)
von: Liang, Ruiming, et al.
Veröffentlicht: (2025)
TamedPUMA: safe and stable imitation learning with geometric fabrics
von: Bakker, Saray, et al.
Veröffentlicht: (2025)
von: Bakker, Saray, et al.
Veröffentlicht: (2025)
RoboMD: Uncovering Robot Vulnerabilities through Semantic Potential Fields
von: Sagar, Som, et al.
Veröffentlicht: (2024)
von: Sagar, Som, et al.
Veröffentlicht: (2024)
SimVLA: A Simple VLA Baseline for Robotic Manipulation
von: Luo, Yuankai, et al.
Veröffentlicht: (2026)
von: Luo, Yuankai, et al.
Veröffentlicht: (2026)
Human locomotor control timescales depend on the environmental context and sensory input modality
von: Wang, Wei-Chen, et al.
Veröffentlicht: (2025)
von: Wang, Wei-Chen, et al.
Veröffentlicht: (2025)
Safe Multi-Agent Reinforcement Learning with Bilevel Optimization in Autonomous Driving
von: Zheng, Zhi, et al.
Veröffentlicht: (2024)
von: Zheng, Zhi, et al.
Veröffentlicht: (2024)
Using large language models for embodied planning introduces systematic safety risks
von: Zhang, Tao, et al.
Veröffentlicht: (2026)
von: Zhang, Tao, et al.
Veröffentlicht: (2026)
Knitting Robots: A Deep Learning Approach for Reverse-Engineering Fabric Patterns
von: Sheng, Haoliang, et al.
Veröffentlicht: (2025)
von: Sheng, Haoliang, et al.
Veröffentlicht: (2025)
Spatial Temporal Attention based Target Vehicle Trajectory Prediction for Internet of Vehicles
von: Huang, Ouhan, et al.
Veröffentlicht: (2025)
von: Huang, Ouhan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Adaptive Autopilot: Constrained DRL for Diverse Driving Behaviors
von: Selvaraj, Dinesh Cyril, et al.
Veröffentlicht: (2024) -
Interpretable DRL-based Maneuver Decision of UCAV Dogfight
von: Han, Haoran, et al.
Veröffentlicht: (2024) -
Evaluating deep learning models for fault diagnosis of a rotating machinery with epistemic and aleatoric uncertainty
von: Jalayer, Reza, et al.
Veröffentlicht: (2024) -
Vision-based DRL Autonomous Driving Agent with Sim2Real Transfer
von: Li, Dianzhao, et al.
Veröffentlicht: (2023) -
BAPR: Bayesian amnesic piecewise-robust reinforcement learning for non-stationary continuous control
von: Zhang, Yifan, et al.
Veröffentlicht: (2026)