Online Bayesian Risk-Averse Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Yuhao, Zhou, Enlu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs
by: Song, Meichen, et al.
Published: (2026)
by: Song, Meichen, et al.
Published: (2026)
Ranking and Selection with Simultaneous Input Data Collection
by: Wang, Yuhao, et al.
Published: (2025)
by: Wang, Yuhao, et al.
Published: (2025)
Bayesian Risk-Sensitive Policy Optimization For MDPs With General Loss Functions
by: Wang, Xiaoshuang, et al.
Published: (2025)
by: Wang, Xiaoshuang, et al.
Published: (2025)
Risk-Averse Total-Reward Reinforcement Learning
by: Su, Xihong, et al.
Published: (2025)
by: Su, Xihong, et al.
Published: (2025)
Risk-Averse Certification of Bayesian Neural Networks
by: Zhang, Xiyue, et al.
Published: (2024)
by: Zhang, Xiyue, et al.
Published: (2024)
Reusing Historical Trajectories in Natural Policy Gradient via Importance Sampling: Convergence and Convergence Rate
by: Lin, Yifan, et al.
Published: (2024)
by: Lin, Yifan, et al.
Published: (2024)
Risk-Averse Reinforcement Learning with Itakura-Saito Loss
by: Udovichenko, Igor, et al.
Published: (2025)
by: Udovichenko, Igor, et al.
Published: (2025)
Diffusion Policies for Risk-Averse Behavior Modeling in Offline Reinforcement Learning
by: Chen, Xiaocong, et al.
Published: (2024)
by: Chen, Xiaocong, et al.
Published: (2024)
Risk-Averse Constrained Reinforcement Learning with Optimized Certainty Equivalents
by: Lee, Jane H., et al.
Published: (2025)
by: Lee, Jane H., et al.
Published: (2025)
Optimistic Exploration for Risk-Averse Constrained Reinforcement Learning
by: McCarthy, James, et al.
Published: (2025)
by: McCarthy, James, et al.
Published: (2025)
Risk-Averse Learning with Varying Risk Levels
by: Wang, Siyi, et al.
Published: (2025)
by: Wang, Siyi, et al.
Published: (2025)
Conflict-Averse Gradient Aggregation for Constrained Multi-Objective Reinforcement Learning
by: Kim, Dohyeong, et al.
Published: (2024)
by: Kim, Dohyeong, et al.
Published: (2024)
Curiosity is Knowledge: Self-Consistent Learning and No-Regret Optimization with Active Inference
by: Li, Yingke, et al.
Published: (2026)
by: Li, Yingke, et al.
Published: (2026)
Pragmatic Curiosity: A Unified Framework for Hybrid Learning and Optimization via Active Inference
by: Li, Yingke, et al.
Published: (2026)
by: Li, Yingke, et al.
Published: (2026)
Bayesian Design Principles for Offline-to-Online Reinforcement Learning
by: Hu, Hao, et al.
Published: (2024)
by: Hu, Hao, et al.
Published: (2024)
RASR: Risk-Averse Soft-Robust MDPs with EVaR and Entropic Risk
by: Hau, Jia Lin, et al.
Published: (2022)
by: Hau, Jia Lin, et al.
Published: (2022)
ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining
by: Bal, Melis Ilayda, et al.
Published: (2025)
by: Bal, Melis Ilayda, et al.
Published: (2025)
Risk-Averse Best Arm Set Identification with Fixed Budget and Fixed Confidence
by: Nonaga, Shunta, et al.
Published: (2025)
by: Nonaga, Shunta, et al.
Published: (2025)
Adaptive Simulation Experiment for LLM Policy Optimization
by: Hu, Mingjie, et al.
Published: (2026)
by: Hu, Mingjie, et al.
Published: (2026)
On the Global Convergence of Risk-Averse Natural Policy Gradient Methods with Expected Conditional Risk Measures
by: Yu, Xian, et al.
Published: (2023)
by: Yu, Xian, et al.
Published: (2023)
DRL-ORA: Distributional Reinforcement Learning with Online Risk Adaption
by: Wu, Yupeng, et al.
Published: (2023)
by: Wu, Yupeng, et al.
Published: (2023)
Conformal Uncertainty Quantification of Electricity Price Predictions for Risk-Averse Storage Arbitrage
by: Alghumayjan, Saud, et al.
Published: (2024)
by: Alghumayjan, Saud, et al.
Published: (2024)
Conflict-Averse Gradient Descent for Multi-task Learning
by: Liu, Bo, et al.
Published: (2021)
by: Liu, Bo, et al.
Published: (2021)
Risk-sensitive Actor-Critic with Static Spectral Risk Measures for Online and Offline Reinforcement Learning
by: Moghimi, Mehrdad, et al.
Published: (2025)
by: Moghimi, Mehrdad, et al.
Published: (2025)
Decision Theoretic Foundations for Conformal Prediction: Optimal Uncertainty Quantification for Risk-Averse Agents
by: Kiyani, Shayan, et al.
Published: (2025)
by: Kiyani, Shayan, et al.
Published: (2025)
Shape Derivative-Informed Neural Operators with Application to Risk-Averse Shape Optimization
by: Gong, Xindi, et al.
Published: (2026)
by: Gong, Xindi, et al.
Published: (2026)
Optimal Computing Budget Allocation for Data-driven Ranking and Selection
by: Wang, Yuhao, et al.
Published: (2022)
by: Wang, Yuhao, et al.
Published: (2022)
Robust Bayesian Dynamic Programming for On-policy Risk-sensitive Reinforcement Learning
by: Han, Shanyu, et al.
Published: (2025)
by: Han, Shanyu, et al.
Published: (2025)
Unsupervised-to-Online Reinforcement Learning
by: Kim, Junsu, et al.
Published: (2024)
by: Kim, Junsu, et al.
Published: (2024)
Efficient RF Passive Components Modeling with Bayesian Online Learning and Uncertainty Aware Sampling
by: Zhang, Huifan, et al.
Published: (2025)
by: Zhang, Huifan, et al.
Published: (2025)
Efficient Online Reinforcement Learning for Diffusion Policy
by: Ma, Haitong, et al.
Published: (2025)
by: Ma, Haitong, et al.
Published: (2025)
Is Risk-Sensitive Reinforcement Learning Properly Resolved?
by: Zhou, Ruiwen, et al.
Published: (2023)
by: Zhou, Ruiwen, et al.
Published: (2023)
Multi-Path Collaborative Reasoning via Reinforcement Learning
by: Lv, Jindi, et al.
Published: (2025)
by: Lv, Jindi, et al.
Published: (2025)
Online Pre-Training for Offline-to-Online Reinforcement Learning
by: Shin, Yongjae, et al.
Published: (2025)
by: Shin, Yongjae, et al.
Published: (2025)
Approximate Bilevel Difference Convex Programming for Bayesian Risk Markov Decision Processes
by: Lin, Yifan, et al.
Published: (2023)
by: Lin, Yifan, et al.
Published: (2023)
The Benefit of Being Bayesian in Online Conformal Prediction
by: Zhang, Zhiyu, et al.
Published: (2024)
by: Zhang, Zhiyu, et al.
Published: (2024)
CAWR: Corruption-Averse Advantage-Weighted Regression for Robust Policy Optimization
by: Hu, Ranting
Published: (2025)
by: Hu, Ranting
Published: (2025)
Driving Through Uncertainty: Risk-Averse Control with LLM Commonsense for Autonomous Driving under Perception Deficits
by: Hu, Yuting, et al.
Published: (2025)
by: Hu, Yuting, et al.
Published: (2025)
Online Robust Reinforcement Learning with General Function Approximation
by: Ghosh, Debamita, et al.
Published: (2025)
by: Ghosh, Debamita, et al.
Published: (2025)
Online Episodic Convex Reinforcement Learning
by: Moreno, Bianca Marin, et al.
Published: (2025)
by: Moreno, Bianca Marin, et al.
Published: (2025)
Similar Items
-
Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs
by: Song, Meichen, et al.
Published: (2026) -
Ranking and Selection with Simultaneous Input Data Collection
by: Wang, Yuhao, et al.
Published: (2025) -
Bayesian Risk-Sensitive Policy Optimization For MDPs With General Loss Functions
by: Wang, Xiaoshuang, et al.
Published: (2025) -
Risk-Averse Total-Reward Reinforcement Learning
by: Su, Xihong, et al.
Published: (2025) -
Risk-Averse Certification of Bayesian Neural Networks
by: Zhang, Xiyue, et al.
Published: (2024)