Reinforcement Learning with Markov Risk Measures and Multipattern Risk Approximation
Fuente:
arXiv
Saved in:
| Main Authors: | Ruszczynski, Andrzej, Zhang, Tiangang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Stochastic dynamic programming under recursive Epstein-Zin preferences
by: Jaśkiewicz, Anna, et al.
Published: (2024)
by: Jaśkiewicz, Anna, et al.
Published: (2024)
Dynamic resource matching in manufacturing using deep reinforcement learning
by: Panda, Saunak Kumar, et al.
Published: (2026)
by: Panda, Saunak Kumar, et al.
Published: (2026)
Multi-level meta-reinforcement learning with skill-based curriculum
by: Yang, Sichen, et al.
Published: (2026)
by: Yang, Sichen, et al.
Published: (2026)
Stochastic convergence of parallel asynchronous adaptive first-order methods
by: Gratton, Serge, et al.
Published: (2026)
by: Gratton, Serge, et al.
Published: (2026)
Adaptive Conditional Forest Sampling for Spectral Risk Optimisation under Decision-Dependent Uncertainty
by: Kurbucz, Marcell T.
Published: (2026)
by: Kurbucz, Marcell T.
Published: (2026)
Convergence Rate of a Functional Learning Method for Contextual Stochastic Optimization
by: Smith, Noel, et al.
Published: (2026)
by: Smith, Noel, et al.
Published: (2026)
SALE-Based Offline Reinforcement Learning with Ensemble Q-Networks
by: Chun, Zheng
Published: (2025)
by: Chun, Zheng
Published: (2025)
Average-Cost MDPs with Infinite State and Action Sets: New Sufficient Conditions for Optimality Inequalities and Equations
by: Feinberg, Eugene A., et al.
Published: (2024)
by: Feinberg, Eugene A., et al.
Published: (2024)
Multi-Objective Optimization and Hyperparameter Tuning With Desirability Functions
by: Bartz-Beielstein, Thomas
Published: (2025)
by: Bartz-Beielstein, Thomas
Published: (2025)
Policy Optimization over General State and Action Spaces
by: Ju, Caleb, et al.
Published: (2022)
by: Ju, Caleb, et al.
Published: (2022)
Variance Reduced Policy Gradient Method for Multi-Objective Reinforcement Learning
by: Guidobene, Davide, et al.
Published: (2025)
by: Guidobene, Davide, et al.
Published: (2025)
Multi-Objective Optimization with Desirability and Morris-Mitchell Criterion
by: Bartz-Beielstein, Thomas, et al.
Published: (2025)
by: Bartz-Beielstein, Thomas, et al.
Published: (2025)
Refining Graphical Neural Network Predictions Using Flow Matching for Optimal Power Flow with Constraint-Satisfaction Guarantee
by: Khanal, Kshitiz
Published: (2025)
by: Khanal, Kshitiz
Published: (2025)
Black-Box Uniform Stability for Non-Euclidean Empirical Risk Minimization
by: Vary, Simon, et al.
Published: (2024)
by: Vary, Simon, et al.
Published: (2024)
Amazon Locker Capacity Management
by: Sethuraman, Samyukta, et al.
Published: (2023)
by: Sethuraman, Samyukta, et al.
Published: (2023)
Singularity-aware Optimization via Randomized Geometric Probing: Towards Stable Non-smooth Optimization
by: Xu, Ruoran, et al.
Published: (2026)
by: Xu, Ruoran, et al.
Published: (2026)
Simplifying Hyperparameter Tuning in Online Machine Learning -- The spotRiverGUI
by: Bartz-Beielstein, Thomas
Published: (2024)
by: Bartz-Beielstein, Thomas
Published: (2024)
Bed-Attached Vibration Sensor System: A Machine Learning Approach for Fall Detection in Nursing Homes
by: Bartz-Beielstein, Thomas, et al.
Published: (2024)
by: Bartz-Beielstein, Thomas, et al.
Published: (2024)
Multi-gear bandits, partial conservation laws, and indexability
by: Niño-Mora, José
Published: (2026)
by: Niño-Mora, José
Published: (2026)
$\texttt{skwdro}$: a library for Wasserstein distributionally robust machine learning
by: Vincent, Florian, et al.
Published: (2024)
by: Vincent, Florian, et al.
Published: (2024)
Robustness of Reinforcement Learning-Based Traffic Signal Control under Incidents: A Comparative Study
by: Nguyen, Dang Viet Anh, et al.
Published: (2025)
by: Nguyen, Dang Viet Anh, et al.
Published: (2025)
Pseudoconvex Problems in Operational Decision Systems: Algorithms for Joint Learning and Optimization
by: Li, Zijun, et al.
Published: (2026)
by: Li, Zijun, et al.
Published: (2026)
Monotone and Conservative Policy Iteration Beyond the Tabular Case
by: Eshwar, S. R., et al.
Published: (2025)
by: Eshwar, S. R., et al.
Published: (2025)
Escaping Saddle Points via Curvature-Calibrated Perturbations: A Complete Analysis with Explicit Constants and Empirical Validation
by: Alpay, Faruk, et al.
Published: (2025)
by: Alpay, Faruk, et al.
Published: (2025)
Autonomous AI Agents for Real-Time Affordable Housing Site Selection: Multi-Objective Reinforcement Learning Under Regulatory Constraints
by: Imanov, Olaf Yunus Laitinen, et al.
Published: (2026)
by: Imanov, Olaf Yunus Laitinen, et al.
Published: (2026)
AI-Enhanced IoT Systems for Predictive Maintenance and Affordability Optimization in Smart Microgrids: A Digital Twin Approach
by: Kushal, Koushik Ahmed, et al.
Published: (2025)
by: Kushal, Koushik Ahmed, et al.
Published: (2025)
Online Statistical Inference of Constant Sample-averaged Q-Learning
by: Panda, Saunak Kumar, et al.
Published: (2026)
by: Panda, Saunak Kumar, et al.
Published: (2026)
Taming Nonconvex Stochastic Mirror Descent with General Bregman Divergence
by: Fatkhullin, Ilyas, et al.
Published: (2024)
by: Fatkhullin, Ilyas, et al.
Published: (2024)
Interior-Point Vanishing Problem in Semidefinite Relaxations for Neural Network Verification
by: Ueda, Ryota, et al.
Published: (2025)
by: Ueda, Ryota, et al.
Published: (2025)
Error Feedback Reloaded: From Quadratic to Arithmetic Mean of Smoothness Constants
by: Richtárik, Peter, et al.
Published: (2024)
by: Richtárik, Peter, et al.
Published: (2024)
Natural Gradient VI: Guarantees for Non-Conjugate Models
by: Sun, Fangyuan, et al.
Published: (2025)
by: Sun, Fangyuan, et al.
Published: (2025)
A Fisher-Rao gradient flow for entropy-regularised Markov decision processes in Polish spaces
by: Kerimkulov, Bekzhan, et al.
Published: (2023)
by: Kerimkulov, Bekzhan, et al.
Published: (2023)
NOVAK: Unified adaptive optimizer for deep neural networks
by: Kavun, Sergii
Published: (2026)
by: Kavun, Sergii
Published: (2026)
A unified convergence theory for adaptive first-order methods in the nonconvex case, including AdaNorm, full and diagonal AdaGrad, Shampoo and Muo
by: Gratton, S., et al.
Published: (2026)
by: Gratton, S., et al.
Published: (2026)
Stochastic dynamic programming with non-linear discounting
by: Bäuerle, Nicole, et al.
Published: (2020)
by: Bäuerle, Nicole, et al.
Published: (2020)
Convex Regularization and Convergence of Policy Gradient Flows under Safety Constraints
by: Malo, Pekka, et al.
Published: (2024)
by: Malo, Pekka, et al.
Published: (2024)
Projection-based Lyapunov method for fully heterogeneous weakly-coupled MDPs
by: Zhang, Xiangcheng, et al.
Published: (2025)
by: Zhang, Xiangcheng, et al.
Published: (2025)
Subdifferentials of Convex Operators Valued in the Space of Integrable Functions with Application to Risk-Averse Optimization
by: Dentcheva, Darinka, et al.
Published: (2025)
by: Dentcheva, Darinka, et al.
Published: (2025)
Semi-Supervised Clustering of Sparse Graphs: Crossing the Information-Theoretic Threshold
by: Sheng, Junda, et al.
Published: (2022)
by: Sheng, Junda, et al.
Published: (2022)
Shuffling the Stochastic Mirror Descent via Dual Lipschitz Continuity and Kernel Conditioning
by: Qiu, Junwen, et al.
Published: (2026)
by: Qiu, Junwen, et al.
Published: (2026)
Similar Items
-
Stochastic dynamic programming under recursive Epstein-Zin preferences
by: Jaśkiewicz, Anna, et al.
Published: (2024) -
Dynamic resource matching in manufacturing using deep reinforcement learning
by: Panda, Saunak Kumar, et al.
Published: (2026) -
Multi-level meta-reinforcement learning with skill-based curriculum
by: Yang, Sichen, et al.
Published: (2026) -
Stochastic convergence of parallel asynchronous adaptive first-order methods
by: Gratton, Serge, et al.
Published: (2026) -
Adaptive Conditional Forest Sampling for Spectral Risk Optimisation under Decision-Dependent Uncertainty
by: Kurbucz, Marcell T.
Published: (2026)