Actor-Critic Algorithm for Dynamic Expectile and CVaR
Fuente:
arXiv
Saved in:
| Main Authors: | Luo, Yudong, Delage, Erick |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Boosting CVaR Policy Optimization with Quantile Gradients
by: Luo, Yudong, et al.
Published: (2026)
by: Luo, Yudong, et al.
Published: (2026)
Reward Redistribution for CVaR MDPs using a Bellman Operator on L-infinity
by: Muni, Aneri, et al.
Published: (2026)
by: Muni, Aneri, et al.
Published: (2026)
A Simple Mixture Policy Parameterization for Improving Sample Efficiency of CVaR Optimization
by: Luo, Yudong, et al.
Published: (2024)
by: Luo, Yudong, et al.
Published: (2024)
Autonomous Sparse Mean-CVaR Portfolio Optimization
by: Lin, Yizun, et al.
Published: (2024)
by: Lin, Yizun, et al.
Published: (2024)
On the Fundamental Limitations of Dual Static CVaR Decompositions in Markov Decision Processes
by: Godbout, Mathieu, et al.
Published: (2025)
by: Godbout, Mathieu, et al.
Published: (2025)
Return Capping: Sample-Efficient CVaR Policy Gradient Optimisation
by: Mead, Harry, et al.
Published: (2025)
by: Mead, Harry, et al.
Published: (2025)
Instantiating Bayesian CVaR lower bounds in Interactive Decision Making Problems
by: Bongole, Raghav, et al.
Published: (2026)
by: Bongole, Raghav, et al.
Published: (2026)
Near-Optimal Sample Complexity for Iterated CVaR Reinforcement Learning with a Generative Model
by: Deng, Zilong, et al.
Published: (2025)
by: Deng, Zilong, et al.
Published: (2025)
Moderate Actor-Critic Methods: Controlling Overestimation Bias via Expectile Loss
by: Hwang, Ukjo, et al.
Published: (2025)
by: Hwang, Ukjo, et al.
Published: (2025)
Statistical Robustness of Interval CVaR Based Regression Models under Perturbation and Contamination
by: You, Yulei, et al.
Published: (2026)
by: You, Yulei, et al.
Published: (2026)
Multi-Agent Regime-Conditioned Diffusion (MARCD) for CVaR-Constrained Portfolio Decisions
by: Alzahrani, Ali Atiah
Published: (2025)
by: Alzahrani, Ali Atiah
Published: (2025)
Adaptive Insurance Reserving with CVaR-Constrained Reinforcement Learning under Macroeconomic Regimes
by: Dong, Stella C.
Published: (2025)
by: Dong, Stella C.
Published: (2025)
The Privacy Price of Tail-Risk Learning: Effective Tail Sample Size in Differentially Private CVaR Optimization
by: Mansouri, El Mustapha
Published: (2026)
by: Mansouri, El Mustapha
Published: (2026)
Stabilized Maximum-Likelihood Iterative Quantum Amplitude Estimation for Structural CVaR under Correlated Random Fields
by: Tabarraei, Alireza
Published: (2026)
by: Tabarraei, Alireza
Published: (2026)
Beyond CVaR: Leveraging Static Spectral Risk Measures for Enhanced Decision-Making in Distributional Reinforcement Learning
by: Moghimi, Mehrdad, et al.
Published: (2025)
by: Moghimi, Mehrdad, et al.
Published: (2025)
Distributionally Robust Safety Verification of Neural Networks via Worst-Case CVaR
by: Kishida, Masako
Published: (2025)
by: Kishida, Masako
Published: (2025)
CVaR-Based Variational Quantum Optimization for User Association in Handoff-Aware Vehicular Networks
by: Yan, Zijiang, et al.
Published: (2025)
by: Yan, Zijiang, et al.
Published: (2025)
Dynamically Augmented CVaR for MDPs
by: Feinberg, Eugene A., et al.
Published: (2022)
by: Feinberg, Eugene A., et al.
Published: (2022)
End-to-end Conditional Robust Optimization
by: Chenreddy, Abhilash, et al.
Published: (2024)
by: Chenreddy, Abhilash, et al.
Published: (2024)
Risk-Aware Decision Making in Restless Bandits: Theory and Algorithms for Planning and Learning
by: Akbarzadeh, Nima, et al.
Published: (2024)
by: Akbarzadeh, Nima, et al.
Published: (2024)
Planning and Learning in Average Risk-aware MDPs
by: Wang, Weikai, et al.
Published: (2025)
by: Wang, Weikai, et al.
Published: (2025)
Epistemic Robust Offline Reinforcement Learning
by: Chenreddy, Abhilash Reddy, et al.
Published: (2026)
by: Chenreddy, Abhilash Reddy, et al.
Published: (2026)
ClauseLens: Clause-Grounded, CVaR-Constrained Reinforcement Learning for Trustworthy Reinsurance Pricing
by: Dong, Stella C., et al.
Published: (2025)
by: Dong, Stella C., et al.
Published: (2025)
Mitigating optimistic bias in entropic risk estimation and optimization
by: Sadana, Utsav, et al.
Published: (2024)
by: Sadana, Utsav, et al.
Published: (2024)
Compatible Gradient Approximations for Actor-Critic Algorithms
by: Saglam, Baturay, et al.
Published: (2024)
by: Saglam, Baturay, et al.
Published: (2024)
Finite-Time Analysis of Three-Timescale Constrained Actor-Critic and Constrained Natural Actor-Critic Algorithms
by: Panda, Prashansa, et al.
Published: (2023)
by: Panda, Prashansa, et al.
Published: (2023)
Robust Data-driven Prescriptiveness Optimization
by: Poursoltani, Mehran, et al.
Published: (2023)
by: Poursoltani, Mehran, et al.
Published: (2023)
Value Improved Actor Critic Algorithms
by: Oren, Yaniv, et al.
Published: (2024)
by: Oren, Yaniv, et al.
Published: (2024)
Actor-Critic without Actor
by: Ki, Donghyeon, et al.
Published: (2025)
by: Ki, Donghyeon, et al.
Published: (2025)
Fair Resource Allocation in Weakly Coupled Markov Decision Processes
by: Tu, Xiaohui, et al.
Published: (2024)
by: Tu, Xiaohui, et al.
Published: (2024)
A Theoretical Justification for Asymmetric Actor-Critic Algorithms
by: Lambrechts, Gaspard, et al.
Published: (2025)
by: Lambrechts, Gaspard, et al.
Published: (2025)
Bootstrapping Expectiles in Reinforcement Learning
by: Clavier, Pierre, et al.
Published: (2024)
by: Clavier, Pierre, et al.
Published: (2024)
An Investigation of Batch Normalization in Off-Policy Actor-Critic Algorithms
by: Wang, Li, et al.
Published: (2025)
by: Wang, Li, et al.
Published: (2025)
Actor-Critic or Critic-Actor? A Tale of Two Time Scales
by: Bhatnagar, Shalabh, et al.
Published: (2022)
by: Bhatnagar, Shalabh, et al.
Published: (2022)
D2 Actor Critic: Diffusion Actor Meets Distributional Critic
by: Zhang, Lunjun, et al.
Published: (2025)
by: Zhang, Lunjun, et al.
Published: (2025)
Actor-Critic Reinforcement Learning with Phased Actor
by: Wu, Ruofan, et al.
Published: (2024)
by: Wu, Ruofan, et al.
Published: (2024)
A Communication-Efficient Decentralized Actor-Critic Algorithm
by: Ren, Xiaoxing, et al.
Published: (2025)
by: Ren, Xiaoxing, et al.
Published: (2025)
An Asymptotic CVaR Measure of Risk for Markov Chains
by: Patel, Shivam, et al.
Published: (2024)
by: Patel, Shivam, et al.
Published: (2024)
Bayesian Optimization for CVaR-based portfolio optimization
by: Millar, Robert, et al.
Published: (2025)
by: Millar, Robert, et al.
Published: (2025)
Generative Actor Critic
by: Qin, Aoyang, et al.
Published: (2025)
by: Qin, Aoyang, et al.
Published: (2025)
Similar Items
-
Boosting CVaR Policy Optimization with Quantile Gradients
by: Luo, Yudong, et al.
Published: (2026) -
Reward Redistribution for CVaR MDPs using a Bellman Operator on L-infinity
by: Muni, Aneri, et al.
Published: (2026) -
A Simple Mixture Policy Parameterization for Improving Sample Efficiency of CVaR Optimization
by: Luo, Yudong, et al.
Published: (2024) -
Autonomous Sparse Mean-CVaR Portfolio Optimization
by: Lin, Yizun, et al.
Published: (2024) -
On the Fundamental Limitations of Dual Static CVaR Decompositions in Markov Decision Processes
by: Godbout, Mathieu, et al.
Published: (2025)