On the Global Convergence of Risk-Averse Natural Policy Gradient Methods with Expected Conditional Risk Measures
Fuente:
arXiv
Guardado en:
| Autores principales: | Yu, Xian, Ying, Lei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Policy Gradient Methods for Risk-Sensitive Distributional Reinforcement Learning with Provable Convergence
por: Xiao, Minheng, et al.
Publicado: (2024)
por: Xiao, Minheng, et al.
Publicado: (2024)
Risk-Averse Learning with Varying Risk Levels
por: Wang, Siyi, et al.
Publicado: (2025)
por: Wang, Siyi, et al.
Publicado: (2025)
Global Convergence of Natural Policy Gradient with Hessian-aided Momentum Variance Reduction
por: Feng, Jie, et al.
Publicado: (2024)
por: Feng, Jie, et al.
Publicado: (2024)
Global Convergence Guarantees for Federated Policy Gradient Methods with Adversaries
por: Ganesh, Swetha, et al.
Publicado: (2024)
por: Ganesh, Swetha, et al.
Publicado: (2024)
Conformal Uncertainty Quantification of Electricity Price Predictions for Risk-Averse Storage Arbitrage
por: Alghumayjan, Saud, et al.
Publicado: (2024)
por: Alghumayjan, Saud, et al.
Publicado: (2024)
Policy Gradient Converges to the Globally Optimal Policy for Nearly Linear-Quadratic Regulators
por: Han, Yinbin, et al.
Publicado: (2023)
por: Han, Yinbin, et al.
Publicado: (2023)
Last-Iterate Convergent Policy Gradient Primal-Dual Methods for Constrained MDPs
por: Ding, Dongsheng, et al.
Publicado: (2023)
por: Ding, Dongsheng, et al.
Publicado: (2023)
Reusing Historical Trajectories in Natural Policy Gradient via Importance Sampling: Convergence and Convergence Rate
por: Lin, Yifan, et al.
Publicado: (2024)
por: Lin, Yifan, et al.
Publicado: (2024)
Linear Convergence of Entropy-Regularized Natural Policy Gradient with Linear Function Approximation
por: Cayci, Semih, et al.
Publicado: (2021)
por: Cayci, Semih, et al.
Publicado: (2021)
Beyond Stationarity: Convergence Analysis of Stochastic Softmax Policy Gradient Methods
por: Klein, Sara, et al.
Publicado: (2023)
por: Klein, Sara, et al.
Publicado: (2023)
Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods
por: Carmona, René, et al.
Publicado: (2019)
por: Carmona, René, et al.
Publicado: (2019)
On the Value of Risk-Averse Multistage Stochastic Programming in Capacity Planning
por: Yu, Xian, et al.
Publicado: (2024)
por: Yu, Xian, et al.
Publicado: (2024)
On the Stochastic (Variance-Reduced) Proximal Gradient Method for Regularized Expected Reward Optimization
por: Liang, Ling, et al.
Publicado: (2024)
por: Liang, Ling, et al.
Publicado: (2024)
Shape Derivative-Informed Neural Operators with Application to Risk-Averse Shape Optimization
por: Gong, Xindi, et al.
Publicado: (2026)
por: Gong, Xindi, et al.
Publicado: (2026)
Stochastic Approximation Methods for Distortion Risk Measure Optimization
por: Jiang, Jinyang, et al.
Publicado: (2025)
por: Jiang, Jinyang, et al.
Publicado: (2025)
Linear Convergence of Independent Natural Policy Gradient in Games with Entropy Regularization
por: Sun, Youbang, et al.
Publicado: (2024)
por: Sun, Youbang, et al.
Publicado: (2024)
Recurrent Natural Policy Gradient for POMDPs
por: Cayci, Semih, et al.
Publicado: (2024)
por: Cayci, Semih, et al.
Publicado: (2024)
On the Last-Iterate Convergence of Shuffling Gradient Methods
por: Liu, Zijian, et al.
Publicado: (2024)
por: Liu, Zijian, et al.
Publicado: (2024)
On the Convergence of Adaptive Gradient Methods for Nonconvex Optimization
por: Zhou, Dongruo, et al.
Publicado: (2018)
por: Zhou, Dongruo, et al.
Publicado: (2018)
Directional Smoothness and Gradient Methods: Convergence and Adaptivity
por: Mishkin, Aaron, et al.
Publicado: (2024)
por: Mishkin, Aaron, et al.
Publicado: (2024)
Gradient Regularized Newton Boosting Trees with Global Convergence
por: Zozoulenko, Nikita, et al.
Publicado: (2026)
por: Zozoulenko, Nikita, et al.
Publicado: (2026)
Soft Robust MDPs and Risk-Sensitive MDPs: Equivalence, Policy Gradient, and Sample Complexity
por: Zhang, Runyu, et al.
Publicado: (2023)
por: Zhang, Runyu, et al.
Publicado: (2023)
Natural Policy Gradient and Actor Critic Methods for Constrained Multi-Task Reinforcement Learning
por: Zeng, Sihan, et al.
Publicado: (2024)
por: Zeng, Sihan, et al.
Publicado: (2024)
Convergence Properties of Natural Gradient Descent for Minimizing KL Divergence
por: Datar, Adwait, et al.
Publicado: (2025)
por: Datar, Adwait, et al.
Publicado: (2025)
GANs as Gradient Flows that Converge
por: Huang, Yu-Jui, et al.
Publicado: (2022)
por: Huang, Yu-Jui, et al.
Publicado: (2022)
Accelerated Gradient Methods with Biased Gradient Estimates: Risk Sensitivity, High-Probability Guarantees, and Large Deviation Bounds
por: Gürbüzbalaban, Mert, et al.
Publicado: (2025)
por: Gürbüzbalaban, Mert, et al.
Publicado: (2025)
Robust Risk-Sensitive Reinforcement Learning with Conditional Value-at-Risk
por: Ni, Xinyi, et al.
Publicado: (2024)
por: Ni, Xinyi, et al.
Publicado: (2024)
Revisiting the Last-Iterate Convergence of Stochastic Gradient Methods
por: Liu, Zijian, et al.
Publicado: (2023)
por: Liu, Zijian, et al.
Publicado: (2023)
Using Taylor-Approximated Gradients to Improve the Frank-Wolfe Method for Empirical Risk Minimization
por: Xiong, Zikai, et al.
Publicado: (2022)
por: Xiong, Zikai, et al.
Publicado: (2022)
Elementary Analysis of Policy Gradient Methods
por: Liu, Jiacai, et al.
Publicado: (2024)
por: Liu, Jiacai, et al.
Publicado: (2024)
A Gradient Method for Risk Averse Control of a PDE-SDE Interconnected System
por: Velho, Gabriel, et al.
Publicado: (2025)
por: Velho, Gabriel, et al.
Publicado: (2025)
In-Expectation Convergence of Stochastic Gradient Methods under Heavy-Tailed Noise
por: Liu, Zijian
Publicado: (2026)
por: Liu, Zijian
Publicado: (2026)
Almost Sure Convergence Analysis of Differentially Private Stochastic Gradient Methods
por: Mukherjee, Amartya, et al.
Publicado: (2025)
por: Mukherjee, Amartya, et al.
Publicado: (2025)
Towards Efficient Risk-Sensitive Policy Gradient: An Iteration Complexity Analysis
por: Liu, Rui, et al.
Publicado: (2024)
por: Liu, Rui, et al.
Publicado: (2024)
Toward Global Convergence of Gradient EM for Over-Parameterized Gaussian Mixture Models
por: Xu, Weihang, et al.
Publicado: (2024)
por: Xu, Weihang, et al.
Publicado: (2024)
Natural Policy Gradient as Doubly Smoothed Policy Iteration: A Bellman-Operator Framework
por: Nanda, Phalguni, et al.
Publicado: (2026)
por: Nanda, Phalguni, et al.
Publicado: (2026)
On the Role of Batch Size in Stochastic Conditional Gradient Methods
por: Islamov, Rustem, et al.
Publicado: (2026)
por: Islamov, Rustem, et al.
Publicado: (2026)
Accelerated Convergence of Stochastic Heavy Ball Method under Anisotropic Gradient Noise
por: Pan, Rui, et al.
Publicado: (2023)
por: Pan, Rui, et al.
Publicado: (2023)
Simple Stepsize for Quasi-Newton Methods with Global Convergence Guarantees
por: Agafonov, Artem, et al.
Publicado: (2025)
por: Agafonov, Artem, et al.
Publicado: (2025)
Improved Last-Iterate Convergence of Shuffling Gradient Methods for Nonsmooth Convex Optimization
por: Liu, Zijian, et al.
Publicado: (2025)
por: Liu, Zijian, et al.
Publicado: (2025)
Ejemplares similares
-
Policy Gradient Methods for Risk-Sensitive Distributional Reinforcement Learning with Provable Convergence
por: Xiao, Minheng, et al.
Publicado: (2024) -
Risk-Averse Learning with Varying Risk Levels
por: Wang, Siyi, et al.
Publicado: (2025) -
Global Convergence of Natural Policy Gradient with Hessian-aided Momentum Variance Reduction
por: Feng, Jie, et al.
Publicado: (2024) -
Global Convergence Guarantees for Federated Policy Gradient Methods with Adversaries
por: Ganesh, Swetha, et al.
Publicado: (2024) -
Conformal Uncertainty Quantification of Electricity Price Predictions for Risk-Averse Storage Arbitrage
por: Alghumayjan, Saud, et al.
Publicado: (2024)