Policy Gradient Methods for Distortion Risk Measures
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Vijayan, Nithia, A, Prashanth L. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2021
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Smoothed functional-based gradient algorithms for off-policy reinforcement learning: A non-asymptotic viewpoint
von: Vijayan, Nithia, et al.
Veröffentlicht: (2021)
von: Vijayan, Nithia, et al.
Veröffentlicht: (2021)
A policy gradient approach for optimization of smooth risk measures
von: Vijayan, Nithia, et al.
Veröffentlicht: (2022)
von: Vijayan, Nithia, et al.
Veröffentlicht: (2022)
Self-Interested Agents in Collaborative Machine Learning: An Incentivized Adaptive Data-Centric Framework
von: Vijayan, Nithia, et al.
Veröffentlicht: (2024)
von: Vijayan, Nithia, et al.
Veröffentlicht: (2024)
Policy Newton methods for Distortion Riskmetrics
von: Pachal, Soumen, et al.
Veröffentlicht: (2025)
von: Pachal, Soumen, et al.
Veröffentlicht: (2025)
Stochastic Approximation Methods for Distortion Risk Measure Optimization
von: Jiang, Jinyang, et al.
Veröffentlicht: (2025)
von: Jiang, Jinyang, et al.
Veröffentlicht: (2025)
On the Global Convergence of Risk-Averse Natural Policy Gradient Methods with Expected Conditional Risk Measures
von: Yu, Xian, et al.
Veröffentlicht: (2023)
von: Yu, Xian, et al.
Veröffentlicht: (2023)
Measures of Variability for Risk-averse Policy Gradient
von: Luo, Yudong, et al.
Veröffentlicht: (2025)
von: Luo, Yudong, et al.
Veröffentlicht: (2025)
Optimizing Shortfall Risk Metric for Learning Regression Models
von: Ramaswamy, Harish G., et al.
Veröffentlicht: (2025)
von: Ramaswamy, Harish G., et al.
Veröffentlicht: (2025)
Concentration Bounds for Optimized Certainty Equivalent Risk Estimation
von: Ghosh, Ayon, et al.
Veröffentlicht: (2024)
von: Ghosh, Ayon, et al.
Veröffentlicht: (2024)
Risk Estimation in a Markov Cost Process: Lower and Upper Bounds
von: Thoppe, Gugan, et al.
Veröffentlicht: (2023)
von: Thoppe, Gugan, et al.
Veröffentlicht: (2023)
When Do Off-Policy and On-Policy Policy Gradient Methods Align?
von: Mambelli, Davide, et al.
Veröffentlicht: (2024)
von: Mambelli, Davide, et al.
Veröffentlicht: (2024)
Learning General Policies with Policy Gradient Methods
von: Ståhlberg, Simon, et al.
Veröffentlicht: (2025)
von: Ståhlberg, Simon, et al.
Veröffentlicht: (2025)
Elementary Analysis of Policy Gradient Methods
von: Liu, Jiacai, et al.
Veröffentlicht: (2024)
von: Liu, Jiacai, et al.
Veröffentlicht: (2024)
Policy Gradient Methods for Risk-Sensitive Distributional Reinforcement Learning with Provable Convergence
von: Xiao, Minheng, et al.
Veröffentlicht: (2024)
von: Xiao, Minheng, et al.
Veröffentlicht: (2024)
Scaling Internal-State Policy-Gradient Methods for POMDPs
von: Aberdeen, Douglas, et al.
Veröffentlicht: (2025)
von: Aberdeen, Douglas, et al.
Veröffentlicht: (2025)
Reevaluating Policy Gradient Methods for Imperfect-Information Games
von: Rudolph, Max, et al.
Veröffentlicht: (2025)
von: Rudolph, Max, et al.
Veröffentlicht: (2025)
Risk-sensitive Bandits: Arm Mixture Optimality and Regret-efficient Algorithms
von: Tatlı, Meltem, et al.
Veröffentlicht: (2025)
von: Tatlı, Meltem, et al.
Veröffentlicht: (2025)
Mollification Effects of Policy Gradient Methods
von: Wang, Tao, et al.
Veröffentlicht: (2024)
von: Wang, Tao, et al.
Veröffentlicht: (2024)
Towards minimax optimal algorithms for Active Simple Hypothesis Testing
von: Vijayan, Sushant
Veröffentlicht: (2025)
von: Vijayan, Sushant
Veröffentlicht: (2025)
Ordering-based Conditions for Global Convergence of Policy Gradient Methods
von: Mei, Jincheng, et al.
Veröffentlicht: (2025)
von: Mei, Jincheng, et al.
Veröffentlicht: (2025)
Logit Dynamics in Softmax Policy Gradient Methods
von: Li, Yingru
Veröffentlicht: (2025)
von: Li, Yingru
Veröffentlicht: (2025)
Beyond Exact Gradients: Convergence of Stochastic Soft-Max Policy Gradient Methods with Entropy Regularization
von: Ding, Yuhao, et al.
Veröffentlicht: (2021)
von: Ding, Yuhao, et al.
Veröffentlicht: (2021)
Distorted Distributional Policy Evaluation for Offline Reinforcement Learning
von: Iwaki, Ryo, et al.
Veröffentlicht: (2026)
von: Iwaki, Ryo, et al.
Veröffentlicht: (2026)
Policy Gradient Methods for Non-Markovian Reinforcement Learning
von: Kar, Avik, et al.
Veröffentlicht: (2026)
von: Kar, Avik, et al.
Veröffentlicht: (2026)
Matrix Low-Rank Approximation For Policy Gradient Methods
von: Rozada, Sergio, et al.
Veröffentlicht: (2024)
von: Rozada, Sergio, et al.
Veröffentlicht: (2024)
Policy Gradient Methods in the Presence of Symmetries and State Abstractions
von: Panangaden, Prakash, et al.
Veröffentlicht: (2023)
von: Panangaden, Prakash, et al.
Veröffentlicht: (2023)
Risk-sensitive reinforcement learning using expectiles, shortfall risk and optimized certainty equivalent risk
von: Gupte, Sumedh, et al.
Veröffentlicht: (2026)
von: Gupte, Sumedh, et al.
Veröffentlicht: (2026)
A Policy Gradient-Based Sequence-to-Sequence Method for Time Series Prediction
von: Sima, Qi, et al.
Veröffentlicht: (2024)
von: Sima, Qi, et al.
Veröffentlicht: (2024)
Multilinear Tensor Low-Rank Approximation for Policy-Gradient Methods in Reinforcement Learning
von: Rozada, Sergio, et al.
Veröffentlicht: (2025)
von: Rozada, Sergio, et al.
Veröffentlicht: (2025)
Group Policy Gradient
von: Chen, Junhua, et al.
Veröffentlicht: (2025)
von: Chen, Junhua, et al.
Veröffentlicht: (2025)
Analysis of On-policy Policy Gradient Methods under the Distribution Mismatch
von: Wang, Weizhen, et al.
Veröffentlicht: (2025)
von: Wang, Weizhen, et al.
Veröffentlicht: (2025)
Off-OAB: Off-Policy Policy Gradient Method with Optimal Action-Dependent Baseline
von: Meng, Wenjia, et al.
Veröffentlicht: (2024)
von: Meng, Wenjia, et al.
Veröffentlicht: (2024)
Zero Collapse: A Failure Mode of Policy Gradient Methods in Discontinuous Reward Environments
von: Kumar, Nishant, et al.
Veröffentlicht: (2026)
von: Kumar, Nishant, et al.
Veröffentlicht: (2026)
Policy Gradient Primal-Dual Method for Safe Reinforcement Learning from Human Feedback
von: Liu, Qiang, et al.
Veröffentlicht: (2026)
von: Liu, Qiang, et al.
Veröffentlicht: (2026)
Revisiting Policy Gradients for Restricted Policy Classes: Escaping Myopic Local Optima with $k$-step Policy Gradients
von: DeWeese, Alex, et al.
Veröffentlicht: (2026)
von: DeWeese, Alex, et al.
Veröffentlicht: (2026)
On-Policy Policy Gradient Reinforcement Learning Without On-Policy Sampling
von: Corrado, Nicholas E., et al.
Veröffentlicht: (2023)
von: Corrado, Nicholas E., et al.
Veröffentlicht: (2023)
Real-Time Privacy Risk Measurement with Privacy Tokens for Gradient Leakage
von: Meng, Jiayang, et al.
Veröffentlicht: (2025)
von: Meng, Jiayang, et al.
Veröffentlicht: (2025)
Stabilizing Policy Gradient Methods via Reward Profiling
von: Ahmed, Shihab, et al.
Veröffentlicht: (2025)
von: Ahmed, Shihab, et al.
Veröffentlicht: (2025)
On the Global Optimality of Policy Gradient Methods in General Utility Reinforcement Learning
von: Barakat, Anas, et al.
Veröffentlicht: (2024)
von: Barakat, Anas, et al.
Veröffentlicht: (2024)
PG-Rainbow: Using Distributional Reinforcement Learning in Policy Gradient Methods
von: Jeon, WooJae, et al.
Veröffentlicht: (2024)
von: Jeon, WooJae, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Smoothed functional-based gradient algorithms for off-policy reinforcement learning: A non-asymptotic viewpoint
von: Vijayan, Nithia, et al.
Veröffentlicht: (2021) -
A policy gradient approach for optimization of smooth risk measures
von: Vijayan, Nithia, et al.
Veröffentlicht: (2022) -
Self-Interested Agents in Collaborative Machine Learning: An Incentivized Adaptive Data-Centric Framework
von: Vijayan, Nithia, et al.
Veröffentlicht: (2024) -
Policy Newton methods for Distortion Riskmetrics
von: Pachal, Soumen, et al.
Veröffentlicht: (2025) -
Stochastic Approximation Methods for Distortion Risk Measure Optimization
von: Jiang, Jinyang, et al.
Veröffentlicht: (2025)