Policy Gradient Algorithms for Robust MDPs with Non-Rectangular Uncertainty Sets
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Mengmeng, Kuhn, Daniel, Sutter, Tobias |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Large Deviations Perspective on Policy Gradient Algorithms
by: Jongeneel, Wouter, et al.
Published: (2023)
by: Jongeneel, Wouter, et al.
Published: (2023)
Sparse Polynomial Optimization with Unbounded Sets
by: Huang, Lei, et al.
Published: (2024)
by: Huang, Lei, et al.
Published: (2024)
A Moment-SOS Hierarchy for Robust Polynomial Matrix Inequality Optimization with SOS-Convexity
by: Guo, Feng, et al.
Published: (2023)
by: Guo, Feng, et al.
Published: (2023)
Gradient Norm Regularization Second-Order Algorithms for Solving Nonconvex-Strongly Concave Minimax Problems
by: Wang, Jun-Lin, et al.
Published: (2024)
by: Wang, Jun-Lin, et al.
Published: (2024)
A Normal Map-Based Proximal Stochastic Gradient Method: Convergence and Identification Properties
by: Qiu, Junwen, et al.
Published: (2023)
by: Qiu, Junwen, et al.
Published: (2023)
Sparse Polynomial Matrix Optimization
by: Miller, Jared, et al.
Published: (2024)
by: Miller, Jared, et al.
Published: (2024)
A KL-based Analysis Framework with Applications to Non-Descent Optimization Methods
by: Qiu, Junwen, et al.
Published: (2024)
by: Qiu, Junwen, et al.
Published: (2024)
Operator-Theoretic Foundations and Policy Gradient Methods for General MDPs with Unbounded Costs
by: Gupta, Abhishek, et al.
Published: (2026)
by: Gupta, Abhishek, et al.
Published: (2026)
Stochastic Gradient Descent Revisited
by: Louzi, Azar
Published: (2024)
by: Louzi, Azar
Published: (2024)
A Fully Parameter-Free Second-Order Algorithm for Convex-Concave Minimax Problems
by: Wang, Junlin, et al.
Published: (2024)
by: Wang, Junlin, et al.
Published: (2024)
An Inexact Halpern Iteration with Application to Distributionally Robust Optimization
by: Liang, Ling, et al.
Published: (2024)
by: Liang, Ling, et al.
Published: (2024)
Global Solutions to Non-Convex Functional Constrained Problems with Hidden Convexity
by: Fatkhullin, Ilyas, et al.
Published: (2025)
by: Fatkhullin, Ilyas, et al.
Published: (2025)
Communication Compression for Byzantine Robust Learning: New Efficient Algorithms and Improved Rates
by: Rammal, Ahmad, et al.
Published: (2023)
by: Rammal, Ahmad, et al.
Published: (2023)
Stochastic First-Order Methods with Non-smooth and Non-Euclidean Proximal Terms for Nonconvex High-Dimensional Stochastic Optimization
by: Xie, Yue, et al.
Published: (2024)
by: Xie, Yue, et al.
Published: (2024)
An Algebraically Converging Stochastic Gradient Descent Algorithm for Global Optimization
by: Engquist, Björn, et al.
Published: (2022)
by: Engquist, Björn, et al.
Published: (2022)
A Gauge Set Framework for Flexible Robustness Design
by: Wei, Ningji, et al.
Published: (2025)
by: Wei, Ningji, et al.
Published: (2025)
Natural Gradient VI: Guarantees for Non-Conjugate Models
by: Sun, Fangyuan, et al.
Published: (2025)
by: Sun, Fangyuan, et al.
Published: (2025)
A New Random Reshuffling Method for Nonsmooth Nonconvex Finite-sum Optimization
by: Qiu, Junwen, et al.
Published: (2023)
by: Qiu, Junwen, et al.
Published: (2023)
Adjustable Robust Nonlinear Network Design Without Controllable Elements under Load Scenario Uncertainties
by: Thürauf, Johannes, et al.
Published: (2024)
by: Thürauf, Johannes, et al.
Published: (2024)
High Probability Guarantees for Random Reshuffling
by: Yu, Hengxu, et al.
Published: (2023)
by: Yu, Hengxu, et al.
Published: (2023)
Optimization without Retraction on the Random Generalized Stiefel Manifold
by: Vary, Simon, et al.
Published: (2024)
by: Vary, Simon, et al.
Published: (2024)
FedSLoP: Memory-Efficient Federated Learning with Low-Rank Gradient Projection
by: He, Yutong, et al.
Published: (2026)
by: He, Yutong, et al.
Published: (2026)
Alternating minimization for square root principal component pursuit
by: Deng, Shengxiang, et al.
Published: (2024)
by: Deng, Shengxiang, et al.
Published: (2024)
Shuffling the Stochastic Mirror Descent via Dual Lipschitz Continuity and Kernel Conditioning
by: Qiu, Junwen, et al.
Published: (2026)
by: Qiu, Junwen, et al.
Published: (2026)
Two trust region type algorithms for solving nonconvex-strongly concave minimax problems
by: Yao, Tongliang, et al.
Published: (2024)
by: Yao, Tongliang, et al.
Published: (2024)
Exact Convex Reformulations of Linear Neural Networks via Completely Positive Lifting
by: Prakhya, Karthik, et al.
Published: (2026)
by: Prakhya, Karthik, et al.
Published: (2026)
A Generalized Version of Chung's Lemma and its Applications
by: Jiang, Li, et al.
Published: (2024)
by: Jiang, Li, et al.
Published: (2024)
Projected Gradient Methods with Momentum
by: Lapucci, Matteo, et al.
Published: (2026)
by: Lapucci, Matteo, et al.
Published: (2026)
SPAM: Stochastic Proximal Point Method with Momentum Variance Reduction for Non-convex Cross-Device Federated Learning
by: Karagulyan, Avetik, et al.
Published: (2024)
by: Karagulyan, Avetik, et al.
Published: (2024)
A Globally Convergent Gradient Method with Momentum
by: Lapucci, Matteo, et al.
Published: (2024)
by: Lapucci, Matteo, et al.
Published: (2024)
A Riemannian Accelerated Proximal Gradient Method
by: Feng, Shuailing, et al.
Published: (2025)
by: Feng, Shuailing, et al.
Published: (2025)
Det-CGD: Compressed Gradient Descent with Matrix Stepsizes for Non-Convex Optimization
by: Li, Hanmin, et al.
Published: (2023)
by: Li, Hanmin, et al.
Published: (2023)
Riemannian Gradient Method with Momentum
by: Leggio, Filippo, et al.
Published: (2026)
by: Leggio, Filippo, et al.
Published: (2026)
Pareto-optimal Trade-offs Between Communication and Computation with Flexible Gradient Tracking
by: Huang, Yan, et al.
Published: (2025)
by: Huang, Yan, et al.
Published: (2025)
Non-Attainment of Minima in Non-Polyhedral Conic Optimization: A Robust SOCP Example
by: Nguyen, Vinh
Published: (2025)
by: Nguyen, Vinh
Published: (2025)
On the Convexity of the Solution Set of Linear Complementarity Problem over Tensor Spaces
by: Sharma, Sonali, et al.
Published: (2026)
by: Sharma, Sonali, et al.
Published: (2026)
Dual Spectral Projected Gradient Method for Generalized Log-det Semidefinite Programming
by: Namchaisiri, Charles, et al.
Published: (2024)
by: Namchaisiri, Charles, et al.
Published: (2024)
Combining Gradient Information and Primitive Directions for High-Performance Mixed-Integer Optimization
by: Lapucci, Matteo, et al.
Published: (2024)
by: Lapucci, Matteo, et al.
Published: (2024)
Effective Front-Descent Algorithms with Convergence Guarantees
by: Lapucci, Matteo, et al.
Published: (2024)
by: Lapucci, Matteo, et al.
Published: (2024)
On a minimization problem of the maximum generalized eigenvalue: properties and algorithms
by: Nishioka, Akatsuki, et al.
Published: (2023)
by: Nishioka, Akatsuki, et al.
Published: (2023)
Similar Items
-
A Large Deviations Perspective on Policy Gradient Algorithms
by: Jongeneel, Wouter, et al.
Published: (2023) -
Sparse Polynomial Optimization with Unbounded Sets
by: Huang, Lei, et al.
Published: (2024) -
A Moment-SOS Hierarchy for Robust Polynomial Matrix Inequality Optimization with SOS-Convexity
by: Guo, Feng, et al.
Published: (2023) -
Gradient Norm Regularization Second-Order Algorithms for Solving Nonconvex-Strongly Concave Minimax Problems
by: Wang, Jun-Lin, et al.
Published: (2024) -
A Normal Map-Based Proximal Stochastic Gradient Method: Convergence and Identification Properties
by: Qiu, Junwen, et al.
Published: (2023)