Operator-Theoretic Foundations and Policy Gradient Methods for General MDPs with Unbounded Costs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gupta, Abhishek, Mahajan, Aditya |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Policy stability and ultimate stationarity in discounted risk-sensitive stochastic control
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2026)
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2026)
Blackwell optimality and policy stability for long-run risk sensitive stochastic control
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2024)
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2024)
Markov Decision Processes of the Third Kind: Learning Distributions by Policy Gradient Descent
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2026)
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2026)
Convex Regularization and Convergence of Policy Gradient Flows under Safety Constraints
von: Malo, Pekka, et al.
Veröffentlicht: (2024)
von: Malo, Pekka, et al.
Veröffentlicht: (2024)
Optimal Control of a Stochastic Power System -- Algorithms and Mathematical Analysis
von: Wang, Zhen, et al.
Veröffentlicht: (2024)
von: Wang, Zhen, et al.
Veröffentlicht: (2024)
Average Cost Optimality of Partially Observed MDPS: Contraction of Non-linear Filters, Optimal Solutions and Approximations
von: Demirci, Yunus Emre, et al.
Veröffentlicht: (2023)
von: Demirci, Yunus Emre, et al.
Veröffentlicht: (2023)
Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation
von: Zheng, Yaowei, et al.
Veröffentlicht: (2026)
von: Zheng, Yaowei, et al.
Veröffentlicht: (2026)
Long run control of nonhomogeneous Markov processes
von: Stettner, Łukasz
Veröffentlicht: (2025)
von: Stettner, Łukasz
Veröffentlicht: (2025)
A Fisher-Rao gradient flow for entropy-regularised Markov decision processes in Polish spaces
von: Kerimkulov, Bekzhan, et al.
Veröffentlicht: (2023)
von: Kerimkulov, Bekzhan, et al.
Veröffentlicht: (2023)
Characterizing nonconvex boundaries via scalarization
von: Ma, Jin, et al.
Veröffentlicht: (2025)
von: Ma, Jin, et al.
Veröffentlicht: (2025)
Existence of bounded solutions to multiplicative Poisson equations under mixing property
von: Pitera, Marcin, et al.
Veröffentlicht: (2023)
von: Pitera, Marcin, et al.
Veröffentlicht: (2023)
Deep Relaxation of Controlled Stochastic Gradient Descent via Singular Perturbations
von: Bardi, Martino, et al.
Veröffentlicht: (2022)
von: Bardi, Martino, et al.
Veröffentlicht: (2022)
Large-Scale Minimization of the Pseudospectral Abscissa
von: Aliyev, Nicat, et al.
Veröffentlicht: (2022)
von: Aliyev, Nicat, et al.
Veröffentlicht: (2022)
Gradient Norm Regularization Second-Order Algorithms for Solving Nonconvex-Strongly Concave Minimax Problems
von: Wang, Jun-Lin, et al.
Veröffentlicht: (2024)
von: Wang, Jun-Lin, et al.
Veröffentlicht: (2024)
Sample Complexity of Policy Gradient for Log-Growth Control
von: Pan, Qiuhua, et al.
Veröffentlicht: (2026)
von: Pan, Qiuhua, et al.
Veröffentlicht: (2026)
Cost-optimal Management of a Residential Heating System With a Geothermal Energy Storage Under Uncertainty
von: Takam, Paul Honore, et al.
Veröffentlicht: (2025)
von: Takam, Paul Honore, et al.
Veröffentlicht: (2025)
Policy Gradient Algorithms for Robust MDPs with Non-Rectangular Uncertainty Sets
von: Li, Mengmeng, et al.
Veröffentlicht: (2023)
von: Li, Mengmeng, et al.
Veröffentlicht: (2023)
A Linear Parameter-Varying Framework for the Analysis of Time-Varying Optimization Algorithms
von: Jakob, Fabian, et al.
Veröffentlicht: (2025)
von: Jakob, Fabian, et al.
Veröffentlicht: (2025)
On the Practical Implementation of a Sequential Quadratic Programming Algorithm for Nonconvex Sum-of-squares Problems
von: Olucak, Jan, et al.
Veröffentlicht: (2026)
von: Olucak, Jan, et al.
Veröffentlicht: (2026)
Measuring dissimilarity between convex cones by means of max-min angles
von: de Oliveira, Welington, et al.
Veröffentlicht: (2025)
von: de Oliveira, Welington, et al.
Veröffentlicht: (2025)
State-Dependent Uncertainty Modeling in Robust Optimal Control Problems through Generalized Semi-Infinite Programming
von: Wehbeh, J., et al.
Veröffentlicht: (2025)
von: Wehbeh, J., et al.
Veröffentlicht: (2025)
Sampled-Data Wasserstein Distributionally Robust Control of Multiplicative Systems: A Convex Relaxation with Performance Guarantees
von: Hsieh, Chung-Han
Veröffentlicht: (2026)
von: Hsieh, Chung-Han
Veröffentlicht: (2026)
Application and issues in abstract convexity
von: Millán, Reinier Díaz, et al.
Veröffentlicht: (2022)
von: Millán, Reinier Díaz, et al.
Veröffentlicht: (2022)
A multiscale Consensus-Based algorithm for multi-level optimization
von: Herty, Michael, et al.
Veröffentlicht: (2024)
von: Herty, Michael, et al.
Veröffentlicht: (2024)
Two trust region type algorithms for solving nonconvex-strongly concave minimax problems
von: Yao, Tongliang, et al.
Veröffentlicht: (2024)
von: Yao, Tongliang, et al.
Veröffentlicht: (2024)
A Fully Parameter-Free Second-Order Algorithm for Convex-Concave Minimax Problems
von: Wang, Junlin, et al.
Veröffentlicht: (2024)
von: Wang, Junlin, et al.
Veröffentlicht: (2024)
Bounding-Focused Discretization Methods for the Global Optimization of Nonconvex Semi-Infinite Programs
von: Turan, Evren M., et al.
Veröffentlicht: (2023)
von: Turan, Evren M., et al.
Veröffentlicht: (2023)
Designing Tractable Piecewise Affine Policies for Multi-Stage Adjustable Robust Optimization
von: Thomä, Simon, et al.
Veröffentlicht: (2022)
von: Thomä, Simon, et al.
Veröffentlicht: (2022)
From Semi-Infinite Constraints to Structured Robust Policies: Optimal Gain Selection for Financial Systems
von: Hsieh, Chung-Han
Veröffentlicht: (2022)
von: Hsieh, Chung-Han
Veröffentlicht: (2022)
On Strategic Measures and Optimality Properties in Discrete-Time Stochastic Control with Universally Measurable Policies
von: Yu, Huizhen
Veröffentlicht: (2022)
von: Yu, Huizhen
Veröffentlicht: (2022)
Multivariate approximation by polynomial and generalised rational functions
von: Millán, R. Díaz, et al.
Veröffentlicht: (2021)
von: Millán, R. Díaz, et al.
Veröffentlicht: (2021)
Average-Cost MDPs with Infinite State and Action Sets: New Sufficient Conditions for Optimality Inequalities and Equations
von: Feinberg, Eugene A., et al.
Veröffentlicht: (2024)
von: Feinberg, Eugene A., et al.
Veröffentlicht: (2024)
Sparse Polynomial Optimization with Unbounded Sets
von: Huang, Lei, et al.
Veröffentlicht: (2024)
von: Huang, Lei, et al.
Veröffentlicht: (2024)
Projected Gradient Methods with Momentum
von: Lapucci, Matteo, et al.
Veröffentlicht: (2026)
von: Lapucci, Matteo, et al.
Veröffentlicht: (2026)
Semi-Infinite Programs for Robust Control and Optimization: Efficient Solutions and Extensions to Existence Constraints
von: Wehbeh, Jad, et al.
Veröffentlicht: (2024)
von: Wehbeh, Jad, et al.
Veröffentlicht: (2024)
Dual Spectral Projected Gradient Method for Generalized Log-det Semidefinite Programming
von: Namchaisiri, Charles, et al.
Veröffentlicht: (2024)
von: Namchaisiri, Charles, et al.
Veröffentlicht: (2024)
Benign landscapes of low-dimensional relaxations for orthogonal synchronization on general graphs
von: McRae, Andrew D., et al.
Veröffentlicht: (2023)
von: McRae, Andrew D., et al.
Veröffentlicht: (2023)
Polarized consensus-based dynamics for optimization and sampling
von: Bungert, Leon, et al.
Veröffentlicht: (2022)
von: Bungert, Leon, et al.
Veröffentlicht: (2022)
Non-Expansive Mappings in Two-Time-Scale Stochastic Approximation: Finite-Time Analysis
von: Chandak, Siddharth
Veröffentlicht: (2025)
von: Chandak, Siddharth
Veröffentlicht: (2025)
A Globally Convergent Gradient Method with Momentum
von: Lapucci, Matteo, et al.
Veröffentlicht: (2024)
von: Lapucci, Matteo, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Policy stability and ultimate stationarity in discounted risk-sensitive stochastic control
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2026) -
Blackwell optimality and policy stability for long-run risk sensitive stochastic control
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2024) -
Markov Decision Processes of the Third Kind: Learning Distributions by Policy Gradient Descent
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2026) -
Convex Regularization and Convergence of Policy Gradient Flows under Safety Constraints
von: Malo, Pekka, et al.
Veröffentlicht: (2024) -
Optimal Control of a Stochastic Power System -- Algorithms and Mathematical Analysis
von: Wang, Zhen, et al.
Veröffentlicht: (2024)