Properties of Turnpike Functions for Discounted Finite MDPs
Fuente:
arXiv
Saved in:
| Main Authors: | Feinberg, Eugene A., He, Gaojin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dynamically Augmented CVaR for MDPs
by: Feinberg, Eugene A., et al.
Published: (2022)
by: Feinberg, Eugene A., et al.
Published: (2022)
Thompson Sampling for Infinite-Horizon Discounted Decision Processes
by: Adelman, Daniel, et al.
Published: (2024)
by: Adelman, Daniel, et al.
Published: (2024)
On Strategic Measures and Optimality Properties in Discrete-Time Stochastic Control with Universally Measurable Policies
by: Yu, Huizhen
Published: (2022)
by: Yu, Huizhen
Published: (2022)
Average-Cost MDPs with Infinite State and Action Sets: New Sufficient Conditions for Optimality Inequalities and Equations
by: Feinberg, Eugene A., et al.
Published: (2024)
by: Feinberg, Eugene A., et al.
Published: (2024)
Stochastic dynamic programming with non-linear discounting
by: Bäuerle, Nicole, et al.
Published: (2020)
by: Bäuerle, Nicole, et al.
Published: (2020)
Blackwell optimality and policy stability for long-run risk sensitive stochastic control
by: Bäuerle, Nicole, et al.
Published: (2024)
by: Bäuerle, Nicole, et al.
Published: (2024)
Policy stability and ultimate stationarity in discounted risk-sensitive stochastic control
by: Bäuerle, Nicole, et al.
Published: (2026)
by: Bäuerle, Nicole, et al.
Published: (2026)
Continuity of Filters for Discrete-Time Control Problems Defined by Explicit Equations
by: Feinberg, Eugene A., et al.
Published: (2023)
by: Feinberg, Eugene A., et al.
Published: (2023)
Continuous-time mean field Markov decision models
by: Bäuerle, Nicole, et al.
Published: (2023)
by: Bäuerle, Nicole, et al.
Published: (2023)
A Note on the Bias and Kemeny's Constant in Markov Reward Processes with an Application to Markov Chain Perturbation
by: Ortner, Ronald
Published: (2024)
by: Ortner, Ronald
Published: (2024)
Optimal Rates of Convergence for Entropy Regularization in Discounted Markov Decision Processes
by: Müller, Johannes, et al.
Published: (2024)
by: Müller, Johannes, et al.
Published: (2024)
Auto-exploration for online reinforcement learning
by: Ju, Caleb, et al.
Published: (2025)
by: Ju, Caleb, et al.
Published: (2025)
Strategy Complexity of Limsup and Liminf Threshold Objectives in Countable MDPs, with Applications to Optimal Expected Payoffs
by: Mayr, Richard, et al.
Published: (2022)
by: Mayr, Richard, et al.
Published: (2022)
Zero-one Laws for a Control Problem with Random Action Sets
by: Flesch, János, et al.
Published: (2024)
by: Flesch, János, et al.
Published: (2024)
Dynamic programming for the stochastic matching model on general graphs: the case of the `N-graph'
by: Jean, Loïc, et al.
Published: (2024)
by: Jean, Loïc, et al.
Published: (2024)
Long-Run Average Reward Maximization of A Regulated Regime-Switching Diffusion Model
by: Zeng, Lingjia, et al.
Published: (2025)
by: Zeng, Lingjia, et al.
Published: (2025)
Lipschitz upper semicontinuity of linear inequality systems under full perturbations
by: Camacho, Jesús, et al.
Published: (2025)
by: Camacho, Jesús, et al.
Published: (2025)
Deep neural networks can provably solve Bellman equations for Markov decision processes without the curse of dimensionality
by: Jentzen, Arnulf, et al.
Published: (2025)
by: Jentzen, Arnulf, et al.
Published: (2025)
Efficient Learning for Entropy-Regularized Markov Decision Processes via Multilevel Monte Carlo
by: Meunier, Matthieu, et al.
Published: (2025)
by: Meunier, Matthieu, et al.
Published: (2025)
Linear programming for finite-horizon vector-valued Markov decision processes
by: Mifrani, Anas, et al.
Published: (2025)
by: Mifrani, Anas, et al.
Published: (2025)
A Representation Optimization Dichotomy, Lie-Algebraic Policy Optimization
by: KC, Sooraj, et al.
Published: (2026)
by: KC, Sooraj, et al.
Published: (2026)
On seeded subgraph-to-subgraph matching: The ssSGM Algorithm and matchability information theory
by: Meng, Lingyao, et al.
Published: (2023)
by: Meng, Lingyao, et al.
Published: (2023)
Asymptotic behaviour of stochastic inertial dynamics incorporating a Tikhonov regularization term
by: Schindler, Chiara
Published: (2025)
by: Schindler, Chiara
Published: (2025)
A Fisher-Rao gradient flow for entropy-regularised Markov decision processes in Polish spaces
by: Kerimkulov, Bekzhan, et al.
Published: (2023)
by: Kerimkulov, Bekzhan, et al.
Published: (2023)
A New Notion of Tykhonov Well-Posedness for Optimization Problems
by: Chen, J. S., et al.
Published: (2023)
by: Chen, J. S., et al.
Published: (2023)
Robust portfolio optimization model for electronic coupon allocation
by: Uehara, Yuki, et al.
Published: (2024)
by: Uehara, Yuki, et al.
Published: (2024)
On first passage time problems of Brownian motion -- The inverse method of images revisited
by: Christensen, Sören, et al.
Published: (2024)
by: Christensen, Sören, et al.
Published: (2024)
A successive difference-of-convex method for a class of two-stage nonconvex nonsmooth stochastic conic program via SVI
by: Zhang, Chao, et al.
Published: (2026)
by: Zhang, Chao, et al.
Published: (2026)
Getting to the Root of the Problem: Sums of Squares for Limits of Trees
by: Brosch, Daniel, et al.
Published: (2024)
by: Brosch, Daniel, et al.
Published: (2024)
A new geometric approach to multiobjective linear programming problems
by: Kaci, Mustapha, et al.
Published: (2022)
by: Kaci, Mustapha, et al.
Published: (2022)
Empirical Evaluation of Policy-Based Reinforcement Learning for Dynamic Service Control in an M/M/1 Queue
by: Walton, Joseph, et al.
Published: (2026)
by: Walton, Joseph, et al.
Published: (2026)
Analysing heavy-tail properties of Stochastic Gradient Descent by means of Stochastic Recurrence Equations
by: Damek, Ewa, et al.
Published: (2024)
by: Damek, Ewa, et al.
Published: (2024)
Reinforcement Learning Methods for the Stochastic Optimal Control of an Industrial Power-to-Heat System
by: Pilling, Eric, et al.
Published: (2024)
by: Pilling, Eric, et al.
Published: (2024)
Large independent sets in recursive Markov random graphs
by: Gupte, Akshay, et al.
Published: (2022)
by: Gupte, Akshay, et al.
Published: (2022)
A third order dynamical system for generalized monotone equation
by: Hai, Pham Viet, et al.
Published: (2024)
by: Hai, Pham Viet, et al.
Published: (2024)
Reinforcement-learning-based Algorithms for Optimization Problems and Applications to Inverse Problems
by: Xu, Chen, et al.
Published: (2023)
by: Xu, Chen, et al.
Published: (2023)
Convex Regularization and Convergence of Policy Gradient Flows under Safety Constraints
by: Malo, Pekka, et al.
Published: (2024)
by: Malo, Pekka, et al.
Published: (2024)
Robust Ergodic Control of Jump-Diffusion Systems under Drift and Intensity Uncertainty
by: Azze, Abel, et al.
Published: (2026)
by: Azze, Abel, et al.
Published: (2026)
Relative Lipschitz-like property of parametric systems via projectional coderivative
by: Yao, Wenfang, et al.
Published: (2022)
by: Yao, Wenfang, et al.
Published: (2022)
Optimal strategies in Markov decision processes with finitely additive evaluations
by: Flesch, János, et al.
Published: (2026)
by: Flesch, János, et al.
Published: (2026)
Similar Items
-
Dynamically Augmented CVaR for MDPs
by: Feinberg, Eugene A., et al.
Published: (2022) -
Thompson Sampling for Infinite-Horizon Discounted Decision Processes
by: Adelman, Daniel, et al.
Published: (2024) -
On Strategic Measures and Optimality Properties in Discrete-Time Stochastic Control with Universally Measurable Policies
by: Yu, Huizhen
Published: (2022) -
Average-Cost MDPs with Infinite State and Action Sets: New Sufficient Conditions for Optimality Inequalities and Equations
by: Feinberg, Eugene A., et al.
Published: (2024) -
Stochastic dynamic programming with non-linear discounting
by: Bäuerle, Nicole, et al.
Published: (2020)