A Note on the Bias and Kemeny's Constant in Markov Reward Processes with an Application to Markov Chain Perturbation
Fuente:
arXiv
Guardado en:
| Autor principal: | Ortner, Ronald |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Continuous-time mean field Markov decision models
por: Bäuerle, Nicole, et al.
Publicado: (2023)
por: Bäuerle, Nicole, et al.
Publicado: (2023)
Dynamic programming for the stochastic matching model on general graphs: the case of the `N-graph'
por: Jean, Loïc, et al.
Publicado: (2024)
por: Jean, Loïc, et al.
Publicado: (2024)
Kemeny's Constant for Markov Processes
por: Fitzsimmons, P. J.
Publicado: (2025)
por: Fitzsimmons, P. J.
Publicado: (2025)
On additive averaging kernels for finite Markov chains
por: Lim, Ryan J. Y., et al.
Publicado: (2026)
por: Lim, Ryan J. Y., et al.
Publicado: (2026)
Optimising two-block averaging kernels to speed up Markov chains
por: Lim, Ryan J. Y., et al.
Publicado: (2026)
por: Lim, Ryan J. Y., et al.
Publicado: (2026)
Zero-one Laws for a Control Problem with Random Action Sets
por: Flesch, János, et al.
Publicado: (2024)
por: Flesch, János, et al.
Publicado: (2024)
Efficient Learning for Entropy-Regularized Markov Decision Processes via Multilevel Monte Carlo
por: Meunier, Matthieu, et al.
Publicado: (2025)
por: Meunier, Matthieu, et al.
Publicado: (2025)
A note on weak compactness of occupation measures for an absorbing Markov decision process
por: Dufour, François, et al.
Publicado: (2024)
por: Dufour, François, et al.
Publicado: (2024)
Large independent sets in recursive Markov random graphs
por: Gupte, Akshay, et al.
Publicado: (2022)
por: Gupte, Akshay, et al.
Publicado: (2022)
Kemeny's constant minimization for reversible Markov chains via structure-preserving perturbations
por: Durastante, Fabio, et al.
Publicado: (2025)
por: Durastante, Fabio, et al.
Publicado: (2025)
Asymptotically Optimal Policies for Weakly Coupled Markov Decision Processes
por: Goldsztajn, Diego, et al.
Publicado: (2024)
por: Goldsztajn, Diego, et al.
Publicado: (2024)
Improving the convergence of Markov chains via permutations and projections
por: Choi, Michael C. H., et al.
Publicado: (2024)
por: Choi, Michael C. H., et al.
Publicado: (2024)
Long run control of nonhomogeneous Markov processes
por: Stettner, Łukasz
Publicado: (2025)
por: Stettner, Łukasz
Publicado: (2025)
Strategy Complexity of Limsup and Liminf Threshold Objectives in Countable MDPs, with Applications to Optimal Expected Payoffs
por: Mayr, Richard, et al.
Publicado: (2022)
por: Mayr, Richard, et al.
Publicado: (2022)
Mean-Field Langevin Diffusions with Density-dependent Temperature
por: Huang, Yu-Jui, et al.
Publicado: (2025)
por: Huang, Yu-Jui, et al.
Publicado: (2025)
Stochastic dynamic programming with non-linear discounting
por: Bäuerle, Nicole, et al.
Publicado: (2020)
por: Bäuerle, Nicole, et al.
Publicado: (2020)
Blackwell optimality and policy stability for long-run risk sensitive stochastic control
por: Bäuerle, Nicole, et al.
Publicado: (2024)
por: Bäuerle, Nicole, et al.
Publicado: (2024)
Policy stability and ultimate stationarity in discounted risk-sensitive stochastic control
por: Bäuerle, Nicole, et al.
Publicado: (2026)
por: Bäuerle, Nicole, et al.
Publicado: (2026)
A Fisher-Rao gradient flow for entropy-regularised Markov decision processes in Polish spaces
por: Kerimkulov, Bekzhan, et al.
Publicado: (2023)
por: Kerimkulov, Bekzhan, et al.
Publicado: (2023)
Nonexpansive Markov Operators and Random Function Iterations for Stochastic Fixed Point Problems
por: Hermer, Neal, et al.
Publicado: (2022)
por: Hermer, Neal, et al.
Publicado: (2022)
Controlling the low-temperature Ising model using spatiotemporal Markov decision theory
por: de Jongh, M. C., et al.
Publicado: (2025)
por: de Jongh, M. C., et al.
Publicado: (2025)
Properties of Turnpike Functions for Discounted Finite MDPs
por: Feinberg, Eugene A., et al.
Publicado: (2025)
por: Feinberg, Eugene A., et al.
Publicado: (2025)
Unichain and Aperiodicity are Sufficient for Asymptotic Optimality of Average-Reward Restless Bandits
por: Hong, Yige, et al.
Publicado: (2024)
por: Hong, Yige, et al.
Publicado: (2024)
Deep neural networks can provably solve Bellman equations for Markov decision processes without the curse of dimensionality
por: Jentzen, Arnulf, et al.
Publicado: (2025)
por: Jentzen, Arnulf, et al.
Publicado: (2025)
Information-theoretic minimax and submodular optimization algorithms for multivariate Markov chains
por: Lai, Zheyuan, et al.
Publicado: (2025)
por: Lai, Zheyuan, et al.
Publicado: (2025)
Long-Run Average Reward Maximization of A Regulated Regime-Switching Diffusion Model
por: Zeng, Lingjia, et al.
Publicado: (2025)
por: Zeng, Lingjia, et al.
Publicado: (2025)
Self-interacting CBO: Existence, uniqueness, and long-time convergence
por: Huang, Hui, et al.
Publicado: (2024)
por: Huang, Hui, et al.
Publicado: (2024)
Asymptotically optimal reinforcement learning in Block Markov Decision Processes
por: van Vuren, Thomas, et al.
Publicado: (2025)
por: van Vuren, Thomas, et al.
Publicado: (2025)
Computable Bounds on Convergence of Markov Chains in Wasserstein Distance via Contractive Drift
por: Qu, Yanlin, et al.
Publicado: (2023)
por: Qu, Yanlin, et al.
Publicado: (2023)
Decoupling for Markov Chains
por: Bou-Rabee, Nawaf, et al.
Publicado: (2025)
por: Bou-Rabee, Nawaf, et al.
Publicado: (2025)
Constructive approaches to concentration inequalities with independent random variables
por: Moucer, Celine, et al.
Publicado: (2024)
por: Moucer, Celine, et al.
Publicado: (2024)
Optimal strategies in Markov decision processes with finitely additive evaluations
por: Flesch, János, et al.
Publicado: (2026)
por: Flesch, János, et al.
Publicado: (2026)
Robust Ergodic Control of Jump-Diffusion Systems under Drift and Intensity Uncertainty
por: Azze, Abel, et al.
Publicado: (2026)
por: Azze, Abel, et al.
Publicado: (2026)
A partition method for bounding continuous-time Markov chain models of general reaction network
por: Ballif, Guillaume, et al.
Publicado: (2025)
por: Ballif, Guillaume, et al.
Publicado: (2025)
Auto-exploration for online reinforcement learning
por: Ju, Caleb, et al.
Publicado: (2025)
por: Ju, Caleb, et al.
Publicado: (2025)
Inverse Problems for Ergodicity of Markov Chains
por: Wei, Zhi-Feng
Publicado: (2020)
por: Wei, Zhi-Feng
Publicado: (2020)
Duality of causal distributionally robust optimization
por: Jiang, Yifan
Publicado: (2024)
por: Jiang, Yifan
Publicado: (2024)
On Strategic Measures and Optimality Properties in Discrete-Time Stochastic Control with Universally Measurable Policies
por: Yu, Huizhen
Publicado: (2022)
por: Yu, Huizhen
Publicado: (2022)
Relaxed Equilibria for Time-Inconsistent Markov Decision Processes
por: Bayraktar, Erhan, et al.
Publicado: (2023)
por: Bayraktar, Erhan, et al.
Publicado: (2023)
Large Deviation Asymptotics for the Supermarket Model with Growing Choices
por: Budhiraja, Amarjit, et al.
Publicado: (2025)
por: Budhiraja, Amarjit, et al.
Publicado: (2025)
Ejemplares similares
-
Continuous-time mean field Markov decision models
por: Bäuerle, Nicole, et al.
Publicado: (2023) -
Dynamic programming for the stochastic matching model on general graphs: the case of the `N-graph'
por: Jean, Loïc, et al.
Publicado: (2024) -
Kemeny's Constant for Markov Processes
por: Fitzsimmons, P. J.
Publicado: (2025) -
On additive averaging kernels for finite Markov chains
por: Lim, Ryan J. Y., et al.
Publicado: (2026) -
Optimising two-block averaging kernels to speed up Markov chains
por: Lim, Ryan J. Y., et al.
Publicado: (2026)