Asymptotically optimal reinforcement learning in Block Markov Decision Processes
Fuente:
arXiv
Saved in:
| Main Authors: | van Vuren, Thomas, Sloothaak, Fiona, Wolf, Maarten G., Sanders, Jaron |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Estimating the number of clusters of a Block Markov Chain
by: van Vuren, Thomas, et al.
Published: (2024)
by: van Vuren, Thomas, et al.
Published: (2024)
Recovering semipermeable barriers from reflected Brownian motion
by: Van Werde, Alexander, et al.
Published: (2024)
by: Van Werde, Alexander, et al.
Published: (2024)
Rough backward SDEs with discontinuous Young drivers
by: Becherer, Dirk, et al.
Published: (2025)
by: Becherer, Dirk, et al.
Published: (2025)
Learning payoffs while routing in skill-based queues
by: van Kempen, Sanne, et al.
Published: (2024)
by: van Kempen, Sanne, et al.
Published: (2024)
Asymptotic Normality of Chatterjee's Rank Correlation
by: Kroll, Marius
Published: (2024)
by: Kroll, Marius
Published: (2024)
Series representations for the characteristic function of the multidimensional Markov random flight
by: Kolesnik, Alexander D.
Published: (2023)
by: Kolesnik, Alexander D.
Published: (2023)
Occupied Processes: Going with the Flow
by: Tissot-Daguette, Valentin
Published: (2023)
by: Tissot-Daguette, Valentin
Published: (2023)
Zero-Coupon Treasury Rates and Returns using the Volatility Index
by: Park, Jihyun, et al.
Published: (2024)
by: Park, Jihyun, et al.
Published: (2024)
Scaling limits of multi-period distributionally robust optimization problems
by: Nendel, Max, et al.
Published: (2025)
by: Nendel, Max, et al.
Published: (2025)
Complete Asymptotic Expansions for the Normalizing Constants of High-Dimensional Matrix Bingham and Matrix Langevin Distributions
by: Bagyan, Armine, et al.
Published: (2024)
by: Bagyan, Armine, et al.
Published: (2024)
Stein's method for the matrix normal distribution
by: Gaunt, Robert E., et al.
Published: (2026)
by: Gaunt, Robert E., et al.
Published: (2026)
A Note on the Bias and Kemeny's Constant in Markov Reward Processes with an Application to Markov Chain Perturbation
by: Ortner, Ronald
Published: (2024)
by: Ortner, Ronald
Published: (2024)
(Almost) complete characterization of stability of a discrete-time Hawkes process with inhibition and memory of length two
by: Costa, Manon, et al.
Published: (2023)
by: Costa, Manon, et al.
Published: (2023)
On the maximal correlation of some stochastic processes
by: Chang, Yinshan, et al.
Published: (2024)
by: Chang, Yinshan, et al.
Published: (2024)
Multi-level reflecting Brownian motion on the half line and its stationary distribution
by: Miyazawa, Masakiyo
Published: (2024)
by: Miyazawa, Masakiyo
Published: (2024)
The stationary distributions of state-dependent diffusions reflected at one and two sides
by: Miyazawa, Masakiyo
Published: (2024)
by: Miyazawa, Masakiyo
Published: (2024)
A Malliavin-Gamma calculus approach to Score Based Diffusion Generative models for random fields
by: Greco, Giacomo
Published: (2025)
by: Greco, Giacomo
Published: (2025)
Optimal stopping involving a diffusion and its running maximum: a generalisation of the maximality principle
by: Rodosthenous, Neofytos, et al.
Published: (2025)
by: Rodosthenous, Neofytos, et al.
Published: (2025)
The exact region and an inequality between Chatterjee's and Spearman's rank correlations
by: Ansari, Jonathan, et al.
Published: (2025)
by: Ansari, Jonathan, et al.
Published: (2025)
Simulating conditioned diffusions on manifolds
by: Corstanje, Marc, et al.
Published: (2024)
by: Corstanje, Marc, et al.
Published: (2024)
The Gapeev-Shiryaev Conjecture
by: Ernst, Philip A., et al.
Published: (2024)
by: Ernst, Philip A., et al.
Published: (2024)
Zero-one Laws for a Control Problem with Random Action Sets
by: Flesch, János, et al.
Published: (2024)
by: Flesch, János, et al.
Published: (2024)
Asymptotics of generalized Pólya urns with non-linear feedback
by: Gottfried, Thomas, et al.
Published: (2023)
by: Gottfried, Thomas, et al.
Published: (2023)
A Geometric Witness Framework for Signed Multivariate Tail-Dependence Compatibility: Asymptotic Structure and Finite-Threshold Synthesis
by: Milek, Janusz
Published: (2026)
by: Milek, Janusz
Published: (2026)
Stochastic dynamic programming with non-linear discounting
by: Bäuerle, Nicole, et al.
Published: (2020)
by: Bäuerle, Nicole, et al.
Published: (2020)
Large Sample Theory for Bures-Wasserstein Barycentres
by: Santoro, Leonardo V., et al.
Published: (2023)
by: Santoro, Leonardo V., et al.
Published: (2023)
Random Matrices and U-Statistics
by: Benaych-Georges, Florent, et al.
Published: (2025)
by: Benaych-Georges, Florent, et al.
Published: (2025)
Long-Term Average Impulse and Singular Control of a Growth Model with Two Revenue Sources
by: Helmes, K. L., et al.
Published: (2026)
by: Helmes, K. L., et al.
Published: (2026)
Stochastic quantization of the three-dimensional polymer measure via the Dirichlet form method
by: Albeverio, Sergio, et al.
Published: (2023)
by: Albeverio, Sergio, et al.
Published: (2023)
Large deviation principle for stochastic differential equations driven by stochastic integrals
by: Takano, Ryoji
Published: (2024)
by: Takano, Ryoji
Published: (2024)
Limit Profiles for Reversible Markov Chains
by: Nestoridi, Evita, et al.
Published: (2020)
by: Nestoridi, Evita, et al.
Published: (2020)
Blackwell optimality and policy stability for long-run risk sensitive stochastic control
by: Bäuerle, Nicole, et al.
Published: (2024)
by: Bäuerle, Nicole, et al.
Published: (2024)
Continuous-time mean field Markov decision models
by: Bäuerle, Nicole, et al.
Published: (2023)
by: Bäuerle, Nicole, et al.
Published: (2023)
High order splitting methods for SDEs satisfying a commutativity condition
by: Foster, James, et al.
Published: (2022)
by: Foster, James, et al.
Published: (2022)
Approximating the signature of Brownian motion for high order SDE simulation
by: Foster, James
Published: (2024)
by: Foster, James
Published: (2024)
Entropy Estimate for Degenerate SDEs with Applications to Nonlinear Kinetic Fokker-Planck Equations
by: Qian, Zhongmin, et al.
Published: (2023)
by: Qian, Zhongmin, et al.
Published: (2023)
Efficient Learning for Entropy-Regularized Markov Decision Processes via Multilevel Monte Carlo
by: Meunier, Matthieu, et al.
Published: (2025)
by: Meunier, Matthieu, et al.
Published: (2025)
Kemeny's constant minimization for reversible Markov chains via structure-preserving perturbations
by: Durastante, Fabio, et al.
Published: (2025)
by: Durastante, Fabio, et al.
Published: (2025)
On joint returns to zero of Bessel processes
by: Berger, Quentin, et al.
Published: (2024)
by: Berger, Quentin, et al.
Published: (2024)
Operator semigroups in the mixed topology and the infinitesimal description of Markov processes
by: Goldys, Ben, et al.
Published: (2022)
by: Goldys, Ben, et al.
Published: (2022)
Similar Items
-
Estimating the number of clusters of a Block Markov Chain
by: van Vuren, Thomas, et al.
Published: (2024) -
Recovering semipermeable barriers from reflected Brownian motion
by: Van Werde, Alexander, et al.
Published: (2024) -
Rough backward SDEs with discontinuous Young drivers
by: Becherer, Dirk, et al.
Published: (2025) -
Learning payoffs while routing in skill-based queues
by: van Kempen, Sanne, et al.
Published: (2024) -
Asymptotic Normality of Chatterjee's Rank Correlation
by: Kroll, Marius
Published: (2024)