Asymptotically optimal reinforcement learning in Block Markov Decision Processes
Fuente:
arXiv
Salvato in:
| Autori principali: | van Vuren, Thomas, Sloothaak, Fiona, Wolf, Maarten G., Sanders, Jaron |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Estimating the number of clusters of a Block Markov Chain
di: van Vuren, Thomas, et al.
Pubblicazione: (2024)
di: van Vuren, Thomas, et al.
Pubblicazione: (2024)
Recovering semipermeable barriers from reflected Brownian motion
di: Van Werde, Alexander, et al.
Pubblicazione: (2024)
di: Van Werde, Alexander, et al.
Pubblicazione: (2024)
Rough backward SDEs with discontinuous Young drivers
di: Becherer, Dirk, et al.
Pubblicazione: (2025)
di: Becherer, Dirk, et al.
Pubblicazione: (2025)
Learning payoffs while routing in skill-based queues
di: van Kempen, Sanne, et al.
Pubblicazione: (2024)
di: van Kempen, Sanne, et al.
Pubblicazione: (2024)
Asymptotic Normality of Chatterjee's Rank Correlation
di: Kroll, Marius
Pubblicazione: (2024)
di: Kroll, Marius
Pubblicazione: (2024)
Series representations for the characteristic function of the multidimensional Markov random flight
di: Kolesnik, Alexander D.
Pubblicazione: (2023)
di: Kolesnik, Alexander D.
Pubblicazione: (2023)
Occupied Processes: Going with the Flow
di: Tissot-Daguette, Valentin
Pubblicazione: (2023)
di: Tissot-Daguette, Valentin
Pubblicazione: (2023)
Zero-Coupon Treasury Rates and Returns using the Volatility Index
di: Park, Jihyun, et al.
Pubblicazione: (2024)
di: Park, Jihyun, et al.
Pubblicazione: (2024)
Scaling limits of multi-period distributionally robust optimization problems
di: Nendel, Max, et al.
Pubblicazione: (2025)
di: Nendel, Max, et al.
Pubblicazione: (2025)
Complete Asymptotic Expansions for the Normalizing Constants of High-Dimensional Matrix Bingham and Matrix Langevin Distributions
di: Bagyan, Armine, et al.
Pubblicazione: (2024)
di: Bagyan, Armine, et al.
Pubblicazione: (2024)
Stein's method for the matrix normal distribution
di: Gaunt, Robert E., et al.
Pubblicazione: (2026)
di: Gaunt, Robert E., et al.
Pubblicazione: (2026)
A Note on the Bias and Kemeny's Constant in Markov Reward Processes with an Application to Markov Chain Perturbation
di: Ortner, Ronald
Pubblicazione: (2024)
di: Ortner, Ronald
Pubblicazione: (2024)
(Almost) complete characterization of stability of a discrete-time Hawkes process with inhibition and memory of length two
di: Costa, Manon, et al.
Pubblicazione: (2023)
di: Costa, Manon, et al.
Pubblicazione: (2023)
On the maximal correlation of some stochastic processes
di: Chang, Yinshan, et al.
Pubblicazione: (2024)
di: Chang, Yinshan, et al.
Pubblicazione: (2024)
Multi-level reflecting Brownian motion on the half line and its stationary distribution
di: Miyazawa, Masakiyo
Pubblicazione: (2024)
di: Miyazawa, Masakiyo
Pubblicazione: (2024)
The stationary distributions of state-dependent diffusions reflected at one and two sides
di: Miyazawa, Masakiyo
Pubblicazione: (2024)
di: Miyazawa, Masakiyo
Pubblicazione: (2024)
A Malliavin-Gamma calculus approach to Score Based Diffusion Generative models for random fields
di: Greco, Giacomo
Pubblicazione: (2025)
di: Greco, Giacomo
Pubblicazione: (2025)
Optimal stopping involving a diffusion and its running maximum: a generalisation of the maximality principle
di: Rodosthenous, Neofytos, et al.
Pubblicazione: (2025)
di: Rodosthenous, Neofytos, et al.
Pubblicazione: (2025)
The exact region and an inequality between Chatterjee's and Spearman's rank correlations
di: Ansari, Jonathan, et al.
Pubblicazione: (2025)
di: Ansari, Jonathan, et al.
Pubblicazione: (2025)
Simulating conditioned diffusions on manifolds
di: Corstanje, Marc, et al.
Pubblicazione: (2024)
di: Corstanje, Marc, et al.
Pubblicazione: (2024)
The Gapeev-Shiryaev Conjecture
di: Ernst, Philip A., et al.
Pubblicazione: (2024)
di: Ernst, Philip A., et al.
Pubblicazione: (2024)
Zero-one Laws for a Control Problem with Random Action Sets
di: Flesch, János, et al.
Pubblicazione: (2024)
di: Flesch, János, et al.
Pubblicazione: (2024)
Asymptotics of generalized Pólya urns with non-linear feedback
di: Gottfried, Thomas, et al.
Pubblicazione: (2023)
di: Gottfried, Thomas, et al.
Pubblicazione: (2023)
A Geometric Witness Framework for Signed Multivariate Tail-Dependence Compatibility: Asymptotic Structure and Finite-Threshold Synthesis
di: Milek, Janusz
Pubblicazione: (2026)
di: Milek, Janusz
Pubblicazione: (2026)
Stochastic dynamic programming with non-linear discounting
di: Bäuerle, Nicole, et al.
Pubblicazione: (2020)
di: Bäuerle, Nicole, et al.
Pubblicazione: (2020)
Large Sample Theory for Bures-Wasserstein Barycentres
di: Santoro, Leonardo V., et al.
Pubblicazione: (2023)
di: Santoro, Leonardo V., et al.
Pubblicazione: (2023)
Random Matrices and U-Statistics
di: Benaych-Georges, Florent, et al.
Pubblicazione: (2025)
di: Benaych-Georges, Florent, et al.
Pubblicazione: (2025)
Long-Term Average Impulse and Singular Control of a Growth Model with Two Revenue Sources
di: Helmes, K. L., et al.
Pubblicazione: (2026)
di: Helmes, K. L., et al.
Pubblicazione: (2026)
Stochastic quantization of the three-dimensional polymer measure via the Dirichlet form method
di: Albeverio, Sergio, et al.
Pubblicazione: (2023)
di: Albeverio, Sergio, et al.
Pubblicazione: (2023)
Large deviation principle for stochastic differential equations driven by stochastic integrals
di: Takano, Ryoji
Pubblicazione: (2024)
di: Takano, Ryoji
Pubblicazione: (2024)
Limit Profiles for Reversible Markov Chains
di: Nestoridi, Evita, et al.
Pubblicazione: (2020)
di: Nestoridi, Evita, et al.
Pubblicazione: (2020)
Blackwell optimality and policy stability for long-run risk sensitive stochastic control
di: Bäuerle, Nicole, et al.
Pubblicazione: (2024)
di: Bäuerle, Nicole, et al.
Pubblicazione: (2024)
Continuous-time mean field Markov decision models
di: Bäuerle, Nicole, et al.
Pubblicazione: (2023)
di: Bäuerle, Nicole, et al.
Pubblicazione: (2023)
High order splitting methods for SDEs satisfying a commutativity condition
di: Foster, James, et al.
Pubblicazione: (2022)
di: Foster, James, et al.
Pubblicazione: (2022)
Approximating the signature of Brownian motion for high order SDE simulation
di: Foster, James
Pubblicazione: (2024)
di: Foster, James
Pubblicazione: (2024)
Entropy Estimate for Degenerate SDEs with Applications to Nonlinear Kinetic Fokker-Planck Equations
di: Qian, Zhongmin, et al.
Pubblicazione: (2023)
di: Qian, Zhongmin, et al.
Pubblicazione: (2023)
Efficient Learning for Entropy-Regularized Markov Decision Processes via Multilevel Monte Carlo
di: Meunier, Matthieu, et al.
Pubblicazione: (2025)
di: Meunier, Matthieu, et al.
Pubblicazione: (2025)
Kemeny's constant minimization for reversible Markov chains via structure-preserving perturbations
di: Durastante, Fabio, et al.
Pubblicazione: (2025)
di: Durastante, Fabio, et al.
Pubblicazione: (2025)
On joint returns to zero of Bessel processes
di: Berger, Quentin, et al.
Pubblicazione: (2024)
di: Berger, Quentin, et al.
Pubblicazione: (2024)
Operator semigroups in the mixed topology and the infinitesimal description of Markov processes
di: Goldys, Ben, et al.
Pubblicazione: (2022)
di: Goldys, Ben, et al.
Pubblicazione: (2022)
Documenti analoghi
-
Estimating the number of clusters of a Block Markov Chain
di: van Vuren, Thomas, et al.
Pubblicazione: (2024) -
Recovering semipermeable barriers from reflected Brownian motion
di: Van Werde, Alexander, et al.
Pubblicazione: (2024) -
Rough backward SDEs with discontinuous Young drivers
di: Becherer, Dirk, et al.
Pubblicazione: (2025) -
Learning payoffs while routing in skill-based queues
di: van Kempen, Sanne, et al.
Pubblicazione: (2024) -
Asymptotic Normality of Chatterjee's Rank Correlation
di: Kroll, Marius
Pubblicazione: (2024)