Measurized Markov Decision Processes
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Adelman, Daniel, Olivares-Nadal, Alba V. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Thompson Sampling for Infinite-Horizon Discounted Decision Processes
von: Adelman, Daniel, et al.
Veröffentlicht: (2024)
von: Adelman, Daniel, et al.
Veröffentlicht: (2024)
Computable Bounds on Convergence of Markov Chains in Wasserstein Distance via Contractive Drift
von: Qu, Yanlin, et al.
Veröffentlicht: (2023)
von: Qu, Yanlin, et al.
Veröffentlicht: (2023)
Relaxed Equilibria for Time-Inconsistent Markov Decision Processes
von: Bayraktar, Erhan, et al.
Veröffentlicht: (2023)
von: Bayraktar, Erhan, et al.
Veröffentlicht: (2023)
Optimizing Treatment Allocation to Maximize the Health of a Population
von: Adelman, Daniel, et al.
Veröffentlicht: (2026)
von: Adelman, Daniel, et al.
Veröffentlicht: (2026)
Dynamic Basis Function Generation for Network Revenue Management
von: Adelman, Daniel, et al.
Veröffentlicht: (2025)
von: Adelman, Daniel, et al.
Veröffentlicht: (2025)
On finding optimal collective variables for complex systems by minimizing the deviation between effective and full dynamics
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
Properties of Turnpike Functions for Discounted Finite MDPs
von: Feinberg, Eugene A., et al.
Veröffentlicht: (2025)
von: Feinberg, Eugene A., et al.
Veröffentlicht: (2025)
A convex approach for Markov chain estimation from aggregate data via inverse optimal transport
von: Mascherpa, Michele, et al.
Veröffentlicht: (2025)
von: Mascherpa, Michele, et al.
Veröffentlicht: (2025)
First is the worst, second is the best? A Markov chain analysis of the basketball game knockout
von: Flatz, Andrew, et al.
Veröffentlicht: (2025)
von: Flatz, Andrew, et al.
Veröffentlicht: (2025)
On Strategic Measures and Optimality Properties in Discrete-Time Stochastic Control with Universally Measurable Policies
von: Yu, Huizhen
Veröffentlicht: (2022)
von: Yu, Huizhen
Veröffentlicht: (2022)
A Note on the Bias and Kemeny's Constant in Markov Reward Processes with an Application to Markov Chain Perturbation
von: Ortner, Ronald
Veröffentlicht: (2024)
von: Ortner, Ronald
Veröffentlicht: (2024)
Housing Decisions under Mobility Risk: A Stochastic Threshold Approach
von: Wu, Hui
Veröffentlicht: (2026)
von: Wu, Hui
Veröffentlicht: (2026)
Empirical Evaluation of Policy-Based Reinforcement Learning for Dynamic Service Control in an M/M/1 Queue
von: Walton, Joseph, et al.
Veröffentlicht: (2026)
von: Walton, Joseph, et al.
Veröffentlicht: (2026)
Jump Processes with Self-Interactions: Large Deviation Asymptotics
von: Budhiraja, Amarjit, et al.
Veröffentlicht: (2025)
von: Budhiraja, Amarjit, et al.
Veröffentlicht: (2025)
Large independent sets in recursive Markov random graphs
von: Gupte, Akshay, et al.
Veröffentlicht: (2022)
von: Gupte, Akshay, et al.
Veröffentlicht: (2022)
Nonexpansive Markov Operators and Random Function Iterations for Stochastic Fixed Point Problems
von: Hermer, Neal, et al.
Veröffentlicht: (2022)
von: Hermer, Neal, et al.
Veröffentlicht: (2022)
Efficient Learning for Entropy-Regularized Markov Decision Processes via Multilevel Monte Carlo
von: Meunier, Matthieu, et al.
Veröffentlicht: (2025)
von: Meunier, Matthieu, et al.
Veröffentlicht: (2025)
Optimization-based One-side Boundary Control of LWR Traffic Models
von: Vaid, Eryn, et al.
Veröffentlicht: (2026)
von: Vaid, Eryn, et al.
Veröffentlicht: (2026)
Information-theoretic minimax and submodular optimization algorithms for multivariate Markov chains
von: Lai, Zheyuan, et al.
Veröffentlicht: (2025)
von: Lai, Zheyuan, et al.
Veröffentlicht: (2025)
Controlled Interacting Branching Diffusion Processes: A Viscosity Approach
von: Ocello, Antonio
Veröffentlicht: (2026)
von: Ocello, Antonio
Veröffentlicht: (2026)
Continuous-time mean field Markov decision models
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2023)
von: Bäuerle, Nicole, et al.
Veröffentlicht: (2023)
Controlled Interacting Branching Diffusion Processes: Relaxed Formulation in the Mean-Field Regime
von: Ocello, Antonio
Veröffentlicht: (2023)
von: Ocello, Antonio
Veröffentlicht: (2023)
Freidlin-Wentzell solutions of discrete Hamilton Jacobi equations
von: Aleandri, Michele, et al.
Veröffentlicht: (2025)
von: Aleandri, Michele, et al.
Veröffentlicht: (2025)
Large Deviation Asymptotics for the Supermarket Model with Growing Choices
von: Budhiraja, Amarjit, et al.
Veröffentlicht: (2025)
von: Budhiraja, Amarjit, et al.
Veröffentlicht: (2025)
A Framework for Exploring Social Interactions in Multiagent Decision-Making for Two-Queue Systems
von: Gaspard, Mallory E., et al.
Veröffentlicht: (2026)
von: Gaspard, Mallory E., et al.
Veröffentlicht: (2026)
Optimal withdrawals in a general diffusion model with control rates subject to a state-dependent upper bound
von: Guérin, Hélène, et al.
Veröffentlicht: (2024)
von: Guérin, Hélène, et al.
Veröffentlicht: (2024)
Towards Optimal Offline Reinforcement Learning
von: Li, Mengmeng, et al.
Veröffentlicht: (2025)
von: Li, Mengmeng, et al.
Veröffentlicht: (2025)
Partially-Supervised Neural Network Model For Quadratic Multiparametric Programming
von: Beylunioglu, Fuat Can, et al.
Veröffentlicht: (2025)
von: Beylunioglu, Fuat Can, et al.
Veröffentlicht: (2025)
Branched harmonic majorants: representations for multidimensional optimal stopping
von: Moriarty, John
Veröffentlicht: (2025)
von: Moriarty, John
Veröffentlicht: (2025)
Analysing heavy-tail properties of Stochastic Gradient Descent by means of Stochastic Recurrence Equations
von: Damek, Ewa, et al.
Veröffentlicht: (2024)
von: Damek, Ewa, et al.
Veröffentlicht: (2024)
Localization for constrained martingale problems and optimal conditions for uniqueness of reflecting diffusions in 2-dimensional domains
von: Costantini, Cristina, et al.
Veröffentlicht: (2022)
von: Costantini, Cristina, et al.
Veröffentlicht: (2022)
Proving the Chow-Rashevskii Theorem à la Rashevskii
von: Giannotti, Cristina, et al.
Veröffentlicht: (2024)
von: Giannotti, Cristina, et al.
Veröffentlicht: (2024)
Bernstein-type Inequalities for Markov Chains and Markov Processes: A Simple and Robust Proof
von: Huang, De, et al.
Veröffentlicht: (2024)
von: Huang, De, et al.
Veröffentlicht: (2024)
Uniform-in-Time Convergence Rates to a Nonlinear Markov Chain for Mean-Field Interacting Jump Processes
von: Cohen, Asaf, et al.
Veröffentlicht: (2025)
von: Cohen, Asaf, et al.
Veröffentlicht: (2025)
On the continuity of optimal stopping surfaces for jump-diffusions
von: Cai, Cheng, et al.
Veröffentlicht: (2021)
von: Cai, Cheng, et al.
Veröffentlicht: (2021)
Occasionally Observed Piecewise-deterministic Markov Processes
von: Gee, Marissa, et al.
Veröffentlicht: (2024)
von: Gee, Marissa, et al.
Veröffentlicht: (2024)
Extended Laplace Principle for Empirical Measures of a Markov Chain
von: Eckstein, Stephan
Veröffentlicht: (2017)
von: Eckstein, Stephan
Veröffentlicht: (2017)
Stackelberg stopping games
von: Zhang, Jingjie, et al.
Veröffentlicht: (2025)
von: Zhang, Jingjie, et al.
Veröffentlicht: (2025)
On the saddle point of a zero-sum stopper vs. singular-controller game
von: Bovo, Andrea, et al.
Veröffentlicht: (2024)
von: Bovo, Andrea, et al.
Veröffentlicht: (2024)
Parallel Affine Transformation Tuning of Markov Chain Monte Carlo
von: Schär, Philip, et al.
Veröffentlicht: (2024)
von: Schär, Philip, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Thompson Sampling for Infinite-Horizon Discounted Decision Processes
von: Adelman, Daniel, et al.
Veröffentlicht: (2024) -
Computable Bounds on Convergence of Markov Chains in Wasserstein Distance via Contractive Drift
von: Qu, Yanlin, et al.
Veröffentlicht: (2023) -
Relaxed Equilibria for Time-Inconsistent Markov Decision Processes
von: Bayraktar, Erhan, et al.
Veröffentlicht: (2023) -
Optimizing Treatment Allocation to Maximize the Health of a Population
von: Adelman, Daniel, et al.
Veröffentlicht: (2026) -
Dynamic Basis Function Generation for Network Revenue Management
von: Adelman, Daniel, et al.
Veröffentlicht: (2025)