Beyond Average Return in Markov Decision Processes
Fuente:
arXiv
Saved in:
| Main Authors: | Marthe, Alexandre, Garivier, Aurélien, Vernade, Claire |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient Risk-sensitive Planning via Entropic Risk Measures
by: Marthe, Alexandre, et al.
Published: (2025)
by: Marthe, Alexandre, et al.
Published: (2025)
Robust $Q$-learning Algorithm for Markov Decision Processes under Wasserstein Uncertainty
by: Neufeld, Ariel, et al.
Published: (2022)
by: Neufeld, Ariel, et al.
Published: (2022)
An Online Multiobjective Policy Gradient for Long-run Average-reward Markov Decision Process
by: Misra, Rahul, et al.
Published: (2025)
by: Misra, Rahul, et al.
Published: (2025)
Lagrangian Index Policy for Restless Bandits with Average Reward
by: Avrachenkov, Konstantin, et al.
Published: (2024)
by: Avrachenkov, Konstantin, et al.
Published: (2024)
Reoptimization Nearly Solves Weakly Coupled Markov Decision Processes
by: Gast, Nicolas, et al.
Published: (2022)
by: Gast, Nicolas, et al.
Published: (2022)
On Pioneering Works of Albert Shiryaev on Markov Decision Processes and Some Later Developments
by: Feinberg, Eugene A.
Published: (2025)
by: Feinberg, Eugene A.
Published: (2025)
Markov Decision Processing Networks
by: Bhambay, Sanidhay, et al.
Published: (2025)
by: Bhambay, Sanidhay, et al.
Published: (2025)
Markov Decision Process and Approximate Dynamic Programming for a Patient Assignment Scheduling problem
by: O'Reilly, Malgorzata M., et al.
Published: (2024)
by: O'Reilly, Malgorzata M., et al.
Published: (2024)
Non-Exchangeable Mean Field Markov Decision Processes with common noise : from Bellman equation to quantitative propagation of chaos
by: Mekkaoui, Samy, et al.
Published: (2026)
by: Mekkaoui, Samy, et al.
Published: (2026)
On Dynamic Programming Decompositions of Static Risk Measures in Markov Decision Processes
by: Hau, Jia Lin, et al.
Published: (2023)
by: Hau, Jia Lin, et al.
Published: (2023)
Graphon Particle Systems, Part II: Dynamics of Distributed Stochastic Continuum Optimization
by: Chen, Yan, et al.
Published: (2024)
by: Chen, Yan, et al.
Published: (2024)
Bounding the Difference between the Values of Robust and Non-Robust Markov Decision Problems
by: Neufeld, Ariel, et al.
Published: (2023)
by: Neufeld, Ariel, et al.
Published: (2023)
Robust Markov Decision Processes: A Place Where AI and Formal Methods Meet
by: Suilen, Marnix, et al.
Published: (2024)
by: Suilen, Marnix, et al.
Published: (2024)
Optimal Risk-Sensitive Scheduling Policies for Remote Estimation of Autoregressive Markov Processes
by: Dutta, Manali, et al.
Published: (2024)
by: Dutta, Manali, et al.
Published: (2024)
Feature-aligned N-BEATS with Sinkhorn divergence
by: Lee, Joonhun, et al.
Published: (2023)
by: Lee, Joonhun, et al.
Published: (2023)
Neural Brownian Motion
by: Qi, Qian
Published: (2025)
by: Qi, Qian
Published: (2025)
Probabilistic Geometric Alignment via Bayesian Latent Transport for Domain-Adaptive Foundation Models
by: Aueawatthanaphisut, Aueaphum, et al.
Published: (2026)
by: Aueawatthanaphisut, Aueaphum, et al.
Published: (2026)
Q-Learning under Finite Model Uncertainty
by: Sester, Julian, et al.
Published: (2024)
by: Sester, Julian, et al.
Published: (2024)
On the consistent reasoning paradox of intelligence and optimal trust in AI: The power of 'I don't know'
by: Bastounis, Alexander, et al.
Published: (2024)
by: Bastounis, Alexander, et al.
Published: (2024)
Robust Reward Design for Markov Decision Processes
by: Wu, Shuo, et al.
Published: (2024)
by: Wu, Shuo, et al.
Published: (2024)
Further remarks on absorbing Markov decision processes
by: Zhang, Yi, et al.
Published: (2024)
by: Zhang, Yi, et al.
Published: (2024)
Langevin dynamics for the probability of finite state Markov processes
by: Li, Wuchen
Published: (2023)
by: Li, Wuchen
Published: (2023)
Markov decision processes: on the convergence of the Monte-Carlo first visit algorithm
by: Delattre, Sylvain, et al.
Published: (2025)
by: Delattre, Sylvain, et al.
Published: (2025)
Nonstationary nonzero-sum Markov games under a probability criterion
by: Guo, Xin, et al.
Published: (2025)
by: Guo, Xin, et al.
Published: (2025)
Markov-Nash equilibria in mean-field games under model uncertainty
by: Langner, Johannes, et al.
Published: (2024)
by: Langner, Johannes, et al.
Published: (2024)
On reachability of Markov decision processes: a novel state-classification-based PI approach
by: Li, Yanyun, et al.
Published: (2023)
by: Li, Yanyun, et al.
Published: (2023)
Asymptotic Optimality in Data-Driven Decision Making
by: Salač, Radek, et al.
Published: (2025)
by: Salač, Radek, et al.
Published: (2025)
Optimal control of McKean-Vlasov systems under partial observation and hidden Markov switching
by: Fuhrman, Marco, et al.
Published: (2026)
by: Fuhrman, Marco, et al.
Published: (2026)
Linear Algebraic Truncation Algorithm with A Posteriori Error Bounds for Computing Markov Chain Equilibrium Gradients
by: Mahdian, Saied, et al.
Published: (2025)
by: Mahdian, Saied, et al.
Published: (2025)
Zero-Sum Games for piecewise deterministic Markov decision processes with risk-sensitive finite-horizon cost criterion
by: Golui, Subrata
Published: (2024)
by: Golui, Subrata
Published: (2024)
The non-linear multiple stopping problem: between the discrete and the continuous time
by: Grigorova, Miryana, et al.
Published: (2025)
by: Grigorova, Miryana, et al.
Published: (2025)
Chance-Constrained Generic Energy Storage Operations under Decision-Dependent Uncertainty
by: Qi, Ning, et al.
Published: (2022)
by: Qi, Ning, et al.
Published: (2022)
Normalizing flow regularization for photoacoustic tomography
by: Wang, Chao, et al.
Published: (2024)
by: Wang, Chao, et al.
Published: (2024)
Time discretization of BSDEs with singular terminal condition using asymptotic expansion
by: Kruse, Thomas, et al.
Published: (2026)
by: Kruse, Thomas, et al.
Published: (2026)
About a Ball Removal Process on Bins
by: Correa, Jose, et al.
Published: (2026)
by: Correa, Jose, et al.
Published: (2026)
The Wasserstein Space of Stochastic Processes in Continuous Time
by: Bartl, Daniel, et al.
Published: (2025)
by: Bartl, Daniel, et al.
Published: (2025)
Synchronous Heterogeneous Exclusion Processes on Open Lattice
by: Yashina, Marina V., et al.
Published: (2024)
by: Yashina, Marina V., et al.
Published: (2024)
Adapted Optimal Transport between Filtered Gaussian Processes
by: Gunasingam, Madhu, et al.
Published: (2026)
by: Gunasingam, Madhu, et al.
Published: (2026)
Matrix Riccati BSDEs with singular terminal condition and stochastic LQ control with linear terminal constraint
by: Ackermann, Julia, et al.
Published: (2026)
by: Ackermann, Julia, et al.
Published: (2026)
Algebraic Reduction of Hidden Markov Models
by: Grigoletto, Tommaso, et al.
Published: (2022)
by: Grigoletto, Tommaso, et al.
Published: (2022)
Similar Items
-
Efficient Risk-sensitive Planning via Entropic Risk Measures
by: Marthe, Alexandre, et al.
Published: (2025) -
Robust $Q$-learning Algorithm for Markov Decision Processes under Wasserstein Uncertainty
by: Neufeld, Ariel, et al.
Published: (2022) -
An Online Multiobjective Policy Gradient for Long-run Average-reward Markov Decision Process
by: Misra, Rahul, et al.
Published: (2025) -
Lagrangian Index Policy for Restless Bandits with Average Reward
by: Avrachenkov, Konstantin, et al.
Published: (2024) -
Reoptimization Nearly Solves Weakly Coupled Markov Decision Processes
by: Gast, Nicolas, et al.
Published: (2022)