Data-driven robust Markov decision processes on Borel spaces: performance guarantees via an axiomatic approach
Fuente:
arXiv
Guardado en:
| Autor principal: | Ramani, Sivaramakrishnan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A safe exploration approach to constrained Markov decision processes
por: Ni, Tingting, et al.
Publicado: (2023)
por: Ni, Tingting, et al.
Publicado: (2023)
Universal generalization guarantees for Wasserstein distributionally robust models
por: Le, Tam, et al.
Publicado: (2024)
por: Le, Tam, et al.
Publicado: (2024)
Stochastic first-order methods for average-reward Markov decision processes
por: Li, Tianjiao, et al.
Publicado: (2022)
por: Li, Tianjiao, et al.
Publicado: (2022)
Logarithmic regret bounds for continuous-time average-reward Markov decision processes
por: Gao, Xuefeng, et al.
Publicado: (2022)
por: Gao, Xuefeng, et al.
Publicado: (2022)
Finite-sample guarantees for data-driven forward-backward operator methods
por: Fabiani, Filippo, et al.
Publicado: (2025)
por: Fabiani, Filippo, et al.
Publicado: (2025)
End-to-end guarantees for indirect data-driven control of bilinear systems with finite stochastic data
por: Chatzikiriakos, Nicolas, et al.
Publicado: (2024)
por: Chatzikiriakos, Nicolas, et al.
Publicado: (2024)
RGNMR: A Gauss-Newton method for robust matrix completion with theoretical guarantees
por: Laufer, Eilon Vaknin, et al.
Publicado: (2025)
por: Laufer, Eilon Vaknin, et al.
Publicado: (2025)
Beyond IID: data-driven decision-making in heterogeneous environments
por: Besbes, Omar, et al.
Publicado: (2022)
por: Besbes, Omar, et al.
Publicado: (2022)
Proximal basin hopping: global optimization with guarantees
por: Lauga, Guillaume, et al.
Publicado: (2026)
por: Lauga, Guillaume, et al.
Publicado: (2026)
A dynamical neural network approach for distributionally robust chance constrained Markov decision process
por: Xia, Tian, et al.
Publicado: (2023)
por: Xia, Tian, et al.
Publicado: (2023)
Scenario theory for multi-criteria data-driven decision making
por: Garatti, Simone, et al.
Publicado: (2026)
por: Garatti, Simone, et al.
Publicado: (2026)
A robust and adaptive MPC formulation for Gaussian process models
por: Dubied, Mathieu, et al.
Publicado: (2025)
por: Dubied, Mathieu, et al.
Publicado: (2025)
A constrained optimization approach to improve robustness of neural networks
por: Zhao, Shudian, et al.
Publicado: (2024)
por: Zhao, Shudian, et al.
Publicado: (2024)
A minimax optimal control approach for robust neural ODEs
por: Cipriani, Cristina, et al.
Publicado: (2023)
por: Cipriani, Cristina, et al.
Publicado: (2023)
Automatic nonlinear MPC approximation with closed-loop guarantees
por: Tokmak, Abdullah, et al.
Publicado: (2023)
por: Tokmak, Abdullah, et al.
Publicado: (2023)
Data-driven approaches to inverse problems
por: Schönlieb, Carola-Bibiane, et al.
Publicado: (2025)
por: Schönlieb, Carola-Bibiane, et al.
Publicado: (2025)
A randomized algorithm for nonconvex minimization with inexact evaluations and complexity guarantees
por: Li, Shuyao, et al.
Publicado: (2023)
por: Li, Shuyao, et al.
Publicado: (2023)
Learning to accelerate Krasnosel'skii-Mann fixed-point iterations with guarantees
por: Martin, Andrea, et al.
Publicado: (2026)
por: Martin, Andrea, et al.
Publicado: (2026)
Learning to optimize with guarantees: a complete characterization of linearly convergent algorithms
por: Martin, Andrea, et al.
Publicado: (2025)
por: Martin, Andrea, et al.
Publicado: (2025)
Randomized multi-class classification under system constraints: a unified approach via post-processing
por: Chzhen, Evgenii, et al.
Publicado: (2025)
por: Chzhen, Evgenii, et al.
Publicado: (2025)
Data/moment-driven approaches for fast predictive control of collective dynamics
por: Albi, Giacomo, et al.
Publicado: (2024)
por: Albi, Giacomo, et al.
Publicado: (2024)
Policy Mirror Descent with Temporal Difference Learning: Sample Complexity under Online Markov Data
por: Li, Wenye, et al.
Publicado: (2025)
por: Li, Wenye, et al.
Publicado: (2025)
A universal policy wrapper with guarantees
por: Bolychev, Anton, et al.
Publicado: (2025)
por: Bolychev, Anton, et al.
Publicado: (2025)
Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form
por: Kitamura, Toshinori, et al.
Publicado: (2024)
por: Kitamura, Toshinori, et al.
Publicado: (2024)
Data augmentation for machine learning of chemical process flowsheets
por: Balhorn, Lukas Schulze, et al.
Publicado: (2023)
por: Balhorn, Lukas Schulze, et al.
Publicado: (2023)
Robust Data-driven Prescriptiveness Optimization
por: Poursoltani, Mehran, et al.
Publicado: (2023)
por: Poursoltani, Mehran, et al.
Publicado: (2023)
Soft decision trees for survival analysis
por: Consolo, Antonio, et al.
Publicado: (2025)
por: Consolo, Antonio, et al.
Publicado: (2025)
Data-driven Mixed Integer Optimization through Probabilistic Multi-variable Branching
por: Chen, Yanguang, et al.
Publicado: (2023)
por: Chen, Yanguang, et al.
Publicado: (2023)
Data-driven Piecewise Affine Decision Rules for Stochastic Programming with Covariate Information
por: Zhang, Yiyang, et al.
Publicado: (2023)
por: Zhang, Yiyang, et al.
Publicado: (2023)
Data-driven Reachable Set Estimation with Tunable Adversarial and Wasserstein Distributional Guarantees
por: Pantazis, Georgios, et al.
Publicado: (2026)
por: Pantazis, Georgios, et al.
Publicado: (2026)
Stochastic set-valued optimization and its application to robust learning
por: Giovannelli, Tommaso, et al.
Publicado: (2026)
por: Giovannelli, Tommaso, et al.
Publicado: (2026)
Decision-calibrated prediction sets for robust power system operations
por: Stratigakos, Akylas, et al.
Publicado: (2026)
por: Stratigakos, Akylas, et al.
Publicado: (2026)
Scalarisation-based risk concepts for robust multi-objective optimisation
por: Tu, Ben, et al.
Publicado: (2024)
por: Tu, Ben, et al.
Publicado: (2024)
Weakly Time-Coupled Approximation of Markov Decision Processes
por: Soheili, Negar, et al.
Publicado: (2026)
por: Soheili, Negar, et al.
Publicado: (2026)
Online Markov Decision Processes with Terminal Law Constraints
por: Moreno, Bianca Marin, et al.
Publicado: (2026)
por: Moreno, Bianca Marin, et al.
Publicado: (2026)
Reinforcement Learning with Function Approximation for Non-Markov Processes
por: Kara, Ali Devran
Publicado: (2026)
por: Kara, Ali Devran
Publicado: (2026)
Adaptive decision-making for stochastic service network design
por: Durán-Micco, Javier, et al.
Publicado: (2026)
por: Durán-Micco, Javier, et al.
Publicado: (2026)
Unregularized limit of stochastic gradient method for Wasserstein distributionally robust optimization
por: Le, Tam
Publicado: (2025)
por: Le, Tam
Publicado: (2025)
Worst-case generation via minimax optimization in Wasserstein space
por: Cheng, Xiuyuan, et al.
Publicado: (2025)
por: Cheng, Xiuyuan, et al.
Publicado: (2025)
Optimal Sample Complexity for Average Reward Markov Decision Processes
por: Wang, Shengbo, et al.
Publicado: (2023)
por: Wang, Shengbo, et al.
Publicado: (2023)
Ejemplares similares
-
A safe exploration approach to constrained Markov decision processes
por: Ni, Tingting, et al.
Publicado: (2023) -
Universal generalization guarantees for Wasserstein distributionally robust models
por: Le, Tam, et al.
Publicado: (2024) -
Stochastic first-order methods for average-reward Markov decision processes
por: Li, Tianjiao, et al.
Publicado: (2022) -
Logarithmic regret bounds for continuous-time average-reward Markov decision processes
por: Gao, Xuefeng, et al.
Publicado: (2022) -
Finite-sample guarantees for data-driven forward-backward operator methods
por: Fabiani, Filippo, et al.
Publicado: (2025)