Solving Robust Markov Decision Processes: Generic, Reliable, Efficient
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Meggendorfer, Tobias, Weininger, Maximilian, Wienhöft, Patrick |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
What Are the Odds? Improving the foundations of Statistical Model Checking
von: Meggendorfer, Tobias, et al.
Veröffentlicht: (2024)
von: Meggendorfer, Tobias, et al.
Veröffentlicht: (2024)
Playing Games with your PET: Extending the Partial Exploration Tool to Stochastic Games
von: Meggendorfer, Tobias, et al.
Veröffentlicht: (2024)
von: Meggendorfer, Tobias, et al.
Veröffentlicht: (2024)
1-2-3-Go! Policy Synthesis for Parameterized Markov Decision Processes via Decision-Tree Learning and Generalization
von: Azeem, Muqsit, et al.
Veröffentlicht: (2024)
von: Azeem, Muqsit, et al.
Veröffentlicht: (2024)
Efficient and Sharp Off-Policy Evaluation in Robust Markov Decision Processes
von: Bennett, Andrew, et al.
Veröffentlicht: (2024)
von: Bennett, Andrew, et al.
Veröffentlicht: (2024)
Policy Gradient for Robust Markov Decision Processes
von: Wang, Qiuhao, et al.
Veröffentlicht: (2024)
von: Wang, Qiuhao, et al.
Veröffentlicht: (2024)
MATE: Solving Contextual Markov Decision Processes with Memory of Accumulated Transition Embeddings
von: Hwang, Himchan, et al.
Veröffentlicht: (2026)
von: Hwang, Himchan, et al.
Veröffentlicht: (2026)
Linear Mixture Distributionally Robust Markov Decision Processes
von: Liu, Zhishuai, et al.
Veröffentlicht: (2025)
von: Liu, Zhishuai, et al.
Veröffentlicht: (2025)
Dual Formulation for Non-Rectangular Lp Robust Markov Decision Processes
von: Kumar, Navdeep, et al.
Veröffentlicht: (2025)
von: Kumar, Navdeep, et al.
Veröffentlicht: (2025)
Sound Statistical Model Checking for Probabilities and Expected Rewards (extended version)
von: Budde, Carlos E., et al.
Veröffentlicht: (2024)
von: Budde, Carlos E., et al.
Veröffentlicht: (2024)
Statistical Model Checking Beyond Means: Quantiles, CVaR, and the DKW Inequality (extended version)
von: Budde, Carlos E., et al.
Veröffentlicht: (2025)
von: Budde, Carlos E., et al.
Veröffentlicht: (2025)
Policy Regularized Distributionally Robust Markov Decision Processes with Linear Function Approximation
von: Gu, Jingwen, et al.
Veröffentlicht: (2025)
von: Gu, Jingwen, et al.
Veröffentlicht: (2025)
Optimal Decision Tree Policies for Markov Decision Processes
von: Vos, Daniël, et al.
Veröffentlicht: (2023)
von: Vos, Daniël, et al.
Veröffentlicht: (2023)
Markov Decision Processes under External Temporal Processes
von: Ayyagari, Ranga Shaarad, et al.
Veröffentlicht: (2023)
von: Ayyagari, Ranga Shaarad, et al.
Veröffentlicht: (2023)
Provably Efficient Reward Transfer in Reinforcement Learning with Discrete Markov Decision Processes
von: Vora, Kevin, et al.
Veröffentlicht: (2025)
von: Vora, Kevin, et al.
Veröffentlicht: (2025)
Robust Lagrangian and Adversarial Policy Gradient for Robust Constrained Markov Decision Processes
von: Bossens, David M.
Veröffentlicht: (2023)
von: Bossens, David M.
Veröffentlicht: (2023)
Act as You Learn: Adaptive Decision-Making in Non-Stationary Markov Decision Processes
von: Luo, Baiting, et al.
Veröffentlicht: (2024)
von: Luo, Baiting, et al.
Veröffentlicht: (2024)
Hierarchical Average-Reward Linearly-solvable Markov Decision Processes
von: Infante, Guillermo, et al.
Veröffentlicht: (2024)
von: Infante, Guillermo, et al.
Veröffentlicht: (2024)
SPOT: Scalable Policy Optimization with Trees for Markov Decision Processes
von: Xiong, Xuyuan, et al.
Veröffentlicht: (2025)
von: Xiong, Xuyuan, et al.
Veröffentlicht: (2025)
Optimistic Regret Bounds for Online Learning in Adversarial Markov Decision Processes
von: Moon, Sang Bin, et al.
Veröffentlicht: (2024)
von: Moon, Sang Bin, et al.
Veröffentlicht: (2024)
Diffusion-Augmented Markov Decision Processes for Maximum Entropy Reinforcement Learning
von: Sanokowski, Sebastian, et al.
Veröffentlicht: (2025)
von: Sanokowski, Sebastian, et al.
Veröffentlicht: (2025)
Homomorphic Mappings for Value-Preserving State Aggregation in Markov Decision Processes
von: Zhao, Shuo, et al.
Veröffentlicht: (2025)
von: Zhao, Shuo, et al.
Veröffentlicht: (2025)
A Unified Theory of Compositionality, Modularity, and Interpretability in Markov Decision Processes
von: Ringstrom, Thomas J., et al.
Veröffentlicht: (2025)
von: Ringstrom, Thomas J., et al.
Veröffentlicht: (2025)
Robust $Q$-learning Algorithm for Markov Decision Processes under Wasserstein Uncertainty
von: Neufeld, Ariel, et al.
Veröffentlicht: (2022)
von: Neufeld, Ariel, et al.
Veröffentlicht: (2022)
OCMDP: Observation-Constrained Markov Decision Process
von: Wang, Taiyi, et al.
Veröffentlicht: (2024)
von: Wang, Taiyi, et al.
Veröffentlicht: (2024)
A Generalized Bisimulation Metric of State Similarity between Markov Decision Processes: From Theoretical Propositions to Applications
von: Tao, Zhenyu, et al.
Veröffentlicht: (2025)
von: Tao, Zhenyu, et al.
Veröffentlicht: (2025)
REValueD: Regularised Ensemble Value-Decomposition for Factorisable Markov Decision Processes
von: Ireland, David, et al.
Veröffentlicht: (2024)
von: Ireland, David, et al.
Veröffentlicht: (2024)
Globally Optimal Hierarchical Reinforcement Learning for Linearly-Solvable Markov Decision Processes
von: Infante, Guillermo, et al.
Veröffentlicht: (2021)
von: Infante, Guillermo, et al.
Veröffentlicht: (2021)
Conformal Off-Policy Evaluation in Markov Decision Processes
von: Foffano, Daniele, et al.
Veröffentlicht: (2023)
von: Foffano, Daniele, et al.
Veröffentlicht: (2023)
A Cantor-Kantorovich Metric Between Markov Decision Processes with Application to Transfer Learning
von: Banse, Adrien, et al.
Veröffentlicht: (2024)
von: Banse, Adrien, et al.
Veröffentlicht: (2024)
Policy Gradient Algorithms with Monte Carlo Tree Learning for Non-Markov Decision Processes
von: Morimura, Tetsuro, et al.
Veröffentlicht: (2022)
von: Morimura, Tetsuro, et al.
Veröffentlicht: (2022)
Reinforcement Learning in Switching Non-Stationary Markov Decision Processes: Algorithms and Convergence Analysis
von: Amiri, Mohsen, et al.
Veröffentlicht: (2025)
von: Amiri, Mohsen, et al.
Veröffentlicht: (2025)
Rethinking Large Language Model Distillation: A Constrained Markov Decision Process Perspective
von: Zimmer, Matthieu, et al.
Veröffentlicht: (2025)
von: Zimmer, Matthieu, et al.
Veröffentlicht: (2025)
Almost Sure Convergence of Differential Temporal Difference Learning for Average Reward Markov Decision Processes
von: Blaser, Ethan, et al.
Veröffentlicht: (2026)
von: Blaser, Ethan, et al.
Veröffentlicht: (2026)
Regret Analysis of Policy Gradient Algorithm for Infinite Horizon Average Reward Markov Decision Processes
von: Bai, Qinbo, et al.
Veröffentlicht: (2023)
von: Bai, Qinbo, et al.
Veröffentlicht: (2023)
Markov Process-Based Graph Convolutional Networks for Entity Classification in Knowledge Graphs
von: Mäkelburg, Johannes, et al.
Veröffentlicht: (2024)
von: Mäkelburg, Johannes, et al.
Veröffentlicht: (2024)
Improved Sample Complexity Analysis of Natural Policy Gradient Algorithm with General Parameterization for Infinite Horizon Discounted Reward Markov Decision Processes
von: Mondal, Washim Uddin, et al.
Veröffentlicht: (2023)
von: Mondal, Washim Uddin, et al.
Veröffentlicht: (2023)
Observation Adaptation via Annealed Importance Resampling for Partially Observable Markov Decision Processes
von: Zhang, Yunuo, et al.
Veröffentlicht: (2025)
von: Zhang, Yunuo, et al.
Veröffentlicht: (2025)
Burning RED: Unlocking Subtask-Driven Reinforcement Learning and Risk-Awareness in Average-Reward Markov Decision Processes
von: Rojas, Juan Sebastian, et al.
Veröffentlicht: (2024)
von: Rojas, Juan Sebastian, et al.
Veröffentlicht: (2024)
Pessimism in the Face of Confounders: Provably Efficient Offline Reinforcement Learning in Partially Observable Markov Decision Processes
von: Lu, Miao, et al.
Veröffentlicht: (2022)
von: Lu, Miao, et al.
Veröffentlicht: (2022)
Learning Algorithms for Verification of Markov Decision Processes
von: Brázdil, Tomáš, et al.
Veröffentlicht: (2024)
von: Brázdil, Tomáš, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
What Are the Odds? Improving the foundations of Statistical Model Checking
von: Meggendorfer, Tobias, et al.
Veröffentlicht: (2024) -
Playing Games with your PET: Extending the Partial Exploration Tool to Stochastic Games
von: Meggendorfer, Tobias, et al.
Veröffentlicht: (2024) -
1-2-3-Go! Policy Synthesis for Parameterized Markov Decision Processes via Decision-Tree Learning and Generalization
von: Azeem, Muqsit, et al.
Veröffentlicht: (2024) -
Efficient and Sharp Off-Policy Evaluation in Robust Markov Decision Processes
von: Bennett, Andrew, et al.
Veröffentlicht: (2024) -
Policy Gradient for Robust Markov Decision Processes
von: Wang, Qiuhao, et al.
Veröffentlicht: (2024)