Optimal Decision Tree Policies for Markov Decision Processes
Fuente:
arXiv
Saved in:
| Main Authors: | Vos, Daniël, Verwer, Sicco |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Optimizing Interpretable Decision Tree Policies for Reinforcement Learning
by: Vos, Daniël, et al.
Published: (2024)
by: Vos, Daniël, et al.
Published: (2024)
Optimal or Greedy Decision Trees? Revisiting their Objectives, Tuning, and Performance
by: van der Linden, Jacobus G. M., et al.
Published: (2024)
by: van der Linden, Jacobus G. M., et al.
Published: (2024)
SPOT: Scalable Policy Optimization with Trees for Markov Decision Processes
by: Xiong, Xuyuan, et al.
Published: (2025)
by: Xiong, Xuyuan, et al.
Published: (2025)
Policy Gradient for Robust Markov Decision Processes
by: Wang, Qiuhao, et al.
Published: (2024)
by: Wang, Qiuhao, et al.
Published: (2024)
Policy Gradient Algorithms with Monte Carlo Tree Learning for Non-Markov Decision Processes
by: Morimura, Tetsuro, et al.
Published: (2022)
by: Morimura, Tetsuro, et al.
Published: (2022)
Efficient and Sharp Off-Policy Evaluation in Robust Markov Decision Processes
by: Bennett, Andrew, et al.
Published: (2024)
by: Bennett, Andrew, et al.
Published: (2024)
Conformal Off-Policy Evaluation in Markov Decision Processes
by: Foffano, Daniele, et al.
Published: (2023)
by: Foffano, Daniele, et al.
Published: (2023)
Globally Optimal Hierarchical Reinforcement Learning for Linearly-Solvable Markov Decision Processes
by: Infante, Guillermo, et al.
Published: (2021)
by: Infante, Guillermo, et al.
Published: (2021)
Policy Regularized Distributionally Robust Markov Decision Processes with Linear Function Approximation
by: Gu, Jingwen, et al.
Published: (2025)
by: Gu, Jingwen, et al.
Published: (2025)
Markov Decision Processes under External Temporal Processes
by: Ayyagari, Ranga Shaarad, et al.
Published: (2023)
by: Ayyagari, Ranga Shaarad, et al.
Published: (2023)
Act as You Learn: Adaptive Decision-Making in Non-Stationary Markov Decision Processes
by: Luo, Baiting, et al.
Published: (2024)
by: Luo, Baiting, et al.
Published: (2024)
Regret Analysis of Policy Gradient Algorithm for Infinite Horizon Average Reward Markov Decision Processes
by: Bai, Qinbo, et al.
Published: (2023)
by: Bai, Qinbo, et al.
Published: (2023)
An Offline Risk-aware Policy Selection Method for Bayesian Markov Decision Processes
by: Angelotti, Giorgio, et al.
Published: (2021)
by: Angelotti, Giorgio, et al.
Published: (2021)
Solving Robust Markov Decision Processes: Generic, Reliable, Efficient
by: Meggendorfer, Tobias, et al.
Published: (2024)
by: Meggendorfer, Tobias, et al.
Published: (2024)
Hierarchical Average-Reward Linearly-solvable Markov Decision Processes
by: Infante, Guillermo, et al.
Published: (2024)
by: Infante, Guillermo, et al.
Published: (2024)
Linear Mixture Distributionally Robust Markov Decision Processes
by: Liu, Zhishuai, et al.
Published: (2025)
by: Liu, Zhishuai, et al.
Published: (2025)
OCMDP: Observation-Constrained Markov Decision Process
by: Wang, Taiyi, et al.
Published: (2024)
by: Wang, Taiyi, et al.
Published: (2024)
On the Convergence of Modified Policy Iteration in Risk Sensitive Exponential Cost Markov Decision Processes
by: Murthy, Yashaswini, et al.
Published: (2023)
by: Murthy, Yashaswini, et al.
Published: (2023)
Robust Lagrangian and Adversarial Policy Gradient for Robust Constrained Markov Decision Processes
by: Bossens, David M.
Published: (2023)
by: Bossens, David M.
Published: (2023)
Learning Optimal and Sample-Efficient Decision Policies with Guarantees
by: Shao, Daqian
Published: (2026)
by: Shao, Daqian
Published: (2026)
Diffusion-Augmented Markov Decision Processes for Maximum Entropy Reinforcement Learning
by: Sanokowski, Sebastian, et al.
Published: (2025)
by: Sanokowski, Sebastian, et al.
Published: (2025)
Homomorphic Mappings for Value-Preserving State Aggregation in Markov Decision Processes
by: Zhao, Shuo, et al.
Published: (2025)
by: Zhao, Shuo, et al.
Published: (2025)
A Unified Theory of Compositionality, Modularity, and Interpretability in Markov Decision Processes
by: Ringstrom, Thomas J., et al.
Published: (2025)
by: Ringstrom, Thomas J., et al.
Published: (2025)
Optimistic Regret Bounds for Online Learning in Adversarial Markov Decision Processes
by: Moon, Sang Bin, et al.
Published: (2024)
by: Moon, Sang Bin, et al.
Published: (2024)
Dual Formulation for Non-Rectangular Lp Robust Markov Decision Processes
by: Kumar, Navdeep, et al.
Published: (2025)
by: Kumar, Navdeep, et al.
Published: (2025)
1-2-3-Go! Policy Synthesis for Parameterized Markov Decision Processes via Decision-Tree Learning and Generalization
by: Azeem, Muqsit, et al.
Published: (2024)
by: Azeem, Muqsit, et al.
Published: (2024)
MATE: Solving Contextual Markov Decision Processes with Memory of Accumulated Transition Embeddings
by: Hwang, Himchan, et al.
Published: (2026)
by: Hwang, Himchan, et al.
Published: (2026)
REValueD: Regularised Ensemble Value-Decomposition for Factorisable Markov Decision Processes
by: Ireland, David, et al.
Published: (2024)
by: Ireland, David, et al.
Published: (2024)
Provably Efficient Reward Transfer in Reinforcement Learning with Discrete Markov Decision Processes
by: Vora, Kevin, et al.
Published: (2025)
by: Vora, Kevin, et al.
Published: (2025)
A Cantor-Kantorovich Metric Between Markov Decision Processes with Application to Transfer Learning
by: Banse, Adrien, et al.
Published: (2024)
by: Banse, Adrien, et al.
Published: (2024)
Reinforcement Learning in Switching Non-Stationary Markov Decision Processes: Algorithms and Convergence Analysis
by: Amiri, Mohsen, et al.
Published: (2025)
by: Amiri, Mohsen, et al.
Published: (2025)
Rethinking Large Language Model Distillation: A Constrained Markov Decision Process Perspective
by: Zimmer, Matthieu, et al.
Published: (2025)
by: Zimmer, Matthieu, et al.
Published: (2025)
Improved Sample Complexity Analysis of Natural Policy Gradient Algorithm with General Parameterization for Infinite Horizon Discounted Reward Markov Decision Processes
by: Mondal, Washim Uddin, et al.
Published: (2023)
by: Mondal, Washim Uddin, et al.
Published: (2023)
PDFA Distillation via String Probability Queries
by: Baumgartner, Robert, et al.
Published: (2024)
by: Baumgartner, Robert, et al.
Published: (2024)
PAC learning PDFA from data streams
by: Baumgartner, Robert, et al.
Published: (2026)
by: Baumgartner, Robert, et al.
Published: (2026)
Almost Sure Convergence of Differential Temporal Difference Learning for Average Reward Markov Decision Processes
by: Blaser, Ethan, et al.
Published: (2026)
by: Blaser, Ethan, et al.
Published: (2026)
Anytime Optimal Decision Tree Learning with Continuous Features
by: Kiossou, Harold, et al.
Published: (2026)
by: Kiossou, Harold, et al.
Published: (2026)
Clustered Policy Decision Ranking
by: Levin, Mark, et al.
Published: (2023)
by: Levin, Mark, et al.
Published: (2023)
Decision Flow Policy Optimization
by: Hu, Jifeng, et al.
Published: (2025)
by: Hu, Jifeng, et al.
Published: (2025)
FlexFringe: Modeling Software Behavior by Learning Probabilistic Automata
by: Verwer, Sicco, et al.
Published: (2022)
by: Verwer, Sicco, et al.
Published: (2022)
Similar Items
-
Optimizing Interpretable Decision Tree Policies for Reinforcement Learning
by: Vos, Daniël, et al.
Published: (2024) -
Optimal or Greedy Decision Trees? Revisiting their Objectives, Tuning, and Performance
by: van der Linden, Jacobus G. M., et al.
Published: (2024) -
SPOT: Scalable Policy Optimization with Trees for Markov Decision Processes
by: Xiong, Xuyuan, et al.
Published: (2025) -
Policy Gradient for Robust Markov Decision Processes
by: Wang, Qiuhao, et al.
Published: (2024) -
Policy Gradient Algorithms with Monte Carlo Tree Learning for Non-Markov Decision Processes
by: Morimura, Tetsuro, et al.
Published: (2022)