Explainable Representation of Finite-Memory Policies for POMDPs using Decision Trees
Fuente:
arXiv
Saved in:
| Main Authors: | Azeem, Muqsit, Chakraborty, Debraj, Kanav, Sudeep, Kretinsky, Jan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
1-2-3-Go! Policy Synthesis for Parameterized Markov Decision Processes via Decision-Tree Learning and Generalization
by: Azeem, Muqsit, et al.
Published: (2024)
by: Azeem, Muqsit, et al.
Published: (2024)
Explaining Control Policies through Predicate Decision Diagrams
by: Chakraborty, Debraj, et al.
Published: (2025)
by: Chakraborty, Debraj, et al.
Published: (2025)
Monitizer: Automating Design and Evaluation of Neural Network Monitors
by: Azeem, Muqsit, et al.
Published: (2024)
by: Azeem, Muqsit, et al.
Published: (2024)
Learning Explainable and Better Performing Representations of POMDP Strategies
by: Bork, Alexander, et al.
Published: (2024)
by: Bork, Alexander, et al.
Published: (2024)
Decision Transformer as a Foundation Model for Partially Observable Continuous Control
by: Zhang, Xiangyuan, et al.
Published: (2024)
by: Zhang, Xiangyuan, et al.
Published: (2024)
Embedding Morphology into Transformers for Cross-Robot Policy Learning
by: Suzuki, Kei, et al.
Published: (2026)
by: Suzuki, Kei, et al.
Published: (2026)
Align and Filter: Improving Performance in Asynchronous On-Policy RL
by: Honari, Homayoun, et al.
Published: (2026)
by: Honari, Homayoun, et al.
Published: (2026)
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving
by: Zhang, Zhihao, et al.
Published: (2025)
by: Zhang, Zhihao, et al.
Published: (2025)
Predictive Red Teaming: Breaking Policies Without Breaking Robots
by: Majumdar, Anirudha, et al.
Published: (2025)
by: Majumdar, Anirudha, et al.
Published: (2025)
Sound Value Iteration for Simple Stochastic Games
by: Azeem, Muqsit, et al.
Published: (2024)
by: Azeem, Muqsit, et al.
Published: (2024)
Safety Optimized Reinforcement Learning via Multi-Objective Policy Optimization
by: Honari, Homayoun, et al.
Published: (2024)
by: Honari, Homayoun, et al.
Published: (2024)
Efficient Policy Optimization in Robust Constrained MDPs with Iteration Complexity Guarantees
by: Ganguly, Sourav, et al.
Published: (2025)
by: Ganguly, Sourav, et al.
Published: (2025)
Optimal Control-Based Baseline for Guided Exploration in Policy Gradient Methods
by: Lyu, Xubo, et al.
Published: (2020)
by: Lyu, Xubo, et al.
Published: (2020)
A Conflicts-free, Speed-lossless KAN-based Reinforcement Learning Decision System for Interactive Driving in Roundabouts
by: Lin, Zhihao, et al.
Published: (2024)
by: Lin, Zhihao, et al.
Published: (2024)
ManyQuadrupeds: Learning a Single Locomotion Policy for Diverse Quadruped Robots
by: Shafiee, Milad, et al.
Published: (2023)
by: Shafiee, Milad, et al.
Published: (2023)
Scaling Learning based Policy Optimization for Temporal Logic Tasks by Controller Network Dropout
by: Hashemi, Navid, et al.
Published: (2024)
by: Hashemi, Navid, et al.
Published: (2024)
Deep Policy Gradient Methods Without Batch Updates, Target Networks, or Replay Buffers
by: Vasan, Gautham, et al.
Published: (2024)
by: Vasan, Gautham, et al.
Published: (2024)
GenSafe: A Generalizable Safety Enhancer for Safe Reinforcement Learning Algorithms Based on Reduced Order Markov Decision Process Model
by: Zhou, Zhehua, et al.
Published: (2024)
by: Zhou, Zhehua, et al.
Published: (2024)
Frugal Actor-Critic: Sample Efficient Off-Policy Deep Reinforcement Learning Using Unique Experiences
by: Singh, Nikhil Kumar, et al.
Published: (2024)
by: Singh, Nikhil Kumar, et al.
Published: (2024)
GUIDEd Agents: Enhancing Navigation Policies through Task-Specific Uncertainty Abstraction in Localization-Limited Environments
by: Puthumanaillam, Gokul, et al.
Published: (2024)
by: Puthumanaillam, Gokul, et al.
Published: (2024)
Adaptive Smooth Tchebycheff Attention for Multi-Objective Policy Optimization
by: Murillo-Gonzalez, Alejandro, et al.
Published: (2026)
by: Murillo-Gonzalez, Alejandro, et al.
Published: (2026)
Multi-CALF: A Policy Combination Approach with Statistical Guarantees
by: Malaniya, Georgiy, et al.
Published: (2025)
by: Malaniya, Georgiy, et al.
Published: (2025)
Model Tensor Planning
by: Le, An T., et al.
Published: (2025)
by: Le, An T., et al.
Published: (2025)
Robust Recovery Controller for a Quadrupedal Robot using Deep Reinforcement Learning
by: Lee, Joonho, et al.
Published: (2019)
by: Lee, Joonho, et al.
Published: (2019)
Statistically Assuring Safety of Control Systems using Ensembles of Safety Filters and Conformal Prediction
by: Tabbara, Ihab, et al.
Published: (2025)
by: Tabbara, Ihab, et al.
Published: (2025)
Global Tensor Motion Planning
by: Le, An T., et al.
Published: (2024)
by: Le, An T., et al.
Published: (2024)
Learning from Imperfect Demonstrations via Temporal Behavior Tree-Guided Trajectory Repair
by: Puranic, Aniruddh G., et al.
Published: (2026)
by: Puranic, Aniruddh G., et al.
Published: (2026)
From Automation to Autonomy in Smart Manufacturing: A Bayesian Optimization Framework for Modeling Multi-Objective Experimentation and Sequential Decision Making
by: Asru, Avijit Saha, et al.
Published: (2025)
by: Asru, Avijit Saha, et al.
Published: (2025)
Together We Rise: Optimizing Real-Time Multi-Robot Task Allocation using Coordinated Heterogeneous Plays
by: Pal, Aritra, et al.
Published: (2025)
by: Pal, Aritra, et al.
Published: (2025)
FLD: Fourier Latent Dynamics for Structured Motion Representation and Learning
by: Li, Chenhao, et al.
Published: (2024)
by: Li, Chenhao, et al.
Published: (2024)
CHyLL: Learning Continuous Neural Representations of Hybrid Systems
by: Teng, Sangli, et al.
Published: (2025)
by: Teng, Sangli, et al.
Published: (2025)
BIDA: A Bi-level Interaction Decision-making Algorithm for Autonomous Vehicles in Dynamic Traffic Scenarios
by: Yu, Liyang, et al.
Published: (2025)
by: Yu, Liyang, et al.
Published: (2025)
RoboKoop: Efficient Control Conditioned Representations from Visual Input in Robotics using Koopman Operator
by: Kumawat, Hemant, et al.
Published: (2024)
by: Kumawat, Hemant, et al.
Published: (2024)
Towards Safe Autonomous Driving Policies using a Neuro-Symbolic Deep Reinforcement Learning Approach
by: Sharifi, Iman, et al.
Published: (2023)
by: Sharifi, Iman, et al.
Published: (2023)
MAGICS: Adversarial RL with Minimax Actors Guided by Implicit Critic Stackelberg for Convergent Neural Synthesis of Robot Safety
by: Wang, Justin, et al.
Published: (2024)
by: Wang, Justin, et al.
Published: (2024)
A Safe Reinforcement Learning driven Weights-varying Model Predictive Control for Autonomous Vehicle Motion Control
by: Zarrouki, Baha, et al.
Published: (2024)
by: Zarrouki, Baha, et al.
Published: (2024)
Deep Reinforcement Learning for Advanced Longitudinal Control and Collision Avoidance in High-Risk Driving Scenarios
by: Chen, Dianwei, et al.
Published: (2024)
by: Chen, Dianwei, et al.
Published: (2024)
One Filter to Deploy Them All: Robust Safety for Quadrupedal Navigation in Unknown Environments
by: Lin, Albert, et al.
Published: (2024)
by: Lin, Albert, et al.
Published: (2024)
Whole-body End-Effector Pose Tracking
by: Portela, Tifanny, et al.
Published: (2024)
by: Portela, Tifanny, et al.
Published: (2024)
Learning Force Control for Legged Manipulation
by: Portela, Tifanny, et al.
Published: (2024)
by: Portela, Tifanny, et al.
Published: (2024)
Similar Items
-
1-2-3-Go! Policy Synthesis for Parameterized Markov Decision Processes via Decision-Tree Learning and Generalization
by: Azeem, Muqsit, et al.
Published: (2024) -
Explaining Control Policies through Predicate Decision Diagrams
by: Chakraborty, Debraj, et al.
Published: (2025) -
Monitizer: Automating Design and Evaluation of Neural Network Monitors
by: Azeem, Muqsit, et al.
Published: (2024) -
Learning Explainable and Better Performing Representations of POMDP Strategies
by: Bork, Alexander, et al.
Published: (2024) -
Decision Transformer as a Foundation Model for Partially Observable Continuous Control
by: Zhang, Xiangyuan, et al.
Published: (2024)