Inverse Reinforcement Learning with Multiple Planning Horizons
Fuente:
arXiv
Saved in:
| Main Authors: | Yao, Jiayu, Pan, Weiwei, Doshi-Velez, Finale, Engelhardt, Barbara E |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On the Effective Horizon of Inverse Reinforcement Learning
by: Xu, Yiqing, et al.
Published: (2023)
by: Xu, Yiqing, et al.
Published: (2023)
Strategically Linked Decisions in Long-Term Planning and Reinforcement Learning
by: Hüyük, Alihan, et al.
Published: (2025)
by: Hüyük, Alihan, et al.
Published: (2025)
Quantifying Potential Observation Missingness in Inverse Reinforcement Learning
by: Benac, Leo, et al.
Published: (2026)
by: Benac, Leo, et al.
Published: (2026)
Towards Model-Agnostic Posterior Approximation for Fast and Accurate Variational Autoencoders
by: Yacoby, Yaniv, et al.
Published: (2024)
by: Yacoby, Yaniv, et al.
Published: (2024)
Reinforcement Learning Interventions on Boundedly Rational Human Agents in Frictionful Tasks
by: Nofshin, Eura, et al.
Published: (2024)
by: Nofshin, Eura, et al.
Published: (2024)
Semi-parametric Expert Bayesian Network Learning with Gaussian Processes and Horseshoe Priors
by: Weng, Yidou, et al.
Published: (2024)
by: Weng, Yidou, et al.
Published: (2024)
Bayesian Inverse Transition Learning: Learning Dynamics From Near-Optimal Trajectories
by: Benac, Leo, et al.
Published: (2024)
by: Benac, Leo, et al.
Published: (2024)
A Sim2Real Approach for Identifying Task-Relevant Properties in Interpretable Machine Learning
by: Nofshin, Eura, et al.
Published: (2024)
by: Nofshin, Eura, et al.
Published: (2024)
What Makes a Good Explanation?: A Harmonized View of Properties of Explanations
by: Chen, Zixi, et al.
Published: (2022)
by: Chen, Zixi, et al.
Published: (2022)
Kernel Density Bayesian Inverse Reinforcement Learning
by: Mandyam, Aishwarya, et al.
Published: (2023)
by: Mandyam, Aishwarya, et al.
Published: (2023)
Decision-Focused Model-based Reinforcement Learning for Reward Transfer
by: Sharma, Abhishek, et al.
Published: (2023)
by: Sharma, Abhishek, et al.
Published: (2023)
Transparent Trade-offs between Properties of Explanations
by: Tadesse, Hiwot Belay, et al.
Published: (2024)
by: Tadesse, Hiwot Belay, et al.
Published: (2024)
Guarantee Regions for Local Explanations
by: Havasi, Marton, et al.
Published: (2024)
by: Havasi, Marton, et al.
Published: (2024)
Diverse Concept Proposals for Concept Bottleneck Models
by: Brown, Katrina, et al.
Published: (2024)
by: Brown, Katrina, et al.
Published: (2024)
Towards Integrating Personal Knowledge into Test-Time Predictions
by: Lage, Isaac, et al.
Published: (2024)
by: Lage, Isaac, et al.
Published: (2024)
Connecting Federated ADMM to Bayes
by: Swaroop, Siddharth, et al.
Published: (2025)
by: Swaroop, Siddharth, et al.
Published: (2025)
Decision-Point Guided Safe Policy Improvement
by: Sharma, Abhishek, et al.
Published: (2024)
by: Sharma, Abhishek, et al.
Published: (2024)
Pruning the Path to Optimal Care: Identifying Systematically Suboptimal Medical Decision-Making with Inverse Reinforcement Learning
by: Bovenzi, Inko, et al.
Published: (2024)
by: Bovenzi, Inko, et al.
Published: (2024)
A Benchmark for Multi-Party Negotiation Games from Real Negotiation Data
by: Benac, Leo, et al.
Published: (2026)
by: Benac, Leo, et al.
Published: (2026)
Federated ADMM from Bayesian Duality
by: Möllenhoff, Thomas, et al.
Published: (2025)
by: Möllenhoff, Thomas, et al.
Published: (2025)
Modeling Behavioral Preferences of Cyber Adversaries Using Inverse Reinforcement Learning
by: Shinde, Aditya, et al.
Published: (2025)
by: Shinde, Aditya, et al.
Published: (2025)
Understanding the Relationship between Prompts and Response Uncertainty in Large Language Models
by: Zhang, Ze Yu, et al.
Published: (2024)
by: Zhang, Ze Yu, et al.
Published: (2024)
Monitoring Fidelity of Online Reinforcement Learning Algorithms in Clinical Trials
by: Trella, Anna L., et al.
Published: (2024)
by: Trella, Anna L., et al.
Published: (2024)
CANDOR: Counterfactual ANnotated DOubly Robust Off-Policy Evaluation
by: Mandyam, Aishwarya, et al.
Published: (2024)
by: Mandyam, Aishwarya, et al.
Published: (2024)
Compositional Q-learning for electrolyte repletion with imbalanced patient sub-populations
by: Mandyam, Aishwarya, et al.
Published: (2021)
by: Mandyam, Aishwarya, et al.
Published: (2021)
Non-Stationary Latent Auto-Regressive Bandits
by: Trella, Anna L., et al.
Published: (2024)
by: Trella, Anna L., et al.
Published: (2024)
Assessing the Distributional Fidelity of Synthetic Chest X-rays using the Embedded Characteristic Score
by: Tam, Edric, et al.
Published: (2025)
by: Tam, Edric, et al.
Published: (2025)
Planning Transformer: Long-Horizon Offline Reinforcement Learning with Planning Tokens
by: Clinton, Joseph, et al.
Published: (2024)
by: Clinton, Joseph, et al.
Published: (2024)
Inverse Design in Distributed Circuits Using Single-Step Reinforcement Learning
by: Li, Jiayu, et al.
Published: (2025)
by: Li, Jiayu, et al.
Published: (2025)
Inverse Reinforcement Learning without Reinforcement Learning
by: Swamy, Gokul, et al.
Published: (2023)
by: Swamy, Gokul, et al.
Published: (2023)
Strict Subgoal Execution: Reliable Long-Horizon Planning in Hierarchical Reinforcement Learning
by: Hwang, Jaebak, et al.
Published: (2025)
by: Hwang, Jaebak, et al.
Published: (2025)
Distributional Inverse Reinforcement Learning
by: Wu, Feiyang, et al.
Published: (2025)
by: Wu, Feiyang, et al.
Published: (2025)
Safe Reinforcement Learning using Finite-Horizon Gradient-based Estimation
by: Dai, Juntao, et al.
Published: (2024)
by: Dai, Juntao, et al.
Published: (2024)
Horizon Generalization in Reinforcement Learning
by: Myers, Vivek, et al.
Published: (2025)
by: Myers, Vivek, et al.
Published: (2025)
On the Importance of Multistability for Horizon Generalization in Reinforcement Learning
by: Bakija, Asad, et al.
Published: (2026)
by: Bakija, Asad, et al.
Published: (2026)
Stochastic Decision Horizons for Constrained Reinforcement Learning
by: Milosevic, Nikola, et al.
Published: (2026)
by: Milosevic, Nikola, et al.
Published: (2026)
Offline Imitation Learning by Controlling the Effective Planning Horizon
by: Ahn, Hee-Jun, et al.
Published: (2024)
by: Ahn, Hee-Jun, et al.
Published: (2024)
Towards Generalized Inverse Reinforcement Learning
by: Dong, Chaosheng, et al.
Published: (2024)
by: Dong, Chaosheng, et al.
Published: (2024)
The Virtues of Pessimism in Inverse Reinforcement Learning
by: Wu, David, et al.
Published: (2024)
by: Wu, David, et al.
Published: (2024)
Horizon Reduction as Information Loss in Offline Reinforcement Learning
by: Nidadala, Uday Kumar, et al.
Published: (2025)
by: Nidadala, Uday Kumar, et al.
Published: (2025)
Similar Items
-
On the Effective Horizon of Inverse Reinforcement Learning
by: Xu, Yiqing, et al.
Published: (2023) -
Strategically Linked Decisions in Long-Term Planning and Reinforcement Learning
by: Hüyük, Alihan, et al.
Published: (2025) -
Quantifying Potential Observation Missingness in Inverse Reinforcement Learning
by: Benac, Leo, et al.
Published: (2026) -
Towards Model-Agnostic Posterior Approximation for Fast and Accurate Variational Autoencoders
by: Yacoby, Yaniv, et al.
Published: (2024) -
Reinforcement Learning Interventions on Boundedly Rational Human Agents in Frictionful Tasks
by: Nofshin, Eura, et al.
Published: (2024)