Strategically Linked Decisions in Long-Term Planning and Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Hüyük, Alihan, Doshi-Velez, Finale |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Quantifying Potential Observation Missingness in Inverse Reinforcement Learning
by: Benac, Leo, et al.
Published: (2026)
by: Benac, Leo, et al.
Published: (2026)
Transparent Trade-offs between Properties of Explanations
by: Tadesse, Hiwot Belay, et al.
Published: (2024)
by: Tadesse, Hiwot Belay, et al.
Published: (2024)
Disentangling Recognition and Decision Regrets in Image-Based Reinforcement Learning
by: Hüyük, Alihan, et al.
Published: (2024)
by: Hüyük, Alihan, et al.
Published: (2024)
Inverse Reinforcement Learning with Multiple Planning Horizons
by: Yao, Jiayu, et al.
Published: (2024)
by: Yao, Jiayu, et al.
Published: (2024)
Decision-Focused Model-based Reinforcement Learning for Reward Transfer
by: Sharma, Abhishek, et al.
Published: (2023)
by: Sharma, Abhishek, et al.
Published: (2023)
On the Effective Horizon of Inverse Reinforcement Learning
by: Xu, Yiqing, et al.
Published: (2023)
by: Xu, Yiqing, et al.
Published: (2023)
Semi-parametric Expert Bayesian Network Learning with Gaussian Processes and Horseshoe Priors
by: Weng, Yidou, et al.
Published: (2024)
by: Weng, Yidou, et al.
Published: (2024)
Decision-Point Guided Safe Policy Improvement
by: Sharma, Abhishek, et al.
Published: (2024)
by: Sharma, Abhishek, et al.
Published: (2024)
Adaptive Experiment Design with Synthetic Controls
by: Hüyük, Alihan, et al.
Published: (2024)
by: Hüyük, Alihan, et al.
Published: (2024)
Guarantee Regions for Local Explanations
by: Havasi, Marton, et al.
Published: (2024)
by: Havasi, Marton, et al.
Published: (2024)
Towards Model-Agnostic Posterior Approximation for Fast and Accurate Variational Autoencoders
by: Yacoby, Yaniv, et al.
Published: (2024)
by: Yacoby, Yaniv, et al.
Published: (2024)
Diverse Concept Proposals for Concept Bottleneck Models
by: Brown, Katrina, et al.
Published: (2024)
by: Brown, Katrina, et al.
Published: (2024)
Towards Integrating Personal Knowledge into Test-Time Predictions
by: Lage, Isaac, et al.
Published: (2024)
by: Lage, Isaac, et al.
Published: (2024)
Reinforcement Learning Interventions on Boundedly Rational Human Agents in Frictionful Tasks
by: Nofshin, Eura, et al.
Published: (2024)
by: Nofshin, Eura, et al.
Published: (2024)
Bayesian Inverse Transition Learning: Learning Dynamics From Near-Optimal Trajectories
by: Benac, Leo, et al.
Published: (2024)
by: Benac, Leo, et al.
Published: (2024)
Query-Dependent Prompt Evaluation and Optimization with Offline Inverse RL
by: Sun, Hao, et al.
Published: (2023)
by: Sun, Hao, et al.
Published: (2023)
Connecting Federated ADMM to Bayes
by: Swaroop, Siddharth, et al.
Published: (2025)
by: Swaroop, Siddharth, et al.
Published: (2025)
Towards Regulatory-Confirmed Adaptive Clinical Trials: Machine Learning Opportunities and Solutions
by: Klein, Omer Noy, et al.
Published: (2025)
by: Klein, Omer Noy, et al.
Published: (2025)
Defining Expertise: Applications to Treatment Effect Estimation
by: Hüyük, Alihan, et al.
Published: (2024)
by: Hüyük, Alihan, et al.
Published: (2024)
A Benchmark for Multi-Party Negotiation Games from Real Negotiation Data
by: Benac, Leo, et al.
Published: (2026)
by: Benac, Leo, et al.
Published: (2026)
A Sim2Real Approach for Identifying Task-Relevant Properties in Interpretable Machine Learning
by: Nofshin, Eura, et al.
Published: (2024)
by: Nofshin, Eura, et al.
Published: (2024)
What Makes a Good Explanation?: A Harmonized View of Properties of Explanations
by: Chen, Zixi, et al.
Published: (2022)
by: Chen, Zixi, et al.
Published: (2022)
Federated ADMM from Bayesian Duality
by: Möllenhoff, Thomas, et al.
Published: (2025)
by: Möllenhoff, Thomas, et al.
Published: (2025)
Pruning the Path to Optimal Care: Identifying Systematically Suboptimal Medical Decision-Making with Inverse Reinforcement Learning
by: Bovenzi, Inko, et al.
Published: (2024)
by: Bovenzi, Inko, et al.
Published: (2024)
Reasoning Elicitation in Language Models via Counterfactual Feedback
by: Hüyük, Alihan, et al.
Published: (2024)
by: Hüyük, Alihan, et al.
Published: (2024)
When is Off-Policy Evaluation (Reward Modeling) Useful in Contextual Bandits? A Data-Centric Perspective
by: Sun, Hao, et al.
Published: (2023)
by: Sun, Hao, et al.
Published: (2023)
Understanding the Relationship between Prompts and Response Uncertainty in Large Language Models
by: Zhang, Ze Yu, et al.
Published: (2024)
by: Zhang, Ze Yu, et al.
Published: (2024)
Monitoring Fidelity of Online Reinforcement Learning Algorithms in Clinical Trials
by: Trella, Anna L., et al.
Published: (2024)
by: Trella, Anna L., et al.
Published: (2024)
Reinforcement Learning with Pairwise Preferences in Long-Term Decision Problems
by: Carr, Jonathan Colaço, et al.
Published: (2026)
by: Carr, Jonathan Colaço, et al.
Published: (2026)
Compositional Causal Reasoning Evaluation in Language Models
by: Maasch, Jacqueline R. M. A., et al.
Published: (2025)
by: Maasch, Jacqueline R. M. A., et al.
Published: (2025)
Non-Stationary Latent Auto-Regressive Bandits
by: Trella, Anna L., et al.
Published: (2024)
by: Trella, Anna L., et al.
Published: (2024)
Modeling Behavioral Preferences of Cyber Adversaries Using Inverse Reinforcement Learning
by: Shinde, Aditya, et al.
Published: (2025)
by: Shinde, Aditya, et al.
Published: (2025)
RLLTE: Long-Term Evolution Project of Reinforcement Learning
by: Yuan, Mingqi, et al.
Published: (2023)
by: Yuan, Mingqi, et al.
Published: (2023)
Handling Long-Term Safety and Uncertainty in Safe Reinforcement Learning
by: Günster, Jonas, et al.
Published: (2024)
by: Günster, Jonas, et al.
Published: (2024)
Open-World Reinforcement Learning over Long Short-Term Imagination
by: Li, Jiajian, et al.
Published: (2024)
by: Li, Jiajian, et al.
Published: (2024)
Semiparametric Double Reinforcement Learning with Applications to Long-Term Causal Inference
by: van der Laan, Lars, et al.
Published: (2025)
by: van der Laan, Lars, et al.
Published: (2025)
Explaining Strategic Decisions in Multi-Agent Reinforcement Learning for Aerial Combat Tactics
by: Selmonaj, Ardian, et al.
Published: (2025)
by: Selmonaj, Ardian, et al.
Published: (2025)
Policy Learning for Balancing Short-Term and Long-Term Rewards
by: Wu, Peng, et al.
Published: (2024)
by: Wu, Peng, et al.
Published: (2024)
Shaping AI's Impact on Billions of Lives
by: Cuéllar, Mariano-Florentino, et al.
Published: (2024)
by: Cuéllar, Mariano-Florentino, et al.
Published: (2024)
Planning Transformer: Long-Horizon Offline Reinforcement Learning with Planning Tokens
by: Clinton, Joseph, et al.
Published: (2024)
by: Clinton, Joseph, et al.
Published: (2024)
Similar Items
-
Quantifying Potential Observation Missingness in Inverse Reinforcement Learning
by: Benac, Leo, et al.
Published: (2026) -
Transparent Trade-offs between Properties of Explanations
by: Tadesse, Hiwot Belay, et al.
Published: (2024) -
Disentangling Recognition and Decision Regrets in Image-Based Reinforcement Learning
by: Hüyük, Alihan, et al.
Published: (2024) -
Inverse Reinforcement Learning with Multiple Planning Horizons
by: Yao, Jiayu, et al.
Published: (2024) -
Decision-Focused Model-based Reinforcement Learning for Reward Transfer
by: Sharma, Abhishek, et al.
Published: (2023)