Decision-Point Guided Safe Policy Improvement
Fuente:
arXiv
Saved in:
| Main Authors: | Sharma, Abhishek, Benac, Leo, Parbhoo, Sonali, Doshi-Velez, Finale |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bayesian Inverse Transition Learning: Learning Dynamics From Near-Optimal Trajectories
by: Benac, Leo, et al.
Published: (2024)
by: Benac, Leo, et al.
Published: (2024)
Decision-Focused Model-based Reinforcement Learning for Reward Transfer
by: Sharma, Abhishek, et al.
Published: (2023)
by: Sharma, Abhishek, et al.
Published: (2023)
Quantifying Potential Observation Missingness in Inverse Reinforcement Learning
by: Benac, Leo, et al.
Published: (2026)
by: Benac, Leo, et al.
Published: (2026)
Guarantee Regions for Local Explanations
by: Havasi, Marton, et al.
Published: (2024)
by: Havasi, Marton, et al.
Published: (2024)
Towards Integrating Personal Knowledge into Test-Time Predictions
by: Lage, Isaac, et al.
Published: (2024)
by: Lage, Isaac, et al.
Published: (2024)
A Benchmark for Multi-Party Negotiation Games from Real Negotiation Data
by: Benac, Leo, et al.
Published: (2026)
by: Benac, Leo, et al.
Published: (2026)
Strategically Linked Decisions in Long-Term Planning and Reinforcement Learning
by: Hüyük, Alihan, et al.
Published: (2025)
by: Hüyük, Alihan, et al.
Published: (2025)
Semi-parametric Expert Bayesian Network Learning with Gaussian Processes and Horseshoe Priors
by: Weng, Yidou, et al.
Published: (2024)
by: Weng, Yidou, et al.
Published: (2024)
Pruning the Path to Optimal Care: Identifying Systematically Suboptimal Medical Decision-Making with Inverse Reinforcement Learning
by: Bovenzi, Inko, et al.
Published: (2024)
by: Bovenzi, Inko, et al.
Published: (2024)
Concept-driven Off Policy Evaluation
by: Majumdar, Ritam, et al.
Published: (2024)
by: Majumdar, Ritam, et al.
Published: (2024)
Tree-Based Leakage Inspection and Control in Concept Bottleneck Models
by: Ragkousis, Angelos, et al.
Published: (2024)
by: Ragkousis, Angelos, et al.
Published: (2024)
Towards Model-Agnostic Posterior Approximation for Fast and Accurate Variational Autoencoders
by: Yacoby, Yaniv, et al.
Published: (2024)
by: Yacoby, Yaniv, et al.
Published: (2024)
Diverse Concept Proposals for Concept Bottleneck Models
by: Brown, Katrina, et al.
Published: (2024)
by: Brown, Katrina, et al.
Published: (2024)
On the Effective Horizon of Inverse Reinforcement Learning
by: Xu, Yiqing, et al.
Published: (2023)
by: Xu, Yiqing, et al.
Published: (2023)
Connecting Federated ADMM to Bayes
by: Swaroop, Siddharth, et al.
Published: (2025)
by: Swaroop, Siddharth, et al.
Published: (2025)
Do regularization methods for shortcut mitigation work as intended?
by: Hong, Haoyang, et al.
Published: (2025)
by: Hong, Haoyang, et al.
Published: (2025)
Inverse Reinforcement Learning with Multiple Planning Horizons
by: Yao, Jiayu, et al.
Published: (2024)
by: Yao, Jiayu, et al.
Published: (2024)
What Makes a Good Explanation?: A Harmonized View of Properties of Explanations
by: Chen, Zixi, et al.
Published: (2022)
by: Chen, Zixi, et al.
Published: (2022)
Federated ADMM from Bayesian Duality
by: Möllenhoff, Thomas, et al.
Published: (2025)
by: Möllenhoff, Thomas, et al.
Published: (2025)
Causal Bayesian Optimization with Unknown Graphs
by: Durand, Jean, et al.
Published: (2025)
by: Durand, Jean, et al.
Published: (2025)
Feature Importance Depends on Properties of the Data: Towards Choosing the Correct Explanations for Your Data and Decision Trees based Models
by: Ayad, Célia Wafa, et al.
Published: (2025)
by: Ayad, Célia Wafa, et al.
Published: (2025)
Transparent Trade-offs between Properties of Explanations
by: Tadesse, Hiwot Belay, et al.
Published: (2024)
by: Tadesse, Hiwot Belay, et al.
Published: (2024)
Reinforcement Learning Interventions on Boundedly Rational Human Agents in Frictionful Tasks
by: Nofshin, Eura, et al.
Published: (2024)
by: Nofshin, Eura, et al.
Published: (2024)
A Sim2Real Approach for Identifying Task-Relevant Properties in Interpretable Machine Learning
by: Nofshin, Eura, et al.
Published: (2024)
by: Nofshin, Eura, et al.
Published: (2024)
Understanding the Relationship between Prompts and Response Uncertainty in Large Language Models
by: Zhang, Ze Yu, et al.
Published: (2024)
by: Zhang, Ze Yu, et al.
Published: (2024)
Improving ARDS Diagnosis Through Context-Aware Concept Bottleneck Models
by: Narain, Anish, et al.
Published: (2025)
by: Narain, Anish, et al.
Published: (2025)
Learning from Failures: Understanding LLM Alignment through Failure-Aware Inverse RL
by: Patel, Nyal, et al.
Published: (2025)
by: Patel, Nyal, et al.
Published: (2025)
The Alignment Auditor: A Bayesian Framework for Verifying and Refining LLM Objectives
by: Bou, Matthieu, et al.
Published: (2025)
by: Bou, Matthieu, et al.
Published: (2025)
Safe Policy Exploration Improvement via Subgoals
by: Angulo, Brian, et al.
Published: (2024)
by: Angulo, Brian, et al.
Published: (2024)
Non-Stationary Latent Auto-Regressive Bandits
by: Trella, Anna L., et al.
Published: (2024)
by: Trella, Anna L., et al.
Published: (2024)
CSPI-MT: Calibrated Safe Policy Improvement with Multiple Testing for Threshold Policies
by: Cho, Brian M, et al.
Published: (2024)
by: Cho, Brian M, et al.
Published: (2024)
SNPL: Simultaneous Policy Learning and Evaluation for Safe Multi-Objective Policy Improvement
by: Cho, Brian, et al.
Published: (2025)
by: Cho, Brian, et al.
Published: (2025)
Deep SPI: Safe Policy Improvement via World Models
by: Delgrange, Florent, et al.
Published: (2025)
by: Delgrange, Florent, et al.
Published: (2025)
SafeAR: Safe Algorithmic Recourse by Risk-Aware Policies
by: Wu, Haochen, et al.
Published: (2023)
by: Wu, Haochen, et al.
Published: (2023)
Monitoring Fidelity of Online Reinforcement Learning Algorithms in Clinical Trials
by: Trella, Anna L., et al.
Published: (2024)
by: Trella, Anna L., et al.
Published: (2024)
Rule-Guided Reinforcement Learning Policy Evaluation and Improvement
by: Tappler, Martin, et al.
Published: (2025)
by: Tappler, Martin, et al.
Published: (2025)
Shaping AI's Impact on Billions of Lives
by: Cuéllar, Mariano-Florentino, et al.
Published: (2024)
by: Cuéllar, Mariano-Florentino, et al.
Published: (2024)
Policy Improvement Reinforcement Learning
by: Wang, Huaiyang, et al.
Published: (2026)
by: Wang, Huaiyang, et al.
Published: (2026)
Accuracy-Time Tradeoffs in AI-Assisted Decision Making under Time Pressure
by: Swaroop, Siddharth, et al.
Published: (2023)
by: Swaroop, Siddharth, et al.
Published: (2023)
Why LLMs Fail at Causal Discovery and How Interventional Agents Escape
by: Roy, Amartya, et al.
Published: (2026)
by: Roy, Amartya, et al.
Published: (2026)
Similar Items
-
Bayesian Inverse Transition Learning: Learning Dynamics From Near-Optimal Trajectories
by: Benac, Leo, et al.
Published: (2024) -
Decision-Focused Model-based Reinforcement Learning for Reward Transfer
by: Sharma, Abhishek, et al.
Published: (2023) -
Quantifying Potential Observation Missingness in Inverse Reinforcement Learning
by: Benac, Leo, et al.
Published: (2026) -
Guarantee Regions for Local Explanations
by: Havasi, Marton, et al.
Published: (2024) -
Towards Integrating Personal Knowledge into Test-Time Predictions
by: Lage, Isaac, et al.
Published: (2024)