Decision-Focused Model-based Reinforcement Learning for Reward Transfer
Fuente:
arXiv
Saved in:
| Main Authors: | Sharma, Abhishek, Parbhoo, Sonali, Gottesman, Omer, Doshi-Velez, Finale |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bayesian Inverse Transition Learning: Learning Dynamics From Near-Optimal Trajectories
by: Benac, Leo, et al.
Published: (2024)
by: Benac, Leo, et al.
Published: (2024)
Decision-Point Guided Safe Policy Improvement
by: Sharma, Abhishek, et al.
Published: (2024)
by: Sharma, Abhishek, et al.
Published: (2024)
Guarantee Regions for Local Explanations
by: Havasi, Marton, et al.
Published: (2024)
by: Havasi, Marton, et al.
Published: (2024)
Towards Integrating Personal Knowledge into Test-Time Predictions
by: Lage, Isaac, et al.
Published: (2024)
by: Lage, Isaac, et al.
Published: (2024)
On the Effective Horizon of Inverse Reinforcement Learning
by: Xu, Yiqing, et al.
Published: (2023)
by: Xu, Yiqing, et al.
Published: (2023)
Reinforcement Learning Interventions on Boundedly Rational Human Agents in Frictionful Tasks
by: Nofshin, Eura, et al.
Published: (2024)
by: Nofshin, Eura, et al.
Published: (2024)
Strategically Linked Decisions in Long-Term Planning and Reinforcement Learning
by: Hüyük, Alihan, et al.
Published: (2025)
by: Hüyük, Alihan, et al.
Published: (2025)
Connecting Federated ADMM to Bayes
by: Swaroop, Siddharth, et al.
Published: (2025)
by: Swaroop, Siddharth, et al.
Published: (2025)
Feature Importance Depends on Properties of the Data: Towards Choosing the Correct Explanations for Your Data and Decision Trees based Models
by: Ayad, Célia Wafa, et al.
Published: (2025)
by: Ayad, Célia Wafa, et al.
Published: (2025)
Quantifying Potential Observation Missingness in Inverse Reinforcement Learning
by: Benac, Leo, et al.
Published: (2026)
by: Benac, Leo, et al.
Published: (2026)
On the Geometry of Reinforcement Learning in Continuous State and Action Spaces
by: Tiwari, Saket, et al.
Published: (2022)
by: Tiwari, Saket, et al.
Published: (2022)
Geometry of Neural Reinforcement Learning in Continuous State and Action Spaces
by: Tiwari, Saket, et al.
Published: (2025)
by: Tiwari, Saket, et al.
Published: (2025)
Concept-driven Off Policy Evaluation
by: Majumdar, Ritam, et al.
Published: (2024)
by: Majumdar, Ritam, et al.
Published: (2024)
Learning Markov State Abstractions for Deep Reinforcement Learning
by: Allen, Cameron, et al.
Published: (2021)
by: Allen, Cameron, et al.
Published: (2021)
A Benchmark for Multi-Party Negotiation Games from Real Negotiation Data
by: Benac, Leo, et al.
Published: (2026)
by: Benac, Leo, et al.
Published: (2026)
Improving ARDS Diagnosis Through Context-Aware Concept Bottleneck Models
by: Narain, Anish, et al.
Published: (2025)
by: Narain, Anish, et al.
Published: (2025)
Monitoring Fidelity of Online Reinforcement Learning Algorithms in Clinical Trials
by: Trella, Anna L., et al.
Published: (2024)
by: Trella, Anna L., et al.
Published: (2024)
Provably Efficient Reward Transfer in Reinforcement Learning with Discrete Markov Decision Processes
by: Vora, Kevin, et al.
Published: (2025)
by: Vora, Kevin, et al.
Published: (2025)
Semi-parametric Expert Bayesian Network Learning with Gaussian Processes and Horseshoe Priors
by: Weng, Yidou, et al.
Published: (2024)
by: Weng, Yidou, et al.
Published: (2024)
Non-Stationary Latent Auto-Regressive Bandits
by: Trella, Anna L., et al.
Published: (2024)
by: Trella, Anna L., et al.
Published: (2024)
Generalized Adaptive Transfer Network: Enhancing Transfer Learning in Reinforcement Learning Across Domains
by: Verma, Abhishek, et al.
Published: (2025)
by: Verma, Abhishek, et al.
Published: (2025)
Modeling Behavioral Preferences of Cyber Adversaries Using Inverse Reinforcement Learning
by: Shinde, Aditya, et al.
Published: (2025)
by: Shinde, Aditya, et al.
Published: (2025)
Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes
by: Trapasso, Alessandro, et al.
Published: (2025)
by: Trapasso, Alessandro, et al.
Published: (2025)
Towards the Transferability of Rewards Recovered via Regularized Inverse Reinforcement Learning
by: Schlaginhaufen, Andreas, et al.
Published: (2024)
by: Schlaginhaufen, Andreas, et al.
Published: (2024)
Reinforcement Learning-based Approach for Vehicle-to-Building Charging with Heterogeneous Agents and Long Term Rewards
by: Liu, Fangqi, et al.
Published: (2025)
by: Liu, Fangqi, et al.
Published: (2025)
Personalized and Context-Aware Transformer Models for Predicting Post-Intervention Physiological Responses from Wearable Sensor Data
by: Brown, Esther, et al.
Published: (2026)
by: Brown, Esther, et al.
Published: (2026)
Residual Reward Models for Preference-based Reinforcement Learning
by: Cao, Chenyang, et al.
Published: (2025)
by: Cao, Chenyang, et al.
Published: (2025)
Inverse Reinforcement Learning with Multiple Planning Horizons
by: Yao, Jiayu, et al.
Published: (2024)
by: Yao, Jiayu, et al.
Published: (2024)
Tree-Based Leakage Inspection and Control in Concept Bottleneck Models
by: Ragkousis, Angelos, et al.
Published: (2024)
by: Ragkousis, Angelos, et al.
Published: (2024)
Centralized Reward Agent for Knowledge Sharing and Transfer in Multi-Task Reinforcement Learning
by: Ma, Haozhe, et al.
Published: (2024)
by: Ma, Haozhe, et al.
Published: (2024)
Diverse Concept Proposals for Concept Bottleneck Models
by: Brown, Katrina, et al.
Published: (2024)
by: Brown, Katrina, et al.
Published: (2024)
Towards Model-Agnostic Posterior Approximation for Fast and Accurate Variational Autoencoders
by: Yacoby, Yaniv, et al.
Published: (2024)
by: Yacoby, Yaniv, et al.
Published: (2024)
Contextual Pre-planning on Reward Machine Abstractions for Enhanced Transfer in Deep Reinforcement Learning
by: Azran, Guy, et al.
Published: (2023)
by: Azran, Guy, et al.
Published: (2023)
Subgoal-based Reward Shaping to Improve Efficiency in Reinforcement Learning
by: Okudo, Takato, et al.
Published: (2021)
by: Okudo, Takato, et al.
Published: (2021)
Listwise Reward Estimation for Offline Preference-based Reinforcement Learning
by: Choi, Heewoong, et al.
Published: (2024)
by: Choi, Heewoong, et al.
Published: (2024)
Reward Models in Deep Reinforcement Learning: A Survey
by: Yu, Rui, et al.
Published: (2025)
by: Yu, Rui, et al.
Published: (2025)
Mitigating Partial Observability in Sequential Decision Processes via the Lambda Discrepancy
by: Allen, Cameron, et al.
Published: (2024)
by: Allen, Cameron, et al.
Published: (2024)
DreamSmooth: Improving Model-based Reinforcement Learning via Reward Smoothing
by: Lee, Vint, et al.
Published: (2023)
by: Lee, Vint, et al.
Published: (2023)
Similarity as Reward Alignment: Robust and Versatile Preference-based Reinforcement Learning
by: Rajaram, Sara, et al.
Published: (2025)
by: Rajaram, Sara, et al.
Published: (2025)
Sample-Efficient Preference-based Reinforcement Learning with Dynamics Aware Rewards
by: Metcalf, Katherine, et al.
Published: (2024)
by: Metcalf, Katherine, et al.
Published: (2024)
Similar Items
-
Bayesian Inverse Transition Learning: Learning Dynamics From Near-Optimal Trajectories
by: Benac, Leo, et al.
Published: (2024) -
Decision-Point Guided Safe Policy Improvement
by: Sharma, Abhishek, et al.
Published: (2024) -
Guarantee Regions for Local Explanations
by: Havasi, Marton, et al.
Published: (2024) -
Towards Integrating Personal Knowledge into Test-Time Predictions
by: Lage, Isaac, et al.
Published: (2024) -
On the Effective Horizon of Inverse Reinforcement Learning
by: Xu, Yiqing, et al.
Published: (2023)