A Unified View on Solving Objective Mismatch in Model-Based Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Ran, Lambert, Nathan, McDonald, Anthony, Garcia, Alfredo, Calandra, Roberto |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Alignment Ceiling: Objective Mismatch in Reinforcement Learning from Human Feedback
by: Lambert, Nathan, et al.
Published: (2023)
by: Lambert, Nathan, et al.
Published: (2023)
A Bayesian Approach to Robust Inverse Reinforcement Learning
by: Wei, Ran, et al.
Published: (2023)
by: Wei, Ran, et al.
Published: (2023)
Learning An Active Inference Model of Driver Perception and Control: Application to Vehicle Car-Following
by: Wei, Ran, et al.
Published: (2023)
by: Wei, Ran, et al.
Published: (2023)
Reinforcement Learning from Human Feedback
by: Lambert, Nathan
Published: (2025)
by: Lambert, Nathan
Published: (2025)
Constrained Reinforcement Learning Under Model Mismatch
by: Sun, Zhongchang, et al.
Published: (2024)
by: Sun, Zhongchang, et al.
Published: (2024)
Robot Arm Control via Cognitive Map Learners
by: McDonald, Nathan, et al.
Published: (2026)
by: McDonald, Nathan, et al.
Published: (2026)
Deep Latent Force Models: ODE-based Process Convolutions for Bayesian Deep Learning
by: Baldwin-McDonald, Thomas, et al.
Published: (2023)
by: Baldwin-McDonald, Thomas, et al.
Published: (2023)
Apple: Toward General Active Perception via Reinforcement Learning
by: Schneider, Tim, et al.
Published: (2025)
by: Schneider, Tim, et al.
Published: (2025)
Learning Gentle Grasping Using Vision, Sound, and Touch
by: Nakahara, Ken, et al.
Published: (2025)
by: Nakahara, Ken, et al.
Published: (2025)
Understanding Uncertainty-based Active Learning Under Model Mismatch
by: Rahmati, Amir Hossein, et al.
Published: (2024)
by: Rahmati, Amir Hossein, et al.
Published: (2024)
From Solving to Verifying: A Unified Objective for Robust Reasoning in LLMs
by: Wang, Xiaoxuan, et al.
Published: (2025)
by: Wang, Xiaoxuan, et al.
Published: (2025)
A Reward-Free Viewpoint on Multi-Objective Reinforcement Learning
by: Chen, Ying-Tu, et al.
Published: (2026)
by: Chen, Ying-Tu, et al.
Published: (2026)
Diagnosing Training Inference Mismatch in LLM Reinforcement Learning
by: Zhong, Tianle, et al.
Published: (2026)
by: Zhong, Tianle, et al.
Published: (2026)
Solving Prior Distribution Mismatch in Diffusion Models via Optimal Transport
by: Wang, Zhanpeng, et al.
Published: (2024)
by: Wang, Zhanpeng, et al.
Published: (2024)
Multi-Objective Reinforcement Learning Based on Decomposition: A Taxonomy and Framework
by: Felten, Florian, et al.
Published: (2023)
by: Felten, Florian, et al.
Published: (2023)
Adaptive RKHS Fourier Features for Compositional Gaussian Process Models
by: Shi, Xinxing, et al.
Published: (2024)
by: Shi, Xinxing, et al.
Published: (2024)
DFedReweighting: A Unified Framework for Objective-Oriented Reweighting in Decentralized Federated Learning
by: Zhang, Kaichuang, et al.
Published: (2025)
by: Zhang, Kaichuang, et al.
Published: (2025)
Constrained Reinforcement Learning with Average Reward Objective: Model-Based and Model-Free Algorithms
by: Aggarwal, Vaneet, et al.
Published: (2024)
by: Aggarwal, Vaneet, et al.
Published: (2024)
Reinforcement Learning with Non-Cumulative Objective
by: Cui, Wei, et al.
Published: (2023)
by: Cui, Wei, et al.
Published: (2023)
Final-Model-Only Data Attribution with a Unifying View of Gradient-Based Methods
by: Wei, Dennis, et al.
Published: (2024)
by: Wei, Dennis, et al.
Published: (2024)
Log-Concave Coupling for Sampling Neural Net Posteriors
by: McDonald, Curtis, et al.
Published: (2024)
by: McDonald, Curtis, et al.
Published: (2024)
OrgFlow: Generative Modeling of Organic Crystal Structures from Molecular Graphs
by: Vahediahmar, Mohammadmahdi, et al.
Published: (2026)
by: Vahediahmar, Mohammadmahdi, et al.
Published: (2026)
Learning to Play Piano in the Real World
by: Zeulner, Yves-Simon, et al.
Published: (2025)
by: Zeulner, Yves-Simon, et al.
Published: (2025)
Polychromic Objectives for Reinforcement Learning
by: Hamid, Jubayer Ibn, et al.
Published: (2025)
by: Hamid, Jubayer Ibn, et al.
Published: (2025)
Pareto Set Learning for Multi-Objective Reinforcement Learning
by: Liu, Erlong, et al.
Published: (2025)
by: Liu, Erlong, et al.
Published: (2025)
Optimizing Language Models for Inference Time Objectives using Reinforcement Learning
by: Tang, Yunhao, et al.
Published: (2025)
by: Tang, Yunhao, et al.
Published: (2025)
On The Expressivity of Objective-Specification Formalisms in Reinforcement Learning
by: Subramani, Rohan, et al.
Published: (2023)
by: Subramani, Rohan, et al.
Published: (2023)
Shift is Good: Mismatched Data Mixing Improves Test Performance
by: Medvedev, Marko, et al.
Published: (2025)
by: Medvedev, Marko, et al.
Published: (2025)
Preference-based Multi-Objective Reinforcement Learning
by: Mu, Ni, et al.
Published: (2025)
by: Mu, Ni, et al.
Published: (2025)
Utility-Based Reinforcement Learning: Unifying Single-objective and Multi-objective Reinforcement Learning
by: Vamplew, Peter, et al.
Published: (2024)
by: Vamplew, Peter, et al.
Published: (2024)
The ATOM Report: Measuring the Open Language Model Ecosystem
by: Lambert, Nathan, et al.
Published: (2026)
by: Lambert, Nathan, et al.
Published: (2026)
Solving Inverse Problems with Model Mismatch using Untrained Neural Networks within Model-based Architectures
by: Guan, Peimeng, et al.
Published: (2024)
by: Guan, Peimeng, et al.
Published: (2024)
Solving the Granularity Mismatch: Hierarchical Preference Learning for Long-Horizon LLM Agents
by: Gao, Heyang, et al.
Published: (2025)
by: Gao, Heyang, et al.
Published: (2025)
Optimistic Reinforcement Learning with Quantile Objectives
by: Alipour-Vaezi, Mohammad, et al.
Published: (2025)
by: Alipour-Vaezi, Mohammad, et al.
Published: (2025)
Zero-shot and Few-shot Generation Strategies for Artificial Clinical Records
by: Frayling, Erlend, et al.
Published: (2024)
by: Frayling, Erlend, et al.
Published: (2024)
On Generalization Across Environments In Multi-Objective Reinforcement Learning
by: Teoh, Jayden, et al.
Published: (2025)
by: Teoh, Jayden, et al.
Published: (2025)
Random Walks with Tweedie: A Unified View of Score-Based Diffusion Models
by: Park, Chicago Y., et al.
Published: (2024)
by: Park, Chicago Y., et al.
Published: (2024)
BOFormer: Learning to Solve Multi-Objective Bayesian Optimization via Non-Markovian RL
by: Hung, Yu-Heng, et al.
Published: (2025)
by: Hung, Yu-Heng, et al.
Published: (2025)
Hybrid Action Based Reinforcement Learning for Multi-Objective Compatible Autonomous Driving
by: Jin, Guizhe, et al.
Published: (2025)
by: Jin, Guizhe, et al.
Published: (2025)
Deep Multi-Objective Reinforcement Learning for Utility-Based Infrastructural Maintenance Optimization
by: van Remmerden, Jesse, et al.
Published: (2024)
by: van Remmerden, Jesse, et al.
Published: (2024)
Similar Items
-
The Alignment Ceiling: Objective Mismatch in Reinforcement Learning from Human Feedback
by: Lambert, Nathan, et al.
Published: (2023) -
A Bayesian Approach to Robust Inverse Reinforcement Learning
by: Wei, Ran, et al.
Published: (2023) -
Learning An Active Inference Model of Driver Perception and Control: Application to Vehicle Car-Following
by: Wei, Ran, et al.
Published: (2023) -
Reinforcement Learning from Human Feedback
by: Lambert, Nathan
Published: (2025) -
Constrained Reinforcement Learning Under Model Mismatch
by: Sun, Zhongchang, et al.
Published: (2024)