Multi-objective Reinforcement Learning With Augmented States Requires Rewards After Deployment
Fuente:
arXiv
Saved in:
| Main Authors: | Vamplew, Peter, Foale, Cameron |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Empirical Investigation of Value-Based Multi-objective Reinforcement Learning for Stochastic Environments
by: Ding, Kewen, et al.
Published: (2024)
by: Ding, Kewen, et al.
Published: (2024)
Issues with Value-Based Multi-objective Reinforcement Learning: Value Function Interference and Overestimation Sensitivity
by: Vamplew, Peter, et al.
Published: (2024)
by: Vamplew, Peter, et al.
Published: (2024)
ES-C51: Expected Sarsa Based C51 Distributional Reinforcement Learning Algorithm
by: Tandon, Rijul, et al.
Published: (2025)
by: Tandon, Rijul, et al.
Published: (2025)
Multi-objective Reinforcement Learning: A Tool for Pluralistic Alignment
by: Vamplew, Peter, et al.
Published: (2024)
by: Vamplew, Peter, et al.
Published: (2024)
Utility-Based Reinforcement Learning: Unifying Single-objective and Multi-objective Reinforcement Learning
by: Vamplew, Peter, et al.
Published: (2024)
by: Vamplew, Peter, et al.
Published: (2024)
On Generalization Across Environments In Multi-Objective Reinforcement Learning
by: Teoh, Jayden, et al.
Published: (2025)
by: Teoh, Jayden, et al.
Published: (2025)
Learning the Value Systems of Societies with Preference-based Multi-objective Reinforcement Learning
by: Holgado-Sánchez, Andrés, et al.
Published: (2026)
by: Holgado-Sánchez, Andrés, et al.
Published: (2026)
Adaptive Alignment: Dynamic Preference Adjustments via Multi-Objective Reinforcement Learning for Pluralistic AI
by: Harland, Hadassah, et al.
Published: (2024)
by: Harland, Hadassah, et al.
Published: (2024)
Solving Richly Constrained Reinforcement Learning through State Augmentation and Reward Penalties
by: Jiang, Hao, et al.
Published: (2023)
by: Jiang, Hao, et al.
Published: (2023)
Guaranteeing Control Requirements via Reward Shaping in Reinforcement Learning
by: De Lellis, Francesco, et al.
Published: (2023)
by: De Lellis, Francesco, et al.
Published: (2023)
Learning After Model Deployment
by: Kaymak, Derda, et al.
Published: (2025)
by: Kaymak, Derda, et al.
Published: (2025)
The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning
by: Hafez, Wael, et al.
Published: (2026)
by: Hafez, Wael, et al.
Published: (2026)
Reinforcement Learning with Exogenous States and Rewards
by: Trimponias, George, et al.
Published: (2023)
by: Trimponias, George, et al.
Published: (2023)
On Feasible Rewards in Multi-Agent Inverse Reinforcement Learning
by: Freihaut, Till, et al.
Published: (2024)
by: Freihaut, Till, et al.
Published: (2024)
DynaSearcher: Dynamic Knowledge Graph Augmented Search Agent via Multi-Reward Reinforcement Learning
by: Hao, Chuzhan, et al.
Published: (2025)
by: Hao, Chuzhan, et al.
Published: (2025)
A Reward-Free Viewpoint on Multi-Objective Reinforcement Learning
by: Chen, Ying-Tu, et al.
Published: (2026)
by: Chen, Ying-Tu, et al.
Published: (2026)
Reward Dimension Reduction for Scalable Multi-Objective Reinforcement Learning
by: Park, Giseung, et al.
Published: (2025)
by: Park, Giseung, et al.
Published: (2025)
Multi Task Inverse Reinforcement Learning for Common Sense Reward
by: Glazer, Neta, et al.
Published: (2024)
by: Glazer, Neta, et al.
Published: (2024)
Learning Markov State Abstractions for Deep Reinforcement Learning
by: Allen, Cameron, et al.
Published: (2021)
by: Allen, Cameron, et al.
Published: (2021)
Dynamic Reward Adjustment in Multi-Reward Reinforcement Learning for Counselor Reflection Generation
by: Min, Do June, et al.
Published: (2024)
by: Min, Do June, et al.
Published: (2024)
Multi-Agent Reinforcement Learning with Submodular Reward
by: Chen, Wenjing, et al.
Published: (2026)
by: Chen, Wenjing, et al.
Published: (2026)
Reward-Conditioned Reinforcement Learning
by: Nauman, Michal, et al.
Published: (2026)
by: Nauman, Michal, et al.
Published: (2026)
Augmenting Offline Reinforcement Learning with State-only Interactions
by: Li, Shangzhe, et al.
Published: (2024)
by: Li, Shangzhe, et al.
Published: (2024)
Steady-State Error Compensation for Reinforcement Learning with Quadratic Rewards
by: Wang, Liyao, et al.
Published: (2024)
by: Wang, Liyao, et al.
Published: (2024)
Multi-Objective and Mixed-Reward Reinforcement Learning via Reward-Decorrelated Policy Optimization
by: Bai, Yang, et al.
Published: (2026)
by: Bai, Yang, et al.
Published: (2026)
General Intelligence Requires Reward-based Pretraining
by: Han, Seungwook, et al.
Published: (2025)
by: Han, Seungwook, et al.
Published: (2025)
From Pixels to Factors: Learning Independently Controllable State Variables for Reinforcement Learning
by: Rodriguez-Sanchez, Rafael, et al.
Published: (2025)
by: Rodriguez-Sanchez, Rafael, et al.
Published: (2025)
Reinforcement Learning from Bagged Reward
by: Tang, Yuting, et al.
Published: (2024)
by: Tang, Yuting, et al.
Published: (2024)
The Value of Reward Lookahead in Reinforcement Learning
by: Merlis, Nadav, et al.
Published: (2024)
by: Merlis, Nadav, et al.
Published: (2024)
Informativeness of Reward Functions in Reinforcement Learning
by: Devidze, Rati, et al.
Published: (2024)
by: Devidze, Rati, et al.
Published: (2024)
Reward Design for Reinforcement Learning Agents
by: Devidze, Rati
Published: (2025)
by: Devidze, Rati
Published: (2025)
To the Max: Reinventing Reward in Reinforcement Learning
by: Veviurko, Grigorii, et al.
Published: (2024)
by: Veviurko, Grigorii, et al.
Published: (2024)
Off-Dynamics Reinforcement Learning via Domain Adaptation and Reward Augmented Imitation
by: Guo, Yihong, et al.
Published: (2024)
by: Guo, Yihong, et al.
Published: (2024)
Ensemble Elastic DQN: A novel multi-step ensemble approach to address overestimation in deep value-based reinforcement learning
by: Ly, Adrian, et al.
Published: (2025)
by: Ly, Adrian, et al.
Published: (2025)
Preference-Guided Learning for Sparse-Reward Multi-Agent Reinforcement Learning
by: Bui, The Viet, et al.
Published: (2025)
by: Bui, The Viet, et al.
Published: (2025)
Reinforcement Learning for Machine Learning Model Deployment: Evaluating Multi-Armed Bandits in ML Ops Environments
by: McClendon, S. Aaron, et al.
Published: (2025)
by: McClendon, S. Aaron, et al.
Published: (2025)
Reward Augmentation in Reinforcement Learning for Testing Distributed Systems
by: Borgarelli, Andrea, et al.
Published: (2024)
by: Borgarelli, Andrea, et al.
Published: (2024)
Focal Reward: Balanced Reinforcement Learning under Rubric-Based Rewards
by: Huang, Yu, et al.
Published: (2026)
by: Huang, Yu, et al.
Published: (2026)
The Distributional Reward Critic Framework for Reinforcement Learning Under Perturbed Rewards
by: Chen, Xi, et al.
Published: (2024)
by: Chen, Xi, et al.
Published: (2024)
Provable Multi-Task Reinforcement Learning: A Representation Learning Framework with Low Rank Rewards
by: Guo, Yaoze, et al.
Published: (2026)
by: Guo, Yaoze, et al.
Published: (2026)
Similar Items
-
An Empirical Investigation of Value-Based Multi-objective Reinforcement Learning for Stochastic Environments
by: Ding, Kewen, et al.
Published: (2024) -
Issues with Value-Based Multi-objective Reinforcement Learning: Value Function Interference and Overestimation Sensitivity
by: Vamplew, Peter, et al.
Published: (2024) -
ES-C51: Expected Sarsa Based C51 Distributional Reinforcement Learning Algorithm
by: Tandon, Rijul, et al.
Published: (2025) -
Multi-objective Reinforcement Learning: A Tool for Pluralistic Alignment
by: Vamplew, Peter, et al.
Published: (2024) -
Utility-Based Reinforcement Learning: Unifying Single-objective and Multi-objective Reinforcement Learning
by: Vamplew, Peter, et al.
Published: (2024)