On Generalization Across Environments In Multi-Objective Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Teoh, Jayden, Varakantham, Pradeep, Vamplew, Peter |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Environment Novelty Quantification for Effective Unsupervised Environment Design
by: Teoh, Jayden, et al.
Published: (2025)
by: Teoh, Jayden, et al.
Published: (2025)
Enhancing the Hierarchical Environment Design via Generative Trajectory Modeling
by: Li, Dexun, et al.
Published: (2023)
by: Li, Dexun, et al.
Published: (2023)
Multi-objective Reinforcement Learning With Augmented States Requires Rewards After Deployment
by: Vamplew, Peter, et al.
Published: (2026)
by: Vamplew, Peter, et al.
Published: (2026)
An Empirical Investigation of Value-Based Multi-objective Reinforcement Learning for Stochastic Environments
by: Ding, Kewen, et al.
Published: (2024)
by: Ding, Kewen, et al.
Published: (2024)
Offline Safe Reinforcement Learning Using Trajectory Classification
by: Gong, Ze, et al.
Published: (2024)
by: Gong, Ze, et al.
Published: (2024)
Imitate the Good and Avoid the Bad: An Incremental Approach to Safe Reinforcement Learning
by: Hoang, Huy, et al.
Published: (2023)
by: Hoang, Huy, et al.
Published: (2023)
Handling Long and Richly Constrained Tasks through Constrained Hierarchical Reinforcement Learning
by: Lu, Yuxiao, et al.
Published: (2023)
by: Lu, Yuxiao, et al.
Published: (2023)
Imitating Cost-Constrained Behaviors in Reinforcement Learning
by: Shao, Qian, et al.
Published: (2024)
by: Shao, Qian, et al.
Published: (2024)
Adaptive Alignment: Dynamic Preference Adjustments via Multi-Objective Reinforcement Learning for Pluralistic AI
by: Harland, Hadassah, et al.
Published: (2024)
by: Harland, Hadassah, et al.
Published: (2024)
Regret-Based Defense in Adversarial Reinforcement Learning
by: Belaire, Roman, et al.
Published: (2023)
by: Belaire, Roman, et al.
Published: (2023)
On Discovering Algorithms for Adversarial Imitation Learning
by: Chirra, Shashank Reddy, et al.
Published: (2025)
by: Chirra, Shashank Reddy, et al.
Published: (2025)
Issues with Value-Based Multi-objective Reinforcement Learning: Value Function Interference and Overestimation Sensitivity
by: Vamplew, Peter, et al.
Published: (2024)
by: Vamplew, Peter, et al.
Published: (2024)
Solving Richly Constrained Reinforcement Learning through State Augmentation and Reward Penalties
by: Jiang, Hao, et al.
Published: (2023)
by: Jiang, Hao, et al.
Published: (2023)
SPRINQL: Sub-optimal Demonstrations driven Offline Imitation Learning
by: Hoang, Huy, et al.
Published: (2024)
by: Hoang, Huy, et al.
Published: (2024)
ES-C51: Expected Sarsa Based C51 Distributional Reinforcement Learning Algorithm
by: Tandon, Rijul, et al.
Published: (2025)
by: Tandon, Rijul, et al.
Published: (2025)
Multi-objective Reinforcement Learning: A Tool for Pluralistic Alignment
by: Vamplew, Peter, et al.
Published: (2024)
by: Vamplew, Peter, et al.
Published: (2024)
Multi-Objective Reinforcement Learning for Water Management
by: Osika, Zuzanna, et al.
Published: (2025)
by: Osika, Zuzanna, et al.
Published: (2025)
Automatic LLM Red Teaming
by: Belaire, Roman, et al.
Published: (2025)
by: Belaire, Roman, et al.
Published: (2025)
UNIQ: Offline Inverse Q-learning for Avoiding Undesirable Demonstrations
by: Hoang, Huy, et al.
Published: (2024)
by: Hoang, Huy, et al.
Published: (2024)
On Minimizing Adversarial Counterfactual Error in Adversarial RL
by: Belaire, Roman, et al.
Published: (2024)
by: Belaire, Roman, et al.
Published: (2024)
On Learning Informative Trajectory Embeddings for Imitation, Classification and Regression
by: Ge, Zichang, et al.
Published: (2025)
by: Ge, Zichang, et al.
Published: (2025)
Learning What to Do and What Not To Do: Offline Imitation from Expert and Undesirable Demonstrations
by: Hoang, Huy, et al.
Published: (2025)
by: Hoang, Huy, et al.
Published: (2025)
Navigating Trade-offs: Policy Summarization for Multi-Objective Reinforcement Learning
by: Osika, Zuzanna, et al.
Published: (2024)
by: Osika, Zuzanna, et al.
Published: (2024)
Exploring Equity of Climate Policies using Multi-Agent Multi-Objective Reinforcement Learning
by: Biswas, Palok, et al.
Published: (2025)
by: Biswas, Palok, et al.
Published: (2025)
Utility-Based Reinforcement Learning: Unifying Single-objective and Multi-objective Reinforcement Learning
by: Vamplew, Peter, et al.
Published: (2024)
by: Vamplew, Peter, et al.
Published: (2024)
Learning the Value Systems of Societies with Preference-based Multi-objective Reinforcement Learning
by: Holgado-Sánchez, Andrés, et al.
Published: (2026)
by: Holgado-Sánchez, Andrés, et al.
Published: (2026)
Multi-Objective Reinforcement Learning for Generating Covalent Inhibitor Candidates
by: Gil, Renee
Published: (2026)
by: Gil, Renee
Published: (2026)
IRL for Restless Multi-Armed Bandits with Applications in Maternal and Child Health
by: Jain, Gauri, et al.
Published: (2024)
by: Jain, Gauri, et al.
Published: (2024)
BEAVER: Building Environments with Assessable Variation for Evaluating Multi-Objective Reinforcement Learning
by: Liu, Ruohong, et al.
Published: (2025)
by: Liu, Ruohong, et al.
Published: (2025)
A Meta-Learning Approach for Multi-Objective Reinforcement Learning in Sustainable Home Environments
by: Lu, Junlin, et al.
Published: (2024)
by: Lu, Junlin, et al.
Published: (2024)
Preference-based Multi-Objective Reinforcement Learning
by: Mu, Ni, et al.
Published: (2025)
by: Mu, Ni, et al.
Published: (2025)
Pareto Set Learning for Multi-Objective Reinforcement Learning
by: Liu, Erlong, et al.
Published: (2025)
by: Liu, Erlong, et al.
Published: (2025)
BoreaRL: A Multi-Objective Reinforcement Learning Environment for Climate-Adaptive Boreal Forest Management
by: Dsouza, Kevin Bradley, et al.
Published: (2025)
by: Dsouza, Kevin Bradley, et al.
Published: (2025)
Amortized Multi-Objective Optimization Across Tasks with Generative Solution Modeling
by: Wei, Tingyang, et al.
Published: (2025)
by: Wei, Tingyang, et al.
Published: (2025)
Scalable Multi-Objective Robot Reinforcement Learning through Gradient Conflict Resolution
by: Munn, Humphrey, et al.
Published: (2025)
by: Munn, Humphrey, et al.
Published: (2025)
Demonstration Guided Multi-Objective Reinforcement Learning
by: Lu, Junlin, et al.
Published: (2024)
by: Lu, Junlin, et al.
Published: (2024)
Reward Dimension Reduction for Scalable Multi-Objective Reinforcement Learning
by: Park, Giseung, et al.
Published: (2025)
by: Park, Giseung, et al.
Published: (2025)
Benchmarking Offline Multi-Objective Reinforcement Learning in Critical Care
by: Bansal, Aryaman, et al.
Published: (2025)
by: Bansal, Aryaman, et al.
Published: (2025)
A Reward-Free Viewpoint on Multi-Objective Reinforcement Learning
by: Chen, Ying-Tu, et al.
Published: (2026)
by: Chen, Ying-Tu, et al.
Published: (2026)
Theoretical Study of Conflict-Avoidant Multi-Objective Reinforcement Learning
by: Wang, Yudan, et al.
Published: (2024)
by: Wang, Yudan, et al.
Published: (2024)
Similar Items
-
Improving Environment Novelty Quantification for Effective Unsupervised Environment Design
by: Teoh, Jayden, et al.
Published: (2025) -
Enhancing the Hierarchical Environment Design via Generative Trajectory Modeling
by: Li, Dexun, et al.
Published: (2023) -
Multi-objective Reinforcement Learning With Augmented States Requires Rewards After Deployment
by: Vamplew, Peter, et al.
Published: (2026) -
An Empirical Investigation of Value-Based Multi-objective Reinforcement Learning for Stochastic Environments
by: Ding, Kewen, et al.
Published: (2024) -
Offline Safe Reinforcement Learning Using Trajectory Classification
by: Gong, Ze, et al.
Published: (2024)