Multi-objective Reinforcement Learning: A Tool for Pluralistic Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Vamplew, Peter, Hayes, Conor F, Foale, Cameron, Dazeley, Richard, Harland, Hadassah |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adaptive Alignment: Dynamic Preference Adjustments via Multi-Objective Reinforcement Learning for Pluralistic AI
von: Harland, Hadassah, et al.
Veröffentlicht: (2024)
von: Harland, Hadassah, et al.
Veröffentlicht: (2024)
An Empirical Investigation of Value-Based Multi-objective Reinforcement Learning for Stochastic Environments
von: Ding, Kewen, et al.
Veröffentlicht: (2024)
von: Ding, Kewen, et al.
Veröffentlicht: (2024)
Issues with Value-Based Multi-objective Reinforcement Learning: Value Function Interference and Overestimation Sensitivity
von: Vamplew, Peter, et al.
Veröffentlicht: (2024)
von: Vamplew, Peter, et al.
Veröffentlicht: (2024)
Multi-objective Reinforcement Learning With Augmented States Requires Rewards After Deployment
von: Vamplew, Peter, et al.
Veröffentlicht: (2026)
von: Vamplew, Peter, et al.
Veröffentlicht: (2026)
Utility-Based Reinforcement Learning: Unifying Single-objective and Multi-objective Reinforcement Learning
von: Vamplew, Peter, et al.
Veröffentlicht: (2024)
von: Vamplew, Peter, et al.
Veröffentlicht: (2024)
ES-C51: Expected Sarsa Based C51 Distributional Reinforcement Learning Algorithm
von: Tandon, Rijul, et al.
Veröffentlicht: (2025)
von: Tandon, Rijul, et al.
Veröffentlicht: (2025)
Learning the Value Systems of Societies with Preference-based Multi-objective Reinforcement Learning
von: Holgado-Sánchez, Andrés, et al.
Veröffentlicht: (2026)
von: Holgado-Sánchez, Andrés, et al.
Veröffentlicht: (2026)
Ensemble Elastic DQN: A novel multi-step ensemble approach to address overestimation in deep value-based reinforcement learning
von: Ly, Adrian, et al.
Veröffentlicht: (2025)
von: Ly, Adrian, et al.
Veröffentlicht: (2025)
Learning Rewards, Not Labels: Adversarial Inverse Reinforcement Learning for Machinery Fault Detection
von: Neupane, Dhiraj, et al.
Veröffentlicht: (2026)
von: Neupane, Dhiraj, et al.
Veröffentlicht: (2026)
Pairwise Calibrated Rewards for Pluralistic Alignment
von: Halpern, Daniel, et al.
Veröffentlicht: (2025)
von: Halpern, Daniel, et al.
Veröffentlicht: (2025)
Pluralistic Alignment Over Time
von: Klassen, Toryn Q., et al.
Veröffentlicht: (2024)
von: Klassen, Toryn Q., et al.
Veröffentlicht: (2024)
AI Apology: A Critical Review of Apology in AI Systems
von: Harland, Hadassah, et al.
Veröffentlicht: (2024)
von: Harland, Hadassah, et al.
Veröffentlicht: (2024)
Pluralistic Alignment for Healthcare: A Role-Driven Framework
von: Zhong, Jiayou, et al.
Veröffentlicht: (2025)
von: Zhong, Jiayou, et al.
Veröffentlicht: (2025)
APPA: Adaptive Preference Pluralistic Alignment for Fair Federated RLHF of LLMs
von: Srewa, Mahmoud, et al.
Veröffentlicht: (2026)
von: Srewa, Mahmoud, et al.
Veröffentlicht: (2026)
Counterfactual Reasoning for Steerable Pluralistic Value Alignment of Large Language Models
von: Guo, Hanze, et al.
Veröffentlicht: (2025)
von: Guo, Hanze, et al.
Veröffentlicht: (2025)
VITAL: A New Dataset for Benchmarking Pluralistic Alignment in Healthcare
von: Shetty, Anudeex, et al.
Veröffentlicht: (2025)
von: Shetty, Anudeex, et al.
Veröffentlicht: (2025)
MixDPO: Modeling Preference Strength for Pluralistic Alignment
von: Imai, Saki, et al.
Veröffentlicht: (2026)
von: Imai, Saki, et al.
Veröffentlicht: (2026)
Data-driven Machinery Fault Diagnosis: A Comprehensive Review
von: Neupane, Dhiraj, et al.
Veröffentlicht: (2024)
von: Neupane, Dhiraj, et al.
Veröffentlicht: (2024)
A Persona-Based Evaluation Framework for Pluralistic Alignment in Generative AI
von: Karagoz, Atahan
Veröffentlicht: (2026)
von: Karagoz, Atahan
Veröffentlicht: (2026)
VISPA: Pluralistic Alignment via Automatic Value Selection and Activation
von: Zheng, Shenyan, et al.
Veröffentlicht: (2026)
von: Zheng, Shenyan, et al.
Veröffentlicht: (2026)
Being Considerate as a Pathway Towards Pluralistic Alignment for Agentic AI
von: Alamdari, Parand A., et al.
Veröffentlicht: (2024)
von: Alamdari, Parand A., et al.
Veröffentlicht: (2024)
On Generalization Across Environments In Multi-Objective Reinforcement Learning
von: Teoh, Jayden, et al.
Veröffentlicht: (2025)
von: Teoh, Jayden, et al.
Veröffentlicht: (2025)
Overcoming Forgetting in LLM Fine-Tuning with Evolution Strategies
von: Schweighofer, Kajetan, et al.
Veröffentlicht: (2026)
von: Schweighofer, Kajetan, et al.
Veröffentlicht: (2026)
Generalization Limits of Reinforcement Learning Alignment
von: Shida, Haruhi, et al.
Veröffentlicht: (2026)
von: Shida, Haruhi, et al.
Veröffentlicht: (2026)
Learning Markov State Abstractions for Deep Reinforcement Learning
von: Allen, Cameron, et al.
Veröffentlicht: (2021)
von: Allen, Cameron, et al.
Veröffentlicht: (2021)
Evolution Strategies at Scale: LLM Fine-Tuning Beyond Reinforcement Learning
von: Qiu, Xin, et al.
Veröffentlicht: (2025)
von: Qiu, Xin, et al.
Veröffentlicht: (2025)
Vision-Based Generic Potential Function for Policy Alignment in Multi-Agent Reinforcement Learning
von: Ma, Hao, et al.
Veröffentlicht: (2025)
von: Ma, Hao, et al.
Veröffentlicht: (2025)
Rethinking Inverse Reinforcement Learning: from Data Alignment to Task Alignment
von: Zhou, Weichao, et al.
Veröffentlicht: (2024)
von: Zhou, Weichao, et al.
Veröffentlicht: (2024)
The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning
von: Hafez, Wael, et al.
Veröffentlicht: (2026)
von: Hafez, Wael, et al.
Veröffentlicht: (2026)
From Pixels to Factors: Learning Independently Controllable State Variables for Reinforcement Learning
von: Rodriguez-Sanchez, Rafael, et al.
Veröffentlicht: (2025)
von: Rodriguez-Sanchez, Rafael, et al.
Veröffentlicht: (2025)
Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning
von: Dong, Guanting, et al.
Veröffentlicht: (2025)
von: Dong, Guanting, et al.
Veröffentlicht: (2025)
Policy-regularized Offline Multi-objective Reinforcement Learning
von: Lin, Qian, et al.
Veröffentlicht: (2024)
von: Lin, Qian, et al.
Veröffentlicht: (2024)
RLFactory: A Plug-and-Play Reinforcement Learning Post-Training Framework for LLM Multi-Turn Tool-Use
von: Chai, Jiajun, et al.
Veröffentlicht: (2025)
von: Chai, Jiajun, et al.
Veröffentlicht: (2025)
Whose View of Safety? A Deep DIVE Dataset for Pluralistic Alignment of Text-to-Image Models
von: Rastogi, Charvi, et al.
Veröffentlicht: (2025)
von: Rastogi, Charvi, et al.
Veröffentlicht: (2025)
Mutual Information Tracks Policy Coherence in Reinforcement Learning
von: Reid, Cameron, et al.
Veröffentlicht: (2025)
von: Reid, Cameron, et al.
Veröffentlicht: (2025)
Pluralistic Behavior Suite: Stress-Testing Multi-Turn Adherence to Custom Behavioral Policies
von: Varshney, Prasoon, et al.
Veröffentlicht: (2025)
von: Varshney, Prasoon, et al.
Veröffentlicht: (2025)
Offline Regularised Reinforcement Learning for Large Language Models Alignment
von: Richemond, Pierre Harvey, et al.
Veröffentlicht: (2024)
von: Richemond, Pierre Harvey, et al.
Veröffentlicht: (2024)
LearnAlign: Data Selection for LLM Reinforcement Learning with Improved Gradient Alignment
von: Li, Shipeng, et al.
Veröffentlicht: (2025)
von: Li, Shipeng, et al.
Veröffentlicht: (2025)
C-MORAL: Controllable Multi-Objective Molecular Optimization with Reinforcement Alignment for LLMs
von: Gao, Rui, et al.
Veröffentlicht: (2026)
von: Gao, Rui, et al.
Veröffentlicht: (2026)
Benchmarking Partial Observability in Reinforcement Learning with a Suite of Memory-Improvable Domains
von: Tao, Ruo Yu, et al.
Veröffentlicht: (2025)
von: Tao, Ruo Yu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Adaptive Alignment: Dynamic Preference Adjustments via Multi-Objective Reinforcement Learning for Pluralistic AI
von: Harland, Hadassah, et al.
Veröffentlicht: (2024) -
An Empirical Investigation of Value-Based Multi-objective Reinforcement Learning for Stochastic Environments
von: Ding, Kewen, et al.
Veröffentlicht: (2024) -
Issues with Value-Based Multi-objective Reinforcement Learning: Value Function Interference and Overestimation Sensitivity
von: Vamplew, Peter, et al.
Veröffentlicht: (2024) -
Multi-objective Reinforcement Learning With Augmented States Requires Rewards After Deployment
von: Vamplew, Peter, et al.
Veröffentlicht: (2026) -
Utility-Based Reinforcement Learning: Unifying Single-objective and Multi-objective Reinforcement Learning
von: Vamplew, Peter, et al.
Veröffentlicht: (2024)