Balancing Multiple Objectives in Urban Traffic Control with Reinforcement Learning from AI Feedback

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: Zhao, Chenyang, Cahill, Vinny, Dusparic, Ivana
Formato: Preprint
Publicado: 2026
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866911464763162624
author Zhao, Chenyang
Cahill, Vinny
Dusparic, Ivana
author_facet Zhao, Chenyang
Cahill, Vinny
Dusparic, Ivana
contents Reward design has been one of the central challenges for real world reinforcement learning (RL) deployment, especially in settings with multiple objectives. Preference-based RL offers an appealing alternative by learning from human preferences over pairs of behavioural outcomes. More recently, RL from AI feedback (RLAIF) has demonstrated that large language models (LLMs) can generate preference labels at scale, mitigating the reliance on human annotators. However, existing RLAIF work typically focuses only on single-objective tasks, leaving the open question of how RLAIF handles systems that involve multiple objectives. In such systems trade-offs among conflicting objectives are difficult to specify, and policies risk collapsing into optimizing for a dominant goal. In this paper, we explore the extension of the RLAIF paradigm to multi-objective self-adaptive systems. We show that multi-objective RLAIF can produce policies that yield balanced trade-offs reflecting different user priorities without laborious reward engineering. We argue that integrating RLAIF into multi-objective RL offers a scalable path toward user-aligned policy learning in domains with inherently conflicting objectives.
format Preprint
id arxiv_https___arxiv_org_abs_2602_20728
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Balancing Multiple Objectives in Urban Traffic Control with Reinforcement Learning from AI Feedback
Zhao, Chenyang
Cahill, Vinny
Dusparic, Ivana
Artificial Intelligence
Reward design has been one of the central challenges for real world reinforcement learning (RL) deployment, especially in settings with multiple objectives. Preference-based RL offers an appealing alternative by learning from human preferences over pairs of behavioural outcomes. More recently, RL from AI feedback (RLAIF) has demonstrated that large language models (LLMs) can generate preference labels at scale, mitigating the reliance on human annotators. However, existing RLAIF work typically focuses only on single-objective tasks, leaving the open question of how RLAIF handles systems that involve multiple objectives. In such systems trade-offs among conflicting objectives are difficult to specify, and policies risk collapsing into optimizing for a dominant goal. In this paper, we explore the extension of the RLAIF paradigm to multi-objective self-adaptive systems. We show that multi-objective RLAIF can produce policies that yield balanced trade-offs reflecting different user priorities without laborious reward engineering. We argue that integrating RLAIF into multi-objective RL offers a scalable path toward user-aligned policy learning in domains with inherently conflicting objectives.
title Balancing Multiple Objectives in Urban Traffic Control with Reinforcement Learning from AI Feedback
topic Artificial Intelligence
url https://arxiv.org/abs/2602.20728