Enhancing Traffic Signal Control through Model-based Reinforcement Learning and Policy Reuse

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Li, Yihong, Zhang, Chengwei, Zhan, Furui, Liu, Wanting, Zhou, Kailing, Zheng, Longji
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866915192309284864
author Li, Yihong
Zhang, Chengwei
Zhan, Furui
Liu, Wanting
Zhou, Kailing
Zheng, Longji
author_facet Li, Yihong
Zhang, Chengwei
Zhan, Furui
Liu, Wanting
Zhou, Kailing
Zheng, Longji
contents Multi-agent reinforcement learning (MARL) has shown significant potential in traffic signal control (TSC). However, current MARL-based methods often suffer from insufficient generalization due to the fixed traffic patterns and road network conditions used during training. This limitation results in poor adaptability to new traffic scenarios, leading to high retraining costs and complex deployment. To address this challenge, we propose two algorithms: PLight and PRLight. PLight employs a model-based reinforcement learning approach, pretraining control policies and environment models using predefined source-domain traffic scenarios. The environment model predicts the state transitions, which facilitates the comparison of environmental features. PRLight further enhances adaptability by adaptively selecting pre-trained PLight agents based on the similarity between the source and target domains to accelerate the learning process in the target domain. We evaluated the algorithms through two transfer settings: (1) adaptability to different traffic scenarios within the same road network, and (2) generalization across different road networks. The results show that PRLight significantly reduces the adaptation time compared to learning from scratch in new TSC scenarios, achieving optimal performance using similarities between available and target scenarios.
format Preprint
id arxiv_https___arxiv_org_abs_2503_08728
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Enhancing Traffic Signal Control through Model-based Reinforcement Learning and Policy Reuse
Li, Yihong
Zhang, Chengwei
Zhan, Furui
Liu, Wanting
Zhou, Kailing
Zheng, Longji
Multiagent Systems
Artificial Intelligence
Multi-agent reinforcement learning (MARL) has shown significant potential in traffic signal control (TSC). However, current MARL-based methods often suffer from insufficient generalization due to the fixed traffic patterns and road network conditions used during training. This limitation results in poor adaptability to new traffic scenarios, leading to high retraining costs and complex deployment. To address this challenge, we propose two algorithms: PLight and PRLight. PLight employs a model-based reinforcement learning approach, pretraining control policies and environment models using predefined source-domain traffic scenarios. The environment model predicts the state transitions, which facilitates the comparison of environmental features. PRLight further enhances adaptability by adaptively selecting pre-trained PLight agents based on the similarity between the source and target domains to accelerate the learning process in the target domain. We evaluated the algorithms through two transfer settings: (1) adaptability to different traffic scenarios within the same road network, and (2) generalization across different road networks. The results show that PRLight significantly reduces the adaptation time compared to learning from scratch in new TSC scenarios, achieving optimal performance using similarities between available and target scenarios.
title Enhancing Traffic Signal Control through Model-based Reinforcement Learning and Policy Reuse
topic Multiagent Systems
Artificial Intelligence
url https://arxiv.org/abs/2503.08728