Staff View: :: Library Catalog

Saved in:

Bibliographic Details
Main Authors:	Ye, Yuyang, Tang, Lu-An, Wang, Haoyu, Yu, Runlong, Yu, Wenchao, He, Erhu, Chen, Haifeng, Xiong, Hui
Format:	Preprint
Published:	2024
Subjects:	Machine Learning Artificial Intelligence
Online Access:	https://arxiv.org/abs/2407.08910
Tags:	Add Tag No Tags, Be the first to tag this record!

_version_	1866916320356859904
author	Ye, Yuyang Tang, Lu-An Wang, Haoyu Yu, Runlong Yu, Wenchao He, Erhu Chen, Haifeng Xiong, Hui
author_facet	Ye, Yuyang Tang, Lu-An Wang, Haoyu Yu, Runlong Yu, Wenchao He, Erhu Chen, Haifeng Xiong, Hui
contents	Achieving carbon neutrality within industrial operations has become increasingly imperative for sustainable development. It is both a significant challenge and a key opportunity for operational optimization in industry 4.0. In recent years, Deep Reinforcement Learning (DRL) based methods offer promising enhancements for sequential optimization processes and can be used for reducing carbon emissions. However, existing DRL methods need a pre-defined reward function to assess the impact of each action on the final sustainable development goals (SDG). In many real applications, such a reward function cannot be given in advance. To address the problem, this study proposes a Performance based Adversarial Imitation Learning (PAIL) engine. It is a novel method to acquire optimal operational policies for carbon neutrality without any pre-defined action rewards. Specifically, PAIL employs a Transformer-based policy generator to encode historical information and predict following actions within a multi-dimensional space. The entire action sequence will be iteratively updated by an environmental simulator. Then PAIL uses a discriminator to minimize the discrepancy between generated sequences and real-world samples of high SDG. In parallel, a Q-learning framework based performance estimator is designed to estimate the impact of each action on SDG. Based on these estimations, PAIL refines generated policies with the rewards from both discriminator and performance estimator. PAIL is evaluated on multiple real-world application cases and datasets. The experiment results demonstrate the effectiveness of PAIL comparing to other state-of-the-art baselines. In addition, PAIL offers meaningful interpretability for the optimization in carbon neutrality.
format	Preprint
id	arxiv_https___arxiv_org_abs_2407_08910
institution	arXiv
publishDate	2024
record_format	arxiv
spellingShingle	PAIL: Performance based Adversarial Imitation Learning Engine for Carbon Neutral Optimization Ye, Yuyang Tang, Lu-An Wang, Haoyu Yu, Runlong Yu, Wenchao He, Erhu Chen, Haifeng Xiong, Hui Machine Learning Artificial Intelligence Achieving carbon neutrality within industrial operations has become increasingly imperative for sustainable development. It is both a significant challenge and a key opportunity for operational optimization in industry 4.0. In recent years, Deep Reinforcement Learning (DRL) based methods offer promising enhancements for sequential optimization processes and can be used for reducing carbon emissions. However, existing DRL methods need a pre-defined reward function to assess the impact of each action on the final sustainable development goals (SDG). In many real applications, such a reward function cannot be given in advance. To address the problem, this study proposes a Performance based Adversarial Imitation Learning (PAIL) engine. It is a novel method to acquire optimal operational policies for carbon neutrality without any pre-defined action rewards. Specifically, PAIL employs a Transformer-based policy generator to encode historical information and predict following actions within a multi-dimensional space. The entire action sequence will be iteratively updated by an environmental simulator. Then PAIL uses a discriminator to minimize the discrepancy between generated sequences and real-world samples of high SDG. In parallel, a Q-learning framework based performance estimator is designed to estimate the impact of each action on SDG. Based on these estimations, PAIL refines generated policies with the rewards from both discriminator and performance estimator. PAIL is evaluated on multiple real-world application cases and datasets. The experiment results demonstrate the effectiveness of PAIL comparing to other state-of-the-art baselines. In addition, PAIL offers meaningful interpretability for the optimization in carbon neutrality.
title	PAIL: Performance based Adversarial Imitation Learning Engine for Carbon Neutral Optimization
topic	Machine Learning Artificial Intelligence
url	https://arxiv.org/abs/2407.08910

Similar Items