A2PO: Towards Effective Offline Reinforcement Learning from an Advantage-aware Perspective

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Qing, Yunpeng, liu, Shunyu, Cong, Jingyuan, Chen, Kaixuan, Zhou, Yihe, Song, Mingli
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!