Trust, But Verify: A Self-Verification Approach to Reinforcement Learning with Verifiable Rewards

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Liu, Xiaoyuan, Liang, Tian, He, Zhiwei, Xu, Jiahao, Wang, Wenxuan, He, Pinjia, Tu, Zhaopeng, Mi, Haitao, Yu, Dong
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!