Decoupling Reasoning and Confidence: Resurrecting Calibration in Reinforcement Learning from Verifiable Rewards

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Ma, Zhengzhao, Wen, Xueru, Cao, Boxi, Lu, Yaojie, Lin, Hongyu, Yang, Jinglin, He, Min, Han, Xianpei, Sun, Le
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!