Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Wang, Shenzhi, Yu, Le, Gao, Chang, Zheng, Chujie, Liu, Shixuan, Lu, Rui, Dang, Kai, Chen, Xionghui, Yang, Jianxin, Zhang, Zhenru, Liu, Yuqiong, Yang, An, Zhao, Andrew, Yue, Yang, Song, Shiji, Yu, Bowen, Huang, Gao, Lin, Junyang
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!