Act Only When It Pays: Efficient Reinforcement Learning for LLM Reasoning via Selective Rollouts

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Zheng, Haizhong, Zhou, Yang, Bartoldson, Brian R., Kailkhura, Bhavya, Lai, Fan, Zhao, Jiawei, Chen, Beidi
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!