Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Bao, Keqin, Chen, Nuo, Li, Xiaoyuan, Hui, Binyuan, Yu, Bowen, Feng, Fuli, He, Xiangnan, Liu, Dayiheng
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!