RecLLM-R1: A Two-Stage Training Paradigm with Reinforcement Learning and Chain-of-Thought v1

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Xie, Yu, Ren, Xingkai, Qi, Ying, Hu, Yao, Shan, Lianlei
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!