Reinforcement Learning vs. Distillation: Understanding Accuracy and Capability in LLM Reasoning

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Kim, Minwu, Shrestha, Anubhav, Shrestha, Safal, Nepal, Aadim, Ross, Keith
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!