Skip to content
Universidad del Mar SIBUMAR Descubridor Institucional UMAR
  • Inicio
  • Búsqueda avanzada
  • Explorar
  • Login
    • English
    • Deutsch
    • Español
    • Français
    • Italiano
Advanced
  • Self-Evolved Reward Learning for LLMs
Cover Image

Self-Evolved Reward Learning for LLMs

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Huang, Chenghua, Fan, Zhizhen, Wang, Lu, Yang, Fangkai, Zhao, Pu, Lin, Zeqi, Lin, Qingwei, Zhang, Dongmei, Rajmohan, Saravan, Zhang, Qi
Format: Preprint
Published: 2024
Subjects:
Computation and Language
Artificial Intelligence
Online Access:
Acceder al recurso
Tags: Add Tag
No Tags, Be the first to tag this record!
  • Cite this
  • Text this
  • Email this
  • Print
  • Export Record
    • Export to RefWorks
    • Export to EndNoteWeb
    • Export to EndNote
  • Save to List
  • Permanent link
  • Holdings
  • Description
  • Comments
  • Similar Items
  • Staff View

Internet

https://arxiv.org/abs/2411.00418

Similar Items

  • Learning to Refine: Self-Refinement of Parallel Reasoning in LLMs
    by: Wang, Qibin, et al.
    Published: (2025)
  • Pretrain Value, Not Reward: Decoupled Value Policy Optimization
    by: Huang, Chenghua, et al.
    Published: (2025)
  • From Reasoning to Answer: Empirical, Attention-Based and Mechanistic Insights into Distilled DeepSeek R1 Models
    by: Zhang, Jue, et al.
    Published: (2025)
  • Distill Not Only Data but Also Rewards: Can Smaller Language Models Surpass Larger Ones?
    by: Zhang, Yudi, et al.
    Published: (2025)
  • AutoRAG-HP: Automatic Online Hyper-Parameter Tuning for Retrieval-Augmented Generation
    by: Fu, Jia, et al.
    Published: (2024)
Universidad del Mar
Universidad del MarSistema Bibliotecario de la Universidad del MarDescubridor Institucional UMARImplementación y desarrollo: Mtro. Carlos Alonso Albores Pérez
InicioBúsqueda avanzadaExplorar
Visitas al Descubridor: 33,245© 2026 Universidad del Mar