Direct Large Language Model Alignment Through Self-Rewarding Contrastive Prompt Distillation

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Liu, Aiwei, Bai, Haoping, Lu, Zhiyun, Kong, Xiang, Wang, Simon, Shan, Jiulong, Cao, Meng, Wen, Lijie
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!