Fine-tuning Language Models with Generative Adversarial Reward Modelling

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Yu, Zhang Ze, Jaw, Lau Jia, Hui, Zhang, Low, Bryan Kian Hsiang
Format: Preprint
Published: 2023
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!