Human Feedback is not Gold Standard
Fuente:
arXiv
Saved in:
| Main Authors: | Hosking, Tom, Blunsom, Phil, Bartolo, Max |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Reward Models with Synthetic Critiques
by: Ye, Zihuiwen, et al.
Published: (2024)
by: Ye, Zihuiwen, et al.
Published: (2024)
Hierarchical Indexing for Retrieval-Augmented Opinion Summarization
by: Hosking, Tom, et al.
Published: (2024)
by: Hosking, Tom, et al.
Published: (2024)
Uncertainty-Aware Step-wise Verification with Generative Reward Models
by: Ye, Zihuiwen, et al.
Published: (2025)
by: Ye, Zihuiwen, et al.
Published: (2025)
Rope to Nope and Back Again: A New Hybrid Attention Strategy
by: Yang, Bowen, et al.
Published: (2025)
by: Yang, Bowen, et al.
Published: (2025)
Fishing for Magikarp: Automatically Detecting Under-trained Tokens in Large Language Models
by: Land, Sander, et al.
Published: (2024)
by: Land, Sander, et al.
Published: (2024)
Learning is Forgetting: LLM Training As Lossy Compression
by: Conklin, Henry C., et al.
Published: (2026)
by: Conklin, Henry C., et al.
Published: (2026)
Reverse Engineering Human Preferences with Reinforcement Learning
by: Alazraki, Lisa, et al.
Published: (2025)
by: Alazraki, Lisa, et al.
Published: (2025)
Benchmarking LLMs' Judgments with No Gold Standard
by: Xu, Shengwei, et al.
Published: (2024)
by: Xu, Shengwei, et al.
Published: (2024)
The PRISM Alignment Dataset: What Participatory, Representative and Individualised Human Feedback Reveals About the Subjective and Multicultural Alignment of Large Language Models
by: Kirk, Hannah Rose, et al.
Published: (2024)
by: Kirk, Hannah Rose, et al.
Published: (2024)
Striking Gold in Advertising: Standardization and Exploration of Ad Text Generation
by: Mita, Masato, et al.
Published: (2023)
by: Mita, Masato, et al.
Published: (2023)
Persian Abstract Meaning Representation: Annotation Guidelines and Gold Standard Dataset
by: Takhshid, Reza, et al.
Published: (2022)
by: Takhshid, Reza, et al.
Published: (2022)
Universal NER: A Gold-Standard Multilingual Named Entity Recognition Benchmark
by: Mayhew, Stephen, et al.
Published: (2023)
by: Mayhew, Stephen, et al.
Published: (2023)
Beyond Gold Standards: Epistemic Ensemble of LLM Judges for Formal Mathematical Reasoning
by: Zhang, Lan, et al.
Published: (2025)
by: Zhang, Lan, et al.
Published: (2025)
Aya 23: Open Weight Releases to Further Multilingual Progress
by: Aryabumi, Viraat, et al.
Published: (2024)
by: Aryabumi, Viraat, et al.
Published: (2024)
Are Lexicon-Based Tools Still the Gold Standard for Valence Analysis in Low-Resource Flemish?
by: Kandala, Ratna, et al.
Published: (2025)
by: Kandala, Ratna, et al.
Published: (2025)
Shaping Shared Languages: Human and Large Language Models' Inductive Biases in Emergent Communication
by: Kouwenhoven, Tom, et al.
Published: (2025)
by: Kouwenhoven, Tom, et al.
Published: (2025)
AutoLibra: Agent Metric Induction from Open-Ended Human Feedback
by: Zhu, Hao, et al.
Published: (2025)
by: Zhu, Hao, et al.
Published: (2025)
Multilingual Controlled Generation And Gold-Standard-Agnostic Evaluation of Code-Mixed Sentences
by: Gupta, Ayushman, et al.
Published: (2024)
by: Gupta, Ayushman, et al.
Published: (2024)
MCQA-Eval: Efficient Confidence Evaluation in NLG with Gold-Standard Correctness Labels
by: Liu, Xiaoou, et al.
Published: (2025)
by: Liu, Xiaoou, et al.
Published: (2025)
A Gold Standard Dataset and Evaluation Framework for Depression Detection and Explanation in Social Media using LLMs
by: Bolegave, Prajval, et al.
Published: (2025)
by: Bolegave, Prajval, et al.
Published: (2025)
Measuring and Reducing LLM Hallucination without Gold-Standard Answers
by: Wei, Jiaheng, et al.
Published: (2024)
by: Wei, Jiaheng, et al.
Published: (2024)
WikiNER-fr-gold: A Gold-Standard NER Corpus
by: Cao, Danrun, et al.
Published: (2024)
by: Cao, Danrun, et al.
Published: (2024)
If You Can't Use Them, Recycle Them: Optimizing Merging at Scale Mitigates Performance Tradeoffs
by: Khalifa, Muhammad, et al.
Published: (2024)
by: Khalifa, Muhammad, et al.
Published: (2024)
Pairing Orthographically Variant Literary Words to Standard Equivalents Using Neural Edit Distance Models
by: Messner, Craig, et al.
Published: (2024)
by: Messner, Craig, et al.
Published: (2024)
Understanding Likelihood Over-optimisation in Direct Alignment Algorithms
by: Shi, Zhengyan, et al.
Published: (2024)
by: Shi, Zhengyan, et al.
Published: (2024)
Searching for Structure: Investigating Emergent Communication with Large Language Models
by: Kouwenhoven, Tom, et al.
Published: (2024)
by: Kouwenhoven, Tom, et al.
Published: (2024)
Curiosity-Driven Reinforcement Learning from Human Feedback
by: Sun, Haoran, et al.
Published: (2025)
by: Sun, Haoran, et al.
Published: (2025)
Reward Modeling from Natural Language Human Feedback
by: Wang, Zongqi, et al.
Published: (2026)
by: Wang, Zongqi, et al.
Published: (2026)
Beyond "Not Novel Enough": Enriching Scholarly Critique with LLM-Assisted Feedback
by: Afzal, Osama Mohammed, et al.
Published: (2025)
by: Afzal, Osama Mohammed, et al.
Published: (2025)
GPT-4o as the Gold Standard: A Scalable and General Purpose Approach to Filter Language Model Pretraining Data
by: Zhang, Jifan, et al.
Published: (2024)
by: Zhang, Jifan, et al.
Published: (2024)
No Need for Explanations: LLMs can implicitly learn from mistakes in-context
by: Alazraki, Lisa, et al.
Published: (2025)
by: Alazraki, Lisa, et al.
Published: (2025)
Aligning Neural Machine Translation Models: Human Feedback in Training and Inference
by: Ramos, Miguel Moura, et al.
Published: (2023)
by: Ramos, Miguel Moura, et al.
Published: (2023)
Pearmut: Human Evaluation of Translation Made Trivial
by: Zouhar, Vilém, et al.
Published: (2026)
by: Zouhar, Vilém, et al.
Published: (2026)
Hybrid Preferences: Learning to Route Instances for Human vs. AI Feedback
by: Miranda, Lester James V., et al.
Published: (2024)
by: Miranda, Lester James V., et al.
Published: (2024)
PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization
by: Tao, Meiling, et al.
Published: (2025)
by: Tao, Meiling, et al.
Published: (2025)
MA-RLHF: Reinforcement Learning from Human Feedback with Macro Actions
by: Chai, Yekun, et al.
Published: (2024)
by: Chai, Yekun, et al.
Published: (2024)
Human or LLM as Standardized Patients? A Comparative Study for Medical Education
by: Zhang, Bingquan, et al.
Published: (2025)
by: Zhang, Bingquan, et al.
Published: (2025)
AI-Assisted Human Evaluation of Machine Translation
by: Zouhar, Vilém, et al.
Published: (2024)
by: Zouhar, Vilém, et al.
Published: (2024)
Generating Planning Feedback for Open-Ended Programming Exercises with LLMs
by: Demirtaş, Mehmet Arif, et al.
Published: (2025)
by: Demirtaş, Mehmet Arif, et al.
Published: (2025)
BiST: A Gold Standard Bangla-English Bilingual Corpus for Sentence Structure and Tense Classification with Inter-Annotator Agreement
by: Shafi, Abdullah Al, et al.
Published: (2026)
by: Shafi, Abdullah Al, et al.
Published: (2026)
Similar Items
-
Improving Reward Models with Synthetic Critiques
by: Ye, Zihuiwen, et al.
Published: (2024) -
Hierarchical Indexing for Retrieval-Augmented Opinion Summarization
by: Hosking, Tom, et al.
Published: (2024) -
Uncertainty-Aware Step-wise Verification with Generative Reward Models
by: Ye, Zihuiwen, et al.
Published: (2025) -
Rope to Nope and Back Again: A New Hybrid Attention Strategy
by: Yang, Bowen, et al.
Published: (2025) -
Fishing for Magikarp: Automatically Detecting Under-trained Tokens in Large Language Models
by: Land, Sander, et al.
Published: (2024)