Improving Classification Performance With Human Feedback: Label a few, we label the rest
Fuente:
arXiv
Saved in:
| Main Authors: | Vidra, Natan, Clifford, Thomas, Jijo, Katherine, Chung, Eden, Zhang, Liang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Enhancing Large Language Model Performance To Answer Questions and Extract Information More Accurately
by: Zhang, Liang, et al.
Published: (2024)
by: Zhang, Liang, et al.
Published: (2024)
Improving Retrieval for RAG based Question Answering Models on Financial Documents
by: Setty, Spurthi, et al.
Published: (2024)
by: Setty, Spurthi, et al.
Published: (2024)
Improving Reinforcement Learning from Human Feedback with Efficient Reward Model Ensemble
by: Zhang, Shun, et al.
Published: (2024)
by: Zhang, Shun, et al.
Published: (2024)
Learning Personalized Agents from Human Feedback
by: Liang, Kaiqu, et al.
Published: (2026)
by: Liang, Kaiqu, et al.
Published: (2026)
RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback
by: Lee, Harrison, et al.
Published: (2023)
by: Lee, Harrison, et al.
Published: (2023)
AlpacaFarm: A Simulation Framework for Methods that Learn from Human Feedback
by: Dubois, Yann, et al.
Published: (2023)
by: Dubois, Yann, et al.
Published: (2023)
Revisiting the Role of Label Smoothing in Enhanced Text Sentiment Classification
by: Gao, Yijie, et al.
Published: (2023)
by: Gao, Yijie, et al.
Published: (2023)
RLBFF: Binary Flexible Feedback to bridge between Human Feedback & Verifiable Rewards
by: Wang, Zhilin, et al.
Published: (2025)
by: Wang, Zhilin, et al.
Published: (2025)
Label-free Node Classification on Graphs with Large Language Models (LLMS)
by: Chen, Zhikai, et al.
Published: (2023)
by: Chen, Zhikai, et al.
Published: (2023)
Hierarchical Multi-Label Classification of Online Vaccine Concerns
by: Zhu, Chloe Qinyu, et al.
Published: (2024)
by: Zhu, Chloe Qinyu, et al.
Published: (2024)
Parameter Efficient Reinforcement Learning from Human Feedback
by: Sidahmed, Hakim, et al.
Published: (2024)
by: Sidahmed, Hakim, et al.
Published: (2024)
ILRR: Inference-Time Steering Method for Masked Diffusion Language Models
by: Avrahami, Eden, et al.
Published: (2026)
by: Avrahami, Eden, et al.
Published: (2026)
RLTHF: Targeted Human Feedback for LLM Alignment
by: Xu, Yifei, et al.
Published: (2025)
by: Xu, Yifei, et al.
Published: (2025)
Distributionally Robust Reinforcement Learning with Human Feedback
by: Mandal, Debmalya, et al.
Published: (2025)
by: Mandal, Debmalya, et al.
Published: (2025)
Instances and Labels: Hierarchy-aware Joint Supervised Contrastive Learning for Hierarchical Multi-Label Text Classification
by: Yu, Simon, et al.
Published: (2023)
by: Yu, Simon, et al.
Published: (2023)
Comparing Specialised Small and General Large Language Models on Text Classification: 100 Labelled Samples to Achieve Break-Even Performance
by: Pecher, Branislav, et al.
Published: (2024)
by: Pecher, Branislav, et al.
Published: (2024)
Personalized Language Modeling from Personalized Human Feedback
by: Li, Xinyu, et al.
Published: (2024)
by: Li, Xinyu, et al.
Published: (2024)
Reinforcement Learning from Human Feedback with High-Confidence Safety Constraints
by: Chittepu, Yaswanth, et al.
Published: (2025)
by: Chittepu, Yaswanth, et al.
Published: (2025)
AutoLibra: Agent Metric Induction from Open-Ended Human Feedback
by: Zhu, Hao, et al.
Published: (2025)
by: Zhu, Hao, et al.
Published: (2025)
DKEC: Domain Knowledge Enhanced Multi-Label Classification for Diagnosis Prediction
by: Ge, Xueren, et al.
Published: (2023)
by: Ge, Xueren, et al.
Published: (2023)
From Haystack to Needle: Label Space Reduction for Zero-shot Classification
by: Vandemoortele, Nathan, et al.
Published: (2025)
by: Vandemoortele, Nathan, et al.
Published: (2025)
Label-semantics Aware Generative Approach for Domain-Agnostic Multilabel Classification
by: Khatuya, Subhendu, et al.
Published: (2025)
by: Khatuya, Subhendu, et al.
Published: (2025)
Automatic Classification of User Requirements from Online Feedback -- A Replication Study
by: Bhatt, Meet, et al.
Published: (2025)
by: Bhatt, Meet, et al.
Published: (2025)
CHAI for LLMs: Improving Code-Mixed Translation in Large Language Models through Reinforcement Learning with AI Feedback
by: Zhang, Wenbo, et al.
Published: (2024)
by: Zhang, Wenbo, et al.
Published: (2024)
Improving Summarization with Human Edits
by: Yao, Zonghai, et al.
Published: (2023)
by: Yao, Zonghai, et al.
Published: (2023)
An Assessment of Human vs. Model Uncertainty in Soft-Label Learning and Calibration
by: Pavlovic, Maja, et al.
Published: (2026)
by: Pavlovic, Maja, et al.
Published: (2026)
Beyond Labels: Aligning Large Language Models with Human-like Reasoning
by: Kabir, Muhammad Rafsan, et al.
Published: (2024)
by: Kabir, Muhammad Rafsan, et al.
Published: (2024)
Automatic Labelling with Open-source LLMs using Dynamic Label Schema Integration
by: Walshe, Thomas, et al.
Published: (2025)
by: Walshe, Thomas, et al.
Published: (2025)
Prompting in the Dark: Assessing Human Performance in Prompt Engineering for Data Labeling When Gold Labels Are Absent
by: He, Zeyu, et al.
Published: (2025)
by: He, Zeyu, et al.
Published: (2025)
Improved Out-of-Scope Intent Classification with Dual Encoding and Threshold-based Re-Classification
by: Zawbaa, Hossam M., et al.
Published: (2024)
by: Zawbaa, Hossam M., et al.
Published: (2024)
Off-Policy Corrected Reward Modeling for Reinforcement Learning from Human Feedback
by: Ackermann, Johannes, et al.
Published: (2025)
by: Ackermann, Johannes, et al.
Published: (2025)
Self-play with Execution Feedback: Improving Instruction-following Capabilities of Large Language Models
by: Dong, Guanting, et al.
Published: (2024)
by: Dong, Guanting, et al.
Published: (2024)
RLHF Deciphered: A Critical Analysis of Reinforcement Learning from Human Feedback for LLMs
by: Chaudhari, Shreyas, et al.
Published: (2024)
by: Chaudhari, Shreyas, et al.
Published: (2024)
Gradient Regularization Prevents Reward Hacking in Reinforcement Learning from Human Feedback and Verifiable Rewards
by: Ackermann, Johannes, et al.
Published: (2026)
by: Ackermann, Johannes, et al.
Published: (2026)
Alignment Tampering: How Reinforcement Learning from Human Feedback Is Exploited to Optimize Misaligned Biases
by: Hahm, Dongyoon, et al.
Published: (2026)
by: Hahm, Dongyoon, et al.
Published: (2026)
Easy2Hard-Bench: Standardized Difficulty Labels for Profiling LLM Performance and Generalization
by: Ding, Mucong, et al.
Published: (2024)
by: Ding, Mucong, et al.
Published: (2024)
Improving Large Models with Small models: Lower Costs and Better Performance
by: Chen, Dong, et al.
Published: (2024)
by: Chen, Dong, et al.
Published: (2024)
MisSynth: Improving MISSCI Logical Fallacies Classification with Synthetic Data
by: Poliakov, Mykhailo, et al.
Published: (2025)
by: Poliakov, Mykhailo, et al.
Published: (2025)
AgentRec: Agent Recommendation Using Sentence Embeddings Aligned to Human Feedback
by: Park, Joshua, et al.
Published: (2025)
by: Park, Joshua, et al.
Published: (2025)
Can we trust the evaluation on ChatGPT?
by: Aiyappa, Rachith, et al.
Published: (2023)
by: Aiyappa, Rachith, et al.
Published: (2023)
Similar Items
-
Enhancing Large Language Model Performance To Answer Questions and Extract Information More Accurately
by: Zhang, Liang, et al.
Published: (2024) -
Improving Retrieval for RAG based Question Answering Models on Financial Documents
by: Setty, Spurthi, et al.
Published: (2024) -
Improving Reinforcement Learning from Human Feedback with Efficient Reward Model Ensemble
by: Zhang, Shun, et al.
Published: (2024) -
Learning Personalized Agents from Human Feedback
by: Liang, Kaiqu, et al.
Published: (2026) -
RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback
by: Lee, Harrison, et al.
Published: (2023)