Annotation-Efficient Language Model Alignment via Diverse and Representative Response Texts
Fuente:
arXiv
Guardado en:
| Autores principales: | Jinnai, Yuu, Honda, Ukyo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Does Cross-Cultural Alignment Change the Commonsense Morality of Language Models?
por: Jinnai, Yuu
Publicado: (2024)
por: Jinnai, Yuu
Publicado: (2024)
Generating Diverse and High-Quality Texts by Minimum Bayes Risk Decoding
por: Jinnai, Yuu, et al.
Publicado: (2024)
por: Jinnai, Yuu, et al.
Publicado: (2024)
Model-Based Minimum Bayes Risk Decoding for Text Generation
por: Jinnai, Yuu, et al.
Publicado: (2023)
por: Jinnai, Yuu, et al.
Publicado: (2023)
On the True Distribution Approximation of Minimum Bayes-Risk Decoding
por: Ohashi, Atsumoto, et al.
Publicado: (2024)
por: Ohashi, Atsumoto, et al.
Publicado: (2024)
Document-Level Text Generation with Minimum Bayes Risk Decoding using Optimal Transport
por: Jinnai, Yuu
Publicado: (2025)
por: Jinnai, Yuu
Publicado: (2025)
Exploring Explanations Improves the Robustness of In-Context Learning
por: Honda, Ukyo, et al.
Publicado: (2025)
por: Honda, Ukyo, et al.
Publicado: (2025)
Revisiting the Capacity Gap in Chain-of-Thought Distillation from a Practical Perspective
por: Kajitsuka, Tokio, et al.
Publicado: (2026)
por: Kajitsuka, Tokio, et al.
Publicado: (2026)
Filtered Direct Preference Optimization
por: Morimura, Tetsuro, et al.
Publicado: (2024)
por: Morimura, Tetsuro, et al.
Publicado: (2024)
Regularized Best-of-N Sampling with Minimum Bayes Risk Objective for Language Model Alignment
por: Jinnai, Yuu, et al.
Publicado: (2024)
por: Jinnai, Yuu, et al.
Publicado: (2024)
A Single Linear Layer Yields Task-Adapted Low-Rank Matrices
por: Kim, Hwichan, et al.
Publicado: (2024)
por: Kim, Hwichan, et al.
Publicado: (2024)
Re-evaluating Minimum Bayes Risk Decoding for Automatic Speech Recognition
por: Jinnai, Yuu
Publicado: (2025)
por: Jinnai, Yuu
Publicado: (2025)
Hyperparameter-Free Approach for Faster Minimum Bayes Risk Decoding
por: Jinnai, Yuu, et al.
Publicado: (2024)
por: Jinnai, Yuu, et al.
Publicado: (2024)
Efficient Text-Attributed Graph Learning through Selective Annotation and Graph Alignment
por: Xie, Huanyi, et al.
Publicado: (2025)
por: Xie, Huanyi, et al.
Publicado: (2025)
RadAnnotate: Large Language Models for Efficient and Reliable Radiology Report Annotation
por: Shetty, Saisha Pradeep, et al.
Publicado: (2026)
por: Shetty, Saisha Pradeep, et al.
Publicado: (2026)
Spread Preference Annotation: Direct Preference Judgment for Efficient LLM Alignment
por: Kim, Dongyoung, et al.
Publicado: (2024)
por: Kim, Dongyoung, et al.
Publicado: (2024)
Language Models Represent Space and Time
por: Gurnee, Wes, et al.
Publicado: (2023)
por: Gurnee, Wes, et al.
Publicado: (2023)
Large Language Model Hacking: Quantifying the Hidden Risks of Using LLMs for Text Annotation
por: Baumann, Joachim, et al.
Publicado: (2025)
por: Baumann, Joachim, et al.
Publicado: (2025)
SAIL: Self-Improving Efficient Online Alignment of Large Language Models
por: Ding, Mucong, et al.
Publicado: (2024)
por: Ding, Mucong, et al.
Publicado: (2024)
Text Quality-Based Pruning for Efficient Training of Language Models
por: Sharma, Vasu, et al.
Publicado: (2024)
por: Sharma, Vasu, et al.
Publicado: (2024)
Efficient Real-time Refinement of Language Model Text Generation
por: Ko, Joonho, et al.
Publicado: (2025)
por: Ko, Joonho, et al.
Publicado: (2025)
Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts
por: Yin, Yueqin, et al.
Publicado: (2024)
por: Yin, Yueqin, et al.
Publicado: (2024)
HelpSteer3-Preference: Open Human-Annotated Preference Data across Diverse Tasks and Languages
por: Wang, Zhilin, et al.
Publicado: (2025)
por: Wang, Zhilin, et al.
Publicado: (2025)
Not Eliminate but Aggregate: Post-Hoc Control over Mixture-of-Experts to Address Shortcut Shifts in Natural Language Understanding
por: Honda, Ukyo, et al.
Publicado: (2024)
por: Honda, Ukyo, et al.
Publicado: (2024)
On the Robustness of Reward Models for Language Model Alignment
por: Hong, Jiwoo, et al.
Publicado: (2025)
por: Hong, Jiwoo, et al.
Publicado: (2025)
Enhancing Time Series Forecasting via Multi-Level Text Alignment with LLMs
por: Zhao, Taibiao, et al.
Publicado: (2025)
por: Zhao, Taibiao, et al.
Publicado: (2025)
Personalized Alignment Revisited: The Necessity and Sufficiency of User Diversity
por: Kang, Enoch Hyunwook
Publicado: (2026)
por: Kang, Enoch Hyunwook
Publicado: (2026)
Assessing Generalization for Subpopulation Representative Modeling via In-Context Learning
por: Simmons, Gabriel, et al.
Publicado: (2024)
por: Simmons, Gabriel, et al.
Publicado: (2024)
Evolutionary Contrastive Distillation for Language Model Alignment
por: Katz-Samuels, Julian, et al.
Publicado: (2024)
por: Katz-Samuels, Julian, et al.
Publicado: (2024)
Pareto Multi-Objective Alignment for Language Models
por: He, Qiang, et al.
Publicado: (2025)
por: He, Qiang, et al.
Publicado: (2025)
On Efficient and Statistical Quality Estimation for Data Annotation
por: Klie, Jan-Christoph, et al.
Publicado: (2024)
por: Klie, Jan-Christoph, et al.
Publicado: (2024)
Self-Training for Sample-Efficient Active Learning for Text Classification with Pre-Trained Language Models
por: Schröder, Christopher, et al.
Publicado: (2024)
por: Schröder, Christopher, et al.
Publicado: (2024)
Investigating Annotator Bias in Large Language Models for Hate Speech Detection
por: Das, Amit, et al.
Publicado: (2024)
por: Das, Amit, et al.
Publicado: (2024)
Neural Diversity Regularizes Hallucinations in Language Models
por: Chakrabarti, Kushal, et al.
Publicado: (2025)
por: Chakrabarti, Kushal, et al.
Publicado: (2025)
Stream Aligner: Efficient Sentence-Level Alignment via Distribution Induction
por: Lou, Hantao, et al.
Publicado: (2025)
por: Lou, Hantao, et al.
Publicado: (2025)
Diversity Boosts AI-Generated Text Detection
por: Basani, Advik Raj, et al.
Publicado: (2025)
por: Basani, Advik Raj, et al.
Publicado: (2025)
TARDiS : Text Augmentation for Refining Diversity and Separability
por: Kim, Kyungmin, et al.
Publicado: (2025)
por: Kim, Kyungmin, et al.
Publicado: (2025)
NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment
por: Shen, Gerald, et al.
Publicado: (2024)
por: Shen, Gerald, et al.
Publicado: (2024)
Languages are Modalities: Cross-Lingual Alignment via Encoder Injection
por: Agarwal, Rajan, et al.
Publicado: (2025)
por: Agarwal, Rajan, et al.
Publicado: (2025)
Binary Classifier Optimization for Large Language Model Alignment
por: Jung, Seungjae, et al.
Publicado: (2024)
por: Jung, Seungjae, et al.
Publicado: (2024)
Accelerated Preference Optimization for Large Language Model Alignment
por: He, Jiafan, et al.
Publicado: (2024)
por: He, Jiafan, et al.
Publicado: (2024)
Ejemplares similares
-
Does Cross-Cultural Alignment Change the Commonsense Morality of Language Models?
por: Jinnai, Yuu
Publicado: (2024) -
Generating Diverse and High-Quality Texts by Minimum Bayes Risk Decoding
por: Jinnai, Yuu, et al.
Publicado: (2024) -
Model-Based Minimum Bayes Risk Decoding for Text Generation
por: Jinnai, Yuu, et al.
Publicado: (2023) -
On the True Distribution Approximation of Minimum Bayes-Risk Decoding
por: Ohashi, Atsumoto, et al.
Publicado: (2024) -
Document-Level Text Generation with Minimum Bayes Risk Decoding using Optimal Transport
por: Jinnai, Yuu
Publicado: (2025)