Salvato in:
| Autori principali: | Zhang, Michael JQ, Wang, Zhilin, Hwang, Jena D., Dong, Yi, Delalleau, Olivier, Choi, Yejin, Choi, Eunsol, Ren, Xiang, Pyatkin, Valentina |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2410.14632 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PlaSma: Making Small Language Models Better Procedural Knowledge Models for (Counterfactual) Planning
di: Brahman, Faeze, et al.
Pubblicazione: (2023)
di: Brahman, Faeze, et al.
Pubblicazione: (2023)
HelpSteer3-Preference: Open Human-Annotated Preference Data across Diverse Tasks and Languages
di: Wang, Zhilin, et al.
Pubblicazione: (2025)
di: Wang, Zhilin, et al.
Pubblicazione: (2025)
HelpSteer2-Preference: Complementing Ratings with Preferences
di: Wang, Zhilin, et al.
Pubblicazione: (2024)
di: Wang, Zhilin, et al.
Pubblicazione: (2024)
Unpacking DPO and PPO: Disentangling Best Practices for Learning from Preference Feedback
di: Ivison, Hamish, et al.
Pubblicazione: (2024)
di: Ivison, Hamish, et al.
Pubblicazione: (2024)
Phenomenal Yet Puzzling: Testing Inductive Reasoning Capabilities of Language Models with Hypothesis Refinement
di: Qiu, Linlu, et al.
Pubblicazione: (2023)
di: Qiu, Linlu, et al.
Pubblicazione: (2023)
Value Kaleidoscope: Engaging AI with Pluralistic Human Values, Rights, and Duties
di: Sorensen, Taylor, et al.
Pubblicazione: (2023)
di: Sorensen, Taylor, et al.
Pubblicazione: (2023)
What Makes it Ok to Set a Fire? Iterative Self-distillation of Contexts and Rationales for Disambiguating Defeasible Social and Moral Situations
di: Rao, Kavel, et al.
Pubblicazione: (2023)
di: Rao, Kavel, et al.
Pubblicazione: (2023)
HelpSteer3: Human-Annotated Feedback and Edit Data to Empower Inference-Time Scaling in Open-Ended General-Domain Tasks
di: Wang, Zhilin, et al.
Pubblicazione: (2025)
di: Wang, Zhilin, et al.
Pubblicazione: (2025)
Symbolic Working Memory Enhances Language Models for Complex Rule Application
di: Wang, Siyuan, et al.
Pubblicazione: (2024)
di: Wang, Siyuan, et al.
Pubblicazione: (2024)
Mitigating Temporal Misalignment by Discarding Outdated Facts
di: Zhang, Michael J. Q., et al.
Pubblicazione: (2023)
di: Zhang, Michael J. Q., et al.
Pubblicazione: (2023)
DiffAdapt: Difficulty-Adaptive Reasoning for Token-Efficient LLM Inference
di: Liu, Xiang, et al.
Pubblicazione: (2025)
di: Liu, Xiang, et al.
Pubblicazione: (2025)
Can Language Models Reason about Individualistic Human Values and Preferences?
di: Jiang, Liwei, et al.
Pubblicazione: (2024)
di: Jiang, Liwei, et al.
Pubblicazione: (2024)
From Distributional to Overton Pluralism: Investigating Large Language Model Alignment
di: Lake, Thom, et al.
Pubblicazione: (2024)
di: Lake, Thom, et al.
Pubblicazione: (2024)
When Annotators Agree but Labels Disagree: The Projection Problem in Stance Detection
di: Zhang, Bowen
Pubblicazione: (2026)
di: Zhang, Bowen
Pubblicazione: (2026)
Can LLMs Reason with Rules? Logic Scaffolding for Stress-Testing and Improving LLMs
di: Wang, Siyuan, et al.
Pubblicazione: (2024)
di: Wang, Siyuan, et al.
Pubblicazione: (2024)
Open-World Evaluation for Retrieving Diverse Perspectives
di: Chen, Hung-Ting, et al.
Pubblicazione: (2024)
di: Chen, Hung-Ting, et al.
Pubblicazione: (2024)
Reward-aware Preference Optimization: A Unified Mathematical Framework for Model Alignment
di: Sun, Shengyang, et al.
Pubblicazione: (2025)
di: Sun, Shengyang, et al.
Pubblicazione: (2025)
CULTURE-GEN: Revealing Global Cultural Perception in Language Models through Natural Language Prompting
di: Li, Huihan, et al.
Pubblicazione: (2024)
di: Li, Huihan, et al.
Pubblicazione: (2024)
Modeling Future Conversation Turns to Teach LLMs to Ask Clarifying Questions
di: Zhang, Michael J. Q., et al.
Pubblicazione: (2024)
di: Zhang, Michael J. Q., et al.
Pubblicazione: (2024)
AmbigDocs: Reasoning across Documents on Different Entities under the Same Name
di: Lee, Yoonsang, et al.
Pubblicazione: (2024)
di: Lee, Yoonsang, et al.
Pubblicazione: (2024)
RefreshKV: Updating Small KV Cache During Long-form Generation
di: Xu, Fangyuan, et al.
Pubblicazione: (2024)
di: Xu, Fangyuan, et al.
Pubblicazione: (2024)
ParaSpeechCLAP: A Dual-Encoder Speech-Text Model for Rich Stylistic Language-Audio Pretraining
di: Diwan, Anuj, et al.
Pubblicazione: (2026)
di: Diwan, Anuj, et al.
Pubblicazione: (2026)
On Language Models' Sensitivity to Suspicious Coincidences
di: Padmanabhan, Sriram, et al.
Pubblicazione: (2025)
di: Padmanabhan, Sriram, et al.
Pubblicazione: (2025)
Improving LLM-as-a-Judge Inference with the Judgment Distribution
di: Wang, Victor, et al.
Pubblicazione: (2025)
di: Wang, Victor, et al.
Pubblicazione: (2025)
User Feedback in Human-LLM Dialogues: A Lens to Understand Users But Noisy as a Learning Signal
di: Liu, Yuhan, et al.
Pubblicazione: (2025)
di: Liu, Yuhan, et al.
Pubblicazione: (2025)
Beyond Single Embeddings: Capturing Diverse Targets with Multi-Query Retrieval
di: Chen, Hung-Ting, et al.
Pubblicazione: (2025)
di: Chen, Hung-Ting, et al.
Pubblicazione: (2025)
DailyDilemmas: Revealing Value Preferences of LLMs with Quandaries of Daily Life
di: Chiu, Yu Ying, et al.
Pubblicazione: (2024)
di: Chiu, Yu Ying, et al.
Pubblicazione: (2024)
WildBench: Benchmarking LLMs with Challenging Tasks from Real Users in the Wild
di: Lin, Bill Yuchen, et al.
Pubblicazione: (2024)
di: Lin, Bill Yuchen, et al.
Pubblicazione: (2024)
Edited Media Understanding Frames: Reasoning About the Intent and Implications of Visual Misinformation
di: Da, Jeff, et al.
Pubblicazione: (2020)
di: Da, Jeff, et al.
Pubblicazione: (2020)
Exploring Design Choices for Building Language-Specific LLMs
di: Tejaswi, Atula, et al.
Pubblicazione: (2024)
di: Tejaswi, Atula, et al.
Pubblicazione: (2024)
UNcommonsense Reasoning: Abductive Reasoning about Uncommon Situations
di: Zhao, Wenting, et al.
Pubblicazione: (2023)
di: Zhao, Wenting, et al.
Pubblicazione: (2023)
Promptly Predicting Structures: The Return of Inference
di: Mehta, Maitrey, et al.
Pubblicazione: (2024)
di: Mehta, Maitrey, et al.
Pubblicazione: (2024)
Will Annotators Disagree? Identifying Subjectivity in Value-Laden Arguments
di: Homayounirad, Amir, et al.
Pubblicazione: (2025)
di: Homayounirad, Amir, et al.
Pubblicazione: (2025)
RVR: Retrieve-Verify-Retrieve for Comprehensive Question Answering
di: Qian, Deniz, et al.
Pubblicazione: (2026)
di: Qian, Deniz, et al.
Pubblicazione: (2026)
RLBFF: Binary Flexible Feedback to bridge between Human Feedback & Verifiable Rewards
di: Wang, Zhilin, et al.
Pubblicazione: (2025)
di: Wang, Zhilin, et al.
Pubblicazione: (2025)
Superlatives in Context: Modeling the Implicit Semantics of Superlatives
di: Pyatkin, Valentina, et al.
Pubblicazione: (2024)
di: Pyatkin, Valentina, et al.
Pubblicazione: (2024)
Rhapsody: A Dataset for Highlight Detection in Podcasts
di: Park, Younghan, et al.
Pubblicazione: (2025)
di: Park, Younghan, et al.
Pubblicazione: (2025)
Crafting In-context Examples according to LMs' Parametric Knowledge
di: Lee, Yoonsang, et al.
Pubblicazione: (2023)
di: Lee, Yoonsang, et al.
Pubblicazione: (2023)
When Annotators Disagree, Topology Explains: Mapper, a Topological Tool for Exploring Text Embedding Geometry and Ambiguity
di: Rair, Nisrine, et al.
Pubblicazione: (2025)
di: Rair, Nisrine, et al.
Pubblicazione: (2025)
PropMEND: Hypernetworks for Knowledge Propagation in LLMs
di: Liu, Zeyu Leo, et al.
Pubblicazione: (2025)
di: Liu, Zeyu Leo, et al.
Pubblicazione: (2025)
Documenti analoghi
-
PlaSma: Making Small Language Models Better Procedural Knowledge Models for (Counterfactual) Planning
di: Brahman, Faeze, et al.
Pubblicazione: (2023) -
HelpSteer3-Preference: Open Human-Annotated Preference Data across Diverse Tasks and Languages
di: Wang, Zhilin, et al.
Pubblicazione: (2025) -
HelpSteer2-Preference: Complementing Ratings with Preferences
di: Wang, Zhilin, et al.
Pubblicazione: (2024) -
Unpacking DPO and PPO: Disentangling Best Practices for Learning from Preference Feedback
di: Ivison, Hamish, et al.
Pubblicazione: (2024) -
Phenomenal Yet Puzzling: Testing Inductive Reasoning Capabilities of Language Models with Hypothesis Refinement
di: Qiu, Linlu, et al.
Pubblicazione: (2023)