Architectural Sweet Spots for Modeling Human Label Variation by the Example of Argument Quality: It's Best to Relate Perspectives!
Fuente:
arXiv
Saved in:
| Main Authors: | Heinisch, Philipp, Orlikowski, Matthias, Romberg, Julia, Cimiano, Philipp |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Ecological Fallacy in Annotation: Modelling Human Label Variation goes beyond Sociodemographics
by: Orlikowski, Matthias, et al.
Published: (2023)
by: Orlikowski, Matthias, et al.
Published: (2023)
Balancing Quality and Variation: Spam Filtering Distorts Data Label Distributions
by: Fleisig, Eve, et al.
Published: (2025)
by: Fleisig, Eve, et al.
Published: (2025)
From Argumentation to Deliberation: Perspectivized Stance Vectors for Fine-grained (Dis)agreement Analysis
by: Plenz, Moritz, et al.
Published: (2025)
by: Plenz, Moritz, et al.
Published: (2025)
PAKT: Perspectivized Argumentation Knowledge Graph and Tool for Deliberation Analysis (with Supplementary Materials)
by: Plenz, Moritz, et al.
Published: (2024)
by: Plenz, Moritz, et al.
Published: (2024)
Beyond Demographics: Fine-tuning Large Language Models to Predict Individuals' Subjective Text Perceptions
by: Orlikowski, Matthias, et al.
Published: (2025)
by: Orlikowski, Matthias, et al.
Published: (2025)
Towards a Perspectivist Turn in Argument Quality Assessment
by: Romberg, Julia, et al.
Published: (2025)
by: Romberg, Julia, et al.
Published: (2025)
Pointing out the Shortcomings of Relation Extraction Models with Semantically Motivated Adversarials
by: Nolano, Gennaro, et al.
Published: (2024)
by: Nolano, Gennaro, et al.
Published: (2024)
Argument Summarization and its Evaluation in the Era of Large Language Models
by: Altemeyer, Moritz, et al.
Published: (2025)
by: Altemeyer, Moritz, et al.
Published: (2025)
A Factorized Probabilistic Model of the Semantics of Vague Temporal Adverbials Relative to Different Event Types
by: Kenneweg, Svenja, et al.
Published: (2025)
by: Kenneweg, Svenja, et al.
Published: (2025)
Modeling the Quality of Dialogical Explanations
by: Alshomary, Milad, et al.
Published: (2024)
by: Alshomary, Milad, et al.
Published: (2024)
TRAVELER: A Benchmark for Evaluating Temporal Reasoning across Vague, Implicit and Explicit References
by: Kenneweg, Svenja, et al.
Published: (2025)
by: Kenneweg, Svenja, et al.
Published: (2025)
CompoST: A Benchmark for Analyzing the Ability of LLMs To Compositionally Interpret Questions in a QALD Setting
by: Schmidt, David Maria, et al.
Published: (2025)
by: Schmidt, David Maria, et al.
Published: (2025)
Hit the Sweet Spot! Span-Level Ensemble for Large Language Models
by: Xu, Yangyifan, et al.
Published: (2024)
by: Xu, Yangyifan, et al.
Published: (2024)
Human Label Variation in Implicit Discourse Relation Recognition
by: Yung, Frances, et al.
Published: (2026)
by: Yung, Frances, et al.
Published: (2026)
Lexicalization Is All You Need: Examining the Impact of Lexical Knowledge in a Compositional QALD System
by: Schmidt, David Maria, et al.
Published: (2024)
by: Schmidt, David Maria, et al.
Published: (2024)
SSL: Sweet Spot Learning for Differentiated Guidance in Agentic Optimization
by: Wu, Jinyang, et al.
Published: (2026)
by: Wu, Jinyang, et al.
Published: (2026)
DOCE: Finding the Sweet Spot for Execution-Based Code Generation
by: Li, Haau-Sing, et al.
Published: (2024)
by: Li, Haau-Sing, et al.
Published: (2024)
Finding the Sweet Spot: Preference Data Construction for Scaling Preference Optimization
by: Xiao, Yao, et al.
Published: (2025)
by: Xiao, Yao, et al.
Published: (2025)
Decoupling the Effect of Chain-of-Thought Reasoning: A Human Label Variation Perspective
by: Chen, Beiduo, et al.
Published: (2026)
by: Chen, Beiduo, et al.
Published: (2026)
On the Interplay between Human Label Variation and Model Fairness
by: Kurniawan, Kemal, et al.
Published: (2025)
by: Kurniawan, Kemal, et al.
Published: (2025)
Revisiting Active Learning under (Human) Label Variation
by: Gruber, Cornelia, et al.
Published: (2025)
by: Gruber, Cornelia, et al.
Published: (2025)
Probing Language Models' Gesture Understanding for Enhanced Human-AI Interaction
by: Wicke, Philipp
Published: (2024)
by: Wicke, Philipp
Published: (2024)
Interpreting Predictive Probabilities: Model Confidence or Human Label Variation?
by: Baan, Joris, et al.
Published: (2024)
by: Baan, Joris, et al.
Published: (2024)
Language Equivalence is Undecidable in VASS with Restricted Nondeterminism
by: Czerwiński, Wojciech, et al.
Published: (2025)
by: Czerwiński, Wojciech, et al.
Published: (2025)
Strategic Persuasion with Trait-Conditioned Multi-Agent Systems for Iterative Legal Argumentation
by: Siedler, Philipp D.
Published: (2026)
by: Siedler, Philipp D.
Published: (2026)
BoNBoN Alignment for Large Language Models and the Sweetness of Best-of-n Sampling
by: Gui, Lin, et al.
Published: (2024)
by: Gui, Lin, et al.
Published: (2024)
MMUTF: Multimodal Multimedia Event Argument Extraction with Unified Template Filling
by: Seeberger, Philipp, et al.
Published: (2024)
by: Seeberger, Philipp, et al.
Published: (2024)
Fine-grained Fallacy Detection with Human Label Variation
by: Ramponi, Alan, et al.
Published: (2025)
by: Ramponi, Alan, et al.
Published: (2025)
Textarium: Entangling Annotation, Abstraction and Argument
by: Proff, Philipp, et al.
Published: (2025)
by: Proff, Philipp, et al.
Published: (2025)
The Best Defense is Attack: Repairing Semantics in Textual Adversarial Examples
by: Yang, Heng, et al.
Published: (2023)
by: Yang, Heng, et al.
Published: (2023)
In-Context Example Ordering Guided by Label Distributions
by: Xu, Zhichao, et al.
Published: (2024)
by: Xu, Zhichao, et al.
Published: (2024)
Argument Quality Assessment in the Age of Instruction-Following Large Language Models
by: Wachsmuth, Henning, et al.
Published: (2024)
by: Wachsmuth, Henning, et al.
Published: (2024)
Contextualizing Argument Quality Assessment with Relevant Knowledge
by: Deshpande, Darshan, et al.
Published: (2023)
by: Deshpande, Darshan, et al.
Published: (2023)
Training and Evaluating with Human Label Variation: An Empirical Study
by: Kurniawan, Kemal, et al.
Published: (2025)
by: Kurniawan, Kemal, et al.
Published: (2025)
SCENE: Self-Labeled Counterfactuals for Extrapolating to Negative Examples
by: Fu, Deqing, et al.
Published: (2023)
by: Fu, Deqing, et al.
Published: (2023)
Comparing Inferential Strategies of Humans and Large Language Models in Deductive Reasoning
by: Mondorf, Philipp, et al.
Published: (2024)
by: Mondorf, Philipp, et al.
Published: (2024)
VariErr NLI: Separating Annotation Error from Human Label Variation
by: Weber-Genzel, Leon, et al.
Published: (2024)
by: Weber-Genzel, Leon, et al.
Published: (2024)
Threading the Needle: Reweaving Chain-of-Thought Reasoning to Explain Human Label Variation
by: Chen, Beiduo, et al.
Published: (2025)
by: Chen, Beiduo, et al.
Published: (2025)
Different Tastes of Entities: Investigating Human Label Variation in Named Entity Annotations
by: Peng, Siyao, et al.
Published: (2024)
by: Peng, Siyao, et al.
Published: (2024)
Reassessing Active Learning Adoption in Contemporary NLP: A Community Survey
by: Romberg, Julia, et al.
Published: (2025)
by: Romberg, Julia, et al.
Published: (2025)
Similar Items
-
The Ecological Fallacy in Annotation: Modelling Human Label Variation goes beyond Sociodemographics
by: Orlikowski, Matthias, et al.
Published: (2023) -
Balancing Quality and Variation: Spam Filtering Distorts Data Label Distributions
by: Fleisig, Eve, et al.
Published: (2025) -
From Argumentation to Deliberation: Perspectivized Stance Vectors for Fine-grained (Dis)agreement Analysis
by: Plenz, Moritz, et al.
Published: (2025) -
PAKT: Perspectivized Argumentation Knowledge Graph and Tool for Deliberation Analysis (with Supplementary Materials)
by: Plenz, Moritz, et al.
Published: (2024) -
Beyond Demographics: Fine-tuning Large Language Models to Predict Individuals' Subjective Text Perceptions
by: Orlikowski, Matthias, et al.
Published: (2025)