Not All Subjectivity Is the Same! Defining Desiderata for the Evaluation of Subjectivity in NLP
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Khurana, Urja, van der Meer, Michiel, Liscio, Enrico, Fokkens, Antske, Murukannaiah, Pradeep K. |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Annotator-Centric Active Learning for Subjective NLP Tasks
par: van der Meer, Michiel, et autres
Publié: (2024)
par: van der Meer, Michiel, et autres
Publié: (2024)
Crowd-Calibrator: Can Annotator Disagreement Inform Calibration in Subjective Tasks?
par: Khurana, Urja, et autres
Publié: (2024)
par: Khurana, Urja, et autres
Publié: (2024)
DefVerify: Do Hate Speech Models Reflect Their Dataset's Definition?
par: Khurana, Urja, et autres
Publié: (2024)
par: Khurana, Urja, et autres
Publié: (2024)
News is More than a Collection of Facts: Moral Frame Preserving News Summarization
par: Liscio, Enrico, et autres
Publié: (2025)
par: Liscio, Enrico, et autres
Publié: (2025)
A Hybrid Intelligence Method for Argument Mining
par: van der Meer, Michiel, et autres
Publié: (2024)
par: van der Meer, Michiel, et autres
Publié: (2024)
Morality is Non-Binary: Building a Pluralist Moral Sentence Embedding Space using Contrastive Learning
par: Park, Jeongwoo, et autres
Publié: (2024)
par: Park, Jeongwoo, et autres
Publié: (2024)
Do Differences in Values Influence Disagreements in Online Discussions?
par: van der Meer, Michiel, et autres
Publié: (2023)
par: van der Meer, Michiel, et autres
Publié: (2023)
Facilitating Opinion Diversity through Hybrid NLP Approaches
par: van der Meer, Michiel
Publié: (2024)
par: van der Meer, Michiel
Publié: (2024)
Reading Between the Signs: Predicting Future Suicidal Ideation from Adolescent Social Media Texts
par: Blum, Paul, et autres
Publié: (2025)
par: Blum, Paul, et autres
Publié: (2025)
An Empirical Analysis of Diversity in Argument Summarization
par: van der Meer, Michiel, et autres
Publié: (2024)
par: van der Meer, Michiel, et autres
Publié: (2024)
Value Preferences Estimation and Disambiguation in Hybrid Participatory Systems
par: Liscio, Enrico, et autres
Publié: (2024)
par: Liscio, Enrico, et autres
Publié: (2024)
Signs of Struggle: Spotting Cognitive Distortions across Language and Register
par: Kuber, Abhishek, et autres
Publié: (2025)
par: Kuber, Abhishek, et autres
Publié: (2025)
Investigating the Robustness of Modelling Decisions for Few-Shot Cross-Topic Stance Detection: A Preregistered Study
par: Reuver, Myrthe, et autres
Publié: (2024)
par: Reuver, Myrthe, et autres
Publié: (2024)
On the Low-Rank Parametrization of Reward Models for Controlled Language Generation
par: Troshin, Sergey, et autres
Publié: (2024)
par: Troshin, Sergey, et autres
Publié: (2024)
Balancing the Scales: Reinforcement Learning for Fair Classification
par: Eshuijs, Leon, et autres
Publié: (2024)
par: Eshuijs, Leon, et autres
Publié: (2024)
Learning from Sufficient Rationales: Analysing the Relationship Between Explanation Faithfulness and Token-level Regularisation Strategies
par: Kamp, Jonathan, et autres
Publié: (2025)
par: Kamp, Jonathan, et autres
Publié: (2025)
The Role of Syntactic Span Preferences in Post-Hoc Explanation Disagreement
par: Kamp, Jonathan, et autres
Publié: (2024)
par: Kamp, Jonathan, et autres
Publié: (2024)
Improving Causal Interventions in Amnesic Probing with Mean Projection or LEACE
par: Dobrzeniecka, Alicja, et autres
Publié: (2025)
par: Dobrzeniecka, Alicja, et autres
Publié: (2025)
Short-circuiting Shortcuts: Mechanistic Investigation of Shortcuts in Text Classification
par: Eshuijs, Leon, et autres
Publié: (2025)
par: Eshuijs, Leon, et autres
Publié: (2025)
Will Annotators Disagree? Identifying Subjectivity in Value-Laden Arguments
par: Homayounirad, Amir, et autres
Publié: (2025)
par: Homayounirad, Amir, et autres
Publié: (2025)
Asking a Language Model for Diverse Responses
par: Troshin, Sergey, et autres
Publié: (2025)
par: Troshin, Sergey, et autres
Publié: (2025)
Right vs. Right: Can LLMs Make Tough Choices?
par: Yuan, Jiaqing, et autres
Publié: (2024)
par: Yuan, Jiaqing, et autres
Publié: (2024)
Traces of Social Competence in Large Language Models
par: Kouwenhoven, Tom, et autres
Publié: (2026)
par: Kouwenhoven, Tom, et autres
Publié: (2026)
Desiderata for the Context Use of Question Answering Systems
par: Shaier, Sagi, et autres
Publié: (2024)
par: Shaier, Sagi, et autres
Publié: (2024)
Control the Temperature: Selective Sampling for Diverse and High-Quality LLM Outputs
par: Troshin, Sergey, et autres
Publié: (2025)
par: Troshin, Sergey, et autres
Publié: (2025)
Subjective Question Generation and Answer Evaluation using NLP
par: Islam, G. M. Refatul, et autres
Publié: (2025)
par: Islam, G. M. Refatul, et autres
Publié: (2025)
Theory of Mind in Action: The Instruction Inference Task in Dynamic Human-Agent Collaboration
par: Saad, Fardin, et autres
Publié: (2025)
par: Saad, Fardin, et autres
Publié: (2025)
Gricean Norms as a Basis for Effective Collaboration
par: Saad, Fardin, et autres
Publié: (2025)
par: Saad, Fardin, et autres
Publié: (2025)
Beyond the Covariance Trap: Unlocking Generalization in Same-Subject Knowledge Editing for Large Language Models
par: Liu, Xiyu, et autres
Publié: (2026)
par: Liu, Xiyu, et autres
Publié: (2026)
Related Knowledge Perturbation Matters: Rethinking Multiple Pieces of Knowledge Editing in Same-Subject
par: Duan, Zenghao, et autres
Publié: (2025)
par: Duan, Zenghao, et autres
Publié: (2025)
Beyond Majority Voting: Agreement-Based Clustering to Model Annotator Perspectives in Subjective NLP Tasks
par: Belay, Tadesse Destaw, et autres
Publié: (2026)
par: Belay, Tadesse Destaw, et autres
Publié: (2026)
New Desiderata for Direct Preference Optimization
par: Hu, Xiangkun, et autres
Publié: (2024)
par: Hu, Xiangkun, et autres
Publié: (2024)
TartuNLP at SemEval-2025 Task 5: Subject Tagging as Two-Stage Information Retrieval
par: Dorkin, Aleksei, et autres
Publié: (2025)
par: Dorkin, Aleksei, et autres
Publié: (2025)
MEMIT-Merge: Addressing MEMIT's Key-Value Conflicts in Same-Subject Batch Editing for LLMs
par: Dong, Zilu, et autres
Publié: (2025)
par: Dong, Zilu, et autres
Publié: (2025)
Beyond Black-Box Labels: Interpretable Criteria for Diagnosing Subjective NLP Tasks
par: Rair, Nisrine, et autres
Publié: (2026)
par: Rair, Nisrine, et autres
Publié: (2026)
XplaiNLP at CheckThat! 2025: Multilingual Subjectivity Detection with Finetuned Transformers and Prompt-Based Inference with Large Language Models
par: Sahitaj, Ariana, et autres
Publié: (2025)
par: Sahitaj, Ariana, et autres
Publié: (2025)
Subject-level Inference for Realistic Text Anonymization Evaluation
par: Oh, Myeong Seok, et autres
Publié: (2026)
par: Oh, Myeong Seok, et autres
Publié: (2026)
Creation of the Estonian Subjectivity Dataset: Assessing the Degree of Subjectivity on a Scale
par: Gailit, Karl Gustav, et autres
Publié: (2025)
par: Gailit, Karl Gustav, et autres
Publié: (2025)
Why Expert Alignment Is Hard: Evidence from Subjective Evaluation
par: Lin, Tzu-Mi, et autres
Publié: (2026)
par: Lin, Tzu-Mi, et autres
Publié: (2026)
Is the Top Still Spinning? Evaluating Subjectivity in Narrative Understanding
par: Subbiah, Melanie, et autres
Publié: (2025)
par: Subbiah, Melanie, et autres
Publié: (2025)
Documents similaires
-
Annotator-Centric Active Learning for Subjective NLP Tasks
par: van der Meer, Michiel, et autres
Publié: (2024) -
Crowd-Calibrator: Can Annotator Disagreement Inform Calibration in Subjective Tasks?
par: Khurana, Urja, et autres
Publié: (2024) -
DefVerify: Do Hate Speech Models Reflect Their Dataset's Definition?
par: Khurana, Urja, et autres
Publié: (2024) -
News is More than a Collection of Facts: Moral Frame Preserving News Summarization
par: Liscio, Enrico, et autres
Publié: (2025) -
A Hybrid Intelligence Method for Argument Mining
par: van der Meer, Michiel, et autres
Publié: (2024)