Interpreting Predictive Probabilities: Model Confidence or Human Label Variation?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Baan, Joris, Fernández, Raquel, Plank, Barbara, Aziz, Wilker |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Clarify, Abstain or Answer? Strategising in Conversation with Belief-Augmented Generation
von: Baan, Joris, et al.
Veröffentlicht: (2026)
von: Baan, Joris, et al.
Veröffentlicht: (2026)
Predict the Next Word: Humans exhibit uncertainty in this task and language models _____
von: Ilia, Evgenia, et al.
Veröffentlicht: (2024)
von: Ilia, Evgenia, et al.
Veröffentlicht: (2024)
Different Tastes of Entities: Investigating Human Label Variation in Named Entity Annotations
von: Peng, Siyao, et al.
Veröffentlicht: (2024)
von: Peng, Siyao, et al.
Veröffentlicht: (2024)
From Noise to Signal to Selbstzweck: Reframing Human Label Variation in the Era of Post-training in NLP
von: Xu, Shanshan, et al.
Veröffentlicht: (2025)
von: Xu, Shanshan, et al.
Veröffentlicht: (2025)
Threading the Needle: Reweaving Chain-of-Thought Reasoning to Explain Human Label Variation
von: Chen, Beiduo, et al.
Veröffentlicht: (2025)
von: Chen, Beiduo, et al.
Veröffentlicht: (2025)
VariErr NLI: Separating Annotation Error from Human Label Variation
von: Weber-Genzel, Leon, et al.
Veröffentlicht: (2024)
von: Weber-Genzel, Leon, et al.
Veröffentlicht: (2024)
Revisiting Active Learning under (Human) Label Variation
von: Gruber, Cornelia, et al.
Veröffentlicht: (2025)
von: Gruber, Cornelia, et al.
Veröffentlicht: (2025)
Decoupling the Effect of Chain-of-Thought Reasoning: A Human Label Variation Perspective
von: Chen, Beiduo, et al.
Veröffentlicht: (2026)
von: Chen, Beiduo, et al.
Veröffentlicht: (2026)
Variability Need Not Imply Error: The Case of Adequate but Semantically Distinct Responses
von: Ilia, Evgenia, et al.
Veröffentlicht: (2024)
von: Ilia, Evgenia, et al.
Veröffentlicht: (2024)
Agree, Disagree, Explain: Decomposing Human Label Variation in NLI through the Lens of Explanations
von: Hong, Pingjun, et al.
Veröffentlicht: (2025)
von: Hong, Pingjun, et al.
Veröffentlicht: (2025)
If Probable, Then Acceptable? Understanding Conditional Acceptability Judgments in Large Language Models
von: Orth, Jasmin, et al.
Veröffentlicht: (2025)
von: Orth, Jasmin, et al.
Veröffentlicht: (2025)
RAcQUEt: Unveiling the Dangers of Overlooked Referential Ambiguity in Visual LLMs
von: Testoni, Alberto, et al.
Veröffentlicht: (2024)
von: Testoni, Alberto, et al.
Veröffentlicht: (2024)
Human Label Variation as Stable Signal: Learning Annotator-Specific Explanation Behavior via Cross-Annotator Preference Optimization
von: Chen, Beiduo, et al.
Veröffentlicht: (2026)
von: Chen, Beiduo, et al.
Veröffentlicht: (2026)
Mind the Uncertainty in Human Disagreement: Evaluating Discrepancies between Model Predictions and Human Responses in VQA
von: Lan, Jian, et al.
Veröffentlicht: (2024)
von: Lan, Jian, et al.
Veröffentlicht: (2024)
Explanation Regularisation through the Lens of Attributions
von: Ferreira, Pedro, et al.
Veröffentlicht: (2024)
von: Ferreira, Pedro, et al.
Veröffentlicht: (2024)
Truthful or Fabricated? Using Causal Attribution to Mitigate Reward Hacking in Explanations
von: Ferreira, Pedro, et al.
Veröffentlicht: (2025)
von: Ferreira, Pedro, et al.
Veröffentlicht: (2025)
MedPath: Multi-Domain Cross-Vocabulary Hierarchical Paths for Biomedical Entity Linking
von: Mishra, Nishant, et al.
Veröffentlicht: (2025)
von: Mishra, Nishant, et al.
Veröffentlicht: (2025)
Comparing Inferential Strategies of Humans and Large Language Models in Deductive Reasoning
von: Mondorf, Philipp, et al.
Veröffentlicht: (2024)
von: Mondorf, Philipp, et al.
Veröffentlicht: (2024)
A Rose by Any Other Name: LLM-Generated Explanations Are Good Proxies for Human Explanations to Collect Label Distributions on NLI
von: Chen, Beiduo, et al.
Veröffentlicht: (2024)
von: Chen, Beiduo, et al.
Veröffentlicht: (2024)
LiTEx: A Linguistic Taxonomy of Explanations for Understanding Within-Label Variation in Natural Language Inference
von: Hong, Pingjun, et al.
Veröffentlicht: (2025)
von: Hong, Pingjun, et al.
Veröffentlicht: (2025)
Liar, Liar, Logical Mire: A Benchmark for Suppositional Reasoning in Large Language Models
von: Mondorf, Philipp, et al.
Veröffentlicht: (2024)
von: Mondorf, Philipp, et al.
Veröffentlicht: (2024)
On the Interplay between Human Label Variation and Model Fairness
von: Kurniawan, Kemal, et al.
Veröffentlicht: (2025)
von: Kurniawan, Kemal, et al.
Veröffentlicht: (2025)
Neural Text Normalization for Luxembourgish using Real-Life Variation Data
von: Lutgen, Anne-Marie, et al.
Veröffentlicht: (2024)
von: Lutgen, Anne-Marie, et al.
Veröffentlicht: (2024)
Variation is the Norm: Embracing Sociolinguistics in NLP
von: Lutgen, Anne-Marie, et al.
Veröffentlicht: (2026)
von: Lutgen, Anne-Marie, et al.
Veröffentlicht: (2026)
Teaching Language Models to Faithfully Express their Uncertainty
von: Eikema, Bryan, et al.
Veröffentlicht: (2025)
von: Eikema, Bryan, et al.
Veröffentlicht: (2025)
Probing LLMs for Multilingual Discourse Generalization Through a Unified Label Set
von: Eichin, Florian, et al.
Veröffentlicht: (2025)
von: Eichin, Florian, et al.
Veröffentlicht: (2025)
Beyond Accuracy: Evaluating the Reasoning Behavior of Large Language Models -- A Survey
von: Mondorf, Philipp, et al.
Veröffentlicht: (2024)
von: Mondorf, Philipp, et al.
Veröffentlicht: (2024)
Describing Images $\textit{Fast and Slow}$: Quantifying and Predicting the Variation in Human Signals during Visuo-Linguistic Processes
von: Takmaz, Ece, et al.
Veröffentlicht: (2024)
von: Takmaz, Ece, et al.
Veröffentlicht: (2024)
Make Every Letter Count: Building Dialect Variation Dictionaries from Monolingual Corpora
von: Litschko, Robert, et al.
Veröffentlicht: (2025)
von: Litschko, Robert, et al.
Veröffentlicht: (2025)
Crossing Domains without Labels: Distant Supervision for Term Extraction
von: Senger, Elena, et al.
Veröffentlicht: (2025)
von: Senger, Elena, et al.
Veröffentlicht: (2025)
"My Answer is C": First-Token Probabilities Do Not Match Text Answers in Instruction-Tuned Language Models
von: Wang, Xinpeng, et al.
Veröffentlicht: (2024)
von: Wang, Xinpeng, et al.
Veröffentlicht: (2024)
Human Label Variation in Implicit Discourse Relation Recognition
von: Yung, Frances, et al.
Veröffentlicht: (2026)
von: Yung, Frances, et al.
Veröffentlicht: (2026)
Fine-grained Fallacy Detection with Human Label Variation
von: Ramponi, Alan, et al.
Veröffentlicht: (2025)
von: Ramponi, Alan, et al.
Veröffentlicht: (2025)
Confidence Under the Hood: An Investigation into the Confidence-Probability Alignment in Large Language Models
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
The Ecological Fallacy in Annotation: Modelling Human Label Variation goes beyond Sociodemographics
von: Orlikowski, Matthias, et al.
Veröffentlicht: (2023)
von: Orlikowski, Matthias, et al.
Veröffentlicht: (2023)
References Matter: Investigating the Impact of Reference Set Variation on Summarization Evaluation
von: Casola, Silvia, et al.
Veröffentlicht: (2025)
von: Casola, Silvia, et al.
Veröffentlicht: (2025)
Copy First, Translate Later: Interpreting Translation Dynamics in Multilingual Pretraining
von: Körner, Felicia, et al.
Veröffentlicht: (2026)
von: Körner, Felicia, et al.
Veröffentlicht: (2026)
Let the Model Distribute Its Doubt: Confidence Estimation through Verbalized Probability Distribution
von: Wang, Ante, et al.
Veröffentlicht: (2025)
von: Wang, Ante, et al.
Veröffentlicht: (2025)
KARRIEREWEGE: A Large Scale Career Path Prediction Dataset
von: Senger, Elena, et al.
Veröffentlicht: (2024)
von: Senger, Elena, et al.
Veröffentlicht: (2024)
A Deep Learning Approach to Language-independent Gender Prediction on Twitter
von: Hashempour, Reyhaneh, et al.
Veröffentlicht: (2024)
von: Hashempour, Reyhaneh, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Clarify, Abstain or Answer? Strategising in Conversation with Belief-Augmented Generation
von: Baan, Joris, et al.
Veröffentlicht: (2026) -
Predict the Next Word: Humans exhibit uncertainty in this task and language models _____
von: Ilia, Evgenia, et al.
Veröffentlicht: (2024) -
Different Tastes of Entities: Investigating Human Label Variation in Named Entity Annotations
von: Peng, Siyao, et al.
Veröffentlicht: (2024) -
From Noise to Signal to Selbstzweck: Reframing Human Label Variation in the Era of Post-training in NLP
von: Xu, Shanshan, et al.
Veröffentlicht: (2025) -
Threading the Needle: Reweaving Chain-of-Thought Reasoning to Explain Human Label Variation
von: Chen, Beiduo, et al.
Veröffentlicht: (2025)