Interpreting Predictive Probabilities: Model Confidence or Human Label Variation?
Fuente:
arXiv
Salvato in:
| Autori principali: | Baan, Joris, Fernández, Raquel, Plank, Barbara, Aziz, Wilker |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Clarify, Abstain or Answer? Strategising in Conversation with Belief-Augmented Generation
di: Baan, Joris, et al.
Pubblicazione: (2026)
di: Baan, Joris, et al.
Pubblicazione: (2026)
Predict the Next Word: Humans exhibit uncertainty in this task and language models _____
di: Ilia, Evgenia, et al.
Pubblicazione: (2024)
di: Ilia, Evgenia, et al.
Pubblicazione: (2024)
Different Tastes of Entities: Investigating Human Label Variation in Named Entity Annotations
di: Peng, Siyao, et al.
Pubblicazione: (2024)
di: Peng, Siyao, et al.
Pubblicazione: (2024)
From Noise to Signal to Selbstzweck: Reframing Human Label Variation in the Era of Post-training in NLP
di: Xu, Shanshan, et al.
Pubblicazione: (2025)
di: Xu, Shanshan, et al.
Pubblicazione: (2025)
Threading the Needle: Reweaving Chain-of-Thought Reasoning to Explain Human Label Variation
di: Chen, Beiduo, et al.
Pubblicazione: (2025)
di: Chen, Beiduo, et al.
Pubblicazione: (2025)
VariErr NLI: Separating Annotation Error from Human Label Variation
di: Weber-Genzel, Leon, et al.
Pubblicazione: (2024)
di: Weber-Genzel, Leon, et al.
Pubblicazione: (2024)
Revisiting Active Learning under (Human) Label Variation
di: Gruber, Cornelia, et al.
Pubblicazione: (2025)
di: Gruber, Cornelia, et al.
Pubblicazione: (2025)
Decoupling the Effect of Chain-of-Thought Reasoning: A Human Label Variation Perspective
di: Chen, Beiduo, et al.
Pubblicazione: (2026)
di: Chen, Beiduo, et al.
Pubblicazione: (2026)
Variability Need Not Imply Error: The Case of Adequate but Semantically Distinct Responses
di: Ilia, Evgenia, et al.
Pubblicazione: (2024)
di: Ilia, Evgenia, et al.
Pubblicazione: (2024)
Agree, Disagree, Explain: Decomposing Human Label Variation in NLI through the Lens of Explanations
di: Hong, Pingjun, et al.
Pubblicazione: (2025)
di: Hong, Pingjun, et al.
Pubblicazione: (2025)
If Probable, Then Acceptable? Understanding Conditional Acceptability Judgments in Large Language Models
di: Orth, Jasmin, et al.
Pubblicazione: (2025)
di: Orth, Jasmin, et al.
Pubblicazione: (2025)
RAcQUEt: Unveiling the Dangers of Overlooked Referential Ambiguity in Visual LLMs
di: Testoni, Alberto, et al.
Pubblicazione: (2024)
di: Testoni, Alberto, et al.
Pubblicazione: (2024)
Human Label Variation as Stable Signal: Learning Annotator-Specific Explanation Behavior via Cross-Annotator Preference Optimization
di: Chen, Beiduo, et al.
Pubblicazione: (2026)
di: Chen, Beiduo, et al.
Pubblicazione: (2026)
Mind the Uncertainty in Human Disagreement: Evaluating Discrepancies between Model Predictions and Human Responses in VQA
di: Lan, Jian, et al.
Pubblicazione: (2024)
di: Lan, Jian, et al.
Pubblicazione: (2024)
Explanation Regularisation through the Lens of Attributions
di: Ferreira, Pedro, et al.
Pubblicazione: (2024)
di: Ferreira, Pedro, et al.
Pubblicazione: (2024)
Truthful or Fabricated? Using Causal Attribution to Mitigate Reward Hacking in Explanations
di: Ferreira, Pedro, et al.
Pubblicazione: (2025)
di: Ferreira, Pedro, et al.
Pubblicazione: (2025)
MedPath: Multi-Domain Cross-Vocabulary Hierarchical Paths for Biomedical Entity Linking
di: Mishra, Nishant, et al.
Pubblicazione: (2025)
di: Mishra, Nishant, et al.
Pubblicazione: (2025)
Comparing Inferential Strategies of Humans and Large Language Models in Deductive Reasoning
di: Mondorf, Philipp, et al.
Pubblicazione: (2024)
di: Mondorf, Philipp, et al.
Pubblicazione: (2024)
A Rose by Any Other Name: LLM-Generated Explanations Are Good Proxies for Human Explanations to Collect Label Distributions on NLI
di: Chen, Beiduo, et al.
Pubblicazione: (2024)
di: Chen, Beiduo, et al.
Pubblicazione: (2024)
LiTEx: A Linguistic Taxonomy of Explanations for Understanding Within-Label Variation in Natural Language Inference
di: Hong, Pingjun, et al.
Pubblicazione: (2025)
di: Hong, Pingjun, et al.
Pubblicazione: (2025)
Liar, Liar, Logical Mire: A Benchmark for Suppositional Reasoning in Large Language Models
di: Mondorf, Philipp, et al.
Pubblicazione: (2024)
di: Mondorf, Philipp, et al.
Pubblicazione: (2024)
On the Interplay between Human Label Variation and Model Fairness
di: Kurniawan, Kemal, et al.
Pubblicazione: (2025)
di: Kurniawan, Kemal, et al.
Pubblicazione: (2025)
Neural Text Normalization for Luxembourgish using Real-Life Variation Data
di: Lutgen, Anne-Marie, et al.
Pubblicazione: (2024)
di: Lutgen, Anne-Marie, et al.
Pubblicazione: (2024)
Variation is the Norm: Embracing Sociolinguistics in NLP
di: Lutgen, Anne-Marie, et al.
Pubblicazione: (2026)
di: Lutgen, Anne-Marie, et al.
Pubblicazione: (2026)
Teaching Language Models to Faithfully Express their Uncertainty
di: Eikema, Bryan, et al.
Pubblicazione: (2025)
di: Eikema, Bryan, et al.
Pubblicazione: (2025)
Probing LLMs for Multilingual Discourse Generalization Through a Unified Label Set
di: Eichin, Florian, et al.
Pubblicazione: (2025)
di: Eichin, Florian, et al.
Pubblicazione: (2025)
Beyond Accuracy: Evaluating the Reasoning Behavior of Large Language Models -- A Survey
di: Mondorf, Philipp, et al.
Pubblicazione: (2024)
di: Mondorf, Philipp, et al.
Pubblicazione: (2024)
Describing Images $\textit{Fast and Slow}$: Quantifying and Predicting the Variation in Human Signals during Visuo-Linguistic Processes
di: Takmaz, Ece, et al.
Pubblicazione: (2024)
di: Takmaz, Ece, et al.
Pubblicazione: (2024)
Make Every Letter Count: Building Dialect Variation Dictionaries from Monolingual Corpora
di: Litschko, Robert, et al.
Pubblicazione: (2025)
di: Litschko, Robert, et al.
Pubblicazione: (2025)
Crossing Domains without Labels: Distant Supervision for Term Extraction
di: Senger, Elena, et al.
Pubblicazione: (2025)
di: Senger, Elena, et al.
Pubblicazione: (2025)
"My Answer is C": First-Token Probabilities Do Not Match Text Answers in Instruction-Tuned Language Models
di: Wang, Xinpeng, et al.
Pubblicazione: (2024)
di: Wang, Xinpeng, et al.
Pubblicazione: (2024)
Human Label Variation in Implicit Discourse Relation Recognition
di: Yung, Frances, et al.
Pubblicazione: (2026)
di: Yung, Frances, et al.
Pubblicazione: (2026)
Fine-grained Fallacy Detection with Human Label Variation
di: Ramponi, Alan, et al.
Pubblicazione: (2025)
di: Ramponi, Alan, et al.
Pubblicazione: (2025)
Confidence Under the Hood: An Investigation into the Confidence-Probability Alignment in Large Language Models
di: Kumar, Abhishek, et al.
Pubblicazione: (2024)
di: Kumar, Abhishek, et al.
Pubblicazione: (2024)
The Ecological Fallacy in Annotation: Modelling Human Label Variation goes beyond Sociodemographics
di: Orlikowski, Matthias, et al.
Pubblicazione: (2023)
di: Orlikowski, Matthias, et al.
Pubblicazione: (2023)
References Matter: Investigating the Impact of Reference Set Variation on Summarization Evaluation
di: Casola, Silvia, et al.
Pubblicazione: (2025)
di: Casola, Silvia, et al.
Pubblicazione: (2025)
Copy First, Translate Later: Interpreting Translation Dynamics in Multilingual Pretraining
di: Körner, Felicia, et al.
Pubblicazione: (2026)
di: Körner, Felicia, et al.
Pubblicazione: (2026)
Let the Model Distribute Its Doubt: Confidence Estimation through Verbalized Probability Distribution
di: Wang, Ante, et al.
Pubblicazione: (2025)
di: Wang, Ante, et al.
Pubblicazione: (2025)
KARRIEREWEGE: A Large Scale Career Path Prediction Dataset
di: Senger, Elena, et al.
Pubblicazione: (2024)
di: Senger, Elena, et al.
Pubblicazione: (2024)
A Deep Learning Approach to Language-independent Gender Prediction on Twitter
di: Hashempour, Reyhaneh, et al.
Pubblicazione: (2024)
di: Hashempour, Reyhaneh, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Clarify, Abstain or Answer? Strategising in Conversation with Belief-Augmented Generation
di: Baan, Joris, et al.
Pubblicazione: (2026) -
Predict the Next Word: Humans exhibit uncertainty in this task and language models _____
di: Ilia, Evgenia, et al.
Pubblicazione: (2024) -
Different Tastes of Entities: Investigating Human Label Variation in Named Entity Annotations
di: Peng, Siyao, et al.
Pubblicazione: (2024) -
From Noise to Signal to Selbstzweck: Reframing Human Label Variation in the Era of Post-training in NLP
di: Xu, Shanshan, et al.
Pubblicazione: (2025) -
Threading the Needle: Reweaving Chain-of-Thought Reasoning to Explain Human Label Variation
di: Chen, Beiduo, et al.
Pubblicazione: (2025)