Aligning LLM Uncertainty with Human Disagreement in Subjectivity Analysis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lu, Junyu, Ji, Deyi, Liu, Xuanyi, Zhu, Lanyun, Xu, Bo, Yang, Liang, Hua, Xian-Sheng, Lin, Hongfei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Is LLM an Overconfident Judge? Unveiling the Capabilities of LLMs in Detecting Offensive Language with Annotation Disagreement
von: Lu, Junyu, et al.
Veröffentlicht: (2025)
von: Lu, Junyu, et al.
Veröffentlicht: (2025)
ARGUS: Policy-Adaptive Ad Governance via Evolving Reinforcement with Adversarial Umpiring
von: Ji, Deyi, et al.
Veröffentlicht: (2026)
von: Ji, Deyi, et al.
Veröffentlicht: (2026)
Self-signals Driven Multi-LLM Debate for Efficient and Accurate Reasoning
von: Chen, Xuhang, et al.
Veröffentlicht: (2025)
von: Chen, Xuhang, et al.
Veröffentlicht: (2025)
Integrating Multi-view Analysis: Multi-view Mixture-of-Expert for Textual Personality Detection
von: Zhu, Haohao, et al.
Veröffentlicht: (2024)
von: Zhu, Haohao, et al.
Veröffentlicht: (2024)
DiscoUQ: Structured Disagreement Analysis for Uncertainty Quantification in LLM Agent Ensembles
von: Jiang, Bo
Veröffentlicht: (2026)
von: Jiang, Bo
Veröffentlicht: (2026)
Tree-of-Table: Unleashing the Power of LLMs for Enhanced Large-Scale Table Understanding
von: Ji, Deyi, et al.
Veröffentlicht: (2024)
von: Ji, Deyi, et al.
Veröffentlicht: (2024)
RAVEN: Robust Advertisement Video Violation Temporal Grounding via Reinforcement Reasoning
von: Ji, Deyi, et al.
Veröffentlicht: (2025)
von: Ji, Deyi, et al.
Veröffentlicht: (2025)
Distinguishing Right from Wrong in Debates: Attribution Analysis of Chinese Harmful Memes
von: Wang, Weiming, et al.
Veröffentlicht: (2026)
von: Wang, Weiming, et al.
Veröffentlicht: (2026)
Take its Essence, Discard its Dross! Debiasing for Toxic Language Detection via Counterfactual Causal Effect
von: Lu, Junyu, et al.
Veröffentlicht: (2024)
von: Lu, Junyu, et al.
Veröffentlicht: (2024)
PclGPT: A Large Language Model for Patronizing and Condescending Language Detection
von: Wang, Hongbo, et al.
Veröffentlicht: (2024)
von: Wang, Hongbo, et al.
Veröffentlicht: (2024)
Towards Comprehensive Detection of Chinese Harmful Memes
von: Lu, Junyu, et al.
Veröffentlicht: (2024)
von: Lu, Junyu, et al.
Veröffentlicht: (2024)
Video-Zero: Self-Evolution Video Understanding
von: Zhang, Ruixu, et al.
Veröffentlicht: (2026)
von: Zhang, Ruixu, et al.
Veröffentlicht: (2026)
Commonality and Individuality! Integrating Humor Commonality with Speaker Individuality for Humor Recognition
von: Zhu, Haohao, et al.
Veröffentlicht: (2025)
von: Zhu, Haohao, et al.
Veröffentlicht: (2025)
Rewiring the Transformer with Depth-Wise LSTMs
von: Xu, Hongfei, et al.
Veröffentlicht: (2020)
von: Xu, Hongfei, et al.
Veröffentlicht: (2020)
Enhancing Textual Personality Detection toward Social Media: Integrating Long-term and Short-term Perspectives
von: Zhu, Haohao, et al.
Veröffentlicht: (2024)
von: Zhu, Haohao, et al.
Veröffentlicht: (2024)
STATE ToxiCN: A Benchmark for Span-level Target-Aware Toxicity Extraction in Chinese Hate Speech Detection
von: Bai, Zewen, et al.
Veröffentlicht: (2025)
von: Bai, Zewen, et al.
Veröffentlicht: (2025)
Towards Patronizing and Condescending Language in Chinese Videos: A Multimodal Dataset and Detector
von: Wang, Hongbo, et al.
Veröffentlicht: (2024)
von: Wang, Hongbo, et al.
Veröffentlicht: (2024)
RAVEN++: Pinpointing Fine-Grained Violations in Advertisement Videos with Active Reinforcement Reasoning
von: Ji, Deyi, et al.
Veröffentlicht: (2025)
von: Ji, Deyi, et al.
Veröffentlicht: (2025)
POPEN: Preference-Based Optimization and Ensemble for LVLM-Based Reasoning Segmentation
von: Zhu, Lanyun, et al.
Veröffentlicht: (2025)
von: Zhu, Lanyun, et al.
Veröffentlicht: (2025)
Harder to Defend: Towards Chinese Toxicity Attacks via Implicit Enhancement and Obfuscation Rewriting
von: Kang, Jingyi, et al.
Veröffentlicht: (2026)
von: Kang, Jingyi, et al.
Veröffentlicht: (2026)
Crowd-Calibrator: Can Annotator Disagreement Inform Calibration in Subjective Tasks?
von: Khurana, Urja, et al.
Veröffentlicht: (2024)
von: Khurana, Urja, et al.
Veröffentlicht: (2024)
When the Majority is Wrong: Modeling Annotator Disagreement for Subjective Tasks
von: Fleisig, Eve, et al.
Veröffentlicht: (2023)
von: Fleisig, Eve, et al.
Veröffentlicht: (2023)
When Disagreements Elicit Robustness: Investigating Self-Repair Capabilities under LLM Multi-Agent Disagreements
von: Ju, Tianjie, et al.
Veröffentlicht: (2025)
von: Ju, Tianjie, et al.
Veröffentlicht: (2025)
Mind the Uncertainty in Human Disagreement: Evaluating Discrepancies between Model Predictions and Human Responses in VQA
von: Lan, Jian, et al.
Veröffentlicht: (2024)
von: Lan, Jian, et al.
Veröffentlicht: (2024)
Quantifying and Predicting Disagreement in Graded Human Ratings
von: Zhang, Leixin, et al.
Veröffentlicht: (2026)
von: Zhang, Leixin, et al.
Veröffentlicht: (2026)
Aligning LLMs with Human Uncertainty: A Beta-Bernoulli Calibrator for LLM Forecasting
von: Dai, Hui, et al.
Veröffentlicht: (2026)
von: Dai, Hui, et al.
Veröffentlicht: (2026)
Investigating Human-Aligned Large Language Model Uncertainty
von: Moore, Kyle, et al.
Veröffentlicht: (2025)
von: Moore, Kyle, et al.
Veröffentlicht: (2025)
Bridging the Gap: In-Context Learning for Modeling Human Disagreement
von: Muscato, Benedetta, et al.
Veröffentlicht: (2025)
von: Muscato, Benedetta, et al.
Veröffentlicht: (2025)
StreamCacheVGGT: Streaming Visual Geometry Transformers with Robust Scoring and Hybrid Cache Compression
von: Liu, Xuanyi, et al.
Veröffentlicht: (2026)
von: Liu, Xuanyi, et al.
Veröffentlicht: (2026)
Leveraging Annotator Disagreement for Text Classification
von: Xu, Jin, et al.
Veröffentlicht: (2024)
von: Xu, Jin, et al.
Veröffentlicht: (2024)
A Dataset for Physical and Abstract Plausibility and Sources of Human Disagreement
von: Eichel, Annerose, et al.
Veröffentlicht: (2024)
von: Eichel, Annerose, et al.
Veröffentlicht: (2024)
IPEval: A Bilingual Intellectual Property Agency Consultation Evaluation Benchmark for Large Language Models
von: Wang, Qiyao, et al.
Veröffentlicht: (2024)
von: Wang, Qiyao, et al.
Veröffentlicht: (2024)
Visual Puns from Idioms: An Iterative LLM-T2IM-MLLM Framework
von: Xiao, Kelaiti, et al.
Veröffentlicht: (2025)
von: Xiao, Kelaiti, et al.
Veröffentlicht: (2025)
Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification
von: Zhu, Lanyun, et al.
Veröffentlicht: (2025)
von: Zhu, Lanyun, et al.
Veröffentlicht: (2025)
LLaFS: When Large Language Models Meet Few-Shot Segmentation
von: Zhu, Lanyun, et al.
Veröffentlicht: (2023)
von: Zhu, Lanyun, et al.
Veröffentlicht: (2023)
LANDeRMT: Detecting and Routing Language-Aware Neurons for Selectively Finetuning LLMs to Machine Translation
von: Zhu, Shaolin, et al.
Veröffentlicht: (2024)
von: Zhu, Shaolin, et al.
Veröffentlicht: (2024)
Estimating the Black-box LLM Uncertainty with Distribution-Aligned Adversarial Distillation
von: Cui, Huizi, et al.
Veröffentlicht: (2026)
von: Cui, Huizi, et al.
Veröffentlicht: (2026)
AlignSum: Data Pyramid Hierarchical Fine-tuning for Aligning with Human Summarization Preference
von: Han, Yang, et al.
Veröffentlicht: (2024)
von: Han, Yang, et al.
Veröffentlicht: (2024)
CrossTrafficLLM: A Human-Centric Framework for Interpretable Traffic Intelligence via Large Language Model
von: Du, Zeming, et al.
Veröffentlicht: (2025)
von: Du, Zeming, et al.
Veröffentlicht: (2025)
JuniperLiu at CoMeDi Shared Task: Models as Annotators in Lexical Semantics Disagreements
von: Liu, Zhu, et al.
Veröffentlicht: (2024)
von: Liu, Zhu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Is LLM an Overconfident Judge? Unveiling the Capabilities of LLMs in Detecting Offensive Language with Annotation Disagreement
von: Lu, Junyu, et al.
Veröffentlicht: (2025) -
ARGUS: Policy-Adaptive Ad Governance via Evolving Reinforcement with Adversarial Umpiring
von: Ji, Deyi, et al.
Veröffentlicht: (2026) -
Self-signals Driven Multi-LLM Debate for Efficient and Accurate Reasoning
von: Chen, Xuhang, et al.
Veröffentlicht: (2025) -
Integrating Multi-view Analysis: Multi-view Mixture-of-Expert for Textual Personality Detection
von: Zhu, Haohao, et al.
Veröffentlicht: (2024) -
DiscoUQ: Structured Disagreement Analysis for Uncertainty Quantification in LLM Agent Ensembles
von: Jiang, Bo
Veröffentlicht: (2026)