Decoupling the Effect of Chain-of-Thought Reasoning: A Human Label Variation Perspective
Fuente:
arXiv
Guardado en:
| Autores principales: | Chen, Beiduo, Hu, Tiancheng, Zhang, Caiqi, Litschko, Robert, Korhonen, Anna, Plank, Barbara |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Threading the Needle: Reweaving Chain-of-Thought Reasoning to Explain Human Label Variation
por: Chen, Beiduo, et al.
Publicado: (2025)
por: Chen, Beiduo, et al.
Publicado: (2025)
"Seeing the Big through the Small": Can LLMs Approximate Human Judgment Distributions on NLI from a Few Explanations?
por: Chen, Beiduo, et al.
Publicado: (2024)
por: Chen, Beiduo, et al.
Publicado: (2024)
A Rose by Any Other Name: LLM-Generated Explanations Are Good Proxies for Human Explanations to Collect Label Distributions on NLI
por: Chen, Beiduo, et al.
Publicado: (2024)
por: Chen, Beiduo, et al.
Publicado: (2024)
Human Label Variation as Stable Signal: Learning Annotator-Specific Explanation Behavior via Cross-Annotator Preference Optimization
por: Chen, Beiduo, et al.
Publicado: (2026)
por: Chen, Beiduo, et al.
Publicado: (2026)
Reasoning that Travels: Dissecting How Chain-of-Thought Transfers Across Models
por: Cheng, Xinyuan, et al.
Publicado: (2026)
por: Cheng, Xinyuan, et al.
Publicado: (2026)
Agree, Disagree, Explain: Decomposing Human Label Variation in NLI through the Lens of Explanations
por: Hong, Pingjun, et al.
Publicado: (2025)
por: Hong, Pingjun, et al.
Publicado: (2025)
LiTEx: A Linguistic Taxonomy of Explanations for Understanding Within-Label Variation in Natural Language Inference
por: Hong, Pingjun, et al.
Publicado: (2025)
por: Hong, Pingjun, et al.
Publicado: (2025)
Resource-Lean Lexicon Induction for German Dialects
por: Litschko, Robert, et al.
Publicado: (2026)
por: Litschko, Robert, et al.
Publicado: (2026)
Make Every Letter Count: Building Dialect Variation Dictionaries from Monolingual Corpora
por: Litschko, Robert, et al.
Publicado: (2025)
por: Litschko, Robert, et al.
Publicado: (2025)
MaiNLP at SemEval-2024 Task 1: Analyzing Source Language Selection in Cross-Lingual Textual Relatedness
por: Zhou, Shijia, et al.
Publicado: (2024)
por: Zhou, Shijia, et al.
Publicado: (2024)
Reason to Rote: Rethinking Memorization in Reasoning
por: Du, Yupei, et al.
Publicado: (2025)
por: Du, Yupei, et al.
Publicado: (2025)
Beyond the Final Layer: Intermediate Representations for Better Multilingual Calibration in Large Language Models
por: Zhou, Ej, et al.
Publicado: (2025)
por: Zhou, Ej, et al.
Publicado: (2025)
Cross-Dialect Information Retrieval: Information Access in Low-Resource and High-Variance Languages
por: Litschko, Robert, et al.
Publicado: (2024)
por: Litschko, Robert, et al.
Publicado: (2024)
MAKIEval: A Multilingual Automatic WiKidata-based Framework for Cultural Awareness Evaluation for LLMs
por: Zhao, Raoyuan, et al.
Publicado: (2025)
por: Zhao, Raoyuan, et al.
Publicado: (2025)
Donkii: Can Annotation Error Detection Methods Find Errors in Instruction-Tuning Datasets?
por: Weber-Genzel, Leon, et al.
Publicado: (2023)
por: Weber-Genzel, Leon, et al.
Publicado: (2023)
Information Asymmetry across Language Varieties: A Case Study on Cantonese-Mandarin and Bavarian-German QA
por: Pei, Renhao, et al.
Publicado: (2026)
por: Pei, Renhao, et al.
Publicado: (2026)
Interpreting Predictive Probabilities: Model Confidence or Human Label Variation?
por: Baan, Joris, et al.
Publicado: (2024)
por: Baan, Joris, et al.
Publicado: (2024)
Disagreeing Rationales: Rethinking Classification and Explainability Evaluation in Hate Speech Detection
por: Muscato, Benedetta, et al.
Publicado: (2026)
por: Muscato, Benedetta, et al.
Publicado: (2026)
Different Tastes of Entities: Investigating Human Label Variation in Named Entity Annotations
por: Peng, Siyao, et al.
Publicado: (2024)
por: Peng, Siyao, et al.
Publicado: (2024)
Evaluating Large Language Models for Cross-Lingual Retrieval
por: Zuo, Longfei, et al.
Publicado: (2025)
por: Zuo, Longfei, et al.
Publicado: (2025)
Exploring Large Language Models for Product Attribute Value Identification
por: Sabeh, Kassem, et al.
Publicado: (2024)
por: Sabeh, Kassem, et al.
Publicado: (2024)
An Empirical Comparison of Generative Approaches for Product Attribute-Value Identification
por: Sabeh, Kassem, et al.
Publicado: (2024)
por: Sabeh, Kassem, et al.
Publicado: (2024)
To Know or Not To Know? Analyzing Self-Consistency of Large Language Models under Ambiguity
por: Sedova, Anastasiia, et al.
Publicado: (2024)
por: Sedova, Anastasiia, et al.
Publicado: (2024)
From Noise to Signal to Selbstzweck: Reframing Human Label Variation in the Era of Post-training in NLP
por: Xu, Shanshan, et al.
Publicado: (2025)
por: Xu, Shanshan, et al.
Publicado: (2025)
VariErr NLI: Separating Annotation Error from Human Label Variation
por: Weber-Genzel, Leon, et al.
Publicado: (2024)
por: Weber-Genzel, Leon, et al.
Publicado: (2024)
Revisiting Active Learning under (Human) Label Variation
por: Gruber, Cornelia, et al.
Publicado: (2025)
por: Gruber, Cornelia, et al.
Publicado: (2025)
TopViewRS: Vision-Language Models as Top-View Spatial Reasoners
por: Li, Chengzu, et al.
Publicado: (2024)
por: Li, Chengzu, et al.
Publicado: (2024)
When Meanings Meet: Investigating the Emergence and Quality of Shared Concept Spaces during Multilingual Language Model Training
por: Körner, Felicia, et al.
Publicado: (2026)
por: Körner, Felicia, et al.
Publicado: (2026)
Value of Information: A Framework for Human-Agent Communication
por: Dong, Yijiang River, et al.
Publicado: (2026)
por: Dong, Yijiang River, et al.
Publicado: (2026)
Latent Chain-of-Thought as Planning: Decoupling Reasoning from Verbalization
por: Wang, Jiecong, et al.
Publicado: (2026)
por: Wang, Jiecong, et al.
Publicado: (2026)
Comparing Inferential Strategies of Humans and Large Language Models in Deductive Reasoning
por: Mondorf, Philipp, et al.
Publicado: (2024)
por: Mondorf, Philipp, et al.
Publicado: (2024)
Putting on the Thinking Hats: A Survey on Chain of Thought Fine-tuning from the Perspective of Human Reasoning Mechanism
por: Chen, Xiaoshu, et al.
Publicado: (2025)
por: Chen, Xiaoshu, et al.
Publicado: (2025)
Reinforcing the Diffusion Chain of Lateral Thought with Diffusion Language Models
por: Huang, Zemin, et al.
Publicado: (2025)
por: Huang, Zemin, et al.
Publicado: (2025)
Liar, Liar, Logical Mire: A Benchmark for Suppositional Reasoning in Large Language Models
por: Mondorf, Philipp, et al.
Publicado: (2024)
por: Mondorf, Philipp, et al.
Publicado: (2024)
Making Reasoning Matter: Measuring and Improving Faithfulness of Chain-of-Thought Reasoning
por: Paul, Debjit, et al.
Publicado: (2024)
por: Paul, Debjit, et al.
Publicado: (2024)
Reasoning Beyond Language: A Comprehensive Survey on Latent Chain-of-Thought Reasoning
por: Chen, Xinghao, et al.
Publicado: (2025)
por: Chen, Xinghao, et al.
Publicado: (2025)
ReGuLaR: Variational Latent Reasoning Guided by Rendered Chain-of-Thought
por: Wang, Fanmeng, et al.
Publicado: (2026)
por: Wang, Fanmeng, et al.
Publicado: (2026)
Visual Thoughts: A Unified Perspective of Understanding Multimodal Chain-of-Thought
por: Cheng, Zihui, et al.
Publicado: (2025)
por: Cheng, Zihui, et al.
Publicado: (2025)
Beyond Accuracy: Evaluating the Reasoning Behavior of Large Language Models -- A Survey
por: Mondorf, Philipp, et al.
Publicado: (2024)
por: Mondorf, Philipp, et al.
Publicado: (2024)
Efficient Reasoning via Chain of Unconscious Thought
por: Gong, Ruihan, et al.
Publicado: (2025)
por: Gong, Ruihan, et al.
Publicado: (2025)
Ejemplares similares
-
Threading the Needle: Reweaving Chain-of-Thought Reasoning to Explain Human Label Variation
por: Chen, Beiduo, et al.
Publicado: (2025) -
"Seeing the Big through the Small": Can LLMs Approximate Human Judgment Distributions on NLI from a Few Explanations?
por: Chen, Beiduo, et al.
Publicado: (2024) -
A Rose by Any Other Name: LLM-Generated Explanations Are Good Proxies for Human Explanations to Collect Label Distributions on NLI
por: Chen, Beiduo, et al.
Publicado: (2024) -
Human Label Variation as Stable Signal: Learning Annotator-Specific Explanation Behavior via Cross-Annotator Preference Optimization
por: Chen, Beiduo, et al.
Publicado: (2026) -
Reasoning that Travels: Dissecting How Chain-of-Thought Transfers Across Models
por: Cheng, Xinyuan, et al.
Publicado: (2026)