Beg to Differ: Understanding Reasoning-Answer Misalignment Across Languages
Fuente:
arXiv
Salvato in:
| Autori principali: | Ovalle, Anaelia, Ross, Candace, Ruder, Sebastian, Williams, Adina, Ullrich, Karen, Ibrahim, Mark, Sagun, Levent |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The Root Shapes the Fruit: On the Persistence of Gender-Exclusive Harms in Aligned Language Models
di: Ovalle, Anaelia, et al.
Pubblicazione: (2024)
di: Ovalle, Anaelia, et al.
Pubblicazione: (2024)
Changing Answer Order Can Decrease MMLU Accuracy
di: Gupta, Vipul, et al.
Pubblicazione: (2024)
di: Gupta, Vipul, et al.
Pubblicazione: (2024)
What's in Common? Multimodal Models Hallucinate When Reasoning Across Scenes
di: Ross, Candace, et al.
Pubblicazione: (2025)
di: Ross, Candace, et al.
Pubblicazione: (2025)
Chained Tuning Leads to Biased Forgetting
di: Ung, Megan, et al.
Pubblicazione: (2024)
di: Ung, Megan, et al.
Pubblicazione: (2024)
What makes a good metric? Evaluating automatic metrics for text-to-image consistency
di: Ross, Candace, et al.
Pubblicazione: (2024)
di: Ross, Candace, et al.
Pubblicazione: (2024)
Arbiters of Ambivalence: Challenges of Using LLMs in No-Consensus Tasks
di: Radharapu, Bhaktipriya, et al.
Pubblicazione: (2025)
di: Radharapu, Bhaktipriya, et al.
Pubblicazione: (2025)
LLM Knowledge is Brittle: Truthfulness Representations Rely on Superficial Resemblance
di: Haller, Patrick, et al.
Pubblicazione: (2025)
di: Haller, Patrick, et al.
Pubblicazione: (2025)
Understanding and Mitigating Language Confusion in LLMs
di: Marchisio, Kelly, et al.
Pubblicazione: (2024)
di: Marchisio, Kelly, et al.
Pubblicazione: (2024)
Task-Dependent Evaluation of LLM Output Homogenization: A Taxonomy-Guided Framework
di: Jain, Shomik, et al.
Pubblicazione: (2025)
di: Jain, Shomik, et al.
Pubblicazione: (2025)
Improving Model Evaluation using SMART Filtering of Benchmark Datasets
di: Gupta, Vipul, et al.
Pubblicazione: (2024)
di: Gupta, Vipul, et al.
Pubblicazione: (2024)
Learning the Wrong Lessons: Syntactic-Domain Spurious Correlations in Language Models
di: Shaib, Chantal, et al.
Pubblicazione: (2025)
di: Shaib, Chantal, et al.
Pubblicazione: (2025)
Are Female Carpenters like Blue Bananas? A Corpus Investigation of Occupation Gender Typicality
di: Ju, Da, et al.
Pubblicazione: (2024)
di: Ju, Da, et al.
Pubblicazione: (2024)
Refining Answer Distributions for Improved Large Language Model Reasoning
di: Pal, Soumyasundar, et al.
Pubblicazione: (2024)
di: Pal, Soumyasundar, et al.
Pubblicazione: (2024)
Understanding and Mitigating Tokenization Bias in Language Models
di: Phan, Buu, et al.
Pubblicazione: (2024)
di: Phan, Buu, et al.
Pubblicazione: (2024)
Can Large Language Models Understand, Reason About, and Generate Code-Switched Text?
di: Winata, Genta Indra, et al.
Pubblicazione: (2026)
di: Winata, Genta Indra, et al.
Pubblicazione: (2026)
A Single Character can Make or Break Your LLM Evals
di: Su, Jingtong, et al.
Pubblicazione: (2025)
di: Su, Jingtong, et al.
Pubblicazione: (2025)
Improving Text-to-Image Consistency via Automatic Prompt Optimization
di: Mañas, Oscar, et al.
Pubblicazione: (2024)
di: Mañas, Oscar, et al.
Pubblicazione: (2024)
[Call for Papers] The 2nd BabyLM Challenge: Sample-efficient pretraining on a developmentally plausible corpus
di: Choshen, Leshem, et al.
Pubblicazione: (2024)
di: Choshen, Leshem, et al.
Pubblicazione: (2024)
Brittlebench: Quantifying LLM robustness via prompt sensitivity
di: Romanou, Angelika, et al.
Pubblicazione: (2026)
di: Romanou, Angelika, et al.
Pubblicazione: (2026)
Findings of the Second BabyLM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora
di: Hu, Michael Y., et al.
Pubblicazione: (2024)
di: Hu, Michael Y., et al.
Pubblicazione: (2024)
Language and Task Arithmetic with Parameter-Efficient Layers for Zero-Shot Summarization
di: Chronopoulou, Alexandra, et al.
Pubblicazione: (2023)
di: Chronopoulou, Alexandra, et al.
Pubblicazione: (2023)
Cross-Tokenizer Likelihood Scoring Algorithms for Language Model Distillation
di: Phan, Buu, et al.
Pubblicazione: (2025)
di: Phan, Buu, et al.
Pubblicazione: (2025)
Reasoning over mathematical objects: on-policy reward modeling and test time aggregation
di: Aggarwal, Pranjal, et al.
Pubblicazione: (2026)
di: Aggarwal, Pranjal, et al.
Pubblicazione: (2026)
On the Role of Speech Data in Reducing Toxicity Detection Bias
di: Bell, Samuel J., et al.
Pubblicazione: (2024)
di: Bell, Samuel J., et al.
Pubblicazione: (2024)
MENLO: From Preferences to Proficiency -- Evaluating and Modeling Native-like Quality Across 47 Languages
di: Whitehouse, Chenxi, et al.
Pubblicazione: (2025)
di: Whitehouse, Chenxi, et al.
Pubblicazione: (2025)
Tokenization Matters: Navigating Data-Scarce Tokenization for Gender Inclusive Language Technologies
di: Ovalle, Anaelia, et al.
Pubblicazione: (2023)
di: Ovalle, Anaelia, et al.
Pubblicazione: (2023)
AL-QASIDA: Analyzing LLM Quality and Accuracy Systematically in Dialectal Arabic
di: Robinson, Nathaniel R., et al.
Pubblicazione: (2024)
di: Robinson, Nathaniel R., et al.
Pubblicazione: (2024)
The Factorization Curse: Which Tokens You Predict Underlie the Reversal Curse and More
di: Kitouni, Ouail, et al.
Pubblicazione: (2024)
di: Kitouni, Ouail, et al.
Pubblicazione: (2024)
A Post-trainer's Guide to Multilingual Training Data: Uncovering Cross-lingual Transfer Dynamics
di: Shimabucoro, Luisa, et al.
Pubblicazione: (2025)
di: Shimabucoro, Luisa, et al.
Pubblicazione: (2025)
A Discriminative Latent-Variable Model for Bilingual Lexicon Induction
di: Ruder, Sebastian, et al.
Pubblicazione: (2018)
di: Ruder, Sebastian, et al.
Pubblicazione: (2018)
Domain Regeneration: How well do LLMs match syntactic properties of text domains?
di: Ju, Da, et al.
Pubblicazione: (2025)
di: Ju, Da, et al.
Pubblicazione: (2025)
Learn Beyond The Answer: Training Language Models with Reflection for Mathematical Reasoning
di: Zhang, Zhihan, et al.
Pubblicazione: (2024)
di: Zhang, Zhihan, et al.
Pubblicazione: (2024)
Assessing the Reliability and Validity of GPT-4 in Annotating Emotion Appraisal Ratings
di: Ruder, Deniss, et al.
Pubblicazione: (2025)
di: Ruder, Deniss, et al.
Pubblicazione: (2025)
Sandwich Reasoning: An Answer-Reasoning-Answer Approach for Low-Latency Query Correction
di: Zhang, Chen, et al.
Pubblicazione: (2026)
di: Zhang, Chen, et al.
Pubblicazione: (2026)
When Thinking Backfires: Mechanistic Insights Into Reasoning-Induced Misalignment
di: Yan, Hanqi, et al.
Pubblicazione: (2025)
di: Yan, Hanqi, et al.
Pubblicazione: (2025)
From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought
di: Tan, Wentao, et al.
Pubblicazione: (2025)
di: Tan, Wentao, et al.
Pubblicazione: (2025)
Weisfeiler and Leman Go Measurement Modeling: Probing the Validity of the WL Test
di: Subramonian, Arjun, et al.
Pubblicazione: (2023)
di: Subramonian, Arjun, et al.
Pubblicazione: (2023)
Are Aligned Large Language Models Still Misaligned?
di: Naseem, Usman, et al.
Pubblicazione: (2026)
di: Naseem, Usman, et al.
Pubblicazione: (2026)
Parallel-SFT: Improving Zero-Shot Cross-Programming-Language Transfer for Code RL
di: Wu, Zhaofeng, et al.
Pubblicazione: (2026)
di: Wu, Zhaofeng, et al.
Pubblicazione: (2026)
Disparities In Negation Understanding Across Languages In Vision-Language Models
di: Moraitaki, Charikleia, et al.
Pubblicazione: (2026)
di: Moraitaki, Charikleia, et al.
Pubblicazione: (2026)
Documenti analoghi
-
The Root Shapes the Fruit: On the Persistence of Gender-Exclusive Harms in Aligned Language Models
di: Ovalle, Anaelia, et al.
Pubblicazione: (2024) -
Changing Answer Order Can Decrease MMLU Accuracy
di: Gupta, Vipul, et al.
Pubblicazione: (2024) -
What's in Common? Multimodal Models Hallucinate When Reasoning Across Scenes
di: Ross, Candace, et al.
Pubblicazione: (2025) -
Chained Tuning Leads to Biased Forgetting
di: Ung, Megan, et al.
Pubblicazione: (2024) -
What makes a good metric? Evaluating automatic metrics for text-to-image consistency
di: Ross, Candace, et al.
Pubblicazione: (2024)