When Does Reasoning Matter? A Controlled Study of Reasoning's Contribution to Model Performance
Fuente:
arXiv
Saved in:
| Main Authors: | Boizard, Nicolas, Gisserot-Boukhlef, Hippolyte, El-Haddad, Kevin, Hudelot, Céline, Colombo, Pierre |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BERT-as-a-Judge: A Robust Alternative to Lexical Methods for Efficient Reference-Based LLM Evaluation
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2026)
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2026)
BidirLM: From Text to Omnimodal Bidirectional Encoders by Adapting and Composing Causal LLMs
by: Boizard, Nicolas, et al.
Published: (2026)
by: Boizard, Nicolas, et al.
Published: (2026)
Towards Trustworthy Reranking: A Simple yet Effective Abstention Mechanism
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2024)
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2024)
Towards Cross-Tokenizer Distillation: the Universal Logit Distillation Loss for LLMs
by: Boizard, Nicolas, et al.
Published: (2024)
by: Boizard, Nicolas, et al.
Published: (2024)
Should We Still Pretrain Encoders with Masked Language Modeling?
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2025)
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2025)
Is Preference Alignment Always the Best Option to Enhance LLM-Based Translation? An Empirical Analysis
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2024)
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2024)
EuroBERT: Scaling Multilingual Encoders for European Languages
by: Boizard, Nicolas, et al.
Published: (2025)
by: Boizard, Nicolas, et al.
Published: (2025)
EuroLLM-22B: Technical Report
by: Ramos, Miguel Moura, et al.
Published: (2026)
by: Ramos, Miguel Moura, et al.
Published: (2026)
How Does Prefix Matter in Reasoning Model Tuning?
by: Tomar, Raj Vardhan, et al.
Published: (2026)
by: Tomar, Raj Vardhan, et al.
Published: (2026)
ColPali: Efficient Document Retrieval with Vision Language Models
by: Faysse, Manuel, et al.
Published: (2024)
by: Faysse, Manuel, et al.
Published: (2024)
ConceptGuard: Neuro-Symbolic Safety Guardrails via Sparse Interpretable Jailbreak Concepts
by: Aswal, Darpan, et al.
Published: (2025)
by: Aswal, Darpan, et al.
Published: (2025)
CroissantLLM: A Truly Bilingual French-English Language Model
by: Faysse, Manuel, et al.
Published: (2024)
by: Faysse, Manuel, et al.
Published: (2024)
Curate-Train-Refine: A Closed-Loop Agentic Framework for Zero Shot Classification
by: Maheshwari, Gaurav, et al.
Published: (2026)
by: Maheshwari, Gaurav, et al.
Published: (2026)
When Silence Matters: The Impact of Irrelevant Audio on Text Reasoning in Large Audio-Language Models
by: Li, Chen-An, et al.
Published: (2025)
by: Li, Chen-An, et al.
Published: (2025)
Does Table Source Matter? Benchmarking and Improving Multimodal Scientific Table Understanding and Reasoning
by: Yang, Bohao, et al.
Published: (2025)
by: Yang, Bohao, et al.
Published: (2025)
Stop When Reasoning Converges: Semantic-Preserving Early Exit for Reasoning Models
by: Min, Dehai, et al.
Published: (2026)
by: Min, Dehai, et al.
Published: (2026)
A Complexity Map of Probabilistic Reasoning for Neurosymbolic Classification Techniques
by: Ledaguenel, Arthur, et al.
Published: (2024)
by: Ledaguenel, Arthur, et al.
Published: (2024)
When Models Reason in Your Language: Controlling Thinking Language Comes at the Cost of Accuracy
by: Qi, Jirui, et al.
Published: (2025)
by: Qi, Jirui, et al.
Published: (2025)
Order Doesn't Matter, But Reasoning Does: Training LLMs with Order-Centric Augmentation
by: He, Qianxi, et al.
Published: (2025)
by: He, Qianxi, et al.
Published: (2025)
Does Rationale Quality Matter? Enhancing Mental Disorder Detection via Selective Reasoning Distillation
by: Song, Hoyun, et al.
Published: (2025)
by: Song, Hoyun, et al.
Published: (2025)
Does Reasoning Introduce Bias? A Study of Social Bias Evaluation and Mitigation in LLM Reasoning
by: Wu, Xuyang, et al.
Published: (2025)
by: Wu, Xuyang, et al.
Published: (2025)
TextReasoningBench: Does Reasoning Really Improve Text Classification in Large Language Models?
by: Guo, Xinyu, et al.
Published: (2026)
by: Guo, Xinyu, et al.
Published: (2026)
When and Why Does Unsupervised RL Succeed in Mathematical Reasoning? A Manifold Envelopment Perspective
by: Zhang, Zelin, et al.
Published: (2026)
by: Zhang, Zelin, et al.
Published: (2026)
Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know?
by: Mei, Zhiting, et al.
Published: (2025)
by: Mei, Zhiting, et al.
Published: (2025)
Does It Tie Out? Towards Autonomous Legal Agents in Venture Capital
by: Colombo, Pierre, et al.
Published: (2025)
by: Colombo, Pierre, et al.
Published: (2025)
Language Matters: How Do Multilingual Input and Reasoning Paths Affect Large Reasoning Models?
by: Tam, Zhi Rui, et al.
Published: (2025)
by: Tam, Zhi Rui, et al.
Published: (2025)
Question-Analysis Prompting Improves LLM Performance in Reasoning Tasks
by: Yugeswardeenoo, Dharunish, et al.
Published: (2024)
by: Yugeswardeenoo, Dharunish, et al.
Published: (2024)
Semantic Deception: When Reasoning Models Can't Compute an Addition
by: de Leeuw, Nathaniël, et al.
Published: (2025)
by: de Leeuw, Nathaniël, et al.
Published: (2025)
Making Reasoning Matter: Measuring and Improving Faithfulness of Chain-of-Thought Reasoning
by: Paul, Debjit, et al.
Published: (2024)
by: Paul, Debjit, et al.
Published: (2024)
Revisiting Anisotropy in Language Transformers: The Geometry of Learning Dynamics
by: Bernas, Raphael, et al.
Published: (2026)
by: Bernas, Raphael, et al.
Published: (2026)
TEGRA: Text Encoding With Graph and Retrieval Augmentation for Misinformation Detection
by: Faye, Géraud, et al.
Published: (2026)
by: Faye, Géraud, et al.
Published: (2026)
When to Think, When to Speak: Learning Disclosure Policies for LLM Reasoning
by: Wei, Jiaqi, et al.
Published: (2026)
by: Wei, Jiaqi, et al.
Published: (2026)
Think Only When You Need with Large Hybrid-Reasoning Models
by: Jiang, Lingjie, et al.
Published: (2025)
by: Jiang, Lingjie, et al.
Published: (2025)
StreaMulT: Streaming Multimodal Transformer for Heterogeneous and Arbitrary Long Sequential Data
by: Pellegrain, Victor, et al.
Published: (2021)
by: Pellegrain, Victor, et al.
Published: (2021)
When Can Large Reasoning Models Save Thinking? Mechanistic Analysis of Behavioral Divergence in Reasoning
by: Zhu, Rongzhi, et al.
Published: (2025)
by: Zhu, Rongzhi, et al.
Published: (2025)
Efficacy of Synthetic Data as a Benchmark
by: Maheshwari, Gaurav, et al.
Published: (2024)
by: Maheshwari, Gaurav, et al.
Published: (2024)
Reasoning Pattern Matters: Learning to Reason without Human Rationales
by: Pang, Chaoxu, et al.
Published: (2025)
by: Pang, Chaoxu, et al.
Published: (2025)
Reasoning Topology Matters: Network-of-Thought for Complex Reasoning Tasks
by: Huang, Fan
Published: (2026)
by: Huang, Fan
Published: (2026)
Can Large Language Models Create New Knowledge for Spatial Reasoning Tasks?
by: Greatrix, Thomas, et al.
Published: (2024)
by: Greatrix, Thomas, et al.
Published: (2024)
Premise Order Matters in Reasoning with Large Language Models
by: Chen, Xinyun, et al.
Published: (2024)
by: Chen, Xinyun, et al.
Published: (2024)
Similar Items
-
BERT-as-a-Judge: A Robust Alternative to Lexical Methods for Efficient Reference-Based LLM Evaluation
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2026) -
BidirLM: From Text to Omnimodal Bidirectional Encoders by Adapting and Composing Causal LLMs
by: Boizard, Nicolas, et al.
Published: (2026) -
Towards Trustworthy Reranking: A Simple yet Effective Abstention Mechanism
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2024) -
Towards Cross-Tokenizer Distillation: the Universal Logit Distillation Loss for LLMs
by: Boizard, Nicolas, et al.
Published: (2024) -
Should We Still Pretrain Encoders with Masked Language Modeling?
by: Gisserot-Boukhlef, Hippolyte, et al.
Published: (2025)