The Relationship Between Reasoning and Performance in Large Language Models -- o3 (mini) Thinks Harder, Not Longer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ballon, Marthe, Algaba, Andres, Ginis, Vincent |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Probing the Trajectories of Reasoning Traces in Large Language Models
von: Ballon, Marthe, et al.
Veröffentlicht: (2026)
von: Ballon, Marthe, et al.
Veröffentlicht: (2026)
Estimating problem difficulty without ground truth using Large Language Model comparisons
von: Ballon, Marthe, et al.
Veröffentlicht: (2025)
von: Ballon, Marthe, et al.
Veröffentlicht: (2025)
Benchmarks Saturate When The Model Gets Smarter Than The Judge
von: Ballon, Marthe, et al.
Veröffentlicht: (2026)
von: Ballon, Marthe, et al.
Veröffentlicht: (2026)
Flexible Counterfactual Explanations with Generative Models
von: Hellemans, Stig, et al.
Veröffentlicht: (2025)
von: Hellemans, Stig, et al.
Veröffentlicht: (2025)
Early Evidence of Vibe-Proving with Consumer LLMs: A Case Study on Spectral Region Characterization with ChatGPT-5.2 (Thinking)
von: Verbeken, Brecht, et al.
Veröffentlicht: (2026)
von: Verbeken, Brecht, et al.
Veröffentlicht: (2026)
Large Language Models Reflect Human Citation Patterns with a Heightened Citation Bias
von: Algaba, Andres, et al.
Veröffentlicht: (2024)
von: Algaba, Andres, et al.
Veröffentlicht: (2024)
How Deep Do Large Language Models Internalize Scientific Literature and Citation Practices?
von: Algaba, Andres, et al.
Veröffentlicht: (2025)
von: Algaba, Andres, et al.
Veröffentlicht: (2025)
Lexical Hints of Accuracy in LLM Reasoning Chains
von: Vanhoyweghen, Arne, et al.
Veröffentlicht: (2025)
von: Vanhoyweghen, Arne, et al.
Veröffentlicht: (2025)
Longer Context, Deeper Thinking: Uncovering the Role of Long-Context Ability in Reasoning
von: Yang, Wang, et al.
Veröffentlicht: (2025)
von: Yang, Wang, et al.
Veröffentlicht: (2025)
AdaReasoner: Adaptive Reasoning Enables More Flexible Thinking in Large Language Models
von: Wang, Xiangqi, et al.
Veröffentlicht: (2025)
von: Wang, Xiangqi, et al.
Veröffentlicht: (2025)
Thinking Deeper, Not Longer: Depth-Recurrent Transformers for Compositional Generalization
von: Chen, Hung-Hsuan
Veröffentlicht: (2026)
von: Chen, Hung-Hsuan
Veröffentlicht: (2026)
DAST: Difficulty-Adaptive Slow-Thinking for Large Reasoning Models
von: Shen, Yi, et al.
Veröffentlicht: (2025)
von: Shen, Yi, et al.
Veröffentlicht: (2025)
To Think or Not to Think: Exploring the Unthinking Vulnerability in Large Reasoning Models
von: Zhu, Zihao, et al.
Veröffentlicht: (2025)
von: Zhu, Zihao, et al.
Veröffentlicht: (2025)
Understanding Reasoning in Thinking Language Models via Steering Vectors
von: Venhoff, Constantin, et al.
Veröffentlicht: (2025)
von: Venhoff, Constantin, et al.
Veröffentlicht: (2025)
AdapThink: Adaptive Thinking Preferences for Reasoning Language Model
von: Wan, Xu, et al.
Veröffentlicht: (2025)
von: Wan, Xu, et al.
Veröffentlicht: (2025)
Probing Graph Neural Network Activation Patterns Through Graph Topology
von: Tori, Floriano, et al.
Veröffentlicht: (2026)
von: Tori, Floriano, et al.
Veröffentlicht: (2026)
Structurally Human, Semantically Biased: Detecting LLM-Generated References with Embeddings and GNNs
von: Mobini, Melika, et al.
Veröffentlicht: (2026)
von: Mobini, Melika, et al.
Veröffentlicht: (2026)
h1: Bootstrapping LLMs to Reason over Longer Horizons via Reinforcement Learning
von: Motwani, Sumeet Ramesh, et al.
Veröffentlicht: (2025)
von: Motwani, Sumeet Ramesh, et al.
Veröffentlicht: (2025)
Near-Lossless Model Compression Enables Longer Context Inference in DNA Large Language Models
von: Zhu, Rui, et al.
Veröffentlicht: (2025)
von: Zhu, Rui, et al.
Veröffentlicht: (2025)
Dynamic Thinking-Token Selection for Efficient Reasoning in Large Reasoning Models
von: Guo, Zhenyuan, et al.
Veröffentlicht: (2026)
von: Guo, Zhenyuan, et al.
Veröffentlicht: (2026)
Can Large Reasoning Models Improve Accuracy on Mathematical Tasks Using Flawed Thinking?
von: Amjith, Saraswathy, et al.
Veröffentlicht: (2025)
von: Amjith, Saraswathy, et al.
Veröffentlicht: (2025)
Don't Think Longer, Think Wisely: Optimizing Thinking Dynamics for Large Reasoning Models
von: An, Sohyun, et al.
Veröffentlicht: (2025)
von: An, Sohyun, et al.
Veröffentlicht: (2025)
Thinking Forward and Backward: Effective Backward Planning with Large Language Models
von: Ren, Allen Z., et al.
Veröffentlicht: (2024)
von: Ren, Allen Z., et al.
Veröffentlicht: (2024)
Thinking About Thinking: SAGE-nano's Inverse Reasoning for Self-Aware Language Models
von: Jha, Basab, et al.
Veröffentlicht: (2025)
von: Jha, Basab, et al.
Veröffentlicht: (2025)
Extending Pretrained 10-Second ECG Foundation Models to Longer Horizons
von: Tang, Wei, et al.
Veröffentlicht: (2026)
von: Tang, Wei, et al.
Veröffentlicht: (2026)
MeTHanol: Modularized Thinking Language Models with Intermediate Layer Thinking, Decoding and Bootstrapping Reasoning
von: Xi, Ningyuan, et al.
Veröffentlicht: (2024)
von: Xi, Ningyuan, et al.
Veröffentlicht: (2024)
DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models
von: Yin, Cheng, et al.
Veröffentlicht: (2025)
von: Yin, Cheng, et al.
Veröffentlicht: (2025)
Are Large-Language Models Graph Algorithmic Reasoners?
von: Taylor, Alexander K, et al.
Veröffentlicht: (2024)
von: Taylor, Alexander K, et al.
Veröffentlicht: (2024)
Base Models Know How to Reason, Thinking Models Learn When
von: Venhoff, Constantin, et al.
Veröffentlicht: (2025)
von: Venhoff, Constantin, et al.
Veröffentlicht: (2025)
Improving Reasoning Performance in Large Language Models via Representation Engineering
von: Højer, Bertram, et al.
Veröffentlicht: (2025)
von: Højer, Bertram, et al.
Veröffentlicht: (2025)
AdaptThink: Reasoning Models Can Learn When to Think
von: Zhang, Jiajie, et al.
Veröffentlicht: (2025)
von: Zhang, Jiajie, et al.
Veröffentlicht: (2025)
Interpretable Physics Reasoning and Performance Taxonomy in Vision-Language Models
von: Pawar, Pranav, et al.
Veröffentlicht: (2025)
von: Pawar, Pranav, et al.
Veröffentlicht: (2025)
Debug Smarter, Not Harder: AI Agents for Error Resolution in Computational Notebooks
von: Grotov, Konstantin, et al.
Veröffentlicht: (2024)
von: Grotov, Konstantin, et al.
Veröffentlicht: (2024)
Temperature-Dependent Performance of Prompting Strategies in Extended Reasoning Large Language Models
von: Salah, Mousa, et al.
Veröffentlicht: (2026)
von: Salah, Mousa, et al.
Veröffentlicht: (2026)
BRiTE: Bootstrapping Reinforced Thinking Process to Enhance Language Model Reasoning
von: Zhong, Han, et al.
Veröffentlicht: (2025)
von: Zhong, Han, et al.
Veröffentlicht: (2025)
ReasoningWeekly: A General Knowledge and Verbal Reasoning Challenge for Large Language Models
von: Wu, Zixuan, et al.
Veröffentlicht: (2025)
von: Wu, Zixuan, et al.
Veröffentlicht: (2025)
Meta-Reasoner: Dynamic Guidance for Optimized Inference-time Reasoning in Large Language Models
von: Sui, Yuan, et al.
Veröffentlicht: (2025)
von: Sui, Yuan, et al.
Veröffentlicht: (2025)
DecepChain: Inducing Deceptive Reasoning in Large Language Models
von: Shen, Wei, et al.
Veröffentlicht: (2025)
von: Shen, Wei, et al.
Veröffentlicht: (2025)
Optimal Self-Consistency for Efficient Reasoning with Large Language Models
von: Feng, Austin, et al.
Veröffentlicht: (2025)
von: Feng, Austin, et al.
Veröffentlicht: (2025)
Reasoning-Enhanced Large Language Models for Molecular Property Prediction
von: Zhuang, Jiaxi, et al.
Veröffentlicht: (2025)
von: Zhuang, Jiaxi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Probing the Trajectories of Reasoning Traces in Large Language Models
von: Ballon, Marthe, et al.
Veröffentlicht: (2026) -
Estimating problem difficulty without ground truth using Large Language Model comparisons
von: Ballon, Marthe, et al.
Veröffentlicht: (2025) -
Benchmarks Saturate When The Model Gets Smarter Than The Judge
von: Ballon, Marthe, et al.
Veröffentlicht: (2026) -
Flexible Counterfactual Explanations with Generative Models
von: Hellemans, Stig, et al.
Veröffentlicht: (2025) -
Early Evidence of Vibe-Proving with Consumer LLMs: A Case Study on Spectral Region Characterization with ChatGPT-5.2 (Thinking)
von: Verbeken, Brecht, et al.
Veröffentlicht: (2026)