Large Language Models' Reasoning Stalls: An Investigation into the Capabilities of Frontier Models
Fuente:
arXiv
Saved in:
| Main Authors: | McGinness, Lachlan, Baumgartner, Peter |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Large Language Models Imitate Logical Reasoning, but at what Cost?
by: McGinness, Lachlan, et al.
Published: (2025)
by: McGinness, Lachlan, et al.
Published: (2025)
Steamroller Problems: An Evaluation of LLM Reasoning Capability with Automated Theorem Prover Strategies
by: McGinness, Lachlan, et al.
Published: (2024)
by: McGinness, Lachlan, et al.
Published: (2024)
Automated Theorem Provers Help Improve Large Language Model Reasoning
by: McGinness, Lachlan, et al.
Published: (2024)
by: McGinness, Lachlan, et al.
Published: (2024)
The AlphaPhysics Term Rewriting System for Marking Algebraic Expressions in Physics Exams
by: Baumgartner, Peter, et al.
Published: (2025)
by: Baumgartner, Peter, et al.
Published: (2025)
Highlighting Case Studies in LLM Literature Review of Interdisciplinary System Science
by: McGinness, Lachlan, et al.
Published: (2025)
by: McGinness, Lachlan, et al.
Published: (2025)
CON-FOLD -- Explainable Machine Learning with Confidence
by: McGinness, Lachlan, et al.
Published: (2024)
by: McGinness, Lachlan, et al.
Published: (2024)
Overview of AI Grading of Physics Olympiad Exams
by: McGinness, Lachlan
Published: (2025)
by: McGinness, Lachlan
Published: (2025)
Can Large Language Models Correctly Interpret Equations with Errors?
by: McGinness, Lachlan, et al.
Published: (2025)
by: McGinness, Lachlan, et al.
Published: (2025)
The Benefits and Challenges of a Quantum Computing Concept Inventory
by: McGinness, Lachlan
Published: (2025)
by: McGinness, Lachlan
Published: (2025)
Exploring the Capabilities of the Frontier Large Language Models for Nuclear Energy Research
by: Almeldein, Ahmed, et al.
Published: (2025)
by: Almeldein, Ahmed, et al.
Published: (2025)
Forecasting Frontier Language Model Agent Capabilities
by: Pimpale, Govind, et al.
Published: (2025)
by: Pimpale, Govind, et al.
Published: (2025)
Reasoning Capabilities of Large Language Models on Dynamic Tasks
by: Wong, Annie, et al.
Published: (2025)
by: Wong, Annie, et al.
Published: (2025)
Evaluating Consistency and Reasoning Capabilities of Large Language Models
by: Saxena, Yash, et al.
Published: (2024)
by: Saxena, Yash, et al.
Published: (2024)
Does RLVR Extend Reasoning Boundaries? Investigating Capability Expansion in Vision-Language Models
by: Shen, Minghe, et al.
Published: (2025)
by: Shen, Minghe, et al.
Published: (2025)
Unlocking Reasoning Capability on Machine Translation in Large Language Models
by: Rajaee, Sara, et al.
Published: (2026)
by: Rajaee, Sara, et al.
Published: (2026)
Frontier Models are Capable of In-context Scheming
by: Meinke, Alexander, et al.
Published: (2024)
by: Meinke, Alexander, et al.
Published: (2024)
Unmasking the Shadows of AI: Investigating Deceptive Capabilities in Large Language Models
by: Guo, Linge
Published: (2024)
by: Guo, Linge
Published: (2024)
Evaluating Interventional Reasoning Capabilities of Large Language Models
by: Kasetty, Tejas, et al.
Published: (2024)
by: Kasetty, Tejas, et al.
Published: (2024)
CRPE: Expanding The Reasoning Capability of Large Language Model for Code Generation
by: Gui, Ningxin, et al.
Published: (2025)
by: Gui, Ningxin, et al.
Published: (2025)
Cooperative Strategic Planning Enhances Reasoning Capabilities in Large Language Models
by: Wang, Danqing, et al.
Published: (2024)
by: Wang, Danqing, et al.
Published: (2024)
Investigating the Robustness of Deductive Reasoning with Large Language Models
by: Hoppe, Fabian, et al.
Published: (2025)
by: Hoppe, Fabian, et al.
Published: (2025)
Jailbroken Frontier Models Retain Their Capabilities
by: Zhu, Daniel, et al.
Published: (2026)
by: Zhu, Daniel, et al.
Published: (2026)
GraphReason: Enhancing Reasoning Capabilities of Large Language Models through A Graph-Based Verification Approach
by: Cao, Lang
Published: (2023)
by: Cao, Lang
Published: (2023)
GraphInstruct: Empowering Large Language Models with Graph Understanding and Reasoning Capability
by: Luo, Zihan, et al.
Published: (2024)
by: Luo, Zihan, et al.
Published: (2024)
On the Modeling Capabilities of Large Language Models for Sequential Decision Making
by: Klissarov, Martin, et al.
Published: (2024)
by: Klissarov, Martin, et al.
Published: (2024)
The Frontier of Data Erasure: Machine Unlearning for Large Language Models
by: Qu, Youyang, et al.
Published: (2024)
by: Qu, Youyang, et al.
Published: (2024)
The Hierarchy of Agentic Capabilities: Evaluating Frontier Models on Realistic RL Environments
by: Ritchie, Logan, et al.
Published: (2026)
by: Ritchie, Logan, et al.
Published: (2026)
A Benchmark for Audio Reasoning Capabilities of Multimodal Large Language Models
by: Christop, Iwona, et al.
Published: (2026)
by: Christop, Iwona, et al.
Published: (2026)
Disentangling Logic: The Role of Context in Large Language Model Reasoning Capabilities
by: Hua, Wenyue, et al.
Published: (2024)
by: Hua, Wenyue, et al.
Published: (2024)
Tree of Agents: Improving Long-Context Capabilities of Large Language Models through Multi-Perspective Reasoning
by: Yu, Song, et al.
Published: (2025)
by: Yu, Song, et al.
Published: (2025)
Revisiting the Travel Planning Capabilities of Large Language Models
by: Zhang, Bo-Wen, et al.
Published: (2026)
by: Zhang, Bo-Wen, et al.
Published: (2026)
PDDL-Mind: Large Language Models are Capable on Belief Reasoning with Reliable State Tracking
by: Zhu, Wang Bill, et al.
Published: (2026)
by: Zhu, Wang Bill, et al.
Published: (2026)
Survey on Reasoning Capabilities and Accessibility of Large Language Models Using Biology-related Questions
by: Ackerman, Michael
Published: (2024)
by: Ackerman, Michael
Published: (2024)
Lost in the Logic: An Evaluation of Large Language Models' Reasoning Capabilities on LSAT Logic Games
by: Malik, Saumya
Published: (2024)
by: Malik, Saumya
Published: (2024)
Distilling Mathematical Reasoning Capabilities into Small Language Models
by: Zhu, Xunyu, et al.
Published: (2024)
by: Zhu, Xunyu, et al.
Published: (2024)
Investigating Advanced Reasoning of Large Language Models via Black-Box Environment Interaction
by: Yin, Congchi, et al.
Published: (2025)
by: Yin, Congchi, et al.
Published: (2025)
Are Frontier Large Language Models Suitable for Q&A in Science Centres?
by: Watson, Jacob, et al.
Published: (2024)
by: Watson, Jacob, et al.
Published: (2024)
Beyond Reasoning Gains: Mitigating General Capabilities Forgetting in Large Reasoning Models
by: Phan, Hoang, et al.
Published: (2025)
by: Phan, Hoang, et al.
Published: (2025)
Imagine in Space: Exploring the Frontier of Spatial Intelligence and Reasoning Efficiency in Vision Language Models
by: Lian, Xiaoxing, et al.
Published: (2025)
by: Lian, Xiaoxing, et al.
Published: (2025)
Exploring the System 1 Thinking Capability of Large Reasoning Models
by: Zhang, Wenyuan, et al.
Published: (2025)
by: Zhang, Wenyuan, et al.
Published: (2025)
Similar Items
-
Large Language Models Imitate Logical Reasoning, but at what Cost?
by: McGinness, Lachlan, et al.
Published: (2025) -
Steamroller Problems: An Evaluation of LLM Reasoning Capability with Automated Theorem Prover Strategies
by: McGinness, Lachlan, et al.
Published: (2024) -
Automated Theorem Provers Help Improve Large Language Model Reasoning
by: McGinness, Lachlan, et al.
Published: (2024) -
The AlphaPhysics Term Rewriting System for Marking Algebraic Expressions in Physics Exams
by: Baumgartner, Peter, et al.
Published: (2025) -
Highlighting Case Studies in LLM Literature Review of Interdisciplinary System Science
by: McGinness, Lachlan, et al.
Published: (2025)