Reasoning Effort and Problem Complexity: A Scaling Analysis in LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Estermann, Benjamin, Wattenhofer, Roger |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PUZZLES: A Benchmark for Neural Algorithmic Reasoning
by: Estermann, Benjamin, et al.
Published: (2024)
by: Estermann, Benjamin, et al.
Published: (2024)
Beyond Interpolation: Extrapolative Reasoning with Reinforcement Learning and Graph Neural Networks
by: Grillo, Niccolò, et al.
Published: (2025)
by: Grillo, Niccolò, et al.
Published: (2025)
SUPClust: Active Learning at the Boundaries
by: Ono, Yuta, et al.
Published: (2024)
by: Ono, Yuta, et al.
Published: (2024)
Bridging Diversity and Uncertainty in Active learning with Self-Supervised Pre-Training
by: Doucet, Paul, et al.
Published: (2024)
by: Doucet, Paul, et al.
Published: (2024)
Towards Learning to Reason: Comparing LLMs with Neuro-Symbolic on Arithmetic Relations in Abstract Reasoning
by: Hersche, Michael, et al.
Published: (2024)
by: Hersche, Michael, et al.
Published: (2024)
GraphARC: A Comprehensive Benchmark for Graph-Based Abstract Reasoning
by: Peltonen, Saku, et al.
Published: (2026)
by: Peltonen, Saku, et al.
Published: (2026)
FLIP Reasoning Challenge
by: Plesner, Andreas, et al.
Published: (2025)
by: Plesner, Andreas, et al.
Published: (2025)
On the Expressive Power of GNNs for Boolean Satisfiability
by: Peltonen, Saku, et al.
Published: (2026)
by: Peltonen, Saku, et al.
Published: (2026)
Can an AI Agent Safely Run a Government? Existence of Probably Approximately Aligned Policies
by: Berdoz, Frédéric, et al.
Published: (2024)
by: Berdoz, Frédéric, et al.
Published: (2024)
Can Large Reasoning Models do Analogical Reasoning under Perceptual Uncertainty?
by: Camposampiero, Giacomo, et al.
Published: (2025)
by: Camposampiero, Giacomo, et al.
Published: (2025)
I-RAVEN-X: Benchmarking Generalization and Robustness of Analogical and Mathematical Reasoning in Large Language and Reasoning Models
by: Camposampiero, Giacomo, et al.
Published: (2025)
by: Camposampiero, Giacomo, et al.
Published: (2025)
Information Fidelity in Tool-Using LLM Agents: A Martingale Analysis of the Model Context Protocol
by: Fan, Flint Xiaofeng, et al.
Published: (2026)
by: Fan, Flint Xiaofeng, et al.
Published: (2026)
Position Paper: Rethinking Privacy in RL for Sequential Decision-making in the Age of LLMs
by: Fan, Flint Xiaofeng, et al.
Published: (2025)
by: Fan, Flint Xiaofeng, et al.
Published: (2025)
Cue Point Estimation using Object Detection
by: Argüello, Giulia, et al.
Published: (2024)
by: Argüello, Giulia, et al.
Published: (2024)
GraphFSA: A Finite State Automaton Framework for Algorithmic Learning on Graphs
by: Grötschla, Florian, et al.
Published: (2024)
by: Grötschla, Florian, et al.
Published: (2024)
Benchmarking Positional Encodings for GNNs and Graph Transformers
by: Grötschla, Florian, et al.
Published: (2024)
by: Grötschla, Florian, et al.
Published: (2024)
Next Level Message-Passing with Hierarchical Support Graphs
by: Vonessen, Carlos, et al.
Published: (2024)
by: Vonessen, Carlos, et al.
Published: (2024)
The Impact of Scaling Training Data on Adversarial Robustness
by: Zimmerli, Marco, et al.
Published: (2025)
by: Zimmerli, Marco, et al.
Published: (2025)
Sybil Detection using Graph Neural Networks
by: Heeb, Stuart, et al.
Published: (2024)
by: Heeb, Stuart, et al.
Published: (2024)
Problem-Solving Logic Guided Curriculum In-Context Learning for LLMs Complex Reasoning
by: Ma, Xuetao, et al.
Published: (2025)
by: Ma, Xuetao, et al.
Published: (2025)
Improving Multimodal LLMs Ability In Geometry Problem Solving, Reasoning, And Multistep Scoring
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
Nondeterministic Polynomial-time Problem Challenge: An Ever-Scaling Reasoning Benchmark for LLMs
by: Yang, Chang, et al.
Published: (2025)
by: Yang, Chang, et al.
Published: (2025)
Towards Learning Abductive Reasoning using VSA Distributed Representations
by: Camposampiero, Giacomo, et al.
Published: (2024)
by: Camposampiero, Giacomo, et al.
Published: (2024)
Improving LLMs' Generalized Reasoning Abilities by Graph Problems
by: Zhang, Qifan, et al.
Published: (2025)
by: Zhang, Qifan, et al.
Published: (2025)
N-vium: Mixture-of-Exits Transformer for Accelerated Exact Generation
by: Lorenc, Aleksander, et al.
Published: (2026)
by: Lorenc, Aleksander, et al.
Published: (2026)
Ares: Adaptive Reasoning Effort Selection for Efficient LLM Agents
by: Yang, Jingbo, et al.
Published: (2026)
by: Yang, Jingbo, et al.
Published: (2026)
On the Empirical Complexity of Reasoning and Planning in LLMs
by: Kang, Liwei, et al.
Published: (2024)
by: Kang, Liwei, et al.
Published: (2024)
Large-Scale Aspect-Based Sentiment Analysis with Reasoning-Infused LLMs
by: Liskowski, Paweł, et al.
Published: (2026)
by: Liskowski, Paweł, et al.
Published: (2026)
An Agentic Framework with LLMs for Solving Complex Vehicle Routing Problems
by: Zhang, Ni, et al.
Published: (2025)
by: Zhang, Ni, et al.
Published: (2025)
Parametric Neural Amp Modeling with Active Learning
by: Grötschla, Florian, et al.
Published: (2025)
by: Grötschla, Florian, et al.
Published: (2025)
e1: Learning Adaptive Control of Reasoning Effort
by: Kleinman, Michael, et al.
Published: (2025)
by: Kleinman, Michael, et al.
Published: (2025)
Beyond LLMs: Advancing the Landscape of Complex Reasoning
by: Chu-Carroll, Jennifer, et al.
Published: (2024)
by: Chu-Carroll, Jennifer, et al.
Published: (2024)
HyperLens: Quantifying Cognitive Effort in LLMs with Fine-grained Confidence Trajectory
by: Lu, Chengda, et al.
Published: (2026)
by: Lu, Chengda, et al.
Published: (2026)
AEye: A Visualization Tool for Image Datasets
by: Grötschla, Florian, et al.
Published: (2024)
by: Grötschla, Florian, et al.
Published: (2024)
Navigating the Labyrinth: Evaluating LLMs' Ability to Reason About Search Problems
by: Borazjanizadeh, Nasim, et al.
Published: (2024)
by: Borazjanizadeh, Nasim, et al.
Published: (2024)
WorldSpeech: A Multilingual Speech Corpus from Around the World
by: Asonitis, Antonis, et al.
Published: (2026)
by: Asonitis, Antonis, et al.
Published: (2026)
The Novelty Bottleneck: A Framework for Understanding Human Effort Scaling in AI-Assisted Work
by: Liang, Jacky
Published: (2026)
by: Liang, Jacky
Published: (2026)
Recommender Systems for Democracy: Toward Adversarial Robustness in Voting Advice Applications
by: Berdoz, Frédéric, et al.
Published: (2025)
by: Berdoz, Frédéric, et al.
Published: (2025)
Seeing Through the Mask: Rethinking Adversarial Examples for CAPTCHAs
by: Jabary, Yahya, et al.
Published: (2024)
by: Jabary, Yahya, et al.
Published: (2024)
Completion $\neq$ Collaboration: Scaling Collaborative Effort with Agents
by: Shen, Shannon Zejiang, et al.
Published: (2025)
by: Shen, Shannon Zejiang, et al.
Published: (2025)
Similar Items
-
PUZZLES: A Benchmark for Neural Algorithmic Reasoning
by: Estermann, Benjamin, et al.
Published: (2024) -
Beyond Interpolation: Extrapolative Reasoning with Reinforcement Learning and Graph Neural Networks
by: Grillo, Niccolò, et al.
Published: (2025) -
SUPClust: Active Learning at the Boundaries
by: Ono, Yuta, et al.
Published: (2024) -
Bridging Diversity and Uncertainty in Active learning with Self-Supervised Pre-Training
by: Doucet, Paul, et al.
Published: (2024) -
Towards Learning to Reason: Comparing LLMs with Neuro-Symbolic on Arithmetic Relations in Abstract Reasoning
by: Hersche, Michael, et al.
Published: (2024)