The Danger of Overthinking: Examining the Reasoning-Action Dilemma in Agentic Tasks
Fuente:
arXiv
Saved in:
| Main Authors: | Cuadron, Alejandro, Li, Dacheng, Ma, Wenjie, Wang, Xingyao, Wang, Yichuan, Zhuang, Siyuan, Liu, Shu, Schroeder, Luis Gaspar, Xia, Tian, Mao, Huanzhi, Thumiger, Nicholas, Desai, Aditya, Stoica, Ion, Klimovic, Ana, Neubig, Graham, Gonzalez, Joseph E. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
vAttention: Verified Sparse Attention
by: Desai, Aditya, et al.
Published: (2025)
by: Desai, Aditya, et al.
Published: (2025)
HashAttention: Semantic Sparsity for Faster Inference
by: Desai, Aditya, et al.
Published: (2024)
by: Desai, Aditya, et al.
Published: (2024)
JudgeBench: A Benchmark for Evaluating LLM-based Judges
by: Tan, Sijun, et al.
Published: (2024)
by: Tan, Sijun, et al.
Published: (2024)
Coding Agents with Multimodal Browsing are Generalist Problem Solvers
by: Soni, Aditya Bharat, et al.
Published: (2025)
by: Soni, Aditya Bharat, et al.
Published: (2025)
A Rubric-Supervised Critic from Sparse Real-World Outcomes
by: Wang, Xingyao, et al.
Published: (2026)
by: Wang, Xingyao, et al.
Published: (2026)
A Statistical Framework for Ranking LLM-Based Chatbots
by: Ameli, Siavash, et al.
Published: (2024)
by: Ameli, Siavash, et al.
Published: (2024)
Chapter 3 The professional audiences of the Hippocratic Epidemics
by: Thumiger, Chiara
Published: (2019)
by: Thumiger, Chiara
Published: (2019)
Some Present-Day Problems of Romanian Library Science
by: Stoica, Ion
Published: (1973)
by: Stoica, Ion
Published: (1973)
The Central University Library, Bucharest. Over Seventy-five Years in the History of a Collection
by: Stoica, Ion
Published: (1972)
by: Stoica, Ion
Published: (1972)
vCache: Verified Semantic Prompt Caching
by: Schroeder, Luis Gaspar, et al.
Published: (2025)
by: Schroeder, Luis Gaspar, et al.
Published: (2025)
Inducing Programmatic Skills for Agentic Tasks
by: Wang, Zora Zhiruo, et al.
Published: (2025)
by: Wang, Zora Zhiruo, et al.
Published: (2025)
TOM-SWE: User Mental Modeling For Software Engineering Agents
by: Zhou, Xuhui, et al.
Published: (2025)
by: Zhou, Xuhui, et al.
Published: (2025)
MPC-Minimized Secure LLM Inference
by: Rathee, Deevashwer, et al.
Published: (2024)
by: Rathee, Deevashwer, et al.
Published: (2024)
Training Software Engineering Agents and Verifiers with SWE-Gym
by: Pan, Jiayi, et al.
Published: (2024)
by: Pan, Jiayi, et al.
Published: (2024)
Revisiting Overthinking in Long Chain-of-Thought from the Perspective of Self-Doubt
by: Peng, Keqin, et al.
Published: (2025)
by: Peng, Keqin, et al.
Published: (2025)
On the Dangers of Overthinking: A Natural Experiment on Self‐Regulatory Thought, Mind‐Wandering and Undergraduate Exam Performance
by: J. Helgi Clayton McClure, et al.
Published: (2025)
by: J. Helgi Clayton McClure, et al.
Published: (2025)
Clawed and Dangerous: Can We Trust Open Agentic Systems?
by: Chen, Shiping, et al.
Published: (2026)
by: Chen, Shiping, et al.
Published: (2026)
Shangri-La, paraíso terrenal / Liu Huanzhi y Zhang Tao
by: Liu Huanzhi
Published: (2008)
by: Liu Huanzhi
Published: (2008)
Campesinos buscan una vida modestamente acomodada / Liu Huanzhi
by: Liu Huanzhi
by: Liu Huanzhi
EDIT-Bench: Evaluating LLM Abilities to Perform Real-World Instructed Code Edits
by: Chi, Wayne, et al.
Published: (2025)
by: Chi, Wayne, et al.
Published: (2025)
SABER: Small Actions, Big Errors -- Safeguarding Mutating Steps in LLM Agents
by: Cuadron, Alejandro, et al.
Published: (2025)
by: Cuadron, Alejandro, et al.
Published: (2025)
GIST: Gauge-Invariant Spectral Transformers for Scalable Graph Neural Operators
by: Rigotti, Mattia, et al.
Published: (2026)
by: Rigotti, Mattia, et al.
Published: (2026)
Training Proactive and Personalized LLM Agents
by: Sun, Weiwei, et al.
Published: (2025)
by: Sun, Weiwei, et al.
Published: (2025)
TroVE: Inducing Verifiable and Efficient Toolboxes for Solving Programmatic Tasks
by: Wang, Zhiruo, et al.
Published: (2024)
by: Wang, Zhiruo, et al.
Published: (2024)
How can we assess human-agent interactions? Case studies in software agent design
by: Chen, Valerie, et al.
Published: (2025)
by: Chen, Valerie, et al.
Published: (2025)
Inference Time Context Sparsity: Illusion or Opportunity?
by: Joshi, Sahil, et al.
Published: (2026)
by: Joshi, Sahil, et al.
Published: (2026)
Efficient-vDiT: Efficient Video Diffusion Transformers With Attention Tile
by: Ding, Hangliang, et al.
Published: (2025)
by: Ding, Hangliang, et al.
Published: (2025)
Go-Browse: Training Web Agents with Structured Exploration
by: Gandhi, Apurva, et al.
Published: (2025)
by: Gandhi, Apurva, et al.
Published: (2025)
BehaviorBox: Automated Discovery of Fine-Grained Performance Differences Between Language Models
by: Tjuatja, Lindia, et al.
Published: (2025)
by: Tjuatja, Lindia, et al.
Published: (2025)
Effective Strategies for Asynchronous Software Engineering Agents
by: Geng, Jiayi, et al.
Published: (2026)
by: Geng, Jiayi, et al.
Published: (2026)
Continuum: Efficient and Robust Multi-Turn LLM Agent Scheduling with KV Cache Time-to-Live
by: Li, Hanchen, et al.
Published: (2025)
by: Li, Hanchen, et al.
Published: (2025)
Locality-aware Fair Scheduling in LLM Serving
by: Cao, Shiyi, et al.
Published: (2025)
by: Cao, Shiyi, et al.
Published: (2025)
Mitigating Overthinking through Reasoning Shaping
by: Song, Feifan, et al.
Published: (2025)
by: Song, Feifan, et al.
Published: (2025)
Tight Streaming Lower Bounds for Deterministic Approximate Counting
by: Wang, Yichuan
Published: (2024)
by: Wang, Yichuan
Published: (2024)
Think, But Don't Overthink: Reproducing Recursive Language Models
by: Wang, Daren
Published: (2026)
by: Wang, Daren
Published: (2026)
From Token to Action: State Machine Reasoning to Mitigate Overthinking in Information Retrieval
by: Lee, Dohyeon, et al.
Published: (2025)
by: Lee, Dohyeon, et al.
Published: (2025)
SVG-EAR: Parameter-Free Linear Compensation for Sparse Video Generation via Error-aware Routing
by: Zhou, Xuanyi, et al.
Published: (2026)
by: Zhou, Xuanyi, et al.
Published: (2026)
Overthinking Reduction with Decoupled Rewards and Curriculum Data Scheduling
by: Jiang, Shuyang, et al.
Published: (2025)
by: Jiang, Shuyang, et al.
Published: (2025)
Examining LLMs' Uncertainty Expression Towards Questions Outside Parametric Knowledge
by: Liu, Genglin, et al.
Published: (2023)
by: Liu, Genglin, et al.
Published: (2023)
AstraFlow: Dataflow-Oriented Reinforcement Learning for Agentic LLMs
by: Zheng, Haizhong, et al.
Published: (2026)
by: Zheng, Haizhong, et al.
Published: (2026)
Similar Items
-
vAttention: Verified Sparse Attention
by: Desai, Aditya, et al.
Published: (2025) -
HashAttention: Semantic Sparsity for Faster Inference
by: Desai, Aditya, et al.
Published: (2024) -
JudgeBench: A Benchmark for Evaluating LLM-based Judges
by: Tan, Sijun, et al.
Published: (2024) -
Coding Agents with Multimodal Browsing are Generalist Problem Solvers
by: Soni, Aditya Bharat, et al.
Published: (2025) -
A Rubric-Supervised Critic from Sparse Real-World Outcomes
by: Wang, Xingyao, et al.
Published: (2026)