Why Reasoning Fails to Plan: A Planning-Centric Analysis of Long-Horizon Decision Making in LLM Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Zehong, Wu, Fang, Wang, Hongru, Tang, Xiangru, Li, Bolian, Yin, Zhenfei, Ma, Yijun, Li, Yiyang, Sun, Weixiang, Chen, Xiusi, Ye, Yanfang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Temporal Graph Pattern Machine
by: Ma, Yijun, et al.
Published: (2026)
by: Ma, Yijun, et al.
Published: (2026)
Making Reliable and Flexible Decisions in Long-tailed Classification
by: Li, Bolian, et al.
Published: (2025)
by: Li, Bolian, et al.
Published: (2025)
LongDA: Benchmarking LLM Agents for Long-Document Data Analysis
by: Li, Yiyang, et al.
Published: (2026)
by: Li, Yiyang, et al.
Published: (2026)
Why Semantic Entropy Fails: Geometry-Aware and Calibrated Uncertainty for Policy Optimization
by: Zhang, Zheyuan, et al.
Published: (2026)
by: Zhang, Zheyuan, et al.
Published: (2026)
The Cognitive Bandwidth Bottleneck: Shifting Long-Horizon Agent from Planning with Actions to Planning with Schemas
by: Xu, Baixuan, et al.
Published: (2025)
by: Xu, Baixuan, et al.
Published: (2025)
Why Do LLM-based Web Agents Fail? A Hierarchical Planning Perspective
by: Aghzal, Mohamed, et al.
Published: (2026)
by: Aghzal, Mohamed, et al.
Published: (2026)
Policy4OOD: A Knowledge-Guided World Model for Policy Intervention Simulation against the Opioid Overdose Crisis
by: Ma, Yijun, et al.
Published: (2026)
by: Ma, Yijun, et al.
Published: (2026)
Beyond Entangled Planning: Task-Decoupled Planning for Long-Horizon Agents
by: Li, Yunfan, et al.
Published: (2026)
by: Li, Yunfan, et al.
Published: (2026)
PreScam: A Benchmark for Predicting Scam Progression from Early Conversations
by: Sun, Weixiang, et al.
Published: (2026)
by: Sun, Weixiang, et al.
Published: (2026)
Horizon-LM: A RAM-Centric Architecture for LLM Training
by: Yuan, Zhengqing, et al.
Published: (2026)
by: Yuan, Zhengqing, et al.
Published: (2026)
DecisionFlow: Advancing Large Language Model as Principled Decision Maker
by: Chen, Xiusi, et al.
Published: (2025)
by: Chen, Xiusi, et al.
Published: (2025)
Plan-and-Act: Improving Planning of Agents for Long-Horizon Tasks
by: Erdogan, Lutfi Eren, et al.
Published: (2025)
by: Erdogan, Lutfi Eren, et al.
Published: (2025)
WebAnchor: Anchoring Agent Planning to Stabilize Long-Horizon Web Reasoning
by: Yu, Xinmiao, et al.
Published: (2026)
by: Yu, Xinmiao, et al.
Published: (2026)
Do Agents Need to Plan Step-by-Step? Rethinking Planning Horizon in Data-Centric Tool Calling
by: Otani, Naoki, et al.
Published: (2026)
by: Otani, Naoki, et al.
Published: (2026)
MOSAIC: A Skill-Centric Algorithmic Framework for Long-Horizon Manipulation Planning
by: Mishani, Itamar, et al.
Published: (2025)
by: Mishani, Itamar, et al.
Published: (2025)
From Personal to Collective: On the Role of Local and Global Memory in LLM Personalization
by: Wang, Zehong, et al.
Published: (2025)
by: Wang, Zehong, et al.
Published: (2025)
RetailBench: Evaluating Long-Horizon Autonomous Decision-Making and Strategy Stability of LLM Agents in Realistic Retail Environments
by: Zhang, Linghua, et al.
Published: (2026)
by: Zhang, Linghua, et al.
Published: (2026)
ELHPlan: Efficient Long-Horizon Task Planning for Multi-Agent Collaboration
by: Ling, Shaobin, et al.
Published: (2025)
by: Ling, Shaobin, et al.
Published: (2025)
AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning
by: Xi, Zhiheng, et al.
Published: (2025)
by: Xi, Zhiheng, et al.
Published: (2025)
Tracing LLM Reasoning Processes with Strategic Games: A Framework for Planning, Revision, and Resource-Constrained Decision Making
by: Yuan, Xiaopeng, et al.
Published: (2025)
by: Yuan, Xiaopeng, et al.
Published: (2025)
When Planning Fails Despite Correct Execution: On Epistemic Calibration for LLM-Based Multi-Agent Systems
by: Wang, Zehao, et al.
Published: (2026)
by: Wang, Zehao, et al.
Published: (2026)
Hypergraph Pattern Machine: Compositional Tokenization for Higher-Order Interactions
by: Zhao, Kyrie, et al.
Published: (2026)
by: Zhao, Kyrie, et al.
Published: (2026)
Senna-2: Aligning VLM and End-to-End Driving Policy for Consistent Decision Making and Planning
by: Song, Yuehao, et al.
Published: (2026)
by: Song, Yuehao, et al.
Published: (2026)
Heterogeneous Decision Making in Mixed Traffic: Uncertainty-aware Planning and Bounded Rationality
by: Wang, Hang, et al.
Published: (2025)
by: Wang, Hang, et al.
Published: (2025)
The Observability Gap: Why Output-Level Human Feedback Fails for LLM Coding Agents
by: Wang, Yinghao, et al.
Published: (2026)
by: Wang, Yinghao, et al.
Published: (2026)
HiMAP-Travel: Hierarchical Multi-Agent Planning for Long-Horizon Constrained Travel
by: Bui, The Viet, et al.
Published: (2026)
by: Bui, The Viet, et al.
Published: (2026)
DeepPlanning: Benchmarking Long-Horizon Agentic Planning with Verifiable Constraints
by: Zhang, Yinger, et al.
Published: (2026)
by: Zhang, Yinger, et al.
Published: (2026)
Food4All: A Multi-Agent Framework for Real-time Free Food Discovery with Integrated Nutritional Metadata
by: Yuan, Zhengqing, et al.
Published: (2025)
by: Yuan, Zhengqing, et al.
Published: (2025)
Co-Evolving LLM Decision and Skill Bank Agents for Long-Horizon Tasks
by: Wu, Xiyang, et al.
Published: (2026)
by: Wu, Xiyang, et al.
Published: (2026)
Why Do Multi-Agent LLM Systems Fail?
by: Cemri, Mert, et al.
Published: (2025)
by: Cemri, Mert, et al.
Published: (2025)
On Reasoning Strength Planning in Large Reasoning Models
by: Sheng, Leheng, et al.
Published: (2025)
by: Sheng, Leheng, et al.
Published: (2025)
Agent+P: Guiding UI Agents via Symbolic Planning
by: Ma, Shang, et al.
Published: (2025)
by: Ma, Shang, et al.
Published: (2025)
Language-Guided Long Horizon Manipulation with LLM-based Planning and Visual Perception
by: Zhou, Changshi, et al.
Published: (2025)
by: Zhou, Changshi, et al.
Published: (2025)
Long-Horizon Manipulation via Trace-Conditioned VLA Planning
by: Liu, Isabella, et al.
Published: (2026)
by: Liu, Isabella, et al.
Published: (2026)
LongSeeker: Elastic Context Orchestration for Long-Horizon Search Agents
by: Lu, Yijun, et al.
Published: (2026)
by: Lu, Yijun, et al.
Published: (2026)
AsyncVoice Agent: Real-Time Explanation for LLM Planning and Reasoning
by: Lin, Yueqian, et al.
Published: (2025)
by: Lin, Yueqian, et al.
Published: (2025)
Strategist: Self-improvement of LLM Decision Making via Bi-Level Tree Search
by: Light, Jonathan, et al.
Published: (2024)
by: Light, Jonathan, et al.
Published: (2024)
OPBench: A Graph Benchmark to Combat the Opioid Crisis
by: Ma, Tianyi, et al.
Published: (2026)
by: Ma, Tianyi, et al.
Published: (2026)
Interpretable Graph-Language Modeling for Detecting Youth Illicit Drug Use
by: Li, Yiyang, et al.
Published: (2025)
by: Li, Yiyang, et al.
Published: (2025)
STMA: A Spatio-Temporal Memory Agent for Long-Horizon Embodied Task Planning
by: Lei, Mingcong, et al.
Published: (2025)
by: Lei, Mingcong, et al.
Published: (2025)
Similar Items
-
Temporal Graph Pattern Machine
by: Ma, Yijun, et al.
Published: (2026) -
Making Reliable and Flexible Decisions in Long-tailed Classification
by: Li, Bolian, et al.
Published: (2025) -
LongDA: Benchmarking LLM Agents for Long-Document Data Analysis
by: Li, Yiyang, et al.
Published: (2026) -
Why Semantic Entropy Fails: Geometry-Aware and Calibrated Uncertainty for Policy Optimization
by: Zhang, Zheyuan, et al.
Published: (2026) -
The Cognitive Bandwidth Bottleneck: Shifting Long-Horizon Agent from Planning with Actions to Planning with Schemas
by: Xu, Baixuan, et al.
Published: (2025)