Does The Way You Plan Matter? An Empirical Study of Planning Representations for LLM Web Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Zambrano, Alejandra, Marjanovic, Sara Vera, Kerboua, Imene, Lù, Xing Han, Kosseim, Leila |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LineRetriever: Planning-Aware Observation Reduction for Web Agents
by: Kerboua, Imene, et al.
Published: (2025)
by: Kerboua, Imene, et al.
Published: (2025)
FocusAgent: Simple Yet Effective Ways of Trimming the Large Context of Web Agents
by: Kerboua, Imene, et al.
Published: (2025)
by: Kerboua, Imene, et al.
Published: (2025)
CLaC at SemEval-2025 Task 6: A Multi-Architecture Approach for Corporate Environmental Promise Verification
by: Turk, Nawar, et al.
Published: (2025)
by: Turk, Nawar, et al.
Published: (2025)
Maintaining MTEB: Towards Long Term Usability and Reproducibility of Embedding Benchmarks
by: Chung, Isaac, et al.
Published: (2025)
by: Chung, Isaac, et al.
Published: (2025)
AI Planning Framework for LLM-Based Web Agents
by: Shahnovsky, Orit, et al.
Published: (2026)
by: Shahnovsky, Orit, et al.
Published: (2026)
Analyzing Persuasive Strategies in Meme Texts: A Fusion of Language Models with Paraphrase Enrichment
by: Nayak, Kota Shamanth Ramanath, et al.
Published: (2024)
by: Nayak, Kota Shamanth Ramanath, et al.
Published: (2024)
Low-Resource Dialect Adaptation of Large Language Models: A French Dialect Case-Study
by: Khan, Eeham, et al.
Published: (2025)
by: Khan, Eeham, et al.
Published: (2025)
SafeArena: Evaluating the Safety of Autonomous Web Agents
by: Tur, Ada Defne, et al.
Published: (2025)
by: Tur, Ada Defne, et al.
Published: (2025)
AgentRewardBench: Evaluating Automatic Evaluations of Web Agent Trajectories
by: Lù, Xing Han, et al.
Published: (2025)
by: Lù, Xing Han, et al.
Published: (2025)
Why Do LLM-based Web Agents Fail? A Hierarchical Planning Perspective
by: Aghzal, Mohamed, et al.
Published: (2026)
by: Aghzal, Mohamed, et al.
Published: (2026)
MTEB-French: Resources for French Sentence Embedding Evaluation and Analysis
by: Ciancone, Mathieu, et al.
Published: (2024)
by: Ciancone, Mathieu, et al.
Published: (2024)
A Multi-Task and Multi-Label Classification Model for Implicit Discourse Relation Recognition
by: Costa, Nelson Filipe, et al.
Published: (2024)
by: Costa, Nelson Filipe, et al.
Published: (2024)
Multi-Lingual Implicit Discourse Relation Recognition with Multi-Label Hierarchical Learning
by: Costa, Nelson Filipe, et al.
Published: (2025)
by: Costa, Nelson Filipe, et al.
Published: (2025)
HiPlan: Hierarchical Planning for LLM-Based Agents with Adaptive Global-Local Guidance
by: Li, Ziyue, et al.
Published: (2025)
by: Li, Ziyue, et al.
Published: (2025)
Robotouille: An Asynchronous Planning Benchmark for LLM Agents
by: Gonzalez-Pumariega, Gonzalo, et al.
Published: (2025)
by: Gonzalez-Pumariega, Gonzalo, et al.
Published: (2025)
Routine: A Structural Planning Framework for LLM Agent System in Enterprise
by: Zeng, Guancheng, et al.
Published: (2025)
by: Zeng, Guancheng, et al.
Published: (2025)
Web Agents Should Adopt the Plan-Then-Execute Paradigm
by: Piet, Julien, et al.
Published: (2026)
by: Piet, Julien, et al.
Published: (2026)
CLaC at DISRPT 2025: Hierarchical Adapters for Cross-Framework Multi-lingual Discourse Relation Classification
by: Turk, Nawar, et al.
Published: (2025)
by: Turk, Nawar, et al.
Published: (2025)
An Empirical Study on Reinforcement Learning for Reasoning-Search Interleaved LLM Agents
by: Jin, Bowen, et al.
Published: (2025)
by: Jin, Bowen, et al.
Published: (2025)
MPO: Boosting LLM Agents with Meta Plan Optimization
by: Xiong, Weimin, et al.
Published: (2025)
by: Xiong, Weimin, et al.
Published: (2025)
Thought-Augmented Planning for LLM-Powered Interactive Recommender Agent
by: Yu, Haocheng, et al.
Published: (2025)
by: Yu, Haocheng, et al.
Published: (2025)
Testing and Understanding Erroneous Planning in LLM Agents through Synthesized User Inputs
by: Ji, Zhenlan, et al.
Published: (2024)
by: Ji, Zhenlan, et al.
Published: (2024)
PlanGenLLMs: A Modern Survey of LLM Planning Capabilities
by: Wei, Hui, et al.
Published: (2025)
by: Wei, Hui, et al.
Published: (2025)
PreAct: Prediction Enhances Agent's Planning Ability
by: Fu, Dayuan, et al.
Published: (2024)
by: Fu, Dayuan, et al.
Published: (2024)
A Real-World WebAgent with Planning, Long Context Understanding, and Program Synthesis
by: Gur, Izzeddin, et al.
Published: (2023)
by: Gur, Izzeddin, et al.
Published: (2023)
Ask-before-Plan: Proactive Language Agents for Real-World Planning
by: Zhang, Xuan, et al.
Published: (2024)
by: Zhang, Xuan, et al.
Published: (2024)
Systematic Analysis of LLM Contributions to Planning: Solver, Verifier, Heuristic
by: Li, Haoming, et al.
Published: (2024)
by: Li, Haoming, et al.
Published: (2024)
Does This Look Familiar to You? Knowledge Analysis via Model Internal Representations
by: Park, Sihyun
Published: (2025)
by: Park, Sihyun
Published: (2025)
Why Reasoning Fails to Plan: A Planning-Centric Analysis of Long-Horizon Decision Making in LLM Agents
by: Wang, Zehong, et al.
Published: (2026)
by: Wang, Zehong, et al.
Published: (2026)
Agents Are All You Need for LLM Unlearning
by: Sanyal, Debdeep, et al.
Published: (2025)
by: Sanyal, Debdeep, et al.
Published: (2025)
Plan Then Action:High-Level Planning Guidance Reinforcement Learning for LLM Reasoning
by: Dou, Zhihao, et al.
Published: (2025)
by: Dou, Zhihao, et al.
Published: (2025)
ChatSOP: An SOP-Guided MCTS Planning Framework for Controllable LLM Dialogue Agents
by: Li, Zhigen, et al.
Published: (2024)
by: Li, Zhigen, et al.
Published: (2024)
On the Influence of Discourse Relations in Persuasive Texts
by: Turk, Nawar, et al.
Published: (2025)
by: Turk, Nawar, et al.
Published: (2025)
LLM-Empowered Embodied Agent for Memory-Augmented Task Planning in Household Robotics
by: Glocker, Marc, et al.
Published: (2025)
by: Glocker, Marc, et al.
Published: (2025)
RAP: Retrieval-Augmented Planning with Contextual Memory for Multimodal LLM Agents
by: Kagaya, Tomoyuki, et al.
Published: (2024)
by: Kagaya, Tomoyuki, et al.
Published: (2024)
CLaC at SemEval-2026 Task 6: Response Clarity Detection in Political Discourse
by: Turk, Nawar, et al.
Published: (2026)
by: Turk, Nawar, et al.
Published: (2026)
Does LLM Alignment Really Need Diversity? An Empirical Study of Adapting RLVR Methods for Moral Reasoning
by: Zhang, Zhaowei, et al.
Published: (2026)
by: Zhang, Zhaowei, et al.
Published: (2026)
Revealing the Barriers of Language Agents in Planning
by: Xie, Jian, et al.
Published: (2024)
by: Xie, Jian, et al.
Published: (2024)
Planning to Explore: Curiosity-Driven Planning for LLM Test Generation
by: Amayuelas, Alfonso, et al.
Published: (2026)
by: Amayuelas, Alfonso, et al.
Published: (2026)
How Much Heavy Lifting Can an Agent Harness Do?: Measuring the LLM's Residual Role in a Planning Agent
by: Jung, Sungwoo, et al.
Published: (2026)
by: Jung, Sungwoo, et al.
Published: (2026)
Similar Items
-
LineRetriever: Planning-Aware Observation Reduction for Web Agents
by: Kerboua, Imene, et al.
Published: (2025) -
FocusAgent: Simple Yet Effective Ways of Trimming the Large Context of Web Agents
by: Kerboua, Imene, et al.
Published: (2025) -
CLaC at SemEval-2025 Task 6: A Multi-Architecture Approach for Corporate Environmental Promise Verification
by: Turk, Nawar, et al.
Published: (2025) -
Maintaining MTEB: Towards Long Term Usability and Reproducibility of Embedding Benchmarks
by: Chung, Isaac, et al.
Published: (2025) -
AI Planning Framework for LLM-Based Web Agents
by: Shahnovsky, Orit, et al.
Published: (2026)