AI Planning Framework for LLM-Based Web Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Shahnovsky, Orit, Dror, Rotem |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Alternative Annotator Test for LLM-as-a-Judge: How to Statistically Justify Replacing Human Annotators with LLMs
by: Calderon, Nitay, et al.
Published: (2025)
by: Calderon, Nitay, et al.
Published: (2025)
Why Do LLM-based Web Agents Fail? A Hierarchical Planning Perspective
by: Aghzal, Mohamed, et al.
Published: (2026)
by: Aghzal, Mohamed, et al.
Published: (2026)
Infogent: An Agent-Based Framework for Web Information Aggregation
by: Reddy, Revanth Gangi, et al.
Published: (2024)
by: Reddy, Revanth Gangi, et al.
Published: (2024)
AgentOccam: A Simple Yet Strong Baseline for LLM-Based Web Agents
by: Yang, Ke, et al.
Published: (2024)
by: Yang, Ke, et al.
Published: (2024)
Routine: A Structural Planning Framework for LLM Agent System in Enterprise
by: Zeng, Guancheng, et al.
Published: (2025)
by: Zeng, Guancheng, et al.
Published: (2025)
Does The Way You Plan Matter? An Empirical Study of Planning Representations for LLM Web Agents
by: Zambrano, Alejandra, et al.
Published: (2026)
by: Zambrano, Alejandra, et al.
Published: (2026)
HiPlan: Hierarchical Planning for LLM-Based Agents with Adaptive Global-Local Guidance
by: Li, Ziyue, et al.
Published: (2025)
by: Li, Ziyue, et al.
Published: (2025)
DynaWeb: Model-Based Reinforcement Learning of Web Agents
by: Ding, Hang, et al.
Published: (2026)
by: Ding, Hang, et al.
Published: (2026)
ChatSOP: An SOP-Guided MCTS Planning Framework for Controllable LLM Dialogue Agents
by: Li, Zhigen, et al.
Published: (2024)
by: Li, Zhigen, et al.
Published: (2024)
Beyond Pixels: Exploring DOM Downsampling for LLM-Based Web Agents
by: Schiepanski, Thassilo M., et al.
Published: (2025)
by: Schiepanski, Thassilo M., et al.
Published: (2025)
Societal AI Research Has Become Less Interdisciplinary
by: Markus, Dror Kris, et al.
Published: (2025)
by: Markus, Dror Kris, et al.
Published: (2025)
Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey
by: Zhu, Jiachen, et al.
Published: (2025)
by: Zhu, Jiachen, et al.
Published: (2025)
Holistic Evaluation and Failure Diagnosis of AI Agents
by: Madvil, Netta, et al.
Published: (2026)
by: Madvil, Netta, et al.
Published: (2026)
TravelAgent: An AI Assistant for Personalized Travel Planning
by: Chen, Aili, et al.
Published: (2024)
by: Chen, Aili, et al.
Published: (2024)
Script-Based Dialog Policy Planning for LLM-Powered Conversational Agents: A Basic Architecture for an "AI Therapist"
by: Wasenmüller, Robert, et al.
Published: (2024)
by: Wasenmüller, Robert, et al.
Published: (2024)
Robotouille: An Asynchronous Planning Benchmark for LLM Agents
by: Gonzalez-Pumariega, Gonzalo, et al.
Published: (2025)
by: Gonzalez-Pumariega, Gonzalo, et al.
Published: (2025)
BrowserAgent: Building Web Agents with Human-Inspired Web Browsing Actions
by: Yu, Tao, et al.
Published: (2025)
by: Yu, Tao, et al.
Published: (2025)
Researchy Questions: A Dataset of Multi-Perspective, Decompositional Questions for LLM Web Agents
by: Rosset, Corby, et al.
Published: (2024)
by: Rosset, Corby, et al.
Published: (2024)
WebXSkill: Skill Learning for Autonomous Web Agents
by: Wang, Zhaoyang, et al.
Published: (2026)
by: Wang, Zhaoyang, et al.
Published: (2026)
WorkForceAgent-R1: Incentivizing Reasoning Capability in LLM-based Web Agents via Reinforcement Learning
by: Zhuang, Yuchen, et al.
Published: (2025)
by: Zhuang, Yuchen, et al.
Published: (2025)
LiteWebAgent: The Open-Source Suite for VLM-Based Web-Agent Applications
by: Zhang, Danqing, et al.
Published: (2025)
by: Zhang, Danqing, et al.
Published: (2025)
WebRollback: Enhancing Web Agents with Explicit Rollback Mechanisms
by: Zhang, Zhisong, et al.
Published: (2025)
by: Zhang, Zhisong, et al.
Published: (2025)
WebSailor: Navigating Super-human Reasoning for Web Agent
by: Li, Kuan, et al.
Published: (2025)
by: Li, Kuan, et al.
Published: (2025)
Can Agent Conquer Web? Exploring the Frontiers of ChatGPT Atlas Agent in Web Games
by: Zhang, Jingran, et al.
Published: (2025)
by: Zhang, Jingran, et al.
Published: (2025)
BackdoorAgent: A Unified Framework for Backdoor Attacks on LLM-based Agents
by: Feng, Yunhao, et al.
Published: (2026)
by: Feng, Yunhao, et al.
Published: (2026)
AutoAgent: A Fully-Automated and Zero-Code Framework for LLM Agents
by: Tang, Jiabin, et al.
Published: (2025)
by: Tang, Jiabin, et al.
Published: (2025)
MIRIX: Multi-Agent Memory System for LLM-Based Agents
by: Wang, Yu, et al.
Published: (2025)
by: Wang, Yu, et al.
Published: (2025)
PlanGEN: A Multi-Agent Framework for Generating Planning and Reasoning Trajectories for Complex Problem Solving
by: Parmar, Mihir, et al.
Published: (2025)
by: Parmar, Mihir, et al.
Published: (2025)
WebLists: Extracting Structured Information From Complex Interactive Websites Using Executable LLM Agents
by: Bohra, Arth, et al.
Published: (2025)
by: Bohra, Arth, et al.
Published: (2025)
WebNovelBench: Placing LLM Novelists on the Web Novel Distribution
by: Lin, Leon, et al.
Published: (2025)
by: Lin, Leon, et al.
Published: (2025)
Region4Web: Rethinking Observation Space Granularity for Web Agents
by: Kwon, Donguk, et al.
Published: (2026)
by: Kwon, Donguk, et al.
Published: (2026)
OptimAI: Optimization from Natural Language Using LLM-Powered AI Agents
by: Thind, Raghav, et al.
Published: (2025)
by: Thind, Raghav, et al.
Published: (2025)
AgentQuest: A Modular Benchmark Framework to Measure Progress and Improve LLM Agents
by: Gioacchini, Luca, et al.
Published: (2024)
by: Gioacchini, Luca, et al.
Published: (2024)
Insight Agents: An LLM-Based Multi-Agent System for Data Insights
by: Bai, Jincheng, et al.
Published: (2026)
by: Bai, Jincheng, et al.
Published: (2026)
Level-Navi Agent: A Framework and benchmark for Chinese Web Search Agents
by: Hu, Chuanrui, et al.
Published: (2024)
by: Hu, Chuanrui, et al.
Published: (2024)
How Much Heavy Lifting Can an Agent Harness Do?: Measuring the LLM's Residual Role in a Planning Agent
by: Jung, Sungwoo, et al.
Published: (2026)
by: Jung, Sungwoo, et al.
Published: (2026)
AURA: A Diagnostic Framework for Tracking User Satisfaction of Interactive Planning Agents
by: Kim, Takyoung, et al.
Published: (2025)
by: Kim, Takyoung, et al.
Published: (2025)
WebCoach: Self-Evolving Web Agents with Cross-Session Memory Guidance
by: Liu, Genglin, et al.
Published: (2025)
by: Liu, Genglin, et al.
Published: (2025)
Web-CogReasoner: Towards Knowledge-Induced Cognitive Reasoning for Web Agents
by: Guo, Yuhan, et al.
Published: (2025)
by: Guo, Yuhan, et al.
Published: (2025)
WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models
by: He, Hongliang, et al.
Published: (2024)
by: He, Hongliang, et al.
Published: (2024)
Similar Items
-
The Alternative Annotator Test for LLM-as-a-Judge: How to Statistically Justify Replacing Human Annotators with LLMs
by: Calderon, Nitay, et al.
Published: (2025) -
Why Do LLM-based Web Agents Fail? A Hierarchical Planning Perspective
by: Aghzal, Mohamed, et al.
Published: (2026) -
Infogent: An Agent-Based Framework for Web Information Aggregation
by: Reddy, Revanth Gangi, et al.
Published: (2024) -
AgentOccam: A Simple Yet Strong Baseline for LLM-Based Web Agents
by: Yang, Ke, et al.
Published: (2024) -
Routine: A Structural Planning Framework for LLM Agent System in Enterprise
by: Zeng, Guancheng, et al.
Published: (2025)