It's a TRAP! Task-Redirecting Agent Persuasion Benchmark for Web Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Korgul, Karolina, Yang, Yushi, Drohomirecki, Arkadiusz, Błaszczyk, Piotr, Howard, Will, Aichberger, Lukas, Russell, Chris, Torr, Philip H. S., Mahdi, Adam, Bibi, Adel |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AgentWebBench: Benchmarking Multi-Agent Coordination in Agentic Web
by: Zhong, Shanshan, et al.
Published: (2026)
by: Zhong, Shanshan, et al.
Published: (2026)
MIP against Agent: Malicious Image Patches Hijacking Multimodal OS Agents
by: Aichberger, Lukas, et al.
Published: (2025)
by: Aichberger, Lukas, et al.
Published: (2025)
Understanding Persuasion in Long-Running Agents
by: Jeong, Hyejun, et al.
Published: (2026)
by: Jeong, Hyejun, et al.
Published: (2026)
Agent Benchmarks Fail Public Sector Requirements
by: Rystrøm, Jonathan, et al.
Published: (2026)
by: Rystrøm, Jonathan, et al.
Published: (2026)
From Grounding to Planning: Benchmarking Bottlenecks in Web Agents
by: Shlomov, Segev, et al.
Published: (2024)
by: Shlomov, Segev, et al.
Published: (2024)
Verifiable Semantics for Agent-to-Agent Communication
by: Schoenegger, Philipp, et al.
Published: (2026)
by: Schoenegger, Philipp, et al.
Published: (2026)
Multi-Agent Synchronization Tasks
by: Fernandez, Rolando, et al.
Published: (2024)
by: Fernandez, Rolando, et al.
Published: (2024)
Semantic Web Technology for Agent Communication Protocols
by: Berges, Idoia, et al.
Published: (2024)
by: Berges, Idoia, et al.
Published: (2024)
Orchestrator: Active Inference for Multi-Agent Systems in Long-Horizon Tasks
by: Beckenbauer, Lukas, et al.
Published: (2025)
by: Beckenbauer, Lukas, et al.
Published: (2025)
AutoGenesisAgent: Self-Generating Multi-Agent Systems for Complex Tasks
by: Harper, Jeremy
Published: (2024)
by: Harper, Jeremy
Published: (2024)
Simulating Persuasive Dialogues on Meat Reduction with Generative Agents
by: Ahnert, Georg, et al.
Published: (2025)
by: Ahnert, Georg, et al.
Published: (2025)
A Tensor Network Implementation of Multi Agent Reinforcement Learning
by: Howard, Sunny
Published: (2024)
by: Howard, Sunny
Published: (2024)
Optimizing Sequential Multi-Step Tasks with Parallel LLM Agents
by: Zhang, Enhao, et al.
Published: (2025)
by: Zhang, Enhao, et al.
Published: (2025)
Knowledge Graph-Based Multi-Agent Path Planning in Dynamic Environments using WAITR
by: Holmberg, Ted Edward, et al.
Published: (2024)
by: Holmberg, Ted Edward, et al.
Published: (2024)
Detecting Multi-Agent Collusion Through Multi-Agent Interpretability
by: Rose, Aaron, et al.
Published: (2026)
by: Rose, Aaron, et al.
Published: (2026)
Strategic Persuasion with Trait-Conditioned Multi-Agent Systems for Iterative Legal Argumentation
by: Siedler, Philipp D.
Published: (2026)
by: Siedler, Philipp D.
Published: (2026)
Beyond Browsing: API-Based Web Agents
by: Song, Yueqi, et al.
Published: (2024)
by: Song, Yueqi, et al.
Published: (2024)
LiteWebAgent: The Open-Source Suite for VLM-Based Web-Agent Applications
by: Zhang, Danqing, et al.
Published: (2025)
by: Zhang, Danqing, et al.
Published: (2025)
AssetOpsBench: Benchmarking AI Agents for Task Automation in Industrial Asset Operations and Maintenance
by: Patel, Dhaval, et al.
Published: (2025)
by: Patel, Dhaval, et al.
Published: (2025)
Stable Task Allocation in Multi-Agent Systems with Lexicographic Preferences
by: Reveliotis, Spyros, et al.
Published: (2025)
by: Reveliotis, Spyros, et al.
Published: (2025)
Convergence dynamics of Agent-to-Agent Interactions with Misaligned objectives
by: Cosentino, Romain, et al.
Published: (2025)
by: Cosentino, Romain, et al.
Published: (2025)
FinDeepForecast: A Live Multi-Agent System for Benchmarking Deep Research Agents in Financial Forecasting
by: Li, Xiangyu, et al.
Published: (2026)
by: Li, Xiangyu, et al.
Published: (2026)
CTFExplorer: Evaluating LLM Offensive Agents Through Multi-Target Web CTF Benchmarking
by: Rani, Nanda, et al.
Published: (2026)
by: Rani, Nanda, et al.
Published: (2026)
BMW Agents -- A Framework For Task Automation Through Multi-Agent Collaboration
by: Crawford, Noel, et al.
Published: (2024)
by: Crawford, Noel, et al.
Published: (2024)
A BDI Agent-Based Task Scheduling Framework for Cloud Computing
by: Yang, Yikun, et al.
Published: (2024)
by: Yang, Yikun, et al.
Published: (2024)
Skill Description Deception Attack against Task Routing in Internet of Agents
by: He, Jiayi, et al.
Published: (2026)
by: He, Jiayi, et al.
Published: (2026)
TACTIC: Task-Agnostic Contrastive pre-Training for Inter-Agent Communication
by: Yu, Peihong, et al.
Published: (2025)
by: Yu, Peihong, et al.
Published: (2025)
Agents that Matter: Optimizing Multi-Agent LLMs via Removal-Based Attribution
by: Lu, Mingyu, et al.
Published: (2026)
by: Lu, Mingyu, et al.
Published: (2026)
MASPRM: Multi-Agent System Process Reward Model
by: Yazdani, Milad, et al.
Published: (2025)
by: Yazdani, Milad, et al.
Published: (2025)
Online Multi-Agent Pickup and Delivery with Task Deadlines
by: Makino, Hiroya, et al.
Published: (2024)
by: Makino, Hiroya, et al.
Published: (2024)
Holos: A Web-Scale LLM-Based Multi-Agent System for the Agentic Web
by: Nie, Xiaohang, et al.
Published: (2026)
by: Nie, Xiaohang, et al.
Published: (2026)
MedAgentBoard: Benchmarking Multi-Agent Collaboration with Conventional Methods for Diverse Medical Tasks
by: Zhu, Yinghao, et al.
Published: (2025)
by: Zhu, Yinghao, et al.
Published: (2025)
The AI Committee: A Multi-Agent Framework for Automated Validation and Remediation of Web-Sourced Data
by: Vallabhaneni, Sunith, et al.
Published: (2025)
by: Vallabhaneni, Sunith, et al.
Published: (2025)
Cognitive Duality for Adaptive Web Agents
by: Liu, Jiarun, et al.
Published: (2025)
by: Liu, Jiarun, et al.
Published: (2025)
DISPATCH -- Decentralized Informed Spatial Planning and Assignment of Tasks for Cooperative Heterogeneous Agents
by: Liu, Yao, et al.
Published: (2025)
by: Liu, Yao, et al.
Published: (2025)
COLA: A Scalable Multi-Agent Framework For Windows UI Task Automation
by: Zhao, Di, et al.
Published: (2025)
by: Zhao, Di, et al.
Published: (2025)
Multi-Agent Cooperative Transportation: Optimal and Efficient Task Allocation and Path Finding
by: Zhou, Ning, et al.
Published: (2026)
by: Zhou, Ning, et al.
Published: (2026)
Distributed Task Allocation for Multi-Agent Systems: A Submodular Optimization Approach
by: Liu, Jing, et al.
Published: (2024)
by: Liu, Jing, et al.
Published: (2024)
Synergizing Logical Reasoning, Knowledge Management and Collaboration in Multi-Agent LLM System
by: Kostka, Adam, et al.
Published: (2025)
by: Kostka, Adam, et al.
Published: (2025)
Multi-Agent Craftax: Benchmarking Open-Ended Multi-Agent Reinforcement Learning at the Hyperscale
by: Omari, Bassel Al, et al.
Published: (2025)
by: Omari, Bassel Al, et al.
Published: (2025)
Similar Items
-
AgentWebBench: Benchmarking Multi-Agent Coordination in Agentic Web
by: Zhong, Shanshan, et al.
Published: (2026) -
MIP against Agent: Malicious Image Patches Hijacking Multimodal OS Agents
by: Aichberger, Lukas, et al.
Published: (2025) -
Understanding Persuasion in Long-Running Agents
by: Jeong, Hyejun, et al.
Published: (2026) -
Agent Benchmarks Fail Public Sector Requirements
by: Rystrøm, Jonathan, et al.
Published: (2026) -
From Grounding to Planning: Benchmarking Bottlenecks in Web Agents
by: Shlomov, Segev, et al.
Published: (2024)