AgentDAM: Privacy Leakage Evaluation for Autonomous Web Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zharmagambetov, Arman, Guo, Chuan, Evtimov, Ivan, Pavlova, Maya, Salakhutdinov, Ruslan, Chaudhuri, Kamalika |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
WASP: Benchmarking Web Agent Security Against Prompt Injection Attacks
von: Evtimov, Ivan, et al.
Veröffentlicht: (2025)
von: Evtimov, Ivan, et al.
Veröffentlicht: (2025)
Safety Alignment of LMs via Non-cooperative Games
von: Paulus, Anselm, et al.
Veröffentlicht: (2025)
von: Paulus, Anselm, et al.
Veröffentlicht: (2025)
RL Is a Hammer and LLMs Are Nails: A Simple Reinforcement Learning Recipe for Strong Prompt Injection
von: Wen, Yuxin, et al.
Veröffentlicht: (2025)
von: Wen, Yuxin, et al.
Veröffentlicht: (2025)
Uncertainty-Based Abstention in LLMs Improves Safety and Reduces Hallucinations
von: Tomani, Christian, et al.
Veröffentlicht: (2024)
von: Tomani, Christian, et al.
Veröffentlicht: (2024)
Meta SecAlign: A Secure Foundation LLM Against Prompt Injection Attacks
von: Chen, Sizhe, et al.
Veröffentlicht: (2025)
von: Chen, Sizhe, et al.
Veröffentlicht: (2025)
InSTA: Towards Internet-Scale Training For Agents
von: Trabucco, Brandon, et al.
Veröffentlicht: (2025)
von: Trabucco, Brandon, et al.
Veröffentlicht: (2025)
OmniACT: A Dataset and Benchmark for Enabling Multimodal Generalist Autonomous Agents for Desktop and Web
von: Kapoor, Raghav, et al.
Veröffentlicht: (2024)
von: Kapoor, Raghav, et al.
Veröffentlicht: (2024)
Automated Red Teaming with GOAT: the Generative Offensive Agent Tester
von: Pavlova, Maya, et al.
Veröffentlicht: (2024)
von: Pavlova, Maya, et al.
Veröffentlicht: (2024)
Privacy-Preserving Retrieval-Augmented Generation with Differential Privacy
von: Koga, Tatsuki, et al.
Veröffentlicht: (2024)
von: Koga, Tatsuki, et al.
Veröffentlicht: (2024)
Tree Search for Language Model Agents
von: Koh, Jing Yu, et al.
Veröffentlicht: (2024)
von: Koh, Jing Yu, et al.
Veröffentlicht: (2024)
AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs
von: Paulus, Anselm, et al.
Veröffentlicht: (2024)
von: Paulus, Anselm, et al.
Veröffentlicht: (2024)
EIA: Environmental Injection Attack on Generalist Web Agents for Privacy Leakage
von: Liao, Zeyi, et al.
Veröffentlicht: (2024)
von: Liao, Zeyi, et al.
Veröffentlicht: (2024)
LeakAgent: RL-based Red-teaming Agent for LLM Privacy Leakage
von: Nie, Yuzhou, et al.
Veröffentlicht: (2024)
von: Nie, Yuzhou, et al.
Veröffentlicht: (2024)
SecAlign: Defending Against Prompt Injection with Preference Optimization
von: Chen, Sizhe, et al.
Veröffentlicht: (2024)
von: Chen, Sizhe, et al.
Veröffentlicht: (2024)
PATHWAYS: Evaluating Investigation and Context Discovery in AI Web Agents
von: Arman, Shifat E., et al.
Veröffentlicht: (2026)
von: Arman, Shifat E., et al.
Veröffentlicht: (2026)
Contrastive Difference Predictive Coding
von: Zheng, Chongyi, et al.
Veröffentlicht: (2023)
von: Zheng, Chongyi, et al.
Veröffentlicht: (2023)
Learning to Rewrite Tool Descriptions for Reliable LLM-Agent Tool Use
von: Guo, Ruocheng, et al.
Veröffentlicht: (2026)
von: Guo, Ruocheng, et al.
Veröffentlicht: (2026)
How Vulnerable Are AI Agents to Indirect Prompt Injections? Insights from a Large-Scale Public Competition
von: Dziemian, Mateusz, et al.
Veröffentlicht: (2026)
von: Dziemian, Mateusz, et al.
Veröffentlicht: (2026)
Privacy Leakage Overshadowed by Views of AI: A Study on Human Oversight of Privacy in Language Model Agent
von: Zhang, Zhiping, et al.
Veröffentlicht: (2024)
von: Zhang, Zhiping, et al.
Veröffentlicht: (2024)
Training a Generally Curious Agent
von: Tajwar, Fahim, et al.
Veröffentlicht: (2025)
von: Tajwar, Fahim, et al.
Veröffentlicht: (2025)
SafeArena: Evaluating the Safety of Autonomous Web Agents
von: Tur, Ada Defne, et al.
Veröffentlicht: (2025)
von: Tur, Ada Defne, et al.
Veröffentlicht: (2025)
AdvPrefix: An Objective for Nuanced LLM Jailbreaks
von: Zhu, Sicheng, et al.
Veröffentlicht: (2024)
von: Zhu, Sicheng, et al.
Veröffentlicht: (2024)
RTBAS: Defending LLM Agents Against Prompt Injection and Privacy Leakage
von: Zhong, Peter Yong, et al.
Veröffentlicht: (2025)
von: Zhong, Peter Yong, et al.
Veröffentlicht: (2025)
WebSP-Eval: Evaluating Web Agents on Website Security and Privacy Tasks
von: Ramesh, Guruprasad Viswanathan, et al.
Veröffentlicht: (2026)
von: Ramesh, Guruprasad Viswanathan, et al.
Veröffentlicht: (2026)
Influence-based Attributions can be Manipulated
von: Yadav, Chhavi, et al.
Veröffentlicht: (2024)
von: Yadav, Chhavi, et al.
Veröffentlicht: (2024)
WebXSkill: Skill Learning for Autonomous Web Agents
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2026)
von: Wang, Zhaoyang, et al.
Veröffentlicht: (2026)
ST-WebAgentBench: A Benchmark for Evaluating Safety and Trustworthiness in Web Agents
von: Levy, Ido, et al.
Veröffentlicht: (2024)
von: Levy, Ido, et al.
Veröffentlicht: (2024)
AgentKit: Structured LLM Reasoning with Dynamic Graphs
von: Wu, Yue, et al.
Veröffentlicht: (2024)
von: Wu, Yue, et al.
Veröffentlicht: (2024)
The BrowserGym Ecosystem for Web Agent Research
von: De Chezelles, Thibault Le Sellier, et al.
Veröffentlicht: (2024)
von: De Chezelles, Thibault Le Sellier, et al.
Veröffentlicht: (2024)
WebUncertainty: Dual-Level Uncertainty Driven Planning and Reasoning For Autonomous Web Agent
von: Zhang, Lingfeng, et al.
Veröffentlicht: (2026)
von: Zhang, Lingfeng, et al.
Veröffentlicht: (2026)
OpenApps: Simulating Environment Variations to Measure UI-Agent Reliability
von: Ullrich, Karen, et al.
Veröffentlicht: (2025)
von: Ullrich, Karen, et al.
Veröffentlicht: (2025)
Autonomous Evaluation and Refinement of Digital Agents
von: Pan, Jiayi, et al.
Veröffentlicht: (2024)
von: Pan, Jiayi, et al.
Veröffentlicht: (2024)
The Behavior Gap: Evaluating Zero-shot LLM Agents in Complex Task-Oriented Dialogs
von: Baidya, Avinash, et al.
Veröffentlicht: (2025)
von: Baidya, Avinash, et al.
Veröffentlicht: (2025)
AbstentionBench: Reasoning LLMs Fail on Unanswerable Questions
von: Kirichenko, Polina, et al.
Veröffentlicht: (2025)
von: Kirichenko, Polina, et al.
Veröffentlicht: (2025)
OpAgent: Operator Agent for Web Navigation
von: Guo, Yuyu, et al.
Veröffentlicht: (2026)
von: Guo, Yuyu, et al.
Veröffentlicht: (2026)
Evaluating Privacy Leakage in Split Learning
von: Qiu, Xinchi, et al.
Veröffentlicht: (2023)
von: Qiu, Xinchi, et al.
Veröffentlicht: (2023)
ResearchGym: Evaluating Language Model Agents on Real-World AI Research
von: Garikaparthi, Aniketh, et al.
Veröffentlicht: (2026)
von: Garikaparthi, Aniketh, et al.
Veröffentlicht: (2026)
Effective Data Augmentation With Diffusion Models
von: Trabucco, Brandon, et al.
Veröffentlicht: (2023)
von: Trabucco, Brandon, et al.
Veröffentlicht: (2023)
DPrivBench: Benchmarking LLMs' Reasoning for Differential Privacy
von: Wang, Erchi, et al.
Veröffentlicht: (2026)
von: Wang, Erchi, et al.
Veröffentlicht: (2026)
WebPilot: A Versatile and Autonomous Multi-Agent System for Web Task Execution with Strategic Exploration
von: Zhang, Yao, et al.
Veröffentlicht: (2024)
von: Zhang, Yao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
WASP: Benchmarking Web Agent Security Against Prompt Injection Attacks
von: Evtimov, Ivan, et al.
Veröffentlicht: (2025) -
Safety Alignment of LMs via Non-cooperative Games
von: Paulus, Anselm, et al.
Veröffentlicht: (2025) -
RL Is a Hammer and LLMs Are Nails: A Simple Reinforcement Learning Recipe for Strong Prompt Injection
von: Wen, Yuxin, et al.
Veröffentlicht: (2025) -
Uncertainty-Based Abstention in LLMs Improves Safety and Reduces Hallucinations
von: Tomani, Christian, et al.
Veröffentlicht: (2024) -
Meta SecAlign: A Secure Foundation LLM Against Prompt Injection Attacks
von: Chen, Sizhe, et al.
Veröffentlicht: (2025)