WebGuard: Building a Generalizable Guardrail for Web Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zheng, Boyuan, Liao, Zeyi, Salisbury, Scott, Liu, Zeyuan, Lin, Michael, Zheng, Qinyuan, Wang, Zifan, Deng, Xiang, Song, Dawn, Sun, Huan, Su, Yu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GPT-4V(ision) is a Generalist Web Agent, if Grounded
von: Zheng, Boyuan, et al.
Veröffentlicht: (2024)
von: Zheng, Boyuan, et al.
Veröffentlicht: (2024)
A Trembling House of Cards? Mapping Adversarial Attacks against Language Agents
von: Mo, Lingbo, et al.
Veröffentlicht: (2024)
von: Mo, Lingbo, et al.
Veröffentlicht: (2024)
Dual-View Visual Contextualization for Web Navigation
von: Kil, Jihyung, et al.
Veröffentlicht: (2024)
von: Kil, Jihyung, et al.
Veröffentlicht: (2024)
RedTeamCUA: Realistic Adversarial Testing of Computer-Use Agents in Hybrid Web-OS Environments
von: Liao, Zeyi, et al.
Veröffentlicht: (2025)
von: Liao, Zeyi, et al.
Veröffentlicht: (2025)
AdvAgent: Controllable Blackbox Red-teaming on Web Agents
von: Xu, Chejian, et al.
Veröffentlicht: (2024)
von: Xu, Chejian, et al.
Veröffentlicht: (2024)
An Illusion of Progress? Assessing the Current State of Web Agents
von: Xue, Tianci, et al.
Veröffentlicht: (2025)
von: Xue, Tianci, et al.
Veröffentlicht: (2025)
Autonomous Continual Learning for Environment Adaptation of Computer-Use Agents
von: Xue, Tianci, et al.
Veröffentlicht: (2026)
von: Xue, Tianci, et al.
Veröffentlicht: (2026)
AmpleGCG: Learning a Universal and Transferable Generative Model of Adversarial Suffixes for Jailbreaking Both Open and Closed LLMs
von: Liao, Zeyi, et al.
Veröffentlicht: (2024)
von: Liao, Zeyi, et al.
Veröffentlicht: (2024)
EIA: Environmental Injection Attack on Generalist Web Agents for Privacy Leakage
von: Liao, Zeyi, et al.
Veröffentlicht: (2024)
von: Liao, Zeyi, et al.
Veröffentlicht: (2024)
WebSentinel: Detecting and Localizing Prompt Injection Attacks for Web Agents
von: Wang, Xilong, et al.
Veröffentlicht: (2026)
von: Wang, Xilong, et al.
Veröffentlicht: (2026)
Mind2Web 2: Evaluating Agentic Search with Agent-as-a-Judge
von: Gou, Boyu, et al.
Veröffentlicht: (2025)
von: Gou, Boyu, et al.
Veröffentlicht: (2025)
SkillWeaver: Web Agents can Self-Improve by Discovering and Honing Skills
von: Zheng, Boyuan, et al.
Veröffentlicht: (2025)
von: Zheng, Boyuan, et al.
Veröffentlicht: (2025)
MolmoWeb: Open Visual Web Agent and Open Data for the Open Web
von: Gupta, Tanmay, et al.
Veröffentlicht: (2026)
von: Gupta, Tanmay, et al.
Veröffentlicht: (2026)
WebGuard++:Interpretable Malicious URL Detection via Bidirectional Fusion of HTML Subgraphs and Multi-Scale Convolutional BERT
von: Tian, Ye, et al.
Veröffentlicht: (2025)
von: Tian, Ye, et al.
Veröffentlicht: (2025)
Building the Web for Agents: A Declarative Framework for Agent-Web Interaction
von: Schultze, Sven, et al.
Veröffentlicht: (2025)
von: Schultze, Sven, et al.
Veröffentlicht: (2025)
BrowserAgent: Building Web Agents with Human-Inspired Web Browsing Actions
von: Yu, Tao, et al.
Veröffentlicht: (2025)
von: Yu, Tao, et al.
Veröffentlicht: (2025)
SafePred: A Predictive Guardrail for Computer-Using Agents via World Models
von: Chen, Yurun, et al.
Veröffentlicht: (2026)
von: Chen, Yurun, et al.
Veröffentlicht: (2026)
Synergy: A Next-Generation General-Purpose Agent for Open Agentic Web
von: Nie, Xiaohang, et al.
Veröffentlicht: (2026)
von: Nie, Xiaohang, et al.
Veröffentlicht: (2026)
AttributionBench: How Hard is Automatic Attribution Evaluation?
von: Li, Yifei, et al.
Veröffentlicht: (2024)
von: Li, Yifei, et al.
Veröffentlicht: (2024)
WebNovelBench: Placing LLM Novelists on the Web Novel Distribution
von: Lin, Leon, et al.
Veröffentlicht: (2025)
von: Lin, Leon, et al.
Veröffentlicht: (2025)
WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models
von: He, Hongliang, et al.
Veröffentlicht: (2024)
von: He, Hongliang, et al.
Veröffentlicht: (2024)
WebArena: A Realistic Web Environment for Building Autonomous Agents
von: Zhou, Shuyan, et al.
Veröffentlicht: (2023)
von: Zhou, Shuyan, et al.
Veröffentlicht: (2023)
AuthGuard: Generalizable Deepfake Detection via Language Guidance
von: Shen, Guangyu, et al.
Veröffentlicht: (2025)
von: Shen, Guangyu, et al.
Veröffentlicht: (2025)
Cognitive Duality for Adaptive Web Agents
von: Liu, Jiarun, et al.
Veröffentlicht: (2025)
von: Liu, Jiarun, et al.
Veröffentlicht: (2025)
MM-WebAgent: A Hierarchical Multimodal Web Agent for Webpage Generation
von: Li, Yan, et al.
Veröffentlicht: (2026)
von: Li, Yan, et al.
Veröffentlicht: (2026)
CodeGuard: Improving LLM Guardrails in CS Education
von: Raihan, Nishat, et al.
Veröffentlicht: (2026)
von: Raihan, Nishat, et al.
Veröffentlicht: (2026)
DuoGuard: A Two-Player RL-Driven Framework for Multilingual LLM Guardrails
von: Deng, Yihe, et al.
Veröffentlicht: (2025)
von: Deng, Yihe, et al.
Veröffentlicht: (2025)
AmpleGCG-Plus: A Strong Generative Model of Adversarial Suffixes to Jailbreak LLMs with Higher Success Rates in Fewer Attempts
von: Kumar, Vishal, et al.
Veröffentlicht: (2024)
von: Kumar, Vishal, et al.
Veröffentlicht: (2024)
MAmmoTH2: Scaling Instructions from the Web
von: Yue, Xiang, et al.
Veröffentlicht: (2024)
von: Yue, Xiang, et al.
Veröffentlicht: (2024)
Navigating the Digital World as Humans Do: Universal Visual Grounding for GUI Agents
von: Gou, Boyu, et al.
Veröffentlicht: (2024)
von: Gou, Boyu, et al.
Veröffentlicht: (2024)
OmniGuard: Unified Omni-Modal Guardrails with Deliberate Reasoning
von: Zhu, Boyu, et al.
Veröffentlicht: (2025)
von: Zhu, Boyu, et al.
Veröffentlicht: (2025)
ConsisGuard: Aligning Safety Deliberation with Policy Enforcement in LLM Guardrails
von: Wang, Yan, et al.
Veröffentlicht: (2026)
von: Wang, Yan, et al.
Veröffentlicht: (2026)
MrGuard: A Multilingual Reasoning Guardrail for Universal LLM Safety
von: Yang, Yahan, et al.
Veröffentlicht: (2025)
von: Yang, Yahan, et al.
Veröffentlicht: (2025)
SentGuard: Sentence-Level Streaming Guardrails for Large Language Models
von: Yu, Jiaqi, et al.
Veröffentlicht: (2026)
von: Yu, Jiaqi, et al.
Veröffentlicht: (2026)
Guarding the Guardrails: A Taxonomy-Driven Approach to Jailbreak Detection
von: Giarrusso, Francesco, et al.
Veröffentlicht: (2025)
von: Giarrusso, Francesco, et al.
Veröffentlicht: (2025)
WALT: Web Agents that Learn Tools
von: Prabhu, Viraj, et al.
Veröffentlicht: (2025)
von: Prabhu, Viraj, et al.
Veröffentlicht: (2025)
WEPO: Web Element Preference Optimization for LLM-based Web Navigation
von: Liu, Jiarun, et al.
Veröffentlicht: (2024)
von: Liu, Jiarun, et al.
Veröffentlicht: (2024)
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents
von: Yang, Rui, et al.
Veröffentlicht: (2026)
von: Yang, Rui, et al.
Veröffentlicht: (2026)
OpenWebVoyager: Building Multimodal Web Agents via Iterative Real-World Exploration, Feedback and Optimization
von: He, Hongliang, et al.
Veröffentlicht: (2024)
von: He, Hongliang, et al.
Veröffentlicht: (2024)
WebAccessVL: Violation-Aware VLM for Web Accessibility
von: Zheng, Amber Yijia, et al.
Veröffentlicht: (2025)
von: Zheng, Amber Yijia, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
GPT-4V(ision) is a Generalist Web Agent, if Grounded
von: Zheng, Boyuan, et al.
Veröffentlicht: (2024) -
A Trembling House of Cards? Mapping Adversarial Attacks against Language Agents
von: Mo, Lingbo, et al.
Veröffentlicht: (2024) -
Dual-View Visual Contextualization for Web Navigation
von: Kil, Jihyung, et al.
Veröffentlicht: (2024) -
RedTeamCUA: Realistic Adversarial Testing of Computer-Use Agents in Hybrid Web-OS Environments
von: Liao, Zeyi, et al.
Veröffentlicht: (2025) -
AdvAgent: Controllable Blackbox Red-teaming on Web Agents
von: Xu, Chejian, et al.
Veröffentlicht: (2024)