MalURLBench: A Benchmark Evaluating Agents' Vulnerabilities When Processing Web URLs
Fuente:
arXiv
Saved in:
| Main Authors: | Kong, Dezhang, Wu, Zhuxi, Liu, Shiqi, Tan, Zhicheng, Lu, Kuichen, Li, Minghao, Liu, Qichen, Chu, Shengyu, Xu, Zhenhua, Liu, Xuan, Han, Meng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fingerprint Vector: Enabling Scalable and Efficient Model Fingerprint Transfer via Vector Addition
by: Xu, Zhenhua, et al.
Published: (2024)
by: Xu, Zhenhua, et al.
Published: (2024)
Web Fraud Attacks Against LLM-Driven Multi-Agent Systems
by: Kong, Dezhang, et al.
Published: (2025)
by: Kong, Dezhang, et al.
Published: (2025)
AttnDiff: Attention-based Differential Fingerprinting for Large Language Models
by: Zhang, Haobo, et al.
Published: (2026)
by: Zhang, Haobo, et al.
Published: (2026)
Scalpel-SAM: A Semi-Supervised Paradigm for Adapting SAM to Infrared Small Object Detection
by: Liu, Zihan, et al.
Published: (2025)
by: Liu, Zihan, et al.
Published: (2025)
Copyright Protection for Large Language Models: A Survey of Methods, Challenges, and Trends
by: Xu, Zhenhua, et al.
Published: (2025)
by: Xu, Zhenhua, et al.
Published: (2025)
DNF: Dual-Layer Nested Fingerprinting for Large Language Model Intellectual Property Protection
by: Xu, Zhenhua, et al.
Published: (2026)
by: Xu, Zhenhua, et al.
Published: (2026)
WebCanvas: Benchmarking Web Agents in Online Environments
by: Pan, Yichen, et al.
Published: (2024)
by: Pan, Yichen, et al.
Published: (2024)
When Agent Markets Arrive
by: Liu, Xuan, et al.
Published: (2026)
by: Liu, Xuan, et al.
Published: (2026)
LLM Agents for Automated Web Vulnerability Reproduction: Are We There Yet?
by: Liu, Bin, et al.
Published: (2025)
by: Liu, Bin, et al.
Published: (2025)
When Bots Take the Bait: Exposing and Mitigating the Emerging Social Engineering Attack in Web Automation Agent
by: Wu, Xinyi, et al.
Published: (2026)
by: Wu, Xinyi, et al.
Published: (2026)
EconWebArena: Benchmarking Autonomous Agents on Economic Tasks in Realistic Web Environments
by: Liu, Zefang, et al.
Published: (2025)
by: Liu, Zefang, et al.
Published: (2025)
A Survey of LLM-Driven AI Agent Communication: Protocols, Security Risks, and Defense Countermeasures
by: Kong, Dezhang, et al.
Published: (2025)
by: Kong, Dezhang, et al.
Published: (2025)
DomURLs_BERT: Pre-trained BERT-based Model for Malicious Domains and URLs Detection and Classification
by: Mahdaouy, Abdelkader El, et al.
Published: (2024)
by: Mahdaouy, Abdelkader El, et al.
Published: (2024)
NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models
by: Zhou, Yi, et al.
Published: (2025)
by: Zhou, Yi, et al.
Published: (2025)
Parsing Millions of URLs per Second
by: Nizipli, Yagiz, et al.
Published: (2023)
by: Nizipli, Yagiz, et al.
Published: (2023)
Keyword Identification within Greek URLs
by: Maria-Alexandra Vonitsanou
Published: (2011)
by: Maria-Alexandra Vonitsanou
Published: (2011)
Unlearning Concepts from Text-to-Video Diffusion Models
by: Liu, Shiqi, et al.
Published: (2024)
by: Liu, Shiqi, et al.
Published: (2024)
ForgetMark: Stealthy Fingerprint Embedding via Targeted Unlearning in Language Models
by: Xu, Zhenhua, et al.
Published: (2026)
by: Xu, Zhenhua, et al.
Published: (2026)
Longitudinal Sampling of URLs From the Wayback Machine
by: Garg, Kritika, et al.
Published: (2025)
by: Garg, Kritika, et al.
Published: (2025)
Automated Phishing Detection Using URLs and Webpages
by: Wang, Huilin, et al.
Published: (2024)
by: Wang, Huilin, et al.
Published: (2024)
GUI-PRA: Process Reward Agent for GUI Tasks
by: Xiong, Tao, et al.
Published: (2025)
by: Xiong, Tao, et al.
Published: (2025)
SecureWebArena: A Holistic Security Evaluation Benchmark for LVLM-based Web Agents
by: Ying, Zonghao, et al.
Published: (2025)
by: Ying, Zonghao, et al.
Published: (2025)
When the Specification Emerges: Benchmarking Faithfulness Loss in Long-Horizon Coding Agents
by: Yan, Lu, et al.
Published: (2026)
by: Yan, Lu, et al.
Published: (2026)
WAInjectBench: Benchmarking Prompt Injection Detections for Web Agents
by: Liu, Yinuo, et al.
Published: (2025)
by: Liu, Yinuo, et al.
Published: (2025)
URLs in the OPAC: Integrating or Disintegrating Research Libraries' Catalogs
by: Burke, Gerald, et al.
Published: (2003)
by: Burke, Gerald, et al.
Published: (2003)
HarmonyGuard: Toward Safety and Utility in Web Agents via Adaptive Policy Enhancement and Dual-Objective Optimization
by: Chen, Yurun, et al.
Published: (2025)
by: Chen, Yurun, et al.
Published: (2025)
CVE-Bench: A Benchmark for AI Agents' Ability to Exploit Real-World Web Application Vulnerabilities
by: Zhu, Yuxuan, et al.
Published: (2025)
by: Zhu, Yuxuan, et al.
Published: (2025)
Quantum Process Tomography of a Thermal Alkali-Metal Vapor
by: Sun, Yujie, et al.
Published: (2025)
by: Sun, Yujie, et al.
Published: (2025)
Regulation of the HMGA2‐SNAI2/CXCR4 axis in atherosclerosis and retinal neovascularization: new therapeutic insights
by: Jianan Li, et al.
Published: (2024)
by: Jianan Li, et al.
Published: (2024)
BrowseComp-ZH: Benchmarking Web Browsing Ability of Large Language Models in Chinese
by: Zhou, Peilin, et al.
Published: (2025)
by: Zhou, Peilin, et al.
Published: (2025)
World-Model-Augmented Web Agents with Action Correction
by: Shen, Zhouzhou, et al.
Published: (2026)
by: Shen, Zhouzhou, et al.
Published: (2026)
Beyond URLs: Metadata Diversity and Position for Efficient LLM Pretraining
by: Fan, Dongyang, et al.
Published: (2025)
by: Fan, Dongyang, et al.
Published: (2025)
Discrete Differential Principle for Continuous Smooth Function Representation
by: Wang, Guoyou, et al.
Published: (2025)
by: Wang, Guoyou, et al.
Published: (2025)
When the Manual Lies: A Realistic Benchmark to Evaluate MCP Poisoning Attacks for LLM Agents
by: Liu, Shi, et al.
Published: (2026)
by: Liu, Shi, et al.
Published: (2026)
Towards Robust and Secure Embodied AI: A Survey on Vulnerabilities and Attacks
by: Xing, Wenpeng, et al.
Published: (2025)
by: Xing, Wenpeng, et al.
Published: (2025)
DailyQA: A Benchmark to Evaluate Web Retrieval Augmented LLMs Based on Capturing Real-World Changes
by: Cheng, Jiehan, et al.
Published: (2025)
by: Cheng, Jiehan, et al.
Published: (2025)
Natto_asMDin_alloXO
by: Liu, Minghao
Published: (2025)
by: Liu, Minghao
Published: (2025)
Improved Monte Carlo Planning via Causal Disentanglement for Structurally-Decomposed Markov Decision Processes
by: Liu, Larkin, et al.
Published: (2024)
by: Liu, Larkin, et al.
Published: (2024)
WebCoach: Self-Evolving Web Agents with Cross-Session Memory Guidance
by: Liu, Genglin, et al.
Published: (2025)
by: Liu, Genglin, et al.
Published: (2025)
Homologically area-minimizing surfaces mod $v$ have at worst codimension 2 singular sets asymptotically
by: Liu, Zhenhua
Published: (2024)
by: Liu, Zhenhua
Published: (2024)
Similar Items
-
Fingerprint Vector: Enabling Scalable and Efficient Model Fingerprint Transfer via Vector Addition
by: Xu, Zhenhua, et al.
Published: (2024) -
Web Fraud Attacks Against LLM-Driven Multi-Agent Systems
by: Kong, Dezhang, et al.
Published: (2025) -
AttnDiff: Attention-based Differential Fingerprinting for Large Language Models
by: Zhang, Haobo, et al.
Published: (2026) -
Scalpel-SAM: A Semi-Supervised Paradigm for Adapting SAM to Infrared Small Object Detection
by: Liu, Zihan, et al.
Published: (2025) -
Copyright Protection for Large Language Models: A Survey of Methods, Challenges, and Trends
by: Xu, Zhenhua, et al.
Published: (2025)