Similar Items
CellSecInspector: Safeguarding Cellular Networks via Automated Security Analysis on Specifications
by: Xie, Ke, et al.
Published: (2025)
by: Xie, Ke, et al.
Published: (2025)
CellularSpecSec-Bench: A Staged Benchmark for Evidence-Grounded Interpretation and Security Reasoning over 3GPP Specifications
by: Xie, Ke, et al.
Published: (2026)
by: Xie, Ke, et al.
Published: (2026)
SecureVibeBench: Benchmarking Secure Vibe Coding of AI Agents via Reconstructing Vulnerability-Introducing Scenarios
by: Chen, Junkai, et al.
Published: (2025)
by: Chen, Junkai, et al.
Published: (2025)
QLPro: Automated Code Vulnerability Discovery via LLM and Static Code Analysis Integration
by: Hu, Junze, et al.
Published: (2025)
by: Hu, Junze, et al.
Published: (2025)
InsightEval: An Expert-Curated Benchmark for Assessing Insight Discovery in LLM-Driven Data Agents
by: Zhu, Zhenghao, et al.
Published: (2025)
by: Zhu, Zhenghao, et al.
Published: (2025)
IRIS: LLM-Assisted Static Analysis for Detecting Security Vulnerabilities
by: Li, Ziyang, et al.
Published: (2024)
by: Li, Ziyang, et al.
Published: (2024)
LLM Agent Swarm for Hypothesis-Driven Drug Discovery
by: Song, Kevin, et al.
Published: (2025)
by: Song, Kevin, et al.
Published: (2025)
GradingAttack: Exposing Security Vulnerabilities in LLM Based Educational Grading Agents
by: Li, Xueyi, et al.
Published: (2026)
by: Li, Xueyi, et al.
Published: (2026)
SecureFixAgent: A Hybrid LLM Agent for Automated Python Static Vulnerability Repair
by: Gajjar, Jugal, et al.
Published: (2025)
by: Gajjar, Jugal, et al.
Published: (2025)
Security Vulnerabilities in Ethereum Smart Contracts: A Systematic Analysis
by: Wu, Jixuan, et al.
Published: (2025)
by: Wu, Jixuan, et al.
Published: (2025)
SecureRAG-RTL: A Retrieval-Augmented, Multi-Agent, Zero-Shot LLM-Driven Framework for Hardware Vulnerability Detection
by: Hasan, Touseef, et al.
Published: (2026)
by: Hasan, Touseef, et al.
Published: (2026)
Quantifying Security Vulnerabilities: A Metric-Driven Security Analysis of Gaps in Current AI Standards
by: Madhavan, Keerthana, et al.
Published: (2025)
by: Madhavan, Keerthana, et al.
Published: (2025)
Co-RedTeam: Orchestrated Security Discovery and Exploitation with LLM Agents
by: He, Pengfei, et al.
Published: (2026)
by: He, Pengfei, et al.
Published: (2026)
Synthesizing Multi-Agent Harnesses for Vulnerability Discovery
by: Liu, Hanzhi, et al.
Published: (2026)
by: Liu, Hanzhi, et al.
Published: (2026)
Reframing LLM Agent Security as an Agent-Human Interaction Problem
by: Wang, Peiran, et al.
Published: (2026)
by: Wang, Peiran, et al.
Published: (2026)
Benchmarking LLMs and LLM-based Agents in Practical Vulnerability Detection for Code Repositories
by: Yildiz, Alperen, et al.
Published: (2025)
by: Yildiz, Alperen, et al.
Published: (2025)
DatasetResearch: Benchmarking Agent Systems for Demand-Driven Dataset Discovery
by: Li, Keyu, et al.
Published: (2025)
by: Li, Keyu, et al.
Published: (2025)
Breaking the Protocol: Security Analysis of the Model Context Protocol Specification and Prompt Injection Vulnerabilities in Tool-Integrated LLM Agents
by: Maloyan, Narek, et al.
Published: (2026)
by: Maloyan, Narek, et al.
Published: (2026)
SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code
by: Li, Xinghang, et al.
Published: (2025)
by: Li, Xinghang, et al.
Published: (2025)
LLM Security: Vulnerabilities, Attacks, Defenses, and Countermeasures
by: Aguilera-Martínez, Francisco, et al.
Published: (2025)
by: Aguilera-Martínez, Francisco, et al.
Published: (2025)
NewtonBench: Benchmarking Generalizable Scientific Law Discovery in LLM Agents
by: Zheng, Tianshi, et al.
Published: (2025)
by: Zheng, Tianshi, et al.
Published: (2025)
Continuous Discovery of Vulnerabilities in LLM Serving Systems with Fuzzing
by: Zhao, Yunze, et al.
Published: (2026)
by: Zhao, Yunze, et al.
Published: (2026)
Bridging AI and Software Security: A Comparative Vulnerability Assessment of LLM Agent Deployment Paradigms
by: Gasmi, Tarek, et al.
Published: (2025)
by: Gasmi, Tarek, et al.
Published: (2025)
Benchmarking LLM-Driven Network Configuration Repair
by: Protogeros, Ioannis, et al.
Published: (2026)
by: Protogeros, Ioannis, et al.
Published: (2026)
FuzzingBrain V2: A Multi-Agent LLM System for Automated Vulnerability Discovery and Reproduction
by: Sheng, Ze, et al.
Published: (2026)
by: Sheng, Ze, et al.
Published: (2026)
Argus: Reorchestrating Static Analysis via a Multi-Agent Ensemble for Full-Chain Security Vulnerability Detection
by: Liang, Zi, et al.
Published: (2026)
by: Liang, Zi, et al.
Published: (2026)
Agent Security Bench (ASB): Formalizing and Benchmarking Attacks and Defenses in LLM-based Agents
by: Zhang, Hanrong, et al.
Published: (2024)
by: Zhang, Hanrong, et al.
Published: (2024)
Cellular-X: An LLM-empowered Cellular Agent for Efficient Base Station Operations
by: Wang, Liujianfu, et al.
Published: (2025)
by: Wang, Liujianfu, et al.
Published: (2025)
A Systematization of Security Vulnerabilities in Computer Use Agents
by: Jones, Daniel, et al.
Published: (2025)
by: Jones, Daniel, et al.
Published: (2025)
Control at Stake: Evaluating the Security Landscape of LLM-Driven Email Agents
by: Wu, Jiangrong, et al.
Published: (2025)
by: Wu, Jiangrong, et al.
Published: (2025)
Automatically Benchmarking LLM Code Agents through Agent-Driven Annotation and Evaluation
by: Fu, Lingyue, et al.
Published: (2025)
by: Fu, Lingyue, et al.
Published: (2025)
Agent-Fence: Mapping Security Vulnerabilities Across Deep Research Agents
by: Puppala, Sai, et al.
Published: (2026)
by: Puppala, Sai, et al.
Published: (2026)
LLM-based Vulnerability Discovery through the Lens of Code Metrics
by: Weissberg, Felix, et al.
Published: (2025)
by: Weissberg, Felix, et al.
Published: (2025)
The System Prompt Is the Attack Surface: How LLM Agent Configuration Shapes Security and Creates Exploitable Vulnerabilities
by: Litvak, Ron
Published: (2026)
by: Litvak, Ron
Published: (2026)
The Trust Paradox in LLM-Based Multi-Agent Systems: When Collaboration Becomes a Security Vulnerability
by: Xu, Zijie, et al.
Published: (2025)
by: Xu, Zijie, et al.
Published: (2025)
Agent Audit: A Security Analysis System for LLM Agent Applications
by: Zhang, Haiyue, et al.
Published: (2026)
by: Zhang, Haiyue, et al.
Published: (2026)
Securing the Management Plane in Intent-based Cellular Networks
by: Mehmood, Kashif, et al.
Published: (2024)
by: Mehmood, Kashif, et al.
Published: (2024)
Why Are Web AI Agents More Vulnerable Than Standalone LLMs? A Security Analysis
by: Chiang, Jeffrey Yang Fan, et al.
Published: (2025)
by: Chiang, Jeffrey Yang Fan, et al.
Published: (2025)
Mozi: Governed Autonomy for Drug Discovery LLM Agents
by: Cao, He, et al.
Published: (2026)
by: Cao, He, et al.
Published: (2026)
Towards Personalizing Secure Programming Education with LLM-Injected Vulnerabilities
by: Frazier, Matthew, et al.
Published: (2026)
by: Frazier, Matthew, et al.
Published: (2026)
Similar Items
-
CellSecInspector: Safeguarding Cellular Networks via Automated Security Analysis on Specifications
by: Xie, Ke, et al.
Published: (2025) -
CellularSpecSec-Bench: A Staged Benchmark for Evidence-Grounded Interpretation and Security Reasoning over 3GPP Specifications
by: Xie, Ke, et al.
Published: (2026) -
SecureVibeBench: Benchmarking Secure Vibe Coding of AI Agents via Reconstructing Vulnerability-Introducing Scenarios
by: Chen, Junkai, et al.
Published: (2025) -
QLPro: Automated Code Vulnerability Discovery via LLM and Static Code Analysis Integration
by: Hu, Junze, et al.
Published: (2025) -
InsightEval: An Expert-Curated Benchmark for Assessing Insight Discovery in LLM-Driven Data Agents
by: Zhu, Zhenghao, et al.
Published: (2025)