$α^3$-SecBench: A Large-Scale Evaluation Suite of Security, Resilience, and Trust for LLM-based UAV Agents over 6G Networks
Fuente:
arXiv
Saved in:
| Main Authors: | Ferrag, Mohamed Amine, Lakas, Abderrahmane, Debbah, Merouane |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
$α^3$-Bench: A Unified Benchmark of Safety, Robustness, and Efficiency for LLM-Based UAV Agents over 6G Networks
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
From Prompt Injections to Protocol Exploits: Threats in LLM-Powered AI Agents Workflows
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
6G-Bench: An Open Benchmark for Semantic Communication and Network-Level Reasoning with Foundation Models in AI-Native 6G Networks
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
6G Needs Agents: Toward Agentic AI-Native Networks for Autonomous Intelligence
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
How Small Can 6G Reason? Scaling Tiny Language Models for AI-Native Networks
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
UAVBench: An Open Benchmark Dataset for Autonomous and Agentic AI UAV Systems via LLM-Generated Flight Scenarios
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
AgentDrive: An Open Benchmark Dataset for Agentic AI Reasoning with LLM-Generated Scenarios in Autonomous Systems
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
CyberMetric: A Benchmark Dataset based on Retrieval-Augmented Generation for Evaluating LLMs in Cybersecurity Knowledge
by: Tihanyi, Norbert, et al.
Published: (2024)
by: Tihanyi, Norbert, et al.
Published: (2024)
LIDSA: Cognitive Arbitration for Signal-Free Autonomous Intersection Management
by: Lakas, Abderrahmane, et al.
Published: (2026)
by: Lakas, Abderrahmane, et al.
Published: (2026)
SecBench: A Comprehensive Multi-Dimensional Benchmarking Dataset for LLMs in Cybersecurity
by: Jing, Pengfei, et al.
Published: (2024)
by: Jing, Pengfei, et al.
Published: (2024)
Edge Learning for 6G-enabled Internet of Things: A Comprehensive Survey of Vulnerabilities, Datasets, and Defenses
by: Ferrag, Mohamed Amine, et al.
Published: (2023)
by: Ferrag, Mohamed Amine, et al.
Published: (2023)
Generative AI in Cybersecurity: A Comprehensive Review of LLM Applications and Vulnerabilities
by: Ferrag, Mohamed Amine, et al.
Published: (2024)
by: Ferrag, Mohamed Amine, et al.
Published: (2024)
Securing Tomorrow's Smart Cities: Investigating Software Security in Internet of Vehicles and Deep Learning Technologies
by: Jain, Ridhi, et al.
Published: (2024)
by: Jain, Ridhi, et al.
Published: (2024)
SecReEvalBench: A Multi-turned Security Resilience Evaluation Benchmark for Large Language Models
by: Cui, Huining, et al.
Published: (2025)
by: Cui, Huining, et al.
Published: (2025)
Revolutionizing Cyber Threat Detection with Large Language Models: A privacy-preserving BERT-based Lightweight Model for IoT/IIoT Devices
by: Ferrag, Mohamed Amine, et al.
Published: (2023)
by: Ferrag, Mohamed Amine, et al.
Published: (2023)
Innovating Augmented Reality Security: Recent E2E Encryption Approaches
by: Alsop, Hamish, et al.
Published: (2025)
by: Alsop, Hamish, et al.
Published: (2025)
SecureFalcon: Are We There Yet in Automated Software Vulnerability Detection with LLMs?
by: Ferrag, Mohamed Amine, et al.
Published: (2023)
by: Ferrag, Mohamed Amine, et al.
Published: (2023)
SecScale: A Scalable and Secure Trusted Execution Environment for Servers
by: Sunny, Ani, et al.
Published: (2024)
by: Sunny, Ani, et al.
Published: (2024)
Sec5GLoc: Securing 5G Indoor Localization via Adversary-Resilient Deep Learning Architecture
by: Alla, Ildi, et al.
Published: (2025)
by: Alla, Ildi, et al.
Published: (2025)
Reliability and Resilience of AI-Driven Critical Network Infrastructure under Cyber-Physical Threats
by: Lizos, Konstantinos A., et al.
Published: (2025)
by: Lizos, Konstantinos A., et al.
Published: (2025)
CellularSpecSec-Bench: A Staged Benchmark for Evidence-Grounded Interpretation and Security Reasoning over 3GPP Specifications
by: Xie, Ke, et al.
Published: (2026)
by: Xie, Ke, et al.
Published: (2026)
SecRepoBench: Benchmarking Code Agents for Secure Code Completion in Real-World Repositories
by: Shen, Chihao, et al.
Published: (2025)
by: Shen, Chihao, et al.
Published: (2025)
The Hidden DNA of LLM-Generated JavaScript: Structural Patterns Enable High-Accuracy Authorship Attribution
by: Tihanyi, Norbert, et al.
Published: (2025)
by: Tihanyi, Norbert, et al.
Published: (2025)
HardSecBench: Benchmarking the Security Awareness of LLMs for Hardware Code Generation
by: Chen, Qirui, et al.
Published: (2026)
by: Chen, Qirui, et al.
Published: (2026)
How secure is AI-generated Code: A Large-Scale Comparison of Large Language Models
by: Tihanyi, Norbert, et al.
Published: (2024)
by: Tihanyi, Norbert, et al.
Published: (2024)
CyberSecEval 2: A Wide-Ranging Cybersecurity Evaluation Suite for Large Language Models
by: Bhatt, Manish, et al.
Published: (2024)
by: Bhatt, Manish, et al.
Published: (2024)
GenDFIR: Advancing Cyber Incident Timeline Analysis Through Retrieval Augmented Generation and Large Language Models
by: Loumachi, Fatma Yasmine, et al.
Published: (2024)
by: Loumachi, Fatma Yasmine, et al.
Published: (2024)
A Blockchain-Based Trust Framework for Resilient Cross-Domain UAV Service Orchestration
by: Wu, Yao, et al.
Published: (2026)
by: Wu, Yao, et al.
Published: (2026)
CySecBench: Generative AI-based CyberSecurity-focused Prompt Dataset for Benchmarking Large Language Models
by: Wahréus, Johan, et al.
Published: (2025)
by: Wahréus, Johan, et al.
Published: (2025)
SecDTD: Dynamic Token Drop for Secure Transformers Inference
by: Cai, Yifei, et al.
Published: (2026)
by: Cai, Yifei, et al.
Published: (2026)
Quantum-Resilient Blockchain for Secure Transactions in UAV-Assisted Smart Agriculture Networks
by: Ahmad, Taimoor
Published: (2025)
by: Ahmad, Taimoor
Published: (2025)
Secure, Robust, and Energy-Efficient Authenticated Data Sharing in UAV-Assisted 6G Networks
by: Ejiyeh, Atefeh Mohseni
Published: (2024)
by: Ejiyeh, Atefeh Mohseni
Published: (2024)
SecIC3: Customizing IC3 for Hardware Security Verification
by: Tan, Qinhan, et al.
Published: (2026)
by: Tan, Qinhan, et al.
Published: (2026)
Critical Infrastructure Protection: Generative AI, Challenges, and Opportunities
by: Yigit, Yagmur, et al.
Published: (2024)
by: Yigit, Yagmur, et al.
Published: (2024)
From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
Meta SecAlign: A Secure Foundation LLM Against Prompt Injection Attacks
by: Chen, Sizhe, et al.
Published: (2025)
by: Chen, Sizhe, et al.
Published: (2025)
The Trust Paradox in LLM-Based Multi-Agent Systems: When Collaboration Becomes a Security Vulnerability
by: Xu, Zijie, et al.
Published: (2025)
by: Xu, Zijie, et al.
Published: (2025)
ESASCF: Expertise Extraction, Generalization and Reply Framework for an Optimized Automation of Network Security Compliance
by: Ghanem, Mohamed C., et al.
Published: (2023)
by: Ghanem, Mohamed C., et al.
Published: (2023)
Agent Security Bench (ASB): Formalizing and Benchmarking Attacks and Defenses in LLM-based Agents
by: Zhang, Hanrong, et al.
Published: (2024)
by: Zhang, Hanrong, et al.
Published: (2024)
λ-SecAgg: Partial Vector Freezing for Lightweight Secure Aggregation in Federated Learning
by: Zhang, Siqing, et al.
Published: (2023)
by: Zhang, Siqing, et al.
Published: (2023)
Similar Items
-
$α^3$-Bench: A Unified Benchmark of Safety, Robustness, and Efficiency for LLM-Based UAV Agents over 6G Networks
by: Ferrag, Mohamed Amine, et al.
Published: (2026) -
From Prompt Injections to Protocol Exploits: Threats in LLM-Powered AI Agents Workflows
by: Ferrag, Mohamed Amine, et al.
Published: (2025) -
6G-Bench: An Open Benchmark for Semantic Communication and Network-Level Reasoning with Foundation Models in AI-Native 6G Networks
by: Ferrag, Mohamed Amine, et al.
Published: (2026) -
6G Needs Agents: Toward Agentic AI-Native Networks for Autonomous Intelligence
by: Ferrag, Mohamed Amine, et al.
Published: (2026) -
How Small Can 6G Reason? Scaling Tiny Language Models for AI-Native Networks
by: Ferrag, Mohamed Amine, et al.
Published: (2026)