$α^3$-Bench: A Unified Benchmark of Safety, Robustness, and Efficiency for LLM-Based UAV Agents over 6G Networks
Fuente:
arXiv
Salvato in:
| Autori principali: | Ferrag, Mohamed Amine, Lakas, Abderrahmane, Debbah, Merouane |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
$α^3$-SecBench: A Large-Scale Evaluation Suite of Security, Resilience, and Trust for LLM-based UAV Agents over 6G Networks
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2026)
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2026)
6G-Bench: An Open Benchmark for Semantic Communication and Network-Level Reasoning with Foundation Models in AI-Native 6G Networks
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2026)
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2026)
6G Needs Agents: Toward Agentic AI-Native Networks for Autonomous Intelligence
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2026)
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2026)
UAVBench: An Open Benchmark Dataset for Autonomous and Agentic AI UAV Systems via LLM-Generated Flight Scenarios
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2025)
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2025)
AgentDrive: An Open Benchmark Dataset for Agentic AI Reasoning with LLM-Generated Scenarios in Autonomous Systems
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2026)
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2026)
How Small Can 6G Reason? Scaling Tiny Language Models for AI-Native Networks
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2026)
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2026)
LIDSA: Cognitive Arbitration for Signal-Free Autonomous Intersection Management
di: Lakas, Abderrahmane, et al.
Pubblicazione: (2026)
di: Lakas, Abderrahmane, et al.
Pubblicazione: (2026)
From Prompt Injections to Protocol Exploits: Threats in LLM-Powered AI Agents Workflows
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2025)
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2025)
From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2025)
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2025)
VARS-FL: Validation-Aligned Client Selection for Non-IID Federated Learning in IoT Systems
di: Lakas, Mohamed, et al.
Pubblicazione: (2026)
di: Lakas, Mohamed, et al.
Pubblicazione: (2026)
Multi-Agent Reinforcement Learning in Wireless Distributed Networks for 6G
di: Zhang, Jiayi, et al.
Pubblicazione: (2025)
di: Zhang, Jiayi, et al.
Pubblicazione: (2025)
Data-driven Energy Efficiency Modelling in Large-scale Networks: An Expert Knowledge and ML-based Approach
di: López-Pérez, David, et al.
Pubblicazione: (2023)
di: López-Pérez, David, et al.
Pubblicazione: (2023)
Reasoning Beyond Limits: Advances and Open Problems for LLMs
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2025)
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2025)
UAV-Assisted 6G Communication Networks for Railways: Technologies, Applications, and Challenges
di: Huroon, Aamer Mohamed, et al.
Pubblicazione: (2026)
di: Huroon, Aamer Mohamed, et al.
Pubblicazione: (2026)
Reach-Avoid Control Synthesis for a Quadrotor UAV with Formal Safety Guarantees
di: Serry, Mohamed, et al.
Pubblicazione: (2024)
di: Serry, Mohamed, et al.
Pubblicazione: (2024)
Goal-Oriented State Information Compression for Linear Dynamical System Control
di: Wang, Li, et al.
Pubblicazione: (2024)
di: Wang, Li, et al.
Pubblicazione: (2024)
CyberMetric: A Benchmark Dataset based on Retrieval-Augmented Generation for Evaluating LLMs in Cybersecurity Knowledge
di: Tihanyi, Norbert, et al.
Pubblicazione: (2024)
di: Tihanyi, Norbert, et al.
Pubblicazione: (2024)
Telecom World Models: Unifying Digital Twins, Foundation Models, and Predictive Planning for 6G
di: Zou, Hang, et al.
Pubblicazione: (2026)
di: Zou, Hang, et al.
Pubblicazione: (2026)
Towards Near-Field 3D Spot Beamfocusing: Possibilities, Challenges, and Use-cases
di: Monemi, Mehdi, et al.
Pubblicazione: (2023)
di: Monemi, Mehdi, et al.
Pubblicazione: (2023)
Robust Safety-Critical Control of Integrator Chains with Mismatched Perturbations via Linear Time-Varying Feedback
di: Labbadi, Imtiaz Ur Rehman Moussa, et al.
Pubblicazione: (2025)
di: Labbadi, Imtiaz Ur Rehman Moussa, et al.
Pubblicazione: (2025)
AI Agent Access (A\^3) Network: An Embodied, Communication-Aware Multi-Agent Framework for 6G Coverage
di: Zeng, Han, et al.
Pubblicazione: (2025)
di: Zeng, Han, et al.
Pubblicazione: (2025)
Robust Multi-Agent Safety via Tube-Based Tightened Exponential Barrier Functions
di: Koulong, Armel, et al.
Pubblicazione: (2025)
di: Koulong, Armel, et al.
Pubblicazione: (2025)
Multi-Agent Inverse Learning for Sensor Networks: Identifying Coordination in UAV Networks
di: Snow, Luke, et al.
Pubblicazione: (2025)
di: Snow, Luke, et al.
Pubblicazione: (2025)
Two-Timescale Optimization Framework for IAB-Enabled Heterogeneous UAV Networks
di: Deng, Jikang, et al.
Pubblicazione: (2025)
di: Deng, Jikang, et al.
Pubblicazione: (2025)
Robust Online Learning over Networks
di: Bastianello, Nicola, et al.
Pubblicazione: (2023)
di: Bastianello, Nicola, et al.
Pubblicazione: (2023)
TREE:Token-Responsive Energy Efficiency Framework For Green AI-Integrated 6G Networks
di: Yu, Tao, et al.
Pubblicazione: (2025)
di: Yu, Tao, et al.
Pubblicazione: (2025)
Hierarchical Debate-Based Large Language Model (LLM) for Complex Task Planning of 6G Network Management
di: Lin, Yuyan, et al.
Pubblicazione: (2025)
di: Lin, Yuyan, et al.
Pubblicazione: (2025)
Market-Based Replanning for Safety-Critical UAV Swarms in Search and Rescue Missions
di: Giacomossi, Luiz, et al.
Pubblicazione: (2026)
di: Giacomossi, Luiz, et al.
Pubblicazione: (2026)
LLM-Based Agentic Negotiation for 6G: Addressing Uncertainty Neglect and Tail-Event Risk
di: Chergui, Hatim, et al.
Pubblicazione: (2025)
di: Chergui, Hatim, et al.
Pubblicazione: (2025)
Multi-Agent Deep Reinforcement Learning for UAV-Assisted 5G Network Slicing: A Comparative Study of MAPPO, MADDPG, and MADQN
di: Bista, Ghoshana, et al.
Pubblicazione: (2025)
di: Bista, Ghoshana, et al.
Pubblicazione: (2025)
From Single to Multi-Functional RIS: Architecture, Key Technologies, Challenges, and Applications
di: Ni, Wanli, et al.
Pubblicazione: (2024)
di: Ni, Wanli, et al.
Pubblicazione: (2024)
Adversarial Attacks and Defenses in 6G Network-Assisted IoT Systems
di: Son, Bui Duc, et al.
Pubblicazione: (2024)
di: Son, Bui Duc, et al.
Pubblicazione: (2024)
Multi-Agent Reinforcement Learning for UAV-Based Chemical Plume Source Localization
di: Li, Zhirun, et al.
Pubblicazione: (2026)
di: Li, Zhirun, et al.
Pubblicazione: (2026)
Grid-Mind: An LLM-Orchestrated Multi-Fidelity Agent for Automated Connection Impact Assessment
di: Shamseldein, Mohamed
Pubblicazione: (2026)
di: Shamseldein, Mohamed
Pubblicazione: (2026)
Robust Safety-Critical Control of Networked SIR Dynamics
di: Samadi, Saba, et al.
Pubblicazione: (2026)
di: Samadi, Saba, et al.
Pubblicazione: (2026)
Semantic-Aware Edge Intelligence for UAV Handover in 6G Networks
di: Al-Hameed, Aubida A., et al.
Pubblicazione: (2025)
di: Al-Hameed, Aubida A., et al.
Pubblicazione: (2025)
Quantum-Resilient Threat Modelling for Secure RIS-Assisted ISAC in 6G UAV Corridors
di: Hafeez, Sana, et al.
Pubblicazione: (2025)
di: Hafeez, Sana, et al.
Pubblicazione: (2025)
MPC-CBF with Adaptive Safety Margins for Safety-critical Teleoperation over Imperfect Network Connections
di: Periotto, Riccardo, et al.
Pubblicazione: (2024)
di: Periotto, Riccardo, et al.
Pubblicazione: (2024)
Safe and Robust Domains of Attraction for Discrete-Time Systems: A Set-Based Characterization and Certifiable Neural Network Estimation
di: Serry, Mohamed, et al.
Pubblicazione: (2026)
di: Serry, Mohamed, et al.
Pubblicazione: (2026)
IntAgent: NWDAF-Based Intent LLM Agent Towards Advanced Next Generation Networks
di: Soliman, Abdelrahman, et al.
Pubblicazione: (2026)
di: Soliman, Abdelrahman, et al.
Pubblicazione: (2026)
Documenti analoghi
-
$α^3$-SecBench: A Large-Scale Evaluation Suite of Security, Resilience, and Trust for LLM-based UAV Agents over 6G Networks
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2026) -
6G-Bench: An Open Benchmark for Semantic Communication and Network-Level Reasoning with Foundation Models in AI-Native 6G Networks
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2026) -
6G Needs Agents: Toward Agentic AI-Native Networks for Autonomous Intelligence
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2026) -
UAVBench: An Open Benchmark Dataset for Autonomous and Agentic AI UAV Systems via LLM-Generated Flight Scenarios
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2025) -
AgentDrive: An Open Benchmark Dataset for Agentic AI Reasoning with LLM-Generated Scenarios in Autonomous Systems
di: Ferrag, Mohamed Amine, et al.
Pubblicazione: (2026)