SecureFalcon: Are We There Yet in Automated Software Vulnerability Detection with LLMs?
Fuente:
arXiv
Saved in:
| Main Authors: | Ferrag, Mohamed Amine, Battah, Ammar, Tihanyi, Norbert, Jain, Ridhi, Maimut, Diana, Alwahedi, Fatima, Lestable, Thierry, Thandi, Narinderjit Singh, Mechri, Abdechakour, Debbah, Merouane, Cordeiro, Lucas C. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Generative AI in Cybersecurity: A Comprehensive Review of LLM Applications and Vulnerabilities
by: Ferrag, Mohamed Amine, et al.
Published: (2024)
by: Ferrag, Mohamed Amine, et al.
Published: (2024)
Revolutionizing Cyber Threat Detection with Large Language Models: A privacy-preserving BERT-based Lightweight Model for IoT/IIoT Devices
by: Ferrag, Mohamed Amine, et al.
Published: (2023)
by: Ferrag, Mohamed Amine, et al.
Published: (2023)
CyberMetric: A Benchmark Dataset based on Retrieval-Augmented Generation for Evaluating LLMs in Cybersecurity Knowledge
by: Tihanyi, Norbert, et al.
Published: (2024)
by: Tihanyi, Norbert, et al.
Published: (2024)
Securing Tomorrow's Smart Cities: Investigating Software Security in Internet of Vehicles and Deep Learning Technologies
by: Jain, Ridhi, et al.
Published: (2024)
by: Jain, Ridhi, et al.
Published: (2024)
Reasoning Beyond Limits: Advances and Open Problems for LLMs
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
How secure is AI-generated Code: A Large-Scale Comparison of Large Language Models
by: Tihanyi, Norbert, et al.
Published: (2024)
by: Tihanyi, Norbert, et al.
Published: (2024)
Vulnerability Detection: From Formal Verification to Large Language Models and Hybrid Approaches: A Comprehensive Overview
by: Tihanyi, Norbert, et al.
Published: (2025)
by: Tihanyi, Norbert, et al.
Published: (2025)
The FormAI Dataset: Generative AI in Software Security Through the Lens of Formal Verification
by: Tihanyi, Norbert, et al.
Published: (2023)
by: Tihanyi, Norbert, et al.
Published: (2023)
A New Era in Software Security: Towards Self-Healing Software via Large Language Models and Formal Verification
by: Tihanyi, Norbert, et al.
Published: (2023)
by: Tihanyi, Norbert, et al.
Published: (2023)
From Prompt Injections to Protocol Exploits: Threats in LLM-Powered AI Agents Workflows
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
Edge Learning for 6G-enabled Internet of Things: A Comprehensive Survey of Vulnerabilities, Datasets, and Defenses
by: Ferrag, Mohamed Amine, et al.
Published: (2023)
by: Ferrag, Mohamed Amine, et al.
Published: (2023)
UAVBench: An Open Benchmark Dataset for Autonomous and Agentic AI UAV Systems via LLM-Generated Flight Scenarios
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
by: Ferrag, Mohamed Amine, et al.
Published: (2025)
$α^3$-Bench: A Unified Benchmark of Safety, Robustness, and Efficiency for LLM-Based UAV Agents over 6G Networks
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
6G-Bench: An Open Benchmark for Semantic Communication and Network-Level Reasoning with Foundation Models in AI-Native 6G Networks
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
6G Needs Agents: Toward Agentic AI-Native Networks for Autonomous Intelligence
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
LIDSA: Cognitive Arbitration for Signal-Free Autonomous Intersection Management
by: Lakas, Abderrahmane, et al.
Published: (2026)
by: Lakas, Abderrahmane, et al.
Published: (2026)
How Small Can 6G Reason? Scaling Tiny Language Models for AI-Native Networks
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
$α^3$-SecBench: A Large-Scale Evaluation Suite of Security, Resilience, and Trust for LLM-based UAV Agents over 6G Networks
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
AgentDrive: An Open Benchmark Dataset for Agentic AI Reasoning with LLM-Generated Scenarios in Autonomous Systems
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
by: Ferrag, Mohamed Amine, et al.
Published: (2026)
Dynamic Intelligence Assessment: Benchmarking LLMs on the Road to AGI with a Focus on Model Confidence
by: Tihanyi, Norbert, et al.
Published: (2024)
by: Tihanyi, Norbert, et al.
Published: (2024)
The Hidden DNA of LLM-Generated JavaScript: Structural Patterns Enable High-Accuracy Authorship Attribution
by: Tihanyi, Norbert, et al.
Published: (2025)
by: Tihanyi, Norbert, et al.
Published: (2025)
CASTLE: Benchmarking Dataset for Static Code Analyzers and LLMs towards CWE Detection
by: Dubniczky, Richard A., et al.
Published: (2025)
by: Dubniczky, Richard A., et al.
Published: (2025)
TelecomGPT: A Framework to Build Telecom-Specfic Large Language Models
by: Zou, Hang, et al.
Published: (2024)
by: Zou, Hang, et al.
Published: (2024)
Adversarial Attacks and Defenses in 6G Network-Assisted IoT Systems
by: Son, Bui Duc, et al.
Published: (2024)
by: Son, Bui Duc, et al.
Published: (2024)
VARS-FL: Validation-Aligned Client Selection for Non-IID Federated Learning in IoT Systems
by: Lakas, Mohamed, et al.
Published: (2026)
by: Lakas, Mohamed, et al.
Published: (2026)
I Know Which LLM Wrote Your Code Last Summer: LLM generated Code Stylometry for Authorship Attribution
by: Bisztray, Tamas, et al.
Published: (2025)
by: Bisztray, Tamas, et al.
Published: (2025)
Assessing Large Language Models in Comprehending and Verifying Concurrent Programs across Memory Models
by: Jain, Ridhi, et al.
Published: (2025)
by: Jain, Ridhi, et al.
Published: (2025)
LLM Agents for Automated Web Vulnerability Reproduction: Are We There Yet?
by: Liu, Bin, et al.
Published: (2025)
by: Liu, Bin, et al.
Published: (2025)
Why Doesn't Microsoft Let Me Sleep? How Automaticity of Windows Updates Impacts User Autonomy
by: Ahuja, Sanju, et al.
Published: (2024)
by: Ahuja, Sanju, et al.
Published: (2024)
Are We Comparing Yet?
by: Saussy, Haun
Published: (2019)
by: Saussy, Haun
Published: (2019)
Communication-Efficient Zero-Order and First-Order Federated Learning Methods over Wireless Networks
by: Assaad, Mohamad, et al.
Published: (2025)
by: Assaad, Mohamad, et al.
Published: (2025)
Non-Identical Diffusion Models in MIMO-OFDM Channel Generation
by: Yang, Yuzhi, et al.
Published: (2025)
by: Yang, Yuzhi, et al.
Published: (2025)
Multi-target DoA estimation with a single Rydberg atomic receiver by spectral analysis of spatially-resolved fluorescence
by: Han, Liangcheng, et al.
Published: (2026)
by: Han, Liangcheng, et al.
Published: (2026)
Immersive Media and Massive Twinning: Advancing Towards the Metaverse
by: Hamidouche, Wassim, et al.
Published: (2023)
by: Hamidouche, Wassim, et al.
Published: (2023)
Per-antenna power constraints: constructing Pareto-optimal precoders with cubic complexity under non-negligible noise conditions
by: Petrov, Sergey, et al.
Published: (2025)
by: Petrov, Sergey, et al.
Published: (2025)
How Bad is Training on Synthetic Data? A Statistical Analysis of Language Model Collapse
by: Seddik, Mohamed El Amine, et al.
Published: (2024)
by: Seddik, Mohamed El Amine, et al.
Published: (2024)
Diffusion-Based Generative Priors for Efficient Beam Alignment in Directional Networks
by: Othman, Esraa Fahmy, et al.
Published: (2026)
by: Othman, Esraa Fahmy, et al.
Published: (2026)
GenDFIR: Advancing Cyber Incident Timeline Analysis Through Retrieval Augmented Generation and Large Language Models
by: Loumachi, Fatma Yasmine, et al.
Published: (2024)
by: Loumachi, Fatma Yasmine, et al.
Published: (2024)
Diversity and Inclusion: Are We Nearly There Yet?
by: Eikhof, Doris Ruth
Published: (2023)
by: Eikhof, Doris Ruth
Published: (2023)
Similar Items
-
Generative AI in Cybersecurity: A Comprehensive Review of LLM Applications and Vulnerabilities
by: Ferrag, Mohamed Amine, et al.
Published: (2024) -
Revolutionizing Cyber Threat Detection with Large Language Models: A privacy-preserving BERT-based Lightweight Model for IoT/IIoT Devices
by: Ferrag, Mohamed Amine, et al.
Published: (2023) -
CyberMetric: A Benchmark Dataset based on Retrieval-Augmented Generation for Evaluating LLMs in Cybersecurity Knowledge
by: Tihanyi, Norbert, et al.
Published: (2024) -
Securing Tomorrow's Smart Cities: Investigating Software Security in Internet of Vehicles and Deep Learning Technologies
by: Jain, Ridhi, et al.
Published: (2024) -
Reasoning Beyond Limits: Advances and Open Problems for LLMs
by: Ferrag, Mohamed Amine, et al.
Published: (2025)