Evaluating and Improving the Robustness of Security Attack Detectors Generated by LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Pasini, Samuele, Kim, Jinhan, Aiello, Tommaso, Lozoya, Rocio Cabrera, Sabetta, Antonino, Tonella, Paolo |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Detecting Trojaned DNNs via Spectral Regression Analysis
by: Pasini, Samuele, et al.
Published: (2026)
by: Pasini, Samuele, et al.
Published: (2026)
Cross-site scripting adversarial attacks based on deep reinforcement learning: Evaluation and extension study
by: Pasini, Samuele, et al.
Published: (2025)
by: Pasini, Samuele, et al.
Published: (2025)
A Taxonomy of System-Level Attacks on Deep Learning Models in Autonomous Vehicles
by: Tehrani, Masoud Jamshidiyan, et al.
Published: (2024)
by: Tehrani, Masoud Jamshidiyan, et al.
Published: (2024)
Beyond Metadata: Code-centric and Usage-based Analysis of Known Vulnerabilities in Open-source Software
by: Ponta, Serena E., et al.
Published: (2018)
by: Ponta, Serena E., et al.
Published: (2018)
Impact assessment for vulnerabilities in open-source software libraries
by: Plate, Henrik, et al.
Published: (2015)
by: Plate, Henrik, et al.
Published: (2015)
Does Teaming-Up LLMs Improve Secure Code Generation? A Comprehensive Evaluation with Multi-LLMSecCodeEval
by: Sabir, Bushra, et al.
Published: (2026)
by: Sabir, Bushra, et al.
Published: (2026)
A Manually-Curated Dataset of Fixes to Vulnerabilities of Open-Source Software
by: Ponta, Serena E., et al.
Published: (2019)
by: Ponta, Serena E., et al.
Published: (2019)
Automated Mapping of Vulnerability Advisories onto their Fix Commits in Open Source Repositories
by: Hommersom, Daan, et al.
Published: (2021)
by: Hommersom, Daan, et al.
Published: (2021)
From LLMs to Agents: A Comparative Evaluation of LLMs and LLM-based Agents in Security Patch Detection
by: Han, Junxiao, et al.
Published: (2025)
by: Han, Junxiao, et al.
Published: (2025)
Give LLMs a Security Course: Securing Retrieval-Augmented Code Generation via Knowledge Injection
by: Lin, Bo, et al.
Published: (2025)
by: Lin, Bo, et al.
Published: (2025)
LLMs for Cyber Security: New Opportunities
by: Divakaran, Dinil Mon, et al.
Published: (2024)
by: Divakaran, Dinil Mon, et al.
Published: (2024)
An Empirical Security Evaluation of LLM-Generated Cryptographic Rust Code
by: Elsayed, Mohamed, et al.
Published: (2026)
by: Elsayed, Mohamed, et al.
Published: (2026)
Mutation-based Evaluation of Cryptographic API Misuse Detectors
by: Ami, Amit Seal, et al.
Published: (2021)
by: Ami, Amit Seal, et al.
Published: (2021)
An In Depth Analysis of a Cyber Attack: Case Study and Security Insights
by: Pakshad, Puya
Published: (2024)
by: Pakshad, Puya
Published: (2024)
Using LLMs for Security Advisory Investigations: How Far Are We?
by: Abdullah, Bayu Fedra, et al.
Published: (2025)
by: Abdullah, Bayu Fedra, et al.
Published: (2025)
RealSec-bench: A Benchmark for Evaluating Secure Code Generation in Real-World Repositories
by: Wang, Yanlin, et al.
Published: (2026)
by: Wang, Yanlin, et al.
Published: (2026)
FLAMES: Fine-tuning LLMs to Synthesize Invariants for Smart Contract Security
by: Eshghie, Mojtaba, et al.
Published: (2025)
by: Eshghie, Mojtaba, et al.
Published: (2025)
A User-centered Security Evaluation of Copilot
by: Asare, Owura, et al.
Published: (2023)
by: Asare, Owura, et al.
Published: (2023)
Improving Smart Contract Security with Contrastive Learning-based Vulnerability Detection
by: Chen, Yizhou, et al.
Published: (2024)
by: Chen, Yizhou, et al.
Published: (2024)
R+R: Reassessing Java Security API Misuse in Current LLMs: A Replication on JCA and JSSE APIs with External Security Knowledge
by: Lu, Tianhe, et al.
Published: (2026)
by: Lu, Tianhe, et al.
Published: (2026)
How Can ChatGPT Support Human Security Testers to Help Mitigate Supply Chain Attacks?
by: Zhang, Ying, et al.
Published: (2023)
by: Zhang, Ying, et al.
Published: (2023)
Evaluating Software Supply Chain Security in Research Software
by: Hegewald, Richard, et al.
Published: (2025)
by: Hegewald, Richard, et al.
Published: (2025)
How Secure is Secure Code Generation? Adversarial Prompts Put LLM Defenses to the Test
by: Tessa, Melissa, et al.
Published: (2026)
by: Tessa, Melissa, et al.
Published: (2026)
Security Evaluation of Android apps in budget African Mobile Devices
by: Diallo, Alioune, et al.
Published: (2025)
by: Diallo, Alioune, et al.
Published: (2025)
VulInstruct: Teaching LLMs Root-Cause Reasoning for Vulnerability Detection via Security Specifications
by: Zhu, Hao, et al.
Published: (2025)
by: Zhu, Hao, et al.
Published: (2025)
Automatic Attack Script Generation: a MDA Approach
by: Goux, Quentin, et al.
Published: (2026)
by: Goux, Quentin, et al.
Published: (2026)
Automated Attack Synthesis for Constant Product Market Makers
by: Han, Sujin, et al.
Published: (2024)
by: Han, Sujin, et al.
Published: (2024)
Towards Robust Detection of Open Source Software Supply Chain Poisoning Attacks in Industry Environments
by: Zheng, Xinyi, et al.
Published: (2024)
by: Zheng, Xinyi, et al.
Published: (2024)
Evaluating the Role of Security Assurance Cases in Agile Medical Device Development
by: Fransson, Max, et al.
Published: (2024)
by: Fransson, Max, et al.
Published: (2024)
Detecting Stealthy Data Poisoning Attacks in AI Code Generators
by: Improta, Cristina
Published: (2025)
by: Improta, Cristina
Published: (2025)
Reinforcement Learning-Driven Adaptation Chains: A Robust Framework for Multi-Cloud Workflow Security
by: Soveizi, Nafiseh, et al.
Published: (2025)
by: Soveizi, Nafiseh, et al.
Published: (2025)
Rethinking the Evaluation of Secure Code Generation
by: Dai, Shih-Chieh, et al.
Published: (2025)
by: Dai, Shih-Chieh, et al.
Published: (2025)
The Power of Words: Generating PowerShell Attacks from Natural Language
by: Liguori, Pietro, et al.
Published: (2024)
by: Liguori, Pietro, et al.
Published: (2024)
Unknown Attack Detection in IoT Networks using Large Language Models: A Robust, Data-efficient Approach
by: Ali, Shan, et al.
Published: (2026)
by: Ali, Shan, et al.
Published: (2026)
LLMs + Security = Trouble
by: Livshits, Benjamin
Published: (2026)
by: Livshits, Benjamin
Published: (2026)
Generating Proof-of-Vulnerability Tests to Help Enhance the Security of Complex Software
by: Kanchi, Shravya, et al.
Published: (2026)
by: Kanchi, Shravya, et al.
Published: (2026)
Generating API Parameter Security Rules with LLM for API Misuse Detection
by: Liu, Jinghua, et al.
Published: (2024)
by: Liu, Jinghua, et al.
Published: (2024)
Exploring the Security Threats of Retriever Backdoors in Retrieval-Augmented Code Generation
by: Li, Tian, et al.
Published: (2025)
by: Li, Tian, et al.
Published: (2025)
FDI: Attack Neural Code Generation Systems through User Feedback Channel
by: Sun, Zhensu, et al.
Published: (2024)
by: Sun, Zhensu, et al.
Published: (2024)
BIDO: An Out-Of-Distribution Resistant Image-based Malware Detector
by: Wang, Wei, et al.
Published: (2025)
by: Wang, Wei, et al.
Published: (2025)
Similar Items
-
Detecting Trojaned DNNs via Spectral Regression Analysis
by: Pasini, Samuele, et al.
Published: (2026) -
Cross-site scripting adversarial attacks based on deep reinforcement learning: Evaluation and extension study
by: Pasini, Samuele, et al.
Published: (2025) -
A Taxonomy of System-Level Attacks on Deep Learning Models in Autonomous Vehicles
by: Tehrani, Masoud Jamshidiyan, et al.
Published: (2024) -
Beyond Metadata: Code-centric and Usage-based Analysis of Known Vulnerabilities in Open-source Software
by: Ponta, Serena E., et al.
Published: (2018) -
Impact assessment for vulnerabilities in open-source software libraries
by: Plate, Henrik, et al.
Published: (2015)