Does Teaming-Up LLMs Improve Secure Code Generation? A Comprehensive Evaluation with Multi-LLMSecCodeEval
Fuente:
arXiv
Saved in:
| Main Authors: | Sabir, Bushra, Liu, Shigang, Jang, Seung Ick, Abuadbba, Sharif, Gao, Yansong, Moore, Kristen, Kim, SangCheol, Kim, Hyoungshick, Nepal, Surya |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
From Solitary Directives to Interactive Encouragement! LLM Secure Code Generation by Natural Language Prompting
by: Liu, Shigang, et al.
Published: (2024)
by: Liu, Shigang, et al.
Published: (2024)
DeepTaster: Adversarial Perturbation-Based Fingerprinting to Identify Proprietary Dataset Use in Deep Neural Networks
by: Park, Seonhye, et al.
Published: (2022)
by: Park, Seonhye, et al.
Published: (2022)
Comprehensive Evaluation of Cloaking Backdoor Attacks on Object Detector in Real-World
by: Ma, Hua, et al.
Published: (2025)
by: Ma, Hua, et al.
Published: (2025)
DeepiSign-G: Generic Watermark to Stamp Hidden DNN Parameters for Self-contained Tracking
by: Abuadbba, Alsharif, et al.
Published: (2024)
by: Abuadbba, Alsharif, et al.
Published: (2024)
Alert-ME: An Explainability-Driven Defense Against Adversarial Examples in Transformer-Based Text Classification
by: Sabir, Bushra, et al.
Published: (2023)
by: Sabir, Bushra, et al.
Published: (2023)
Token-Modification Adversarial Attacks for Natural Language Processing: A Survey
by: Roth, Tom, et al.
Published: (2021)
by: Roth, Tom, et al.
Published: (2021)
LLMSecCode: Evaluating Large Language Models for Secure Coding
by: Rydén, Anton, et al.
Published: (2024)
by: Rydén, Anton, et al.
Published: (2024)
Contextual Chart Generation for Cyber Deception
by: Nguyen, David D., et al.
Published: (2024)
by: Nguyen, David D., et al.
Published: (2024)
A Systematic Evaluation of Parameter-Efficient Fine-Tuning Methods for the Security of Code LLMs
by: Lee, Kiho, et al.
Published: (2025)
by: Lee, Kiho, et al.
Published: (2025)
Human Society-Inspired Approaches to Agentic AI Security: The 4C Framework
by: Abuadbba, Alsharif, et al.
Published: (2026)
by: Abuadbba, Alsharif, et al.
Published: (2026)
Systematic Literature Review of AI-enabled Spectrum Management in 6G and Future Networks
by: Sabir, Bushra, et al.
Published: (2024)
by: Sabir, Bushra, et al.
Published: (2024)
Adversarial Attacks Against Automated Fact-Checking: A Survey
by: Liu, Fanzhen, et al.
Published: (2025)
by: Liu, Fanzhen, et al.
Published: (2025)
ThreatModeling-LLM: Automating Threat Modeling using Large Language Models for Banking System
by: Wu, Tingmin, et al.
Published: (2024)
by: Wu, Tingmin, et al.
Published: (2024)
SoK: Can Trajectory Generation Combine Privacy and Utility?
by: Buchholz, Erik, et al.
Published: (2024)
by: Buchholz, Erik, et al.
Published: (2024)
Towards Faithful Class-level Self-explainability in Graph Neural Networks by Subgraph Dependencies
by: Liu, Fanzhen, et al.
Published: (2025)
by: Liu, Fanzhen, et al.
Published: (2025)
APT-Agent: Automated Penetration Testing using Large Language Models
by: Li, William Guanting, et al.
Published: (2026)
by: Li, William Guanting, et al.
Published: (2026)
From Promise to Peril: Rethinking Cybersecurity Red and Blue Teaming in the Age of LLMs
by: Abuadbba, Alsharif, et al.
Published: (2025)
by: Abuadbba, Alsharif, et al.
Published: (2025)
What is the Cost of Differential Privacy for Deep Learning-Based Trajectory Generation?
by: Buchholz, Erik, et al.
Published: (2025)
by: Buchholz, Erik, et al.
Published: (2025)
Watch Out! Simple Horizontal Class Backdoor Can Trivially Evade Defense
by: Ma, Hua, et al.
Published: (2023)
by: Ma, Hua, et al.
Published: (2023)
A Login Page Transparency and Visual Similarity Based Zero Day Phishing Defense Protocol
by: Varshney, Gaurav, et al.
Published: (2025)
by: Varshney, Gaurav, et al.
Published: (2025)
A2C: A Modular Multi-stage Collaborative Decision Framework for Human-AI Teams
by: Tariq, Shahroz, et al.
Published: (2024)
by: Tariq, Shahroz, et al.
Published: (2024)
5G LDPC Codes as Root LDPC Codes via Diversity Alignment
by: Ahn, Hyuntae, et al.
Published: (2026)
by: Ahn, Hyuntae, et al.
Published: (2026)
SoK: Systematization and Benchmarking of Deepfake Detectors in a Unified Framework
by: Le, Binh M., et al.
Published: (2024)
by: Le, Binh M., et al.
Published: (2024)
Security in the Era of Perceptive Networks: A Comprehensive Taxonomic Framework for Integrated Sensing and Communication Security
by: Thapa, Chandra, et al.
Published: (2026)
by: Thapa, Chandra, et al.
Published: (2026)
Future G Network's New Reality: Opportunities and Security Challenges
by: Thapa, Chandra, et al.
Published: (2025)
by: Thapa, Chandra, et al.
Published: (2025)
When the Abyss Looks Back: Unveiling Evolving Dark Patterns in Cookie Consent Banners
by: Singh, Nivedita, et al.
Published: (2026)
by: Singh, Nivedita, et al.
Published: (2026)
A Study on the Influential Neighbors to Maximize Information Diffusion in Online Social Networks
by: Kim, Hyoungshick, et al.
Published: (2015)
by: Kim, Hyoungshick, et al.
Published: (2015)
Blind-Touch: Homomorphic Encryption-Based Distributed Neural Network Inference for Privacy-Preserving Fingerprint Authentication
by: Choi, Hyunmin, et al.
Published: (2023)
by: Choi, Hyunmin, et al.
Published: (2023)
Deep Learning-Based Out-of-distribution Source Code Data Identification: How Far Have We Gone?
by: Nguyen, Van, et al.
Published: (2024)
by: Nguyen, Van, et al.
Published: (2024)
Can Current Detectors Catch Face-to-Voice Deepfake Attacks?
by: Nguyen, Nguyen Linh Bao, et al.
Published: (2025)
by: Nguyen, Nguyen Linh Bao, et al.
Published: (2025)
Task-Specific Audio Coding for Machines: Machine-Learned Latent Features Are Codes for That Machine
by: Kuznetsova, Anastasia, et al.
Published: (2025)
by: Kuznetsova, Anastasia, et al.
Published: (2025)
Design of Root Protograph LDPC Codes Simultaneously Achieving Full Diversity and High Coding Gain
by: Kim, Inki, et al.
Published: (2026)
by: Kim, Inki, et al.
Published: (2026)
Framework for evaluating code generation ability of large language models
by: Sangyeop Yeo, et al.
Published: (2024)
by: Sangyeop Yeo, et al.
Published: (2024)
LLMSecConfig: An LLM-Based Approach for Fixing Software Container Misconfigurations
by: Ye, Ziyang, et al.
Published: (2025)
by: Ye, Ziyang, et al.
Published: (2025)
EPhishCADE: A Privacy-Aware Multi-Dimensional Framework for Email Phishing Campaign Detection
by: Kang, Wei, et al.
Published: (2025)
by: Kang, Wei, et al.
Published: (2025)
Large Language Model Adversarial Landscape Through the Lens of Attack Objectives
by: Wang, Nan, et al.
Published: (2025)
by: Wang, Nan, et al.
Published: (2025)
ArchCode: Incorporating Software Requirements in Code Generation with Large Language Models
by: Han, Hojae, et al.
Published: (2024)
by: Han, Hojae, et al.
Published: (2024)
An Investigation into Misuse of Java Security APIs by Large Language Models
by: Mousavi, Zahra, et al.
Published: (2024)
by: Mousavi, Zahra, et al.
Published: (2024)
Detecting Misuse of Security APIs: A Systematic Review
by: Mousavi, Zahra, et al.
Published: (2023)
by: Mousavi, Zahra, et al.
Published: (2023)
ContractEval: A Benchmark for Evaluating Contract-Satisfying Assertions in Code Generation
by: Lim, Soohan, et al.
Published: (2025)
by: Lim, Soohan, et al.
Published: (2025)
Similar Items
-
From Solitary Directives to Interactive Encouragement! LLM Secure Code Generation by Natural Language Prompting
by: Liu, Shigang, et al.
Published: (2024) -
DeepTaster: Adversarial Perturbation-Based Fingerprinting to Identify Proprietary Dataset Use in Deep Neural Networks
by: Park, Seonhye, et al.
Published: (2022) -
Comprehensive Evaluation of Cloaking Backdoor Attacks on Object Detector in Real-World
by: Ma, Hua, et al.
Published: (2025) -
DeepiSign-G: Generic Watermark to Stamp Hidden DNN Parameters for Self-contained Tracking
by: Abuadbba, Alsharif, et al.
Published: (2024) -
Alert-ME: An Explainability-Driven Defense Against Adversarial Examples in Transformer-Based Text Classification
by: Sabir, Bushra, et al.
Published: (2023)