RealSec-bench: A Benchmark for Evaluating Secure Code Generation in Real-World Repositories
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yanlin, Zhang, Ziyao, Wang, Chong, Xu, Xinyi, Liu, Mingwei, Wang, Yong, Chen, Jiachi, Zheng, Zibin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MCGMark: An Encodable and Robust Online Watermark for Tracing LLM-Generated Malicious Code
von: Ning, Kaiwen, et al.
Veröffentlicht: (2024)
von: Ning, Kaiwen, et al.
Veröffentlicht: (2024)
Unity is Strength: Enhancing Precision in Reentrancy Vulnerability Detection of Smart Contract Analysis Tools
von: Wang, Zexu, et al.
Veröffentlicht: (2024)
von: Wang, Zexu, et al.
Veröffentlicht: (2024)
SecCodePRM: A Process Reward Model for Code Security
von: Yu, Weichen, et al.
Veröffentlicht: (2026)
von: Yu, Weichen, et al.
Veröffentlicht: (2026)
SmartOracle: Generating Smart Contract Oracle via Fine-Grained Invariant Detection
von: Su, Jianzhong, et al.
Veröffentlicht: (2024)
von: Su, Jianzhong, et al.
Veröffentlicht: (2024)
Uncover the Premeditated Attacks: Detecting Exploitable Reentrancy Vulnerabilities by Identifying Attacker Contracts
von: Yang, Shuo, et al.
Veröffentlicht: (2024)
von: Yang, Shuo, et al.
Veröffentlicht: (2024)
ProSec: Fortifying Code LLMs with Proactive Security Alignment
von: Xu, Xiangzhe, et al.
Veröffentlicht: (2024)
von: Xu, Xiangzhe, et al.
Veröffentlicht: (2024)
PATCHEVAL: A New Benchmark for Evaluating LLMs on Patching Real-World Vulnerabilities
von: Wei, Zichao, et al.
Veröffentlicht: (2025)
von: Wei, Zichao, et al.
Veröffentlicht: (2025)
An Empirical Study on the Security Vulnerabilities of GPTs
von: Wu, Tong, et al.
Veröffentlicht: (2025)
von: Wu, Tong, et al.
Veröffentlicht: (2025)
RiskTagger: An LLM-based Agent for Automatic Annotation of Web3 Crypto Money Laundering Behaviors
von: Lin, Dan, et al.
Veröffentlicht: (2025)
von: Lin, Dan, et al.
Veröffentlicht: (2025)
Demystifying and Detecting Cryptographic Defects in Ethereum Smart Contracts
von: Zhang, Jiashuo, et al.
Veröffentlicht: (2024)
von: Zhang, Jiashuo, et al.
Veröffentlicht: (2024)
LLM Hallucinations in Practical Code Generation: Phenomena, Mechanism, and Mitigation
von: Zhang, Ziyao, et al.
Veröffentlicht: (2024)
von: Zhang, Ziyao, et al.
Veröffentlicht: (2024)
One Signature, Multiple Payments: Demystifying and Detecting Signature Replay Vulnerabilities in Smart Contracts
von: Wang, Zexu, et al.
Veröffentlicht: (2025)
von: Wang, Zexu, et al.
Veröffentlicht: (2025)
SecMLOps: A Comprehensive Framework for Integrating Security Throughout the MLOps Lifecycle
von: Zhang, Xinrui, et al.
Veröffentlicht: (2026)
von: Zhang, Xinrui, et al.
Veröffentlicht: (2026)
KVerus: Scalable and Resilient Formal Verification Proof Generation for Rust Code
von: Liu, Yuwei, et al.
Veröffentlicht: (2026)
von: Liu, Yuwei, et al.
Veröffentlicht: (2026)
SecDOAR: A Software Reference Architecture for Security Data Orchestration, Analysis and Reporting
von: Chauhan, Muhammad Aufeef, et al.
Veröffentlicht: (2024)
von: Chauhan, Muhammad Aufeef, et al.
Veröffentlicht: (2024)
Integrating Log-Based Security Analytics in Agile Workflows: A Real-World Experience Report
von: Thool, Arpit, et al.
Veröffentlicht: (2026)
von: Thool, Arpit, et al.
Veröffentlicht: (2026)
An Empirical Security Evaluation of LLM-Generated Cryptographic Rust Code
von: Elsayed, Mohamed, et al.
Veröffentlicht: (2026)
von: Elsayed, Mohamed, et al.
Veröffentlicht: (2026)
Exploring the Security Threats of Retriever Backdoors in Retrieval-Augmented Code Generation
von: Li, Tian, et al.
Veröffentlicht: (2025)
von: Li, Tian, et al.
Veröffentlicht: (2025)
Give LLMs a Security Course: Securing Retrieval-Augmented Code Generation via Knowledge Injection
von: Lin, Bo, et al.
Veröffentlicht: (2025)
von: Lin, Bo, et al.
Veröffentlicht: (2025)
Towards Understanding and Characterizing Vulnerabilities in Intelligent Connected Vehicles through Real-World Exploits
von: Wang, Yuelin, et al.
Veröffentlicht: (2026)
von: Wang, Yuelin, et al.
Veröffentlicht: (2026)
ContractTinker: LLM-Empowered Vulnerability Repair for Real-World Smart Contracts
von: Wang, Che, et al.
Veröffentlicht: (2024)
von: Wang, Che, et al.
Veröffentlicht: (2024)
VulEval: Towards Repository-Level Evaluation of Software Vulnerability Detection
von: Wen, Xin-Cheng, et al.
Veröffentlicht: (2024)
von: Wen, Xin-Cheng, et al.
Veröffentlicht: (2024)
Exploring the Security Threats of Knowledge Base Poisoning in Retrieval-Augmented Code Generation
von: Lin, Bo, et al.
Veröffentlicht: (2025)
von: Lin, Bo, et al.
Veröffentlicht: (2025)
Repository-Level Graph Representation Learning for Enhanced Security Patch Detection
von: Wen, Xin-Cheng, et al.
Veröffentlicht: (2024)
von: Wen, Xin-Cheng, et al.
Veröffentlicht: (2024)
Benchmark of Benchmarks: Unpacking Influence and Code Repository Quality in LLM Safety Benchmarks
von: Chu, Junjie, et al.
Veröffentlicht: (2026)
von: Chu, Junjie, et al.
Veröffentlicht: (2026)
SecCodeBench-V2 Technical Report
von: Chen, Longfei, et al.
Veröffentlicht: (2026)
von: Chen, Longfei, et al.
Veröffentlicht: (2026)
An Empirical Study on Oculus Virtual Reality Applications: Security and Privacy Perspectives
von: Guo, Hanyang, et al.
Veröffentlicht: (2024)
von: Guo, Hanyang, et al.
Veröffentlicht: (2024)
Decoding Secret Memorization in Code LLMs Through Token-Level Characterization
von: Nie, Yuqing, et al.
Veröffentlicht: (2024)
von: Nie, Yuqing, et al.
Veröffentlicht: (2024)
Models Are Codes: Towards Measuring Malicious Code Poisoning Attacks on Pre-trained Model Hubs
von: Zhao, Jian, et al.
Veröffentlicht: (2024)
von: Zhao, Jian, et al.
Veröffentlicht: (2024)
Does Teaming-Up LLMs Improve Secure Code Generation? A Comprehensive Evaluation with Multi-LLMSecCodeEval
von: Sabir, Bushra, et al.
Veröffentlicht: (2026)
von: Sabir, Bushra, et al.
Veröffentlicht: (2026)
DUALGUAGE: Automated Joint Security-Functionality Benchmarking for Secure Code Generation
von: Pathak, Abhijeet, et al.
Veröffentlicht: (2025)
von: Pathak, Abhijeet, et al.
Veröffentlicht: (2025)
Towards Secure Logging: Characterizing and Benchmarking Logging Code Security Issues with LLMs
von: Yuan, He Yang, et al.
Veröffentlicht: (2026)
von: Yuan, He Yang, et al.
Veröffentlicht: (2026)
Rethinking the Evaluation of Secure Code Generation
von: Dai, Shih-Chieh, et al.
Veröffentlicht: (2025)
von: Dai, Shih-Chieh, et al.
Veröffentlicht: (2025)
How Secure is Secure Code Generation? Adversarial Prompts Put LLM Defenses to the Test
von: Tessa, Melissa, et al.
Veröffentlicht: (2026)
von: Tessa, Melissa, et al.
Veröffentlicht: (2026)
RepoTransBench: A Real-World Multilingual Benchmark for Repository-Level Code Translation
von: Wang, Yanli, et al.
Veröffentlicht: (2024)
von: Wang, Yanli, et al.
Veröffentlicht: (2024)
Uncovering Hidden Inclusions of Vulnerable Dependencies in Real-World Java Projects
von: Schott, Stefan, et al.
Veröffentlicht: (2026)
von: Schott, Stefan, et al.
Veröffentlicht: (2026)
Real-World Usability of Vulnerability Proof-of-Concepts: A Comprehensive Study
von: Dang, Wenjing, et al.
Veröffentlicht: (2025)
von: Dang, Wenjing, et al.
Veröffentlicht: (2025)
SCAFFOLD-CEGIS: Preventing Latent Security Degradation in LLM-Driven Iterative Code Refinement
von: Chen, Yi, et al.
Veröffentlicht: (2026)
von: Chen, Yi, et al.
Veröffentlicht: (2026)
ReposVul: A Repository-Level High-Quality Vulnerability Dataset
von: Wang, Xinchen, et al.
Veröffentlicht: (2024)
von: Wang, Xinchen, et al.
Veröffentlicht: (2024)
RustRepoTrans: Repository-level Code Translation Benchmark Targeting Rust
von: Ou, Guangsheng, et al.
Veröffentlicht: (2024)
von: Ou, Guangsheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MCGMark: An Encodable and Robust Online Watermark for Tracing LLM-Generated Malicious Code
von: Ning, Kaiwen, et al.
Veröffentlicht: (2024) -
Unity is Strength: Enhancing Precision in Reentrancy Vulnerability Detection of Smart Contract Analysis Tools
von: Wang, Zexu, et al.
Veröffentlicht: (2024) -
SecCodePRM: A Process Reward Model for Code Security
von: Yu, Weichen, et al.
Veröffentlicht: (2026) -
SmartOracle: Generating Smart Contract Oracle via Fine-Grained Invariant Detection
von: Su, Jianzhong, et al.
Veröffentlicht: (2024) -
Uncover the Premeditated Attacks: Detecting Exploitable Reentrancy Vulnerabilities by Identifying Attacker Contracts
von: Yang, Shuo, et al.
Veröffentlicht: (2024)