BackportBench: A Multilingual Benchmark for Automated Backporting of Patches
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhong, Zhiqing, Huang, Jiaming, He, Pinjia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AutoTestForge: A Multidimensional Automated Testing Framework for Natural Language Processing Models
von: Xing, Hengrui, et al.
Veröffentlicht: (2025)
von: Xing, Hengrui, et al.
Veröffentlicht: (2025)
Jailbreak Distillation: Renewable Safety Benchmarking
von: Zhang, Jingyu, et al.
Veröffentlicht: (2025)
von: Zhang, Jingyu, et al.
Veröffentlicht: (2025)
SecureVibeBench: Benchmarking Secure Vibe Coding of AI Agents via Reconstructing Vulnerability-Introducing Scenarios
von: Chen, Junkai, et al.
Veröffentlicht: (2025)
von: Chen, Junkai, et al.
Veröffentlicht: (2025)
Automated Trustworthiness Oracle Generation for Machine Learning Text Classifiers
von: Tung, Lam Nguyen, et al.
Veröffentlicht: (2024)
von: Tung, Lam Nguyen, et al.
Veröffentlicht: (2024)
PortGPT: Towards Automated Backporting Using Large Language Models
von: Li, Zhaoyang, et al.
Veröffentlicht: (2025)
von: Li, Zhaoyang, et al.
Veröffentlicht: (2025)
A Systematic Study of LLM-Based Architectures for Automated Patching
von: Xu, Qingxiao, et al.
Veröffentlicht: (2026)
von: Xu, Qingxiao, et al.
Veröffentlicht: (2026)
Large Language Models Cannot Reliably Detect Vulnerabilities in JavaScript: The First Systematic Benchmark and Evaluation
von: Fei, Qingyuan, et al.
Veröffentlicht: (2025)
von: Fei, Qingyuan, et al.
Veröffentlicht: (2025)
Argus: Reorchestrating Static Analysis via a Multi-Agent Ensemble for Full-Chain Security Vulnerability Detection
von: Liang, Zi, et al.
Veröffentlicht: (2026)
von: Liang, Zi, et al.
Veröffentlicht: (2026)
PATCHEVAL: A New Benchmark for Evaluating LLMs on Patching Real-World Vulnerabilities
von: Wei, Zichao, et al.
Veröffentlicht: (2025)
von: Wei, Zichao, et al.
Veröffentlicht: (2025)
Automated Repair of TEE Partitioning Issues via DSL-Guided and LLM-Assisted Patching
von: Ma, Chengyan, et al.
Veröffentlicht: (2026)
von: Ma, Chengyan, et al.
Veröffentlicht: (2026)
APPATCH: Automated Adaptive Prompting Large Language Models for Real-World Software Vulnerability Patching
von: Nong, Yu, et al.
Veröffentlicht: (2024)
von: Nong, Yu, et al.
Veröffentlicht: (2024)
ProSec: Fortifying Code LLMs with Proactive Security Alignment
von: Xu, Xiangzhe, et al.
Veröffentlicht: (2024)
von: Xu, Xiangzhe, et al.
Veröffentlicht: (2024)
AC4: Algebraic Computation Checker for Circuit Constraints in ZKPs
von: Yang, Qizhe, et al.
Veröffentlicht: (2024)
von: Yang, Qizhe, et al.
Veröffentlicht: (2024)
Evaluation of the Programming Skills of Large Language Models
von: Heitz, Luc Bryan, et al.
Veröffentlicht: (2024)
von: Heitz, Luc Bryan, et al.
Veröffentlicht: (2024)
PoLLMgraph: Unraveling Hallucinations in Large Language Models via State Transition Dynamics
von: Zhu, Derui, et al.
Veröffentlicht: (2024)
von: Zhu, Derui, et al.
Veröffentlicht: (2024)
Minimal Prompt Perturbations Lead to Code Vulnerabilities: Prompt Fragility and Hidden-State Signals in Coding LLMs
von: Sternfeld, Alexander, et al.
Veröffentlicht: (2026)
von: Sternfeld, Alexander, et al.
Veröffentlicht: (2026)
PatchFuzz: Patch Fuzzing for JavaScript Engines
von: Wang, Junjie, et al.
Veröffentlicht: (2025)
von: Wang, Junjie, et al.
Veröffentlicht: (2025)
TOSSS: a CVE-based Software Security Benchmark for Large Language Models
von: Damie, Marc, et al.
Veröffentlicht: (2026)
von: Damie, Marc, et al.
Veröffentlicht: (2026)
SafeToolBench: Pioneering a Prospective Benchmark to Evaluating Tool Utilization Safety in LLMs
von: Xia, Hongfei, et al.
Veröffentlicht: (2025)
von: Xia, Hongfei, et al.
Veröffentlicht: (2025)
Revisiting Vulnerability Patch Identification on Data in the Wild
von: Irsan, Ivana Clairine, et al.
Veröffentlicht: (2026)
von: Irsan, Ivana Clairine, et al.
Veröffentlicht: (2026)
Mono: Is Your "Clean" Vulnerability Dataset Really Solvable? Exposing and Trapping Undecidable Patches and Beyond
von: Gao, Zeyu, et al.
Veröffentlicht: (2025)
von: Gao, Zeyu, et al.
Veröffentlicht: (2025)
Similar but Patched Code Considered Harmful -- The Impact of Similar but Patched Code on Recurring Vulnerability Detection and How to Remove Them
von: Tan, Zixuan, et al.
Veröffentlicht: (2024)
von: Tan, Zixuan, et al.
Veröffentlicht: (2024)
Static Semantics Reconstruction for Enhancing JavaScript-WebAssembly Multilingual Malware Detection
von: Xia, Yifan, et al.
Veröffentlicht: (2023)
von: Xia, Yifan, et al.
Veröffentlicht: (2023)
A Slicing-Based Approach for Detecting and Patching Vulnerable Code Clones
von: Alomari, Hakam, et al.
Veröffentlicht: (2025)
von: Alomari, Hakam, et al.
Veröffentlicht: (2025)
An Investigation of Patch Porting Practices of the Linux Kernel Ecosystem
von: Li, Xingyu, et al.
Veröffentlicht: (2024)
von: Li, Xingyu, et al.
Veröffentlicht: (2024)
PPT4J: Patch Presence Test for Java Binaries
von: Pan, Zhiyuan, et al.
Veröffentlicht: (2023)
von: Pan, Zhiyuan, et al.
Veröffentlicht: (2023)
Artemis: Toward Accurate Detection of Server-Side Request Forgeries through LLM-Assisted Inter-Procedural Path-Sensitive Taint Analysis
von: Ji, Yuchen, et al.
Veröffentlicht: (2025)
von: Ji, Yuchen, et al.
Veröffentlicht: (2025)
Online Safety Analysis for LLMs: a Benchmark, an Assessment, and a Path Forward
von: Xie, Xuan, et al.
Veröffentlicht: (2024)
von: Xie, Xuan, et al.
Veröffentlicht: (2024)
Safety Interventions against Adversarial Patches in an Open-Source Driver Assistance System
von: Chen, Cheng, et al.
Veröffentlicht: (2025)
von: Chen, Cheng, et al.
Veröffentlicht: (2025)
From LLMs to Agents: A Comparative Evaluation of LLMs and LLM-based Agents in Security Patch Detection
von: Han, Junxiao, et al.
Veröffentlicht: (2025)
von: Han, Junxiao, et al.
Veröffentlicht: (2025)
Vulnerability Patching Across Software Products and Software Components: A Case Study of Red Hat's Product Portfolio
von: Ruohonen, Jukka, et al.
Veröffentlicht: (2025)
von: Ruohonen, Jukka, et al.
Veröffentlicht: (2025)
Patch-to-PoC: A Systematic Study of Agentic LLM Systems for Linux Kernel N-Day Reproduction
von: Pu, Juefei, et al.
Veröffentlicht: (2026)
von: Pu, Juefei, et al.
Veröffentlicht: (2026)
SCRIBE: Practical Static Binary Patching via Binary-Aware Recompilation of Decompiled Code
von: Dai, Han, et al.
Veröffentlicht: (2026)
von: Dai, Han, et al.
Veröffentlicht: (2026)
StriderSPD: Structure-Guided Joint Representation Learning for Binary Security Patch Detection
von: Li, Qingyuan, et al.
Veröffentlicht: (2026)
von: Li, Qingyuan, et al.
Veröffentlicht: (2026)
FuzzingBrain V2: A Multi-Agent LLM System for Automated Vulnerability Discovery and Reproduction
von: Sheng, Ze, et al.
Veröffentlicht: (2026)
von: Sheng, Ze, et al.
Veröffentlicht: (2026)
HYDRA: A Hybrid Heuristic-Guided Deep Representation Architecture for Predicting Latent Zero-Day Vulnerabilities in Patched Functions
von: Farhad, Mohammad, et al.
Veröffentlicht: (2025)
von: Farhad, Mohammad, et al.
Veröffentlicht: (2025)
When "Correct" Is Not Safe: Can We Trust Functionally Correct Patches Generated by Code Agents?
von: Peng, Yibo, et al.
Veröffentlicht: (2025)
von: Peng, Yibo, et al.
Veröffentlicht: (2025)
PatchSeeker: Mapping NVD Records to their Vulnerability-fixing Commits with LLM Generated Commits and Embeddings
von: Nguyen, Huu Hung, et al.
Veröffentlicht: (2025)
von: Nguyen, Huu Hung, et al.
Veröffentlicht: (2025)
Match & Mend: Minimally Invasive Local Reassembly for Patching N-day Vulnerabilities in ARM Binaries
von: Jänich, Sebastian, et al.
Veröffentlicht: (2025)
von: Jänich, Sebastian, et al.
Veröffentlicht: (2025)
A Systematic Approach to Predict the Impact of Cybersecurity Vulnerabilities Using LLMs
von: Høst, Anders Mølmen, et al.
Veröffentlicht: (2025)
von: Høst, Anders Mølmen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
AutoTestForge: A Multidimensional Automated Testing Framework for Natural Language Processing Models
von: Xing, Hengrui, et al.
Veröffentlicht: (2025) -
Jailbreak Distillation: Renewable Safety Benchmarking
von: Zhang, Jingyu, et al.
Veröffentlicht: (2025) -
SecureVibeBench: Benchmarking Secure Vibe Coding of AI Agents via Reconstructing Vulnerability-Introducing Scenarios
von: Chen, Junkai, et al.
Veröffentlicht: (2025) -
Automated Trustworthiness Oracle Generation for Machine Learning Text Classifiers
von: Tung, Lam Nguyen, et al.
Veröffentlicht: (2024) -
PortGPT: Towards Automated Backporting Using Large Language Models
von: Li, Zhaoyang, et al.
Veröffentlicht: (2025)