Coding-PTMs: How to Find Optimal Code Pre-trained Models for Code Embedding in Vulnerability Detection?
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Yu, Gong, Lina, Huang, Zhiqiu, Wang, Yongwei, Wei, Mingqiang, Wu, Fei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DeMuVGN: Effective Software Defect Prediction Model by Learning Multi-view Software Dependency via Graph Neural Networks
by: Qiao, Yu, et al.
Published: (2024)
by: Qiao, Yu, et al.
Published: (2024)
Optimizing Code Embeddings and ML Classifiers for Python Source Code Vulnerability Detection
by: Farasat, Talaya, et al.
Published: (2025)
by: Farasat, Talaya, et al.
Published: (2025)
Models Are Codes: Towards Measuring Malicious Code Poisoning Attacks on Pre-trained Model Hubs
by: Zhao, Jian, et al.
Published: (2024)
by: Zhao, Jian, et al.
Published: (2024)
Code Membership Inference for Detecting Unauthorized Data Use in Code Pre-trained Language Models
by: Zhang, Sheng, et al.
Published: (2023)
by: Zhang, Sheng, et al.
Published: (2023)
StagedVulBERT: Multi-Granular Vulnerability Detection with a Novel Pre-trained Code Model
by: Jiang, Yuan, et al.
Published: (2024)
by: Jiang, Yuan, et al.
Published: (2024)
Directional Diffusion-Style Code Editing Pre-training
by: Liang, Qingyuan, et al.
Published: (2025)
by: Liang, Qingyuan, et al.
Published: (2025)
On the Effect of Token Merging on Pre-trained Models for Code
by: Saad, Mootez, et al.
Published: (2025)
by: Saad, Mootez, et al.
Published: (2025)
You Only Train Once: A Flexible Training Framework for Code Vulnerability Detection Driven by Vul-Vector
by: Tian, Bowen, et al.
Published: (2025)
by: Tian, Bowen, et al.
Published: (2025)
Inducing Vulnerable Code Generation in LLM Coding Assistants
by: Zeng, Binqi, et al.
Published: (2025)
by: Zeng, Binqi, et al.
Published: (2025)
Vulnerability Detection with Code Language Models: How Far Are We?
by: Ding, Yangruibo, et al.
Published: (2024)
by: Ding, Yangruibo, et al.
Published: (2024)
CodeCSE: A Simple Multilingual Model for Code and Comment Sentence Embeddings
by: Varkey, Anthony, et al.
Published: (2024)
by: Varkey, Anthony, et al.
Published: (2024)
Natural Is The Best: Model-Agnostic Code Simplification for Pre-trained Large Language Models
by: Wang, Yan, et al.
Published: (2024)
by: Wang, Yan, et al.
Published: (2024)
CoderEval: A Benchmark of Pragmatic Code Generation with Generative Pre-trained Models
by: Yu, Hao, et al.
Published: (2023)
by: Yu, Hao, et al.
Published: (2023)
CodeImprove: Program Adaptation for Deep Code Models
by: Rathnasuriya, Ravishka, et al.
Published: (2025)
by: Rathnasuriya, Ravishka, et al.
Published: (2025)
Security Is Relative: Training-Free Vulnerability Detection via Multi-Agent Behavioral Contract Synthesis
by: Wang, Yongchao, et al.
Published: (2026)
by: Wang, Yongchao, et al.
Published: (2026)
A Vulnerability Code Intent Summary Dataset
by: Huang, Yifan, et al.
Published: (2025)
by: Huang, Yifan, et al.
Published: (2025)
Line-level Semantic Structure Learning for Code Vulnerability Detection
by: Wang, Ziliang, et al.
Published: (2024)
by: Wang, Ziliang, et al.
Published: (2024)
Bridge and Hint: Extending Pre-trained Language Models for Long-Range Code
by: Chen, Yujia, et al.
Published: (2024)
by: Chen, Yujia, et al.
Published: (2024)
Ensembling Large Language Models for Code Vulnerability Detection: An Empirical Evaluation
by: Sun, Zhihong, et al.
Published: (2025)
by: Sun, Zhihong, et al.
Published: (2025)
How Code Representation Shapes False-Positive Dynamics in Cross-Language LLM Vulnerability Detection
by: Chen, Maofei, et al.
Published: (2026)
by: Chen, Maofei, et al.
Published: (2026)
Build Code is Still Code: Finding the Antidote for Pipeline Poisoning
by: Pappas, Brent, et al.
Published: (2026)
by: Pappas, Brent, et al.
Published: (2026)
M2CVD: Enhancing Vulnerability Semantic through Multi-Model Collaboration for Code Vulnerability Detection
by: Wang, Ziliang, et al.
Published: (2024)
by: Wang, Ziliang, et al.
Published: (2024)
Unsupervised Binary Code Translation with Application to Code Similarity Detection and Vulnerability Discovery
by: Ahmad, Iftakhar, et al.
Published: (2024)
by: Ahmad, Iftakhar, et al.
Published: (2024)
K-ASTRO: Structure-Aware Adaptation of LLMs for Code Vulnerability Detection
by: Zhang, Yifan, et al.
Published: (2022)
by: Zhang, Yifan, et al.
Published: (2022)
Keep It Simple: Towards Accurate Vulnerability Detection for Large Code Graphs
by: Peng, Xin, et al.
Published: (2024)
by: Peng, Xin, et al.
Published: (2024)
Similar but Patched Code Considered Harmful -- The Impact of Similar but Patched Code on Recurring Vulnerability Detection and How to Remove Them
by: Tan, Zixuan, et al.
Published: (2024)
by: Tan, Zixuan, et al.
Published: (2024)
Learning to Focus: Context Extraction for Efficient Code Vulnerability Detection with Language Models
by: Zheng, Xinran, et al.
Published: (2025)
by: Zheng, Xinran, et al.
Published: (2025)
Code Vulnerability Detection: A Comparative Analysis of Emerging Large Language Models
by: Sultana, Shaznin, et al.
Published: (2024)
by: Sultana, Shaznin, et al.
Published: (2024)
Structure-aware Fine-tuning for Code Pre-trained Models
by: Wu, Jiayi, et al.
Published: (2024)
by: Wu, Jiayi, et al.
Published: (2024)
LEANCODE: Understanding Models Better for Code Simplification of Pre-trained Large Language Models
by: Wang, Yan, et al.
Published: (2025)
by: Wang, Yan, et al.
Published: (2025)
Evaluating Small-Scale Code Models for Code Clone Detection
by: Martinez-Gil, Jorge
Published: (2025)
by: Martinez-Gil, Jorge
Published: (2025)
Is Vibe Coding Safe? Benchmarking Vulnerability of Agent-Generated Code in Real-World Tasks
by: Zhao, Songwen, et al.
Published: (2025)
by: Zhao, Songwen, et al.
Published: (2025)
Studying Vulnerable Code Entities in R
by: Zhao, Zixiao, et al.
Published: (2024)
by: Zhao, Zixiao, et al.
Published: (2024)
Learning in the Wild: Towards Leveraging Unlabeled Data for Effectively Tuning Pre-trained Code Models
by: Gao, Shuzheng, et al.
Published: (2024)
by: Gao, Shuzheng, et al.
Published: (2024)
Improving Automated Secure Code Reviews: A Synthetic Dataset for Code Vulnerability Flaws
by: Centellas-Claros, Leonardo, et al.
Published: (2025)
by: Centellas-Claros, Leonardo, et al.
Published: (2025)
CodeT5-RNN: Reinforcing Contextual Embeddings for Enhanced Code Comprehension
by: Rahman, Md Mostafizer, et al.
Published: (2026)
by: Rahman, Md Mostafizer, et al.
Published: (2026)
UITrans: Seamless UI Translation from Android to HarmonyOS
by: Gong, Lina, et al.
Published: (2024)
by: Gong, Lina, et al.
Published: (2024)
GraphCoder: Enhancing Repository-Level Code Completion via Code Context Graph-based Retrieval and Language Model
by: Liu, Wei, et al.
Published: (2024)
by: Liu, Wei, et al.
Published: (2024)
VFDelta: A Framework for Detecting Silent Vulnerability Fixes by Enhancing Code Change Learning
by: Yang, Xu, et al.
Published: (2024)
by: Yang, Xu, et al.
Published: (2024)
Exploring Multi-Lingual Bias of Large Code Models in Code Generation
by: Wang, Chaozheng, et al.
Published: (2024)
by: Wang, Chaozheng, et al.
Published: (2024)
Similar Items
-
DeMuVGN: Effective Software Defect Prediction Model by Learning Multi-view Software Dependency via Graph Neural Networks
by: Qiao, Yu, et al.
Published: (2024) -
Optimizing Code Embeddings and ML Classifiers for Python Source Code Vulnerability Detection
by: Farasat, Talaya, et al.
Published: (2025) -
Models Are Codes: Towards Measuring Malicious Code Poisoning Attacks on Pre-trained Model Hubs
by: Zhao, Jian, et al.
Published: (2024) -
Code Membership Inference for Detecting Unauthorized Data Use in Code Pre-trained Language Models
by: Zhang, Sheng, et al.
Published: (2023) -
StagedVulBERT: Multi-Granular Vulnerability Detection with a Novel Pre-trained Code Model
by: Jiang, Yuan, et al.
Published: (2024)