A Survey of Trojans in Neural Models of Source Code: Taxonomy and Techniques
Fuente:
arXiv
Saved in:
| Main Authors: | Hussain, Aftab, Rabin, Md Rafiqul Islam, Ahmed, Toufique, Ayoobi, Navid, Xu, Bowen, Devanbu, Prem, Alipour, Mohammad Amin |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Trojans in Large Language Models of Code: A Critical Review through a Trigger-Based Taxonomy
by: Hussain, Aftab, et al.
Published: (2024)
by: Hussain, Aftab, et al.
Published: (2024)
Measuring Impacts of Poisoning on Model Parameters and Neuron Activations: A Case Study of Poisoning CodeBERT
by: Hussain, Aftab, et al.
Published: (2024)
by: Hussain, Aftab, et al.
Published: (2024)
On Trojan Signatures in Large Language Models of Code
by: Hussain, Aftab, et al.
Published: (2024)
by: Hussain, Aftab, et al.
Published: (2024)
Unlearning Trojans in Large Language Models: A Comparison Between Natural Language and Source Code
by: Kazemi, Mahdi, et al.
Published: (2024)
by: Kazemi, Mahdi, et al.
Published: (2024)
Measuring Impacts of Poisoning on Model Parameters and Embeddings for Large Language Models of Code
by: Hussain, Aftab, et al.
Published: (2024)
by: Hussain, Aftab, et al.
Published: (2024)
Capturing the Effects of Quantization on Trojans in Code LLMs
by: Hussain, Aftab, et al.
Published: (2025)
by: Hussain, Aftab, et al.
Published: (2025)
Calibration and Correctness of Language Models for Code
by: Spiess, Claudio, et al.
Published: (2024)
by: Spiess, Claudio, et al.
Published: (2024)
Harnessing the Power of LLMs: Automating Unit Test Generation for High-Performance Computing
by: Karanjai, Rabimba, et al.
Published: (2024)
by: Karanjai, Rabimba, et al.
Published: (2024)
Calibration of Large Language Models on Code Summarization
by: Virk, Yuvraj, et al.
Published: (2024)
by: Virk, Yuvraj, et al.
Published: (2024)
CoDocBench: A Dataset for Code-Documentation Alignment in Software Maintenance
by: Pai, Kunal, et al.
Published: (2025)
by: Pai, Kunal, et al.
Published: (2025)
Towards Understanding What Code Language Models Learned
by: Ahmed, Toufique, et al.
Published: (2023)
by: Ahmed, Toufique, et al.
Published: (2023)
Localized Calibrated Uncertainty in Code Language Models
by: Gros, David, et al.
Published: (2025)
by: Gros, David, et al.
Published: (2025)
Automatic Semantic Augmentation of Language Model Prompts (for Code Summarization)
by: Ahmed, Toufique, et al.
Published: (2023)
by: Ahmed, Toufique, et al.
Published: (2023)
Studying LLM Performance on Closed- and Open-source Data
by: Ahmed, Toufique, et al.
Published: (2024)
by: Ahmed, Toufique, et al.
Published: (2024)
Can LLMs Replace Manual Annotation of Software Engineering Artifacts?
by: Ahmed, Toufique, et al.
Published: (2024)
by: Ahmed, Toufique, et al.
Published: (2024)
On LLMs' Internal Representation of Code Correctness
by: Ribeiro, Francisco, et al.
Published: (2025)
by: Ribeiro, Francisco, et al.
Published: (2025)
A First Look at the Self-Admitted Technical Debt in Test Code: Taxonomy and Detection
by: Islam, Shahidul, et al.
Published: (2025)
by: Islam, Shahidul, et al.
Published: (2025)
How Robustly do LLMs Understand Execution Semantics?
by: Spiess, Claudio, et al.
Published: (2026)
by: Spiess, Claudio, et al.
Published: (2026)
Model See, Model Do? Exposure-Aware Evaluation of Bug-vs-Fix Preference in Code LLMs
by: Al-Kaswan, Ali, et al.
Published: (2026)
by: Al-Kaswan, Ali, et al.
Published: (2026)
Ecosystem of Large Language Models for Code
by: Yang, Zhou, et al.
Published: (2024)
by: Yang, Zhou, et al.
Published: (2024)
A Taxonomy of Inefficiencies in LLM-Generated Python Code
by: Abbassi, Altaf Allah, et al.
Published: (2025)
by: Abbassi, Altaf Allah, et al.
Published: (2025)
Code Refactoring with LLM: A Comprehensive Evaluation With Few-Shot Settings
by: Tapader, Md. Raihan, et al.
Published: (2025)
by: Tapader, Md. Raihan, et al.
Published: (2025)
Heterogeneous Prompting and Execution Feedback for SWE Issue Test Generation and Selection
by: Ahmed, Toufique, et al.
Published: (2025)
by: Ahmed, Toufique, et al.
Published: (2025)
Reproduction Test Generation for Java SWE Issues
by: Ahmed, Toufique, et al.
Published: (2026)
by: Ahmed, Toufique, et al.
Published: (2026)
Can Old Tests Do New Tricks for Resolving SWE Issues?
by: Chen, Yang, et al.
Published: (2025)
by: Chen, Yang, et al.
Published: (2025)
CodeT5-RNN: Reinforcing Contextual Embeddings for Enhanced Code Comprehension
by: Rahman, Md Mostafizer, et al.
Published: (2026)
by: Rahman, Md Mostafizer, et al.
Published: (2026)
Robustness, Security, Privacy, Explainability, Efficiency, and Usability of Large Language Models for Code
by: Yang, Zhou, et al.
Published: (2024)
by: Yang, Zhou, et al.
Published: (2024)
HistoryFinder: Advancing Method-Level Source Code History Generation with Accurate Oracles and Enhanced Algorithm
by: Islam, Shahidul, et al.
Published: (2025)
by: Islam, Shahidul, et al.
Published: (2025)
Enhanced LLM-Based Framework for Predicting Null Pointer Dereference in Source Code
by: Sultan, Md. Fahim, et al.
Published: (2024)
by: Sultan, Md. Fahim, et al.
Published: (2024)
Secret Breach Detection in Source Code with Large Language Models
by: Rahman, Md Nafiu, et al.
Published: (2025)
by: Rahman, Md Nafiu, et al.
Published: (2025)
Investigating Autonomous Agent Contributions in the Wild: Activity Patterns and Code Change over Time
by: Popescu, Razvan Mihai, et al.
Published: (2026)
by: Popescu, Razvan Mihai, et al.
Published: (2026)
Prompting and Fine-tuning Large Language Models for Automated Code Review Comment Generation
by: Haider, Md. Asif, et al.
Published: (2024)
by: Haider, Md. Asif, et al.
Published: (2024)
Context Before Code: An Experience Report on Vibe Coding in Practice
by: Shuvo, Md Nasir Uddin, et al.
Published: (2026)
by: Shuvo, Md Nasir Uddin, et al.
Published: (2026)
Does In-IDE Calibration of Large Language Models work at Scale?
by: Koohestani, Roham, et al.
Published: (2025)
by: Koohestani, Roham, et al.
Published: (2025)
Investigating Test Overfitting on SWE-bench
by: Ahmed, Toufique, et al.
Published: (2025)
by: Ahmed, Toufique, et al.
Published: (2025)
Understanding the Issue Types in Open Source Blockchain-based Software Projects with the Transformer-based BERTopic
by: Opu, Md Nahidul Islam, et al.
Published: (2025)
by: Opu, Md Nahidul Islam, et al.
Published: (2025)
Do Automatic Comment Generation Techniques Fall Short? Exploring the Influence of Method Dependencies on Code Understanding
by: Billah, Md Mustakim, et al.
Published: (2025)
by: Billah, Md Mustakim, et al.
Published: (2025)
The Repeat Offenders: Characterizing and Predicting Extremely Bug-Prone Source Methods
by: Friesen, Ethan, et al.
Published: (2025)
by: Friesen, Ethan, et al.
Published: (2025)
Error Understanding in Program Code With LLM-DL for Multi-label Classification
by: Amin, Md Faizul Ibne, et al.
Published: (2026)
by: Amin, Md Faizul Ibne, et al.
Published: (2026)
How Quantization Impacts Privacy Risk on LLMs for Code?
by: Haque, Md Nazmul, et al.
Published: (2025)
by: Haque, Md Nazmul, et al.
Published: (2025)
Similar Items
-
Trojans in Large Language Models of Code: A Critical Review through a Trigger-Based Taxonomy
by: Hussain, Aftab, et al.
Published: (2024) -
Measuring Impacts of Poisoning on Model Parameters and Neuron Activations: A Case Study of Poisoning CodeBERT
by: Hussain, Aftab, et al.
Published: (2024) -
On Trojan Signatures in Large Language Models of Code
by: Hussain, Aftab, et al.
Published: (2024) -
Unlearning Trojans in Large Language Models: A Comparison Between Natural Language and Source Code
by: Kazemi, Mahdi, et al.
Published: (2024) -
Measuring Impacts of Poisoning on Model Parameters and Embeddings for Large Language Models of Code
by: Hussain, Aftab, et al.
Published: (2024)