A Metamorphic Testing Perspective on Knowledge Distillation for Language Models of Code: Does the Student Deeply Mimic the Teacher?
Fuente:
arXiv
Saved in:
| Main Authors: | Awal, Md. Abdul, Rochan, Mrigank, Roy, Chanchal K. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MoEKD: Mixture-of-Experts Knowledge Distillation for Robust and High-Performing Compressed Code Models
by: Awal, Md. Abdul, et al.
Published: (2026)
by: Awal, Md. Abdul, et al.
Published: (2026)
Model Compression vs. Adversarial Robustness: An Empirical Study on Language Models for Code
by: Awal, Md. Abdul, et al.
Published: (2025)
by: Awal, Md. Abdul, et al.
Published: (2025)
Large Language Models as Robust Data Generators in Software Analytics: Are We There Yet?
by: Awal, Md. Abdul, et al.
Published: (2024)
by: Awal, Md. Abdul, et al.
Published: (2024)
Investigating Adversarial Attacks in Software Analytics via Machine Learning Explainability
by: Awal, MD Abdul, et al.
Published: (2024)
by: Awal, MD Abdul, et al.
Published: (2024)
EvaluateXAI: A Framework to Evaluate the Reliability and Consistency of Rule-based XAI Techniques for Software Analytics Tasks
by: Awal, Md Abdul, et al.
Published: (2024)
by: Awal, Md Abdul, et al.
Published: (2024)
Carbon-Taxed Transformers: A Green Compression Pipeline for Overgrown Language Models
by: Alam, Ajmain Inqiad, et al.
Published: (2026)
by: Alam, Ajmain Inqiad, et al.
Published: (2026)
Quantum vs. Classical Machine Learning Algorithms for Software Defect Prediction: Challenges and Opportunities
by: Nadim, Md, et al.
Published: (2024)
by: Nadim, Md, et al.
Published: (2024)
Evaluating the Performance of a D-Wave Quantum Annealing System for Feature Subset Selection in Software Defect Prediction
by: Mandal, Ashis Kumar, et al.
Published: (2024)
by: Mandal, Ashis Kumar, et al.
Published: (2024)
Comparative Analysis of Quantum and Classical Support Vector Classifiers for Software Bug Prediction: An Exploratory Study
by: Nadim, Md, et al.
Published: (2025)
by: Nadim, Md, et al.
Published: (2025)
AKD : Adversarial Knowledge Distillation For Large Language Models Alignment on Coding tasks
by: Oulkadda, Ilyas, et al.
Published: (2025)
by: Oulkadda, Ilyas, et al.
Published: (2025)
AlphaTrans: A Neuro-Symbolic Compositional Approach for Repository-Level Code Translation and Validation
by: Ibrahimzada, Ali Reza, et al.
Published: (2024)
by: Ibrahimzada, Ali Reza, et al.
Published: (2024)
Agile Story-Point Estimation: Is RAG a Better Way to Go?
by: Maha, Lamyea, et al.
Published: (2026)
by: Maha, Lamyea, et al.
Published: (2026)
Calibration and Correctness of Language Models for Code
by: Spiess, Claudio, et al.
Published: (2024)
by: Spiess, Claudio, et al.
Published: (2024)
Does Editing Improve Answer Quality on Stack Overflow? A Data-Driven Investigation
by: Mondal, Saikat, et al.
Published: (2025)
by: Mondal, Saikat, et al.
Published: (2025)
Standing on the Shoulders of Giants: Stabilized Knowledge Distillation for Cross--Language Code Clone Detection
by: Khajezade, Mohamad, et al.
Published: (2026)
by: Khajezade, Mohamad, et al.
Published: (2026)
LGMT: Logic-Grounded Metamorphic Testing for Evaluating the Reasoning Reliability of LLMs
by: Zhou, Zenghui, et al.
Published: (2026)
by: Zhou, Zenghui, et al.
Published: (2026)
Unlearning Trojans in Large Language Models: A Comparison Between Natural Language and Source Code
by: Kazemi, Mahdi, et al.
Published: (2024)
by: Kazemi, Mahdi, et al.
Published: (2024)
GENCNIPPET: Automated Generation of Code Snippets for Supporting Programming Questions
by: Mondal, Saikat, et al.
Published: (2025)
by: Mondal, Saikat, et al.
Published: (2025)
On the Use of Deep Learning Models for Semantic Clone Detection
by: Pinku, Subroto Nag, et al.
Published: (2024)
by: Pinku, Subroto Nag, et al.
Published: (2024)
Relative Positioning Based Code Chunking Method For Rich Context Retrieval In Repository Level Code Completion Task With Code Language Model
by: Rahman, Imranur, et al.
Published: (2025)
by: Rahman, Imranur, et al.
Published: (2025)
On Trojan Signatures in Large Language Models of Code
by: Hussain, Aftab, et al.
Published: (2024)
by: Hussain, Aftab, et al.
Published: (2024)
Trojans in Large Language Models of Code: A Critical Review through a Trigger-Based Taxonomy
by: Hussain, Aftab, et al.
Published: (2024)
by: Hussain, Aftab, et al.
Published: (2024)
VeriContest: A Competitive-Programming Benchmark for Verifiable Code Generation
by: Xie, Zichen, et al.
Published: (2026)
by: Xie, Zichen, et al.
Published: (2026)
Learning-Based Testing for Deep Learning: Enhancing Model Robustness with Adversarial Input Prioritization
by: Rahman, Sheikh Md Mushfiqur, et al.
Published: (2025)
by: Rahman, Sheikh Md Mushfiqur, et al.
Published: (2025)
Model Cascading for Code: A Cascaded Black-Box Multi-Model Framework for Cost-Efficient Code Completion with Self-Testing
by: Chen, Boyuan, et al.
Published: (2024)
by: Chen, Boyuan, et al.
Published: (2024)
CodeIF: Benchmarking the Instruction-Following Capabilities of Large Language Models for Code Generation
by: Yan, Kaiwen, et al.
Published: (2025)
by: Yan, Kaiwen, et al.
Published: (2025)
Trained Without My Consent: Detecting Code Inclusion In Language Models Trained on Code
by: Majdinasab, Vahid, et al.
Published: (2024)
by: Majdinasab, Vahid, et al.
Published: (2024)
Automatic Semantic Augmentation of Language Model Prompts (for Code Summarization)
by: Ahmed, Toufique, et al.
Published: (2023)
by: Ahmed, Toufique, et al.
Published: (2023)
Can Code Language Models Learn Clarification-Seeking Behaviors?
by: Wu, Jie JW, et al.
Published: (2025)
by: Wu, Jie JW, et al.
Published: (2025)
FlakyFix: Using Large Language Models for Predicting Flaky Test Fix Categories and Test Code Repair
by: Fatima, Sakina, et al.
Published: (2023)
by: Fatima, Sakina, et al.
Published: (2023)
Can We Identify Stack Overflow Questions Requiring Code Snippets? Investigating the Cause & Effect of Missing Code Snippets
by: Mondal, Saikat, et al.
Published: (2024)
by: Mondal, Saikat, et al.
Published: (2024)
Context-Augmented Code Generation Using Programming Knowledge Graphs
by: Seddik, Shahd, et al.
Published: (2026)
by: Seddik, Shahd, et al.
Published: (2026)
CA2: Code-Aware Agent for Automated Game Testing
by: Adaikkappan, Valliappan Chidambaram, et al.
Published: (2026)
by: Adaikkappan, Valliappan Chidambaram, et al.
Published: (2026)
Enhancing Large Language Models with Faster Code Preprocessing for Vulnerability Detection
by: Gonçalves, José, et al.
Published: (2025)
by: Gonçalves, José, et al.
Published: (2025)
Language Models are Better Bug Detector Through Code-Pair Classification
by: Alrashedy, Kamel, et al.
Published: (2023)
by: Alrashedy, Kamel, et al.
Published: (2023)
LEANCODE: Understanding Models Better for Code Simplification of Pre-trained Large Language Models
by: Wang, Yan, et al.
Published: (2025)
by: Wang, Yan, et al.
Published: (2025)
LLMORPH: Automated Metamorphic Testing of Large Language Models
by: Cho, Steven, et al.
Published: (2026)
by: Cho, Steven, et al.
Published: (2026)
LoRA-MME: Multi-Model Ensemble of LoRA-Tuned Encoders for Code Comment Classification
by: Haider, Md Akib, et al.
Published: (2026)
by: Haider, Md Akib, et al.
Published: (2026)
Assessing the Impact of Code Changes on the Fault Localizability of Large Language Models
by: Haroon, Sabaat, et al.
Published: (2025)
by: Haroon, Sabaat, et al.
Published: (2025)
Using Large Language Models to Generate JUnit Tests: An Empirical Study
by: Siddiq, Mohammed Latif, et al.
Published: (2023)
by: Siddiq, Mohammed Latif, et al.
Published: (2023)
Similar Items
-
MoEKD: Mixture-of-Experts Knowledge Distillation for Robust and High-Performing Compressed Code Models
by: Awal, Md. Abdul, et al.
Published: (2026) -
Model Compression vs. Adversarial Robustness: An Empirical Study on Language Models for Code
by: Awal, Md. Abdul, et al.
Published: (2025) -
Large Language Models as Robust Data Generators in Software Analytics: Are We There Yet?
by: Awal, Md. Abdul, et al.
Published: (2024) -
Investigating Adversarial Attacks in Software Analytics via Machine Learning Explainability
by: Awal, MD Abdul, et al.
Published: (2024) -
EvaluateXAI: A Framework to Evaluate the Reliability and Consistency of Rule-based XAI Techniques for Software Analytics Tasks
by: Awal, Md Abdul, et al.
Published: (2024)