DistillSeq: A Framework for Safety Alignment Testing in Large Language Models using Knowledge Distillation
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Yang, Mingke, Chen, Yuqi, Liu, Yi, Shi, Ling |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
AKD : Adversarial Knowledge Distillation For Large Language Models Alignment on Coding tasks
par: Oulkadda, Ilyas, et autres
Publié: (2025)
par: Oulkadda, Ilyas, et autres
Publié: (2025)
Drowzee: Metamorphic Testing for Fact-Conflicting Hallucination Detection in Large Language Models
par: Li, Ningke, et autres
Publié: (2024)
par: Li, Ningke, et autres
Publié: (2024)
A Metamorphic Testing Perspective on Knowledge Distillation for Language Models of Code: Does the Student Deeply Mimic the Teacher?
par: Awal, Md. Abdul, et autres
Publié: (2025)
par: Awal, Md. Abdul, et autres
Publié: (2025)
Smaller but Better: Self-Paced Knowledge Distillation for Lightweight yet Effective LCMs
par: Chen, Yujia, et autres
Publié: (2024)
par: Chen, Yujia, et autres
Publié: (2024)
Efficient Detection of Toxic Prompts in Large Language Models
par: Liu, Yi, et autres
Publié: (2024)
par: Liu, Yi, et autres
Publié: (2024)
Testing Framework Migration with Large Language Models
par: Alves, Altino, et autres
Publié: (2026)
par: Alves, Altino, et autres
Publié: (2026)
Enhancing Semantic Understanding in Pointer Analysis using Large Language Models
par: Cheng, Baijun, et autres
Publié: (2025)
par: Cheng, Baijun, et autres
Publié: (2025)
An Empirical Study of Knowledge Distillation for Code Understanding Tasks
par: Wang, Ruiqi, et autres
Publié: (2025)
par: Wang, Ruiqi, et autres
Publié: (2025)
Large Language Models for Unit Test Generation: Achievements, Challenges, and Opportunities
par: Chu, Bei, et autres
Publié: (2025)
par: Chu, Bei, et autres
Publié: (2025)
Redefining Crowdsourced Test Report Prioritization: An Innovative Approach with Large Language Model
par: Ling, Yuchen, et autres
Publié: (2024)
par: Ling, Yuchen, et autres
Publié: (2024)
On the Evaluation of Large Language Models in Unit Test Generation
par: Yang, Lin, et autres
Publié: (2024)
par: Yang, Lin, et autres
Publié: (2024)
MoEKD: Mixture-of-Experts Knowledge Distillation for Robust and High-Performing Compressed Code Models
par: Awal, Md. Abdul, et autres
Publié: (2026)
par: Awal, Md. Abdul, et autres
Publié: (2026)
ASTRAL: Automated Safety Testing of Large Language Models
par: Ugarte, Miriam, et autres
Publié: (2025)
par: Ugarte, Miriam, et autres
Publié: (2025)
LLMorpheus: Mutation Testing using Large Language Models
par: Tip, Frank, et autres
Publié: (2024)
par: Tip, Frank, et autres
Publié: (2024)
BootstrapAgent: Distilling Repository Setup into Reusable Agent Knowledge
par: Fu, Sihan, et autres
Publié: (2026)
par: Fu, Sihan, et autres
Publié: (2026)
Semantic-Enhanced Indirect Call Analysis with Large Language Models
par: Cheng, Baijun, et autres
Publié: (2024)
par: Cheng, Baijun, et autres
Publié: (2024)
SeqTG: Scalable Combinatorial Test Generation via Sequential Integer Linear Programming
par: Yang, Sitong, et autres
Publié: (2026)
par: Yang, Sitong, et autres
Publié: (2026)
STELLAR: A Search-Based Testing Framework for Large Language Model Applications
par: Sorokin, Lev, et autres
Publié: (2026)
par: Sorokin, Lev, et autres
Publié: (2026)
Improving the Learning of Code Review Successive Tasks with Cross-Task Knowledge Distillation
par: Sghaier, Oussama Ben, et autres
Publié: (2024)
par: Sghaier, Oussama Ben, et autres
Publié: (2024)
Groot: Adversarial Testing for Generative Text-to-Image Models with Tree-based Semantic Transformation
par: Liu, Yi, et autres
Publié: (2024)
par: Liu, Yi, et autres
Publié: (2024)
Adversarial Attack Classification and Robustness Testing for Large Language Models for Code
par: Liu, Yang, et autres
Publié: (2025)
par: Liu, Yang, et autres
Publié: (2025)
Large Language Models for Software Testing Education: an Experience Report
par: Yang, Peng, et autres
Publié: (2026)
par: Yang, Peng, et autres
Publié: (2026)
Understanding the Effectiveness of Coverage Criteria for Large Language Models: A Special Angle from Jailbreak Attacks
par: Zhou, Shide, et autres
Publié: (2024)
par: Zhou, Shide, et autres
Publié: (2024)
Jailbreak Distillation: Renewable Safety Benchmarking
par: Zhang, Jingyu, et autres
Publié: (2025)
par: Zhang, Jingyu, et autres
Publié: (2025)
SimpleDevQA: Benchmarking Large Language Models on Development Knowledge QA
par: Zhang, Jing, et autres
Publié: (2025)
par: Zhang, Jing, et autres
Publié: (2025)
Towards Extracting Software Requirements from App Reviews using Seq2seq Framework
par: Sorathiya, Aakash, et autres
Publié: (2025)
par: Sorathiya, Aakash, et autres
Publié: (2025)
Software Testing with Large Language Models: Survey, Landscape, and Vision
par: Wang, Junjie, et autres
Publié: (2023)
par: Wang, Junjie, et autres
Publié: (2023)
Reflective Unit Test Generation for Precise Type Error Detection with Large Language Models
par: Yang, Chen, et autres
Publié: (2025)
par: Yang, Chen, et autres
Publié: (2025)
DCE-LLM: Dead Code Elimination with Large Language Models
par: Chen, Minyu, et autres
Publié: (2025)
par: Chen, Minyu, et autres
Publié: (2025)
TickIt: Leveraging Large Language Models for Automated Ticket Escalation
par: Liu, Fengrui, et autres
Publié: (2025)
par: Liu, Fengrui, et autres
Publié: (2025)
Improving the Readability of Automatically Generated Tests using Large Language Models
par: Biagiola, Matteo, et autres
Publié: (2024)
par: Biagiola, Matteo, et autres
Publié: (2024)
Automated Unit Test Improvement using Large Language Models at Meta
par: Alshahwan, Nadia, et autres
Publié: (2024)
par: Alshahwan, Nadia, et autres
Publié: (2024)
TESTEVAL: Benchmarking Large Language Models for Test Case Generation
par: Wang, Wenhan, et autres
Publié: (2024)
par: Wang, Wenhan, et autres
Publié: (2024)
Distilled GPT for Source Code Summarization
par: Su, Chia-Yi, et autres
Publié: (2023)
par: Su, Chia-Yi, et autres
Publié: (2023)
AMR-Evol: Adaptive Modular Response Evolution Elicits Better Knowledge Distillation for Large Language Models in Code Generation
par: Luo, Ziyang, et autres
Publié: (2024)
par: Luo, Ziyang, et autres
Publié: (2024)
Automated Control Logic Test Case Generation using Large Language Models
par: Koziolek, Heiko, et autres
Publié: (2024)
par: Koziolek, Heiko, et autres
Publié: (2024)
Automatic High-Level Test Case Generation using Large Language Models
par: Hasan, Navid Bin, et autres
Publié: (2025)
par: Hasan, Navid Bin, et autres
Publié: (2025)
Large Language Models for Unit Testing: A Systematic Literature Review
par: Zhang, Quanjun, et autres
Publié: (2025)
par: Zhang, Quanjun, et autres
Publié: (2025)
A Large-scale Empirical Study on Fine-tuning Large Language Models for Unit Testing
par: Shang, Ye, et autres
Publié: (2024)
par: Shang, Ye, et autres
Publié: (2024)
Ecosystem of Large Language Models for Code
par: Yang, Zhou, et autres
Publié: (2024)
par: Yang, Zhou, et autres
Publié: (2024)
Documents similaires
-
AKD : Adversarial Knowledge Distillation For Large Language Models Alignment on Coding tasks
par: Oulkadda, Ilyas, et autres
Publié: (2025) -
Drowzee: Metamorphic Testing for Fact-Conflicting Hallucination Detection in Large Language Models
par: Li, Ningke, et autres
Publié: (2024) -
A Metamorphic Testing Perspective on Knowledge Distillation for Language Models of Code: Does the Student Deeply Mimic the Teacher?
par: Awal, Md. Abdul, et autres
Publié: (2025) -
Smaller but Better: Self-Paced Knowledge Distillation for Lightweight yet Effective LCMs
par: Chen, Yujia, et autres
Publié: (2024) -
Efficient Detection of Toxic Prompts in Large Language Models
par: Liu, Yi, et autres
Publié: (2024)