DistillSeq: A Framework for Safety Alignment Testing in Large Language Models using Knowledge Distillation
Fuente:
arXiv
Salvato in:
| Autori principali: | Yang, Mingke, Chen, Yuqi, Liu, Yi, Shi, Ling |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AKD : Adversarial Knowledge Distillation For Large Language Models Alignment on Coding tasks
di: Oulkadda, Ilyas, et al.
Pubblicazione: (2025)
di: Oulkadda, Ilyas, et al.
Pubblicazione: (2025)
Drowzee: Metamorphic Testing for Fact-Conflicting Hallucination Detection in Large Language Models
di: Li, Ningke, et al.
Pubblicazione: (2024)
di: Li, Ningke, et al.
Pubblicazione: (2024)
A Metamorphic Testing Perspective on Knowledge Distillation for Language Models of Code: Does the Student Deeply Mimic the Teacher?
di: Awal, Md. Abdul, et al.
Pubblicazione: (2025)
di: Awal, Md. Abdul, et al.
Pubblicazione: (2025)
Smaller but Better: Self-Paced Knowledge Distillation for Lightweight yet Effective LCMs
di: Chen, Yujia, et al.
Pubblicazione: (2024)
di: Chen, Yujia, et al.
Pubblicazione: (2024)
Efficient Detection of Toxic Prompts in Large Language Models
di: Liu, Yi, et al.
Pubblicazione: (2024)
di: Liu, Yi, et al.
Pubblicazione: (2024)
Testing Framework Migration with Large Language Models
di: Alves, Altino, et al.
Pubblicazione: (2026)
di: Alves, Altino, et al.
Pubblicazione: (2026)
Enhancing Semantic Understanding in Pointer Analysis using Large Language Models
di: Cheng, Baijun, et al.
Pubblicazione: (2025)
di: Cheng, Baijun, et al.
Pubblicazione: (2025)
An Empirical Study of Knowledge Distillation for Code Understanding Tasks
di: Wang, Ruiqi, et al.
Pubblicazione: (2025)
di: Wang, Ruiqi, et al.
Pubblicazione: (2025)
Large Language Models for Unit Test Generation: Achievements, Challenges, and Opportunities
di: Chu, Bei, et al.
Pubblicazione: (2025)
di: Chu, Bei, et al.
Pubblicazione: (2025)
Redefining Crowdsourced Test Report Prioritization: An Innovative Approach with Large Language Model
di: Ling, Yuchen, et al.
Pubblicazione: (2024)
di: Ling, Yuchen, et al.
Pubblicazione: (2024)
On the Evaluation of Large Language Models in Unit Test Generation
di: Yang, Lin, et al.
Pubblicazione: (2024)
di: Yang, Lin, et al.
Pubblicazione: (2024)
MoEKD: Mixture-of-Experts Knowledge Distillation for Robust and High-Performing Compressed Code Models
di: Awal, Md. Abdul, et al.
Pubblicazione: (2026)
di: Awal, Md. Abdul, et al.
Pubblicazione: (2026)
ASTRAL: Automated Safety Testing of Large Language Models
di: Ugarte, Miriam, et al.
Pubblicazione: (2025)
di: Ugarte, Miriam, et al.
Pubblicazione: (2025)
LLMorpheus: Mutation Testing using Large Language Models
di: Tip, Frank, et al.
Pubblicazione: (2024)
di: Tip, Frank, et al.
Pubblicazione: (2024)
BootstrapAgent: Distilling Repository Setup into Reusable Agent Knowledge
di: Fu, Sihan, et al.
Pubblicazione: (2026)
di: Fu, Sihan, et al.
Pubblicazione: (2026)
Semantic-Enhanced Indirect Call Analysis with Large Language Models
di: Cheng, Baijun, et al.
Pubblicazione: (2024)
di: Cheng, Baijun, et al.
Pubblicazione: (2024)
SeqTG: Scalable Combinatorial Test Generation via Sequential Integer Linear Programming
di: Yang, Sitong, et al.
Pubblicazione: (2026)
di: Yang, Sitong, et al.
Pubblicazione: (2026)
STELLAR: A Search-Based Testing Framework for Large Language Model Applications
di: Sorokin, Lev, et al.
Pubblicazione: (2026)
di: Sorokin, Lev, et al.
Pubblicazione: (2026)
Improving the Learning of Code Review Successive Tasks with Cross-Task Knowledge Distillation
di: Sghaier, Oussama Ben, et al.
Pubblicazione: (2024)
di: Sghaier, Oussama Ben, et al.
Pubblicazione: (2024)
Groot: Adversarial Testing for Generative Text-to-Image Models with Tree-based Semantic Transformation
di: Liu, Yi, et al.
Pubblicazione: (2024)
di: Liu, Yi, et al.
Pubblicazione: (2024)
Adversarial Attack Classification and Robustness Testing for Large Language Models for Code
di: Liu, Yang, et al.
Pubblicazione: (2025)
di: Liu, Yang, et al.
Pubblicazione: (2025)
Large Language Models for Software Testing Education: an Experience Report
di: Yang, Peng, et al.
Pubblicazione: (2026)
di: Yang, Peng, et al.
Pubblicazione: (2026)
Understanding the Effectiveness of Coverage Criteria for Large Language Models: A Special Angle from Jailbreak Attacks
di: Zhou, Shide, et al.
Pubblicazione: (2024)
di: Zhou, Shide, et al.
Pubblicazione: (2024)
Jailbreak Distillation: Renewable Safety Benchmarking
di: Zhang, Jingyu, et al.
Pubblicazione: (2025)
di: Zhang, Jingyu, et al.
Pubblicazione: (2025)
SimpleDevQA: Benchmarking Large Language Models on Development Knowledge QA
di: Zhang, Jing, et al.
Pubblicazione: (2025)
di: Zhang, Jing, et al.
Pubblicazione: (2025)
Towards Extracting Software Requirements from App Reviews using Seq2seq Framework
di: Sorathiya, Aakash, et al.
Pubblicazione: (2025)
di: Sorathiya, Aakash, et al.
Pubblicazione: (2025)
Software Testing with Large Language Models: Survey, Landscape, and Vision
di: Wang, Junjie, et al.
Pubblicazione: (2023)
di: Wang, Junjie, et al.
Pubblicazione: (2023)
Reflective Unit Test Generation for Precise Type Error Detection with Large Language Models
di: Yang, Chen, et al.
Pubblicazione: (2025)
di: Yang, Chen, et al.
Pubblicazione: (2025)
DCE-LLM: Dead Code Elimination with Large Language Models
di: Chen, Minyu, et al.
Pubblicazione: (2025)
di: Chen, Minyu, et al.
Pubblicazione: (2025)
TickIt: Leveraging Large Language Models for Automated Ticket Escalation
di: Liu, Fengrui, et al.
Pubblicazione: (2025)
di: Liu, Fengrui, et al.
Pubblicazione: (2025)
Improving the Readability of Automatically Generated Tests using Large Language Models
di: Biagiola, Matteo, et al.
Pubblicazione: (2024)
di: Biagiola, Matteo, et al.
Pubblicazione: (2024)
Automated Unit Test Improvement using Large Language Models at Meta
di: Alshahwan, Nadia, et al.
Pubblicazione: (2024)
di: Alshahwan, Nadia, et al.
Pubblicazione: (2024)
TESTEVAL: Benchmarking Large Language Models for Test Case Generation
di: Wang, Wenhan, et al.
Pubblicazione: (2024)
di: Wang, Wenhan, et al.
Pubblicazione: (2024)
Distilled GPT for Source Code Summarization
di: Su, Chia-Yi, et al.
Pubblicazione: (2023)
di: Su, Chia-Yi, et al.
Pubblicazione: (2023)
AMR-Evol: Adaptive Modular Response Evolution Elicits Better Knowledge Distillation for Large Language Models in Code Generation
di: Luo, Ziyang, et al.
Pubblicazione: (2024)
di: Luo, Ziyang, et al.
Pubblicazione: (2024)
Automated Control Logic Test Case Generation using Large Language Models
di: Koziolek, Heiko, et al.
Pubblicazione: (2024)
di: Koziolek, Heiko, et al.
Pubblicazione: (2024)
Automatic High-Level Test Case Generation using Large Language Models
di: Hasan, Navid Bin, et al.
Pubblicazione: (2025)
di: Hasan, Navid Bin, et al.
Pubblicazione: (2025)
Large Language Models for Unit Testing: A Systematic Literature Review
di: Zhang, Quanjun, et al.
Pubblicazione: (2025)
di: Zhang, Quanjun, et al.
Pubblicazione: (2025)
A Large-scale Empirical Study on Fine-tuning Large Language Models for Unit Testing
di: Shang, Ye, et al.
Pubblicazione: (2024)
di: Shang, Ye, et al.
Pubblicazione: (2024)
Ecosystem of Large Language Models for Code
di: Yang, Zhou, et al.
Pubblicazione: (2024)
di: Yang, Zhou, et al.
Pubblicazione: (2024)
Documenti analoghi
-
AKD : Adversarial Knowledge Distillation For Large Language Models Alignment on Coding tasks
di: Oulkadda, Ilyas, et al.
Pubblicazione: (2025) -
Drowzee: Metamorphic Testing for Fact-Conflicting Hallucination Detection in Large Language Models
di: Li, Ningke, et al.
Pubblicazione: (2024) -
A Metamorphic Testing Perspective on Knowledge Distillation for Language Models of Code: Does the Student Deeply Mimic the Teacher?
di: Awal, Md. Abdul, et al.
Pubblicazione: (2025) -
Smaller but Better: Self-Paced Knowledge Distillation for Lightweight yet Effective LCMs
di: Chen, Yujia, et al.
Pubblicazione: (2024) -
Efficient Detection of Toxic Prompts in Large Language Models
di: Liu, Yi, et al.
Pubblicazione: (2024)