Understanding the Effectiveness of Coverage Criteria for Large Language Models: A Special Angle from Jailbreak Attacks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Shide, Li, Tianlin, Wang, Kailong, Huang, Yihao, Shi, Ling, Liu, Yang, Wang, Haoyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
NeuSemSlice: Towards Effective DNN Model Maintenance via Neuron-level Semantic Slicing
von: Zhou, Shide, et al.
Veröffentlicht: (2024)
von: Zhou, Shide, et al.
Veröffentlicht: (2024)
Enhancing Semantic Understanding in Pointer Analysis using Large Language Models
von: Cheng, Baijun, et al.
Veröffentlicht: (2025)
von: Cheng, Baijun, et al.
Veröffentlicht: (2025)
Drowzee: Metamorphic Testing for Fact-Conflicting Hallucination Detection in Large Language Models
von: Li, Ningke, et al.
Veröffentlicht: (2024)
von: Li, Ningke, et al.
Veröffentlicht: (2024)
Glitch Tokens in Large Language Models: Categorization Taxonomy and Effective Detection
von: Li, Yuxi, et al.
Veröffentlicht: (2024)
von: Li, Yuxi, et al.
Veröffentlicht: (2024)
Semantic-Enhanced Indirect Call Analysis with Large Language Models
von: Cheng, Baijun, et al.
Veröffentlicht: (2024)
von: Cheng, Baijun, et al.
Veröffentlicht: (2024)
MeTMaP: Metamorphic Testing for Detecting False Vector Matching Problems in LLM Augmented Generation
von: Wang, Guanyu, et al.
Veröffentlicht: (2024)
von: Wang, Guanyu, et al.
Veröffentlicht: (2024)
Advancing Code Coverage: Incorporating Program Analysis with Large Language Models
von: Yang, Chen, et al.
Veröffentlicht: (2024)
von: Yang, Chen, et al.
Veröffentlicht: (2024)
Large Language Models for Software Engineering: A Systematic Literature Review
von: Hou, Xinyi, et al.
Veröffentlicht: (2023)
von: Hou, Xinyi, et al.
Veröffentlicht: (2023)
Understanding Large Language Model Supply Chain: Structure, Domain, and Vulnerabilities
von: Hu, Yanzhe, et al.
Veröffentlicht: (2025)
von: Hu, Yanzhe, et al.
Veröffentlicht: (2025)
Boosting Pointer Analysis With LLM-Enhanced Allocation Function Detection
von: Cheng, Baijun, et al.
Veröffentlicht: (2025)
von: Cheng, Baijun, et al.
Veröffentlicht: (2025)
Exploring the Power of Diffusion Large Language Models for Software Engineering: An Empirical Investigation
von: Zhang, Jingyao, et al.
Veröffentlicht: (2025)
von: Zhang, Jingyao, et al.
Veröffentlicht: (2025)
SPOLRE: Semantic Preserving Object Layout Reconstruction for Image Captioning System Testing
von: Liu, Yi, et al.
Veröffentlicht: (2024)
von: Liu, Yi, et al.
Veröffentlicht: (2024)
Large Language Models are overconfident and amplify human bias
von: Sun, Fengfei, et al.
Veröffentlicht: (2025)
von: Sun, Fengfei, et al.
Veröffentlicht: (2025)
CodeMorph: Mitigating Data Leakage in Large Language Model Assessment
von: Rao, Hongzhou, et al.
Veröffentlicht: (2025)
von: Rao, Hongzhou, et al.
Veröffentlicht: (2025)
Models Are Codes: Towards Measuring Malicious Code Poisoning Attacks on Pre-trained Model Hubs
von: Zhao, Jian, et al.
Veröffentlicht: (2024)
von: Zhao, Jian, et al.
Veröffentlicht: (2024)
Generalized Coverage Criteria for Combinatorial Sequence Testing
von: Elyasaf, Achiya, et al.
Veröffentlicht: (2022)
von: Elyasaf, Achiya, et al.
Veröffentlicht: (2022)
Testing Agentic Workflows with Structural Coverage Criteria
von: Kahani, Nafiseh, et al.
Veröffentlicht: (2026)
von: Kahani, Nafiseh, et al.
Veröffentlicht: (2026)
Large Language Model Supply Chain: A Research Agenda
von: Wang, Shenao, et al.
Veröffentlicht: (2024)
von: Wang, Shenao, et al.
Veröffentlicht: (2024)
Towards Robust Detection of Open Source Software Supply Chain Poisoning Attacks in Industry Environments
von: Zheng, Xinyi, et al.
Veröffentlicht: (2024)
von: Zheng, Xinyi, et al.
Veröffentlicht: (2024)
Software Development Life Cycle Perspective: A Survey of Benchmarks for Code Large Language Models and Agents
von: Wang, Kaixin, et al.
Veröffentlicht: (2025)
von: Wang, Kaixin, et al.
Veröffentlicht: (2025)
Coverage Goal Selector for Combining Multiple Criteria in Search-Based Unit Test Generation
von: Zhou, Zhichao, et al.
Veröffentlicht: (2023)
von: Zhou, Zhichao, et al.
Veröffentlicht: (2023)
WaDec: Decompiling WebAssembly Using Large Language Model
von: She, Xinyu, et al.
Veröffentlicht: (2024)
von: She, Xinyu, et al.
Veröffentlicht: (2024)
MiniScope: Automated UI Exploration and Privacy Inconsistency Detection of MiniApps via Two-phase Iterative Hybrid Analysis
von: Wang, Shenao, et al.
Veröffentlicht: (2024)
von: Wang, Shenao, et al.
Veröffentlicht: (2024)
Towards Understanding the Characteristics of Code Generation Errors Made by Large Language Models
von: Wang, Zhijie, et al.
Veröffentlicht: (2024)
von: Wang, Zhijie, et al.
Veröffentlicht: (2024)
DistillSeq: A Framework for Safety Alignment Testing in Large Language Models using Knowledge Distillation
von: Yang, Mingke, et al.
Veröffentlicht: (2024)
von: Yang, Mingke, et al.
Veröffentlicht: (2024)
Annotating Control-Flow Graphs for Formalized Test Coverage Criteria
von: Kauffman, Sean, et al.
Veröffentlicht: (2024)
von: Kauffman, Sean, et al.
Veröffentlicht: (2024)
Beyond Correctness: Exposing LLM-generated Logical Flaws in Reasoning via Multi-step Automated Theorem Proving
von: Zheng, Xinyi, et al.
Veröffentlicht: (2025)
von: Zheng, Xinyi, et al.
Veröffentlicht: (2025)
Smoke and Mirrors: Jailbreaking LLM-based Code Generation via Implicit Malicious Prompts
von: Ouyang, Sheng, et al.
Veröffentlicht: (2025)
von: Ouyang, Sheng, et al.
Veröffentlicht: (2025)
Adversarial Attack Classification and Robustness Testing for Large Language Models for Code
von: Liu, Yang, et al.
Veröffentlicht: (2025)
von: Liu, Yang, et al.
Veröffentlicht: (2025)
Ecosystem of Large Language Models for Code
von: Yang, Zhou, et al.
Veröffentlicht: (2024)
von: Yang, Zhou, et al.
Veröffentlicht: (2024)
LEANCODE: Understanding Models Better for Code Simplification of Pre-trained Large Language Models
von: Wang, Yan, et al.
Veröffentlicht: (2025)
von: Wang, Yan, et al.
Veröffentlicht: (2025)
Exposing the Ghost in the Transformer: Abnormal Detection for Large Language Models via Hidden State Forensics
von: Zhou, Shide, et al.
Veröffentlicht: (2025)
von: Zhou, Shide, et al.
Veröffentlicht: (2025)
Towards an Understanding of Large Language Models in Software Engineering Tasks
von: Zheng, Zibin, et al.
Veröffentlicht: (2023)
von: Zheng, Zibin, et al.
Veröffentlicht: (2023)
Decoding Secret Memorization in Code LLMs Through Token-Level Characterization
von: Nie, Yuqing, et al.
Veröffentlicht: (2024)
von: Nie, Yuqing, et al.
Veröffentlicht: (2024)
Large Language Models-Aided Program Debloating
von: Lin, Bo, et al.
Veröffentlicht: (2025)
von: Lin, Bo, et al.
Veröffentlicht: (2025)
Towards Trustworthy LLMs for Code: A Data-Centric Synergistic Auditing Framework
von: Wang, Chong, et al.
Veröffentlicht: (2024)
von: Wang, Chong, et al.
Veröffentlicht: (2024)
SliceLocator: Locating Vulnerable Statements with Graph-based Detectors
von: Cheng, Baijun, et al.
Veröffentlicht: (2024)
von: Cheng, Baijun, et al.
Veröffentlicht: (2024)
Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study
von: Liu, Yi, et al.
Veröffentlicht: (2023)
von: Liu, Yi, et al.
Veröffentlicht: (2023)
Characterizing and Evaluating the Reliability of LLMs against Jailbreak Attacks
von: Chen, Kexin, et al.
Veröffentlicht: (2024)
von: Chen, Kexin, et al.
Veröffentlicht: (2024)
Software Testing with Large Language Models: Survey, Landscape, and Vision
von: Wang, Junjie, et al.
Veröffentlicht: (2023)
von: Wang, Junjie, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
NeuSemSlice: Towards Effective DNN Model Maintenance via Neuron-level Semantic Slicing
von: Zhou, Shide, et al.
Veröffentlicht: (2024) -
Enhancing Semantic Understanding in Pointer Analysis using Large Language Models
von: Cheng, Baijun, et al.
Veröffentlicht: (2025) -
Drowzee: Metamorphic Testing for Fact-Conflicting Hallucination Detection in Large Language Models
von: Li, Ningke, et al.
Veröffentlicht: (2024) -
Glitch Tokens in Large Language Models: Categorization Taxonomy and Effective Detection
von: Li, Yuxi, et al.
Veröffentlicht: (2024) -
Semantic-Enhanced Indirect Call Analysis with Large Language Models
von: Cheng, Baijun, et al.
Veröffentlicht: (2024)