MeTMaP: Metamorphic Testing for Detecting False Vector Matching Problems in LLM Augmented Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Guanyu, Li, Yuekang, Liu, Yi, Deng, Gelei, Li, Tianlin, Xu, Guosheng, Liu, Yang, Wang, Haoyu, Wang, Kailong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Drowzee: Metamorphic Testing for Fact-Conflicting Hallucination Detection in Large Language Models
di: Li, Ningke, et al.
Pubblicazione: (2024)
di: Li, Ningke, et al.
Pubblicazione: (2024)
SPOLRE: Semantic Preserving Object Layout Reconstruction for Image Captioning System Testing
di: Liu, Yi, et al.
Pubblicazione: (2024)
di: Liu, Yi, et al.
Pubblicazione: (2024)
Glitch Tokens in Large Language Models: Categorization Taxonomy and Effective Detection
di: Li, Yuxi, et al.
Pubblicazione: (2024)
di: Li, Yuxi, et al.
Pubblicazione: (2024)
Prompt Injection attack against LLM-integrated Applications
di: Liu, Yi, et al.
Pubblicazione: (2023)
di: Liu, Yi, et al.
Pubblicazione: (2023)
MiniScope: Automated UI Exploration and Privacy Inconsistency Detection of MiniApps via Two-phase Iterative Hybrid Analysis
di: Wang, Shenao, et al.
Pubblicazione: (2024)
di: Wang, Shenao, et al.
Pubblicazione: (2024)
What Makes a Good LLM Agent for Real-world Penetration Testing?
di: Deng, Gelei, et al.
Pubblicazione: (2026)
di: Deng, Gelei, et al.
Pubblicazione: (2026)
PentestGPT: An LLM-empowered Automatic Penetration Testing Tool
di: Deng, Gelei, et al.
Pubblicazione: (2023)
di: Deng, Gelei, et al.
Pubblicazione: (2023)
Understanding the Effectiveness of Coverage Criteria for Large Language Models: A Special Angle from Jailbreak Attacks
di: Zhou, Shide, et al.
Pubblicazione: (2024)
di: Zhou, Shide, et al.
Pubblicazione: (2024)
Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study
di: Liu, Yi, et al.
Pubblicazione: (2023)
di: Liu, Yi, et al.
Pubblicazione: (2023)
Pandora: Jailbreak GPTs by Retrieval Augmented Generation Poisoning
di: Deng, Gelei, et al.
Pubblicazione: (2024)
di: Deng, Gelei, et al.
Pubblicazione: (2024)
NeuSemSlice: Towards Effective DNN Model Maintenance via Neuron-level Semantic Slicing
di: Zhou, Shide, et al.
Pubblicazione: (2024)
di: Zhou, Shide, et al.
Pubblicazione: (2024)
Decoding Secret Memorization in Code LLMs Through Token-Level Characterization
di: Nie, Yuqing, et al.
Pubblicazione: (2024)
di: Nie, Yuqing, et al.
Pubblicazione: (2024)
Validating LLM-Generated Programs with Metamorphic Prompt Testing
di: Wang, Xiaoyin, et al.
Pubblicazione: (2024)
di: Wang, Xiaoyin, et al.
Pubblicazione: (2024)
Agent Skills in the Wild: An Empirical Study of Security Vulnerabilities at Scale
di: Liu, Yi, et al.
Pubblicazione: (2026)
di: Liu, Yi, et al.
Pubblicazione: (2026)
Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment
di: Li, Yuxi, et al.
Pubblicazione: (2024)
di: Li, Yuxi, et al.
Pubblicazione: (2024)
Digger: Detecting Copyright Content Mis-usage in Large Language Model Training
di: Li, Haodong, et al.
Pubblicazione: (2024)
di: Li, Haodong, et al.
Pubblicazione: (2024)
Exploring and Lifting the Robustness of LLM-powered Automated Program Repair with Metamorphic Testing
di: Xue, Pengyu, et al.
Pubblicazione: (2024)
di: Xue, Pengyu, et al.
Pubblicazione: (2024)
Overeager Coding Agents: Measuring Out-of-Scope Actions on Benign Tasks
di: Qu, Yubin, et al.
Pubblicazione: (2026)
di: Qu, Yubin, et al.
Pubblicazione: (2026)
SegTest: Metamorphic Testing of Image Segmentation via Guided Instance‐Level Test Data Augmentation
di: Zhonghao Hou, et al.
Pubblicazione: (2024)
di: Zhonghao Hou, et al.
Pubblicazione: (2024)
Boosting Pointer Analysis With LLM-Enhanced Allocation Function Detection
di: Cheng, Baijun, et al.
Pubblicazione: (2025)
di: Cheng, Baijun, et al.
Pubblicazione: (2025)
Test Adequacy for Metamorphic Testing: Criteria, Measurement, and Implication
di: Fu, An, et al.
Pubblicazione: (2024)
di: Fu, An, et al.
Pubblicazione: (2024)
AI-Augmented Metamorphic Testing for Comprehensive Validation of Autonomous Vehicles
di: Zhang, Tony, et al.
Pubblicazione: (2025)
di: Zhang, Tony, et al.
Pubblicazione: (2025)
Towards Reliable Vector Database Management Systems: A Software Testing Roadmap for 2030
di: Wang, Shenao, et al.
Pubblicazione: (2025)
di: Wang, Shenao, et al.
Pubblicazione: (2025)
PentestEval: Benchmarking LLM-based Penetration Testing with Modular and Stage-Level Design
di: Yang, Ruozhao, et al.
Pubblicazione: (2025)
di: Yang, Ruozhao, et al.
Pubblicazione: (2025)
Hallucination Detection for LLM-based Text-to-SQL Generation via Two-Stage Metamorphic Testing
di: Yang, Bo, et al.
Pubblicazione: (2025)
di: Yang, Bo, et al.
Pubblicazione: (2025)
MasterKey: Automated Jailbreak Across Multiple Large Language Model Chatbots
di: Deng, Gelei, et al.
Pubblicazione: (2023)
di: Deng, Gelei, et al.
Pubblicazione: (2023)
Metamorphic Testing for Audio Content Moderation Software
di: Wang, Wenxuan, et al.
Pubblicazione: (2025)
di: Wang, Wenxuan, et al.
Pubblicazione: (2025)
AutoMT: A Multi-Agent LLM Framework for Automated Metamorphic Testing of Autonomous Driving Systems
di: Liang, Linfeng, et al.
Pubblicazione: (2025)
di: Liang, Linfeng, et al.
Pubblicazione: (2025)
Latent Imitator: Generating Natural Individual Discriminatory Instances for Black-Box Fairness Testing
di: Xiao, Yisong, et al.
Pubblicazione: (2023)
di: Xiao, Yisong, et al.
Pubblicazione: (2023)
Beyond Correctness: Exposing LLM-generated Logical Flaws in Reasoning via Multi-step Automated Theorem Proving
di: Zheng, Xinyi, et al.
Pubblicazione: (2025)
di: Zheng, Xinyi, et al.
Pubblicazione: (2025)
Enhancing Semantic Understanding in Pointer Analysis using Large Language Models
di: Cheng, Baijun, et al.
Pubblicazione: (2025)
di: Cheng, Baijun, et al.
Pubblicazione: (2025)
Source Code Summarization in the Era of Large Language Models
di: Sun, Weisong, et al.
Pubblicazione: (2024)
di: Sun, Weisong, et al.
Pubblicazione: (2024)
SliceLocator: Locating Vulnerable Statements with Graph-based Detectors
di: Cheng, Baijun, et al.
Pubblicazione: (2024)
di: Cheng, Baijun, et al.
Pubblicazione: (2024)
Metamorphic Testing of Image Captioning Systems via Image-Level Reduction
di: Xie, Xiaoyuan, et al.
Pubblicazione: (2023)
di: Xie, Xiaoyuan, et al.
Pubblicazione: (2023)
Semantic-Enhanced Indirect Call Analysis with Large Language Models
di: Cheng, Baijun, et al.
Pubblicazione: (2024)
di: Cheng, Baijun, et al.
Pubblicazione: (2024)
Continuous Embedding Attacks via Clipped Inputs in Jailbreaking Large Language Models
di: Xu, Zihao, et al.
Pubblicazione: (2024)
di: Xu, Zihao, et al.
Pubblicazione: (2024)
CodeChemist: Functional Knowledge Transfer for Low-Resource Code Generation via Test-Time Scaling
di: Wang, Kaixin, et al.
Pubblicazione: (2025)
di: Wang, Kaixin, et al.
Pubblicazione: (2025)
MORTAR: Multi-turn Metamorphic Testing for LLM-based Dialogue Systems
di: Guo, Guoxiang, et al.
Pubblicazione: (2024)
di: Guo, Guoxiang, et al.
Pubblicazione: (2024)
FidelityGPT: Correcting Decompilation Distortions with Retrieval Augmented Generation
di: Zhou, Zhiping, et al.
Pubblicazione: (2025)
di: Zhou, Zhiping, et al.
Pubblicazione: (2025)
Scenario‐Driven Metamorphic Testing for Autonomous Driving Simulators
di: Yifan Zhang, et al.
Pubblicazione: (2024)
di: Yifan Zhang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Drowzee: Metamorphic Testing for Fact-Conflicting Hallucination Detection in Large Language Models
di: Li, Ningke, et al.
Pubblicazione: (2024) -
SPOLRE: Semantic Preserving Object Layout Reconstruction for Image Captioning System Testing
di: Liu, Yi, et al.
Pubblicazione: (2024) -
Glitch Tokens in Large Language Models: Categorization Taxonomy and Effective Detection
di: Li, Yuxi, et al.
Pubblicazione: (2024) -
Prompt Injection attack against LLM-integrated Applications
di: Liu, Yi, et al.
Pubblicazione: (2023) -
MiniScope: Automated UI Exploration and Privacy Inconsistency Detection of MiniApps via Two-phase Iterative Hybrid Analysis
di: Wang, Shenao, et al.
Pubblicazione: (2024)