Empirical Insights of Test Selection Metrics under Multiple Testing Objectives and Distribution Shifts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Jingyu, Wang, Fan, Keung, Jacky, Liao, Yihan, Xiao, Yan, Ma, Lei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Exposing and Defending Membership Leakage in Vulnerability Prediction Models
von: Liao, Yihan, et al.
Veröffentlicht: (2025)
von: Liao, Yihan, et al.
Veröffentlicht: (2025)
UniAda: Universal Adaptive Multi-objective Adversarial Attack for End-to-End Autonomous Driving Systems
von: Zhang, Jingyu, et al.
Veröffentlicht: (2026)
von: Zhang, Jingyu, et al.
Veröffentlicht: (2026)
FedLAD: A Modular and Adaptive Testbed for Federated Log Anomaly Detection
von: Liao, Yihan, et al.
Veröffentlicht: (2025)
von: Liao, Yihan, et al.
Veröffentlicht: (2025)
Delving into Parameter-Efficient Fine-Tuning in Code Change Learning: An Empirical Study
von: Liu, Shuo, et al.
Veröffentlicht: (2024)
von: Liu, Shuo, et al.
Veröffentlicht: (2024)
An Empirical Study of Perceptions of General LLMs and Multimodal LLMs on Hugging Face
von: Liu, Yujian, et al.
Veröffentlicht: (2026)
von: Liu, Yujian, et al.
Veröffentlicht: (2026)
Chart2Code-MoLA: Efficient Multi-Modal Code Generation via Adaptive Expert Routing
von: Wang, Yifei, et al.
Veröffentlicht: (2025)
von: Wang, Yifei, et al.
Veröffentlicht: (2025)
SysTradeBench: An Iterative Build-Test-Patch Benchmark for Strategy-to-Code Trading Systems with Drift-Aware Diagnostics
von: Cao, Yuchen, et al.
Veröffentlicht: (2026)
von: Cao, Yuchen, et al.
Veröffentlicht: (2026)
On the Influence of Data Resampling for Deep Learning-Based Log Anomaly Detection: Insights and Recommendations
von: Ma, Xiaoxue, et al.
Veröffentlicht: (2024)
von: Ma, Xiaoxue, et al.
Veröffentlicht: (2024)
A Comprehensive Study of Bugs in Modern Distributed Deep Learning Systems
von: Ma, Xiaoxue, et al.
Veröffentlicht: (2025)
von: Ma, Xiaoxue, et al.
Veröffentlicht: (2025)
Towards Understanding Bugs in Distributed Training and Inference Frameworks for Large Language Models
von: Yu, Xiao, et al.
Veröffentlicht: (2025)
von: Yu, Xiao, et al.
Veröffentlicht: (2025)
Fight Fire with Fire: How Much Can We Trust ChatGPT on Source Code-Related Tasks?
von: Yu, Xiao, et al.
Veröffentlicht: (2024)
von: Yu, Xiao, et al.
Veröffentlicht: (2024)
R2Code: A Self-Reflective LLM Framework for Requirements-to-Code Traceability
von: Wang, Yifei, et al.
Veröffentlicht: (2026)
von: Wang, Yifei, et al.
Veröffentlicht: (2026)
Large Language Models for Automated Web-Form-Test Generation: An Empirical Study
von: Li, Tao, et al.
Veröffentlicht: (2024)
von: Li, Tao, et al.
Veröffentlicht: (2024)
Hybrid Privacy Policy-Code Consistency Check using Knowledge Graphs and LLMs
von: Mao, Zhenyu, et al.
Veröffentlicht: (2025)
von: Mao, Zhenyu, et al.
Veröffentlicht: (2025)
Towards Engineering Multi-Agent LLMs: A Protocol-Driven Approach
von: Mao, Zhenyu, et al.
Veröffentlicht: (2025)
von: Mao, Zhenyu, et al.
Veröffentlicht: (2025)
ABFS: Natural Robustness Testing for LLM-based NLP Software
von: Xiao, Mingxuan, et al.
Veröffentlicht: (2025)
von: Xiao, Mingxuan, et al.
Veröffentlicht: (2025)
Assessing the Robustness of LLM-based NLP Software via Automated Testing
von: Xiao, Mingxuan, et al.
Veröffentlicht: (2024)
von: Xiao, Mingxuan, et al.
Veröffentlicht: (2024)
Data Preparation for Deep Learning based Code Smell Detection: A Systematic Literature Review
von: Zhang, Fengji, et al.
Veröffentlicht: (2024)
von: Zhang, Fengji, et al.
Veröffentlicht: (2024)
Quality Assurance for LLM-RAG Systems: Empirical Insights from Tourism Application Testing
von: Ahmed, Bestoun S., et al.
Veröffentlicht: (2025)
von: Ahmed, Bestoun S., et al.
Veröffentlicht: (2025)
Lightweight Model Editing for LLMs to Correct Deprecated API Recommendations
von: Lin, Guancheng, et al.
Veröffentlicht: (2025)
von: Lin, Guancheng, et al.
Veröffentlicht: (2025)
An Empirical Study: MEMS as a Static Performance Metric
von: Zhang, Liwei, et al.
Veröffentlicht: (2025)
von: Zhang, Liwei, et al.
Veröffentlicht: (2025)
Practitioners' Expectations on Log Anomaly Detection
von: Ma, Xiaoxue, et al.
Veröffentlicht: (2024)
von: Ma, Xiaoxue, et al.
Veröffentlicht: (2024)
When Fine-Tuning LLMs Meets Data Privacy: An Empirical Study of Federated Learning in LLM-Based Program Repair
von: Luo, Wenqiang, et al.
Veröffentlicht: (2024)
von: Luo, Wenqiang, et al.
Veröffentlicht: (2024)
Towards Requirements Engineering for GenAI-Enabled Software: Bridging Responsibility Gaps through Human Oversight Requirements
von: Mao, Zhenyu, et al.
Veröffentlicht: (2025)
von: Mao, Zhenyu, et al.
Veröffentlicht: (2025)
Testing with AI Agents: An Empirical Study of Test Generation Frequency, Quality, and Coverage
von: Yoshimoto, Suzuka, et al.
Veröffentlicht: (2026)
von: Yoshimoto, Suzuka, et al.
Veröffentlicht: (2026)
Testing the Untestable? An Empirical Study on the Testing Process of LLM-Powered Software Systems
von: Magalhaes, Cleyton, et al.
Veröffentlicht: (2025)
von: Magalhaes, Cleyton, et al.
Veröffentlicht: (2025)
Fine-grained Testing for Autonomous Driving Software: a Study on Autoware with LLM-driven Unit Testing
von: Wang, Wenhan, et al.
Veröffentlicht: (2025)
von: Wang, Wenhan, et al.
Veröffentlicht: (2025)
Empirical Derivations from an Evolving Test Suite
von: Ruohonen, Jukka, et al.
Veröffentlicht: (2025)
von: Ruohonen, Jukka, et al.
Veröffentlicht: (2025)
MultiTest: Physical-Aware Object Insertion for Testing Multi-sensor Fusion Perception Systems
von: Gao, Xinyu, et al.
Veröffentlicht: (2024)
von: Gao, Xinyu, et al.
Veröffentlicht: (2024)
Beyond Code: Empirical Insights into How Team Dynamics Influence OSS Project Selection
von: Nirmani, Shashiwadana, et al.
Veröffentlicht: (2026)
von: Nirmani, Shashiwadana, et al.
Veröffentlicht: (2026)
Can We Classify Flaky Tests Using Only Test Code? An LLM-Based Empirical Study
von: Berndt, Alexander, et al.
Veröffentlicht: (2026)
von: Berndt, Alexander, et al.
Veröffentlicht: (2026)
On Test Sequence Generation using Multi-Objective Particle Swarm Optimization
von: Iqbal, Zain, et al.
Veröffentlicht: (2024)
von: Iqbal, Zain, et al.
Veröffentlicht: (2024)
Fix the Tests: Augmenting LLMs to Repair Test Cases with Static Collector and Neural Reranker
von: Liu, Jun, et al.
Veröffentlicht: (2024)
von: Liu, Jun, et al.
Veröffentlicht: (2024)
MoDitector: Module-Directed Testing for Autonomous Driving Systems
von: Wang, Renzhi, et al.
Veröffentlicht: (2025)
von: Wang, Renzhi, et al.
Veröffentlicht: (2025)
Understanding Bug-Reproducing Tests: A First Empirical Study
von: Hora, Andre, et al.
Veröffentlicht: (2026)
von: Hora, Andre, et al.
Veröffentlicht: (2026)
Are Coding Agents Generating Over-Mocked Tests? An Empirical Study
von: Hora, Andre, et al.
Veröffentlicht: (2026)
von: Hora, Andre, et al.
Veröffentlicht: (2026)
Understanding API Usage and Testing: An Empirical Study of C Libraries
von: Zaki, Ahmed, et al.
Veröffentlicht: (2025)
von: Zaki, Ahmed, et al.
Veröffentlicht: (2025)
AutoTest: Evolutionary Code Solution Selection with Test Cases
von: Duan, Zhihua, et al.
Veröffentlicht: (2024)
von: Duan, Zhihua, et al.
Veröffentlicht: (2024)
R2ComSync: Improving Code-Comment Synchronization with In-Context Learning and Reranking
von: Yang, Zhen, et al.
Veröffentlicht: (2025)
von: Yang, Zhen, et al.
Veröffentlicht: (2025)
BASFuzz: Towards Robustness Evaluation of LLM-based NLP Software via Automated Fuzz Testing
von: Xiao, Mingxuan, et al.
Veröffentlicht: (2025)
von: Xiao, Mingxuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Exposing and Defending Membership Leakage in Vulnerability Prediction Models
von: Liao, Yihan, et al.
Veröffentlicht: (2025) -
UniAda: Universal Adaptive Multi-objective Adversarial Attack for End-to-End Autonomous Driving Systems
von: Zhang, Jingyu, et al.
Veröffentlicht: (2026) -
FedLAD: A Modular and Adaptive Testbed for Federated Log Anomaly Detection
von: Liao, Yihan, et al.
Veröffentlicht: (2025) -
Delving into Parameter-Efficient Fine-Tuning in Code Change Learning: An Empirical Study
von: Liu, Shuo, et al.
Veröffentlicht: (2024) -
An Empirical Study of Perceptions of General LLMs and Multimodal LLMs on Hugging Face
von: Liu, Yujian, et al.
Veröffentlicht: (2026)