Rethinking Diversity in Deep Neural Network Testing
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wang, Zi, Choi, Jihye, Wang, Ke, Jha, Somesh |
|---|---|
| Format: | Preprint |
| Publié: |
2023
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Auto-SPT: Automating Semantic Preserving Transformations for Code
par: Hooda, Ashish, et autres
Publié: (2025)
par: Hooda, Ashish, et autres
Publié: (2025)
Faster Configuration Performance Bug Testing with Neural Dual-level Prioritization
par: Ma, Youpeng, et autres
Publié: (2025)
par: Ma, Youpeng, et autres
Publié: (2025)
QuanTest: Entanglement-Guided Testing of Quantum Neural Network Systems
par: Shi, Jinjing, et autres
Publié: (2024)
par: Shi, Jinjing, et autres
Publié: (2024)
QuanForge: A Mutation Testing Framework for Quantum Neural Networks
par: Shao, Minqi, et autres
Publié: (2026)
par: Shao, Minqi, et autres
Publié: (2026)
Deep Learning Library Testing: Definition, Methods and Challenges
par: Zhang, Xiaoyu, et autres
Publié: (2024)
par: Zhang, Xiaoyu, et autres
Publié: (2024)
Rethinking Testing for LLM Applications: Characteristics, Challenges, and a Lightweight Interaction Protocol
par: Ma, Wei, et autres
Publié: (2025)
par: Ma, Wei, et autres
Publié: (2025)
Rethinking the Value of Agent-Generated Tests for LLM-Based Software Engineering Agents
par: Chen, Zhi, et autres
Publié: (2026)
par: Chen, Zhi, et autres
Publié: (2026)
SACTOR: LLM-Driven Correct and Idiomatic C to Rust Translation with Static Analysis and FFI-Based Verification
par: Zhou, Tianyang, et autres
Publié: (2025)
par: Zhou, Tianyang, et autres
Publié: (2025)
An Experience Report on Regression-Free Repair of Deep Neural Network Model
par: Nakagawa, Takao, et autres
Publié: (2025)
par: Nakagawa, Takao, et autres
Publié: (2025)
FuzzAug: Data Augmentation by Coverage-guided Fuzzing for Neural Test Generation
par: He, Yifeng, et autres
Publié: (2024)
par: He, Yifeng, et autres
Publié: (2024)
Rethinking the effects of data contamination in Code Intelligence
par: Yang, Zhen, et autres
Publié: (2025)
par: Yang, Zhen, et autres
Publié: (2025)
Beyond Accuracy: An Empirical Study on Unit Testing in Open-source Deep Learning Projects
par: Wang, Han, et autres
Publié: (2024)
par: Wang, Han, et autres
Publié: (2024)
AutoSafeCoder: A Multi-Agent Framework for Securing LLM Code Generation through Static Analysis and Fuzz Testing
par: Nunez, Ana, et autres
Publié: (2024)
par: Nunez, Ana, et autres
Publié: (2024)
Adaptive Testing for LLM-Based Applications: A Diversity-based Approach
par: Yoon, Juyeon, et autres
Publié: (2025)
par: Yoon, Juyeon, et autres
Publié: (2025)
AutoTest: Evolutionary Code Solution Selection with Test Cases
par: Duan, Zhihua, et autres
Publié: (2024)
par: Duan, Zhihua, et autres
Publié: (2024)
Unify and Triumph: Polyglot, Diverse, and Self-Consistent Generation of Unit Tests with LLMs
par: Khelladi, Djamel Eddine, et autres
Publié: (2025)
par: Khelladi, Djamel Eddine, et autres
Publié: (2025)
Rethinking Scientific Modeling: Toward Physically Consistent and Simulation-Executable Programmatic Generation
par: Jiang, Yongqing, et autres
Publié: (2026)
par: Jiang, Yongqing, et autres
Publié: (2026)
Do Large Code Models Understand Programming Concepts? Counterfactual Analysis for Code Predicates
par: Hooda, Ashish, et autres
Publié: (2024)
par: Hooda, Ashish, et autres
Publié: (2024)
DIVE: Scaling Diversity in Agentic Task Synthesis for Generalizable Tool Use
par: Chen, Aili, et autres
Publié: (2026)
par: Chen, Aili, et autres
Publié: (2026)
Debugging and Runtime Analysis of Neural Networks with VLMs (A Case Study)
par: Hu, Boyue Caroline, et autres
Publié: (2025)
par: Hu, Boyue Caroline, et autres
Publié: (2025)
EvoGPT: Leveraging LLM-Driven Seed Diversity to Improve Search-Based Test Suite Generation
par: Broide, Lior, et autres
Publié: (2025)
par: Broide, Lior, et autres
Publié: (2025)
Metamorphic Testing of Deep Code Models: A Systematic Literature Review
par: Asgari, Ali, et autres
Publié: (2025)
par: Asgari, Ali, et autres
Publié: (2025)
InfCode: Adversarial Iterative Refinement of Tests and Patches for Reliable Software Issue Resolution
par: Li, KeFan, et autres
Publié: (2025)
par: Li, KeFan, et autres
Publié: (2025)
Towards Reliable Evaluation of Neural Program Repair with Natural Robustness Testing
par: Le-Cong, Thanh, et autres
Publié: (2024)
par: Le-Cong, Thanh, et autres
Publié: (2024)
From Human Interfaces to Agent Interfaces: Rethinking Software Design in the Age of AI-Native Systems
par: Wang, Shaolin, et autres
Publié: (2026)
par: Wang, Shaolin, et autres
Publié: (2026)
Validating LLM-Generated Programs with Metamorphic Prompt Testing
par: Wang, Xiaoyin, et autres
Publié: (2024)
par: Wang, Xiaoyin, et autres
Publié: (2024)
MAFT: Efficient Model-Agnostic Fairness Testing for Deep Neural Networks via Zero-Order Gradient Search
par: Wang, Zhaohui, et autres
Publié: (2024)
par: Wang, Zhaohui, et autres
Publié: (2024)
Clarifying Semantics of In-Context Examples for Unit Test Generation
par: Yang, Chen, et autres
Publié: (2025)
par: Yang, Chen, et autres
Publié: (2025)
RBT4DNN: Requirements-based Testing of Neural Networks
par: Mozumder, Nusrat Jahan, et autres
Publié: (2025)
par: Mozumder, Nusrat Jahan, et autres
Publié: (2025)
Rethinking the Role of Entropy in Optimizing Tool-Use Behaviors for Large Language Model Agents
par: Li, Zeping, et autres
Publié: (2026)
par: Li, Zeping, et autres
Publié: (2026)
Klear-CodeTest: Scalable Test Case Generation for Code Reinforcement Learning
par: Fu, Jia, et autres
Publié: (2025)
par: Fu, Jia, et autres
Publié: (2025)
DeMuVGN: Effective Software Defect Prediction Model by Learning Multi-view Software Dependency via Graph Neural Networks
par: Qiao, Yu, et autres
Publié: (2024)
par: Qiao, Yu, et autres
Publié: (2024)
Ethics Testing: Proactive Identification of Generative AI System Harms
par: Tan, Shin Hwei, et autres
Publié: (2026)
par: Tan, Shin Hwei, et autres
Publié: (2026)
PBT-Bench: Benchmarking AI Agents on Property-Based Testing
par: Jing, Lucas, et autres
Publié: (2026)
par: Jing, Lucas, et autres
Publié: (2026)
Automated Unit Test Case Generation: A Systematic Literature Review
par: Wang, Jason, et autres
Publié: (2025)
par: Wang, Jason, et autres
Publié: (2025)
Domain Adaptation for Code Model-based Unit Test Case Generation
par: Shin, Jiho, et autres
Publié: (2023)
par: Shin, Jiho, et autres
Publié: (2023)
Retrieval-Augmented Test Generation: How Far Are We?
par: Shin, Jiho, et autres
Publié: (2024)
par: Shin, Jiho, et autres
Publié: (2024)
The Debugging Decay Index: Rethinking Debugging Strategies for Code LLMs
par: Adnan, Muntasir, et autres
Publié: (2025)
par: Adnan, Muntasir, et autres
Publié: (2025)
Rethinking IDE Customization for Enhanced HAX: A Hyperdimensional Perspective
par: Koohestani, Roham, et autres
Publié: (2025)
par: Koohestani, Roham, et autres
Publié: (2025)
LLM-as-a-Judge for Scalable Test Coverage Evaluation: Accuracy, Operational Reliability, and Cost
par: Huang, Donghao, et autres
Publié: (2025)
par: Huang, Donghao, et autres
Publié: (2025)
Documents similaires
-
Auto-SPT: Automating Semantic Preserving Transformations for Code
par: Hooda, Ashish, et autres
Publié: (2025) -
Faster Configuration Performance Bug Testing with Neural Dual-level Prioritization
par: Ma, Youpeng, et autres
Publié: (2025) -
QuanTest: Entanglement-Guided Testing of Quantum Neural Network Systems
par: Shi, Jinjing, et autres
Publié: (2024) -
QuanForge: A Mutation Testing Framework for Quantum Neural Networks
par: Shao, Minqi, et autres
Publié: (2026) -
Deep Learning Library Testing: Definition, Methods and Challenges
par: Zhang, Xiaoyu, et autres
Publié: (2024)