Deep Learning Library Testing: Definition, Methods and Challenges
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Xiaoyu, Jiang, Weipeng, Shen, Chao, Li, Qi, Wang, Qian, Lin, Chenhao, Guan, Xiaohong |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
The Foundation Cracks: A Comprehensive Study on Bugs and Testing Practices in LLM Libraries
par: Jiang, Weipeng, et autres
Publié: (2025)
par: Jiang, Weipeng, et autres
Publié: (2025)
AutoEmpirical: LLM-Based Automated Research for Empirical Software Fault Analysis
par: Yu, Jiongchi, et autres
Publié: (2025)
par: Yu, Jiongchi, et autres
Publié: (2025)
Efficient DNN-Powered Software with Fair Sparse Models
par: Gao, Xuanqi, et autres
Publié: (2024)
par: Gao, Xuanqi, et autres
Publié: (2024)
The Invisible Hand: Unveiling Provider Bias in Large Language Models for Code Generation
par: Zhang, Xiaoyu, et autres
Publié: (2025)
par: Zhang, Xiaoyu, et autres
Publié: (2025)
False Friends in the Shell: Unveiling the Emoticon Semantic Confusion in Large Language Models
par: Jiang, Weipeng, et autres
Publié: (2026)
par: Jiang, Weipeng, et autres
Publié: (2026)
Deep Learning-Based Identification of Inconsistent Method Names: How Far Are We?
par: Wang, Taiming, et autres
Publié: (2025)
par: Wang, Taiming, et autres
Publié: (2025)
DREAM: Debugging and Repairing AutoML Pipelines
par: Zhang, Xiaoyu, et autres
Publié: (2023)
par: Zhang, Xiaoyu, et autres
Publié: (2023)
Explicating Tacit Regulatory Knowledge from LLMs to Auto-Formalize Requirements for Compliance Test Case Generation
par: Xue, Zhiyi, et autres
Publié: (2026)
par: Xue, Zhiyi, et autres
Publié: (2026)
Klear-CodeTest: Scalable Test Case Generation for Code Reinforcement Learning
par: Fu, Jia, et autres
Publié: (2025)
par: Fu, Jia, et autres
Publié: (2025)
Navigating the Labyrinth: Path-Sensitive Unit Test Generation with Large Language Models
par: Liao, Dianshu, et autres
Publié: (2025)
par: Liao, Dianshu, et autres
Publié: (2025)
Rethinking Diversity in Deep Neural Network Testing
par: Wang, Zi, et autres
Publié: (2023)
par: Wang, Zi, et autres
Publié: (2023)
ASSURE: Metamorphic Testing for AI-powered Browser Extensions
par: Gao, Xuanqi, et autres
Publié: (2025)
par: Gao, Xuanqi, et autres
Publié: (2025)
Rethinking Testing for LLM Applications: Characteristics, Challenges, and a Lightweight Interaction Protocol
par: Ma, Wei, et autres
Publié: (2025)
par: Ma, Wei, et autres
Publié: (2025)
Beyond Accuracy: An Empirical Study on Unit Testing in Open-source Deep Learning Projects
par: Wang, Han, et autres
Publié: (2024)
par: Wang, Han, et autres
Publié: (2024)
Keeping Deep Learning Models in Check: A History-Based Approach to Mitigate Overfitting
par: Li, Hao, et autres
Publié: (2024)
par: Li, Hao, et autres
Publié: (2024)
Open-Source AI-based SE Tools: Opportunities and Challenges of Collaborative Software Learning
par: Lin, Zhihao, et autres
Publié: (2024)
par: Lin, Zhihao, et autres
Publié: (2024)
When Fuzzing Meets LLMs: Challenges and Opportunities
par: Jiang, Yu, et autres
Publié: (2024)
par: Jiang, Yu, et autres
Publié: (2024)
Addressing Quality Challenges in Deep Learning: The Role of MLOps and Domain Knowledge
par: del Rey, Santiago, et autres
Publié: (2025)
par: del Rey, Santiago, et autres
Publié: (2025)
Software Development Life Cycle Perspective: A Survey of Benchmarks for Code Large Language Models and Agents
par: Wang, Kaixin, et autres
Publié: (2025)
par: Wang, Kaixin, et autres
Publié: (2025)
Effective Code Membership Inference for Code Completion Models via Adversarial Prompts
par: Jiang, Yuan, et autres
Publié: (2025)
par: Jiang, Yuan, et autres
Publié: (2025)
The Hitchhiker's Guide to Program Analysis, Part II: Deep Thoughts by LLMs
par: Li, Haonan, et autres
Publié: (2025)
par: Li, Haonan, et autres
Publié: (2025)
Tool-integrated Reinforcement Learning for Repo Deep Search
par: Ma, Zexiong, et autres
Publié: (2025)
par: Ma, Zexiong, et autres
Publié: (2025)
Enhancing High-Quality Code Generation in Large Language Models with Comparative Prefix-Tuning
par: Jiang, Yuan, et autres
Publié: (2025)
par: Jiang, Yuan, et autres
Publié: (2025)
DeepCode: Open Agentic Coding
par: Li, Zongwei, et autres
Publié: (2025)
par: Li, Zongwei, et autres
Publié: (2025)
Rethinking Technology Stack Selection with AI Coding Proficiency
par: Zhang, Xiaoyu, et autres
Publié: (2025)
par: Zhang, Xiaoyu, et autres
Publié: (2025)
QuanTest: Entanglement-Guided Testing of Quantum Neural Network Systems
par: Shi, Jinjing, et autres
Publié: (2024)
par: Shi, Jinjing, et autres
Publié: (2024)
Testing Storage-System Correctness: Challenges, Fuzzing Limitations, and AI-Augmented Opportunities
par: Wang, Ying, et autres
Publié: (2026)
par: Wang, Ying, et autres
Publié: (2026)
TENET: Leveraging Tests Beyond Validation for Code Generation
par: Hu, Yiran, et autres
Publié: (2025)
par: Hu, Yiran, et autres
Publié: (2025)
FVSpec: Real-World Property-Based Tests as Lean Challenges
par: Dougherty, Quinn, et autres
Publié: (2026)
par: Dougherty, Quinn, et autres
Publié: (2026)
Rethinking the Value of Agent-Generated Tests for LLM-Based Software Engineering Agents
par: Chen, Zhi, et autres
Publié: (2026)
par: Chen, Zhi, et autres
Publié: (2026)
DeepKnowledge: Generalisation-Driven Deep Learning Testing
par: Missaoui, Sondess, et autres
Publié: (2024)
par: Missaoui, Sondess, et autres
Publié: (2024)
FlaKat: A Machine Learning-Based Categorization Framework for Flaky Tests
par: Lin, Shizhe, et autres
Publié: (2024)
par: Lin, Shizhe, et autres
Publié: (2024)
DialogAgent: An Auto-engagement Agent for Code Question Answering Data Production
par: Liang, Xiaoyun, et autres
Publié: (2024)
par: Liang, Xiaoyun, et autres
Publié: (2024)
Just-In-Time Software Defect Prediction via Bi-modal Change Representation Learning
par: Jiang, Yuze, et autres
Publié: (2024)
par: Jiang, Yuze, et autres
Publié: (2024)
Empowering AI to Generate Better AI Code: Guided Generation of Deep Learning Projects with LLMs
par: Xie, Chen, et autres
Publié: (2025)
par: Xie, Chen, et autres
Publié: (2025)
Challenges in Testing Large Language Model Based Software: A Faceted Taxonomy
par: Dobslaw, Felix, et autres
Publié: (2025)
par: Dobslaw, Felix, et autres
Publié: (2025)
Clarifying Semantics of In-Context Examples for Unit Test Generation
par: Yang, Chen, et autres
Publié: (2025)
par: Yang, Chen, et autres
Publié: (2025)
VSRQ: Quantitative Assessment Method for Safety Risk of Vehicle Intelligent Connected System
par: Zhang, Tian, et autres
Publié: (2023)
par: Zhang, Tian, et autres
Publié: (2023)
Explore-Construct-Filter: An Automated Framework for Rich and Reliable API Knowledge Graph Construction
par: Sun, Yanbang, et autres
Publié: (2025)
par: Sun, Yanbang, et autres
Publié: (2025)
Metamorphic Testing of Deep Code Models: A Systematic Literature Review
par: Asgari, Ali, et autres
Publié: (2025)
par: Asgari, Ali, et autres
Publié: (2025)
Documents similaires
-
The Foundation Cracks: A Comprehensive Study on Bugs and Testing Practices in LLM Libraries
par: Jiang, Weipeng, et autres
Publié: (2025) -
AutoEmpirical: LLM-Based Automated Research for Empirical Software Fault Analysis
par: Yu, Jiongchi, et autres
Publié: (2025) -
Efficient DNN-Powered Software with Fair Sparse Models
par: Gao, Xuanqi, et autres
Publié: (2024) -
The Invisible Hand: Unveiling Provider Bias in Large Language Models for Code Generation
par: Zhang, Xiaoyu, et autres
Publié: (2025) -
False Friends in the Shell: Unveiling the Emoticon Semantic Confusion in Large Language Models
par: Jiang, Weipeng, et autres
Publié: (2026)