DeepSample: DNN sampling-based testing for operational accuracy assessment
Fuente:
arXiv
Saved in:
| Main Authors: | Guerriero, Antonio, Pietrantuono, Roberto, Russo, Stefano |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Operation is the hardest teacher: estimating DNN accuracy looking for mispredictions
by: Guerriero, Antonio, et al.
Published: (2021)
by: Guerriero, Antonio, et al.
Published: (2021)
Iterative Assessment and Improvement of DNN Operational Accuracy
by: Guerriero, Antonio, et al.
Published: (2023)
by: Guerriero, Antonio, et al.
Published: (2023)
Causal Reasoning in Software Quality Assurance: A Systematic Review
by: Giamattei, Luca, et al.
Published: (2024)
by: Giamattei, Luca, et al.
Published: (2024)
Reasoning-Based Software Testing
by: Giamattei, Luca, et al.
Published: (2023)
by: Giamattei, Luca, et al.
Published: (2023)
Causal Software Engineering: A Vision and Roadmap
by: Pietrantuono, Roberto, et al.
Published: (2026)
by: Pietrantuono, Roberto, et al.
Published: (2026)
Reinforcement Learning for Online Testing of Autonomous Driving Systems: a Replication and Extension Study
by: Giamattei, Luca, et al.
Published: (2024)
by: Giamattei, Luca, et al.
Published: (2024)
Efficient DNN-Powered Software with Fair Sparse Models
by: Gao, Xuanqi, et al.
Published: (2024)
by: Gao, Xuanqi, et al.
Published: (2024)
GAN-enhanced Simulation-driven DNN Testing in Absence of Ground Truth
by: Attaoui, Mohammed, et al.
Published: (2025)
by: Attaoui, Mohammed, et al.
Published: (2025)
More Is Different: Toward a Theory of Emergence in AI-Native Software Ecosystems
by: Russo, Daniel
Published: (2026)
by: Russo, Daniel
Published: (2026)
RBT4DNN: Requirements-based Testing of Neural Networks
by: Mozumder, Nusrat Jahan, et al.
Published: (2025)
by: Mozumder, Nusrat Jahan, et al.
Published: (2025)
GenAI-based test case generation and execution in SDV platform
by: Zyberaj, Denesa, et al.
Published: (2025)
by: Zyberaj, Denesa, et al.
Published: (2025)
Modeling Resilience of Collaborative AI Systems
by: Rimawi, Diaeddin, et al.
Published: (2024)
by: Rimawi, Diaeddin, et al.
Published: (2024)
DeepQuali: Initial results of a study on the use of large language models for assessing the quality of user stories
by: Trendowicz, Adam, et al.
Published: (2026)
by: Trendowicz, Adam, et al.
Published: (2026)
A3Rank: Augmentation Alignment Analysis for Prioritizing Overconfident Failing Samples for Deep Learning Models
by: Wei, Zhengyuan, et al.
Published: (2024)
by: Wei, Zhengyuan, et al.
Published: (2024)
Do Prompt Patterns Affect Code Quality? A First Empirical Assessment of ChatGPT-Generated Code
by: Della Porta, Antonio, et al.
Published: (2025)
by: Della Porta, Antonio, et al.
Published: (2025)
CMSA algorithm for solving the prioritized pairwise test data generation problem in software product lines
by: Ferrer, Javier, et al.
Published: (2024)
by: Ferrer, Javier, et al.
Published: (2024)
Design choices made by LLM-based test generators prevent them from finding bugs
by: Mathews, Noble Saji, et al.
Published: (2024)
by: Mathews, Noble Saji, et al.
Published: (2024)
An empirical study of LoRA-based fine-tuning of large language models for automated test case generation
by: Moradi, Milad, et al.
Published: (2026)
by: Moradi, Milad, et al.
Published: (2026)
A Systematic Literature Review on Explainability for Machine/Deep Learning-based Software Engineering Research
by: Cao, Sicong, et al.
Published: (2024)
by: Cao, Sicong, et al.
Published: (2024)
Enhancing AI-based Generation of Software Exploits with Contextual Information
by: Liguori, Pietro, et al.
Published: (2024)
by: Liguori, Pietro, et al.
Published: (2024)
o3-mini vs DeepSeek-R1: Which One is Safer?
by: Arrieta, Aitor, et al.
Published: (2025)
by: Arrieta, Aitor, et al.
Published: (2025)
Are LLMs Ready for TOON? Benchmarking Structural Correctness-Sustainability Trade-offs in Novel Structured Output Formats
by: Masciari, Elio, et al.
Published: (2026)
by: Masciari, Elio, et al.
Published: (2026)
Artificial intelligence for context-aware visual change detection in software test automation
by: Moradi, Milad, et al.
Published: (2024)
by: Moradi, Milad, et al.
Published: (2024)
The importance of visual modelling languages in generative software engineering
by: Rossi, Roberto
Published: (2024)
by: Rossi, Roberto
Published: (2024)
Metamorphic Testing of Large Language Models for Natural Language Processing
by: Cho, Steven, et al.
Published: (2025)
by: Cho, Steven, et al.
Published: (2025)
Document Retrieval Augmented Fine-Tuning (DRAFT) for safety-critical software assessments
by: Bolton, Regan, et al.
Published: (2025)
by: Bolton, Regan, et al.
Published: (2025)
Lemur: Log Parsing with Entropy Sampling and Chain-of-Thought Merging
by: Zhang, Wei, et al.
Published: (2024)
by: Zhang, Wei, et al.
Published: (2024)
DeepCode: Open Agentic Coding
by: Li, Zongwei, et al.
Published: (2025)
by: Li, Zongwei, et al.
Published: (2025)
Machine Learning Experiences: A story of learning AI for use in enterprise software testing that can be used by anyone
by: Cohoon, Michael, et al.
Published: (2025)
by: Cohoon, Michael, et al.
Published: (2025)
On Security Weaknesses and Vulnerabilities in Deep Learning Systems
by: Lai, Zhongzheng, et al.
Published: (2024)
by: Lai, Zhongzheng, et al.
Published: (2024)
Rethinking Diversity in Deep Neural Network Testing
by: Wang, Zi, et al.
Published: (2023)
by: Wang, Zi, et al.
Published: (2023)
When the Code Autopilot Breaks: Why LLMs Falter in Embedded Machine Learning
by: Morabito, Roberto, et al.
Published: (2025)
by: Morabito, Roberto, et al.
Published: (2025)
Evaluating Reliability Gaps in Large Language Model Safety via Repeated Prompt Sampling
by: Broadwater, Keita
Published: (2026)
by: Broadwater, Keita
Published: (2026)
Multi-Sample Prompting and Actor-Critic Prompt Optimization for Diverse Synthetic Data Generation
by: El-Hajjami, Abdelkarim, et al.
Published: (2025)
by: El-Hajjami, Abdelkarim, et al.
Published: (2025)
Toward Patch Robustness Certification and Detection for Deep Learning Systems Beyond Consistent Samples
by: Zhou, Qilin, et al.
Published: (2025)
by: Zhou, Qilin, et al.
Published: (2025)
Pushing the Boundary: Specialising Deep Configuration Performance Learning
by: Gong, Jingzhi
Published: (2024)
by: Gong, Jingzhi
Published: (2024)
Deep Learning Library Testing: Definition, Methods and Challenges
by: Zhang, Xiaoyu, et al.
Published: (2024)
by: Zhang, Xiaoyu, et al.
Published: (2024)
Unraveling Code Clone Dynamics in Deep Learning Frameworks
by: Assi, Maram, et al.
Published: (2024)
by: Assi, Maram, et al.
Published: (2024)
Tool-integrated Reinforcement Learning for Repo Deep Search
by: Ma, Zexiong, et al.
Published: (2025)
by: Ma, Zexiong, et al.
Published: (2025)
Deep Learning for Code Intelligence: Survey, Benchmark and Toolkit
by: Wan, Yao, et al.
Published: (2023)
by: Wan, Yao, et al.
Published: (2023)
Similar Items
-
Operation is the hardest teacher: estimating DNN accuracy looking for mispredictions
by: Guerriero, Antonio, et al.
Published: (2021) -
Iterative Assessment and Improvement of DNN Operational Accuracy
by: Guerriero, Antonio, et al.
Published: (2023) -
Causal Reasoning in Software Quality Assurance: A Systematic Review
by: Giamattei, Luca, et al.
Published: (2024) -
Reasoning-Based Software Testing
by: Giamattei, Luca, et al.
Published: (2023) -
Causal Software Engineering: A Vision and Roadmap
by: Pietrantuono, Roberto, et al.
Published: (2026)