muPRL: A Mutation Testing Pipeline for Deep Reinforcement Learning based on Real Faults
Fuente:
arXiv
Salvato in:
| Autori principali: | Thomas, Deepak-George, Biagiola, Matteo, Humbatova, Nargiz, Wardat, Mohammad, Jahangirova, Gunel, Rajan, Hridesh, Tonella, Paolo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Real Faults in Deep Learning Fault Benchmarks: How Real Are They?
di: Jahangirova, Gunel, et al.
Pubblicazione: (2024)
di: Jahangirova, Gunel, et al.
Pubblicazione: (2024)
MuFF: Stable and Sensitive Post-training Mutation Testing for Deep Learning
di: Kim, Jinhan, et al.
Pubblicazione: (2025)
di: Kim, Jinhan, et al.
Pubblicazione: (2025)
An Empirical Study of Fault Localisation Techniques for Deep Learning
di: Humbatova, Nargiz, et al.
Pubblicazione: (2024)
di: Humbatova, Nargiz, et al.
Pubblicazione: (2024)
New Formulation of DNN Statistical Mutation Killing for Ensuring Monotonicity: A Technical Report
di: Kim, Jinhan, et al.
Pubblicazione: (2025)
di: Kim, Jinhan, et al.
Pubblicazione: (2025)
Fault Localisation and Repair for DL Systems: An Empirical Study with LLMs
di: Kim, Jinhan, et al.
Pubblicazione: (2025)
di: Kim, Jinhan, et al.
Pubblicazione: (2025)
Revisiting "Revisiting Neuron Coverage for DNN Testing: A Layer-Wise and Distribution-Aware Criterion": A Critical Review and Implications on DNN Coverage Testing
di: Kim, Jinhan, et al.
Pubblicazione: (2026)
di: Kim, Jinhan, et al.
Pubblicazione: (2026)
Testing of Deep Reinforcement Learning Agents with Surrogate Models
di: Biagiola, Matteo, et al.
Pubblicazione: (2023)
di: Biagiola, Matteo, et al.
Pubblicazione: (2023)
TopoMap: A Feature-based Semantic Discriminator of the Topographical Regions in the Test Input Space
di: De Vita, Gianmarco, et al.
Pubblicazione: (2025)
di: De Vita, Gianmarco, et al.
Pubblicazione: (2025)
Boundary State Generation for Testing and Improvement of Autonomous Driving Systems
di: Biagiola, Matteo, et al.
Pubblicazione: (2023)
di: Biagiola, Matteo, et al.
Pubblicazione: (2023)
Mock Deep Testing: Toward Separate Development of Data and Models for Deep Learning
di: Manke, Ruchira, et al.
Pubblicazione: (2025)
di: Manke, Ruchira, et al.
Pubblicazione: (2025)
Adaptive Random Testing with Q-grams: The Illusion Comes True
di: Biagiola, Matteo, et al.
Pubblicazione: (2024)
di: Biagiola, Matteo, et al.
Pubblicazione: (2024)
Improving the Readability of Automatically Generated Tests using Large Language Models
di: Biagiola, Matteo, et al.
Pubblicazione: (2024)
di: Biagiola, Matteo, et al.
Pubblicazione: (2024)
A Taxonomy of Real Faults in Hybrid Quantum-Classical Architectures
di: Bensoussan, Avner, et al.
Pubblicazione: (2025)
di: Bensoussan, Avner, et al.
Pubblicazione: (2025)
Leveraging Data Characteristics for Bug Localization in Deep Learning Programs
di: Manke, Ruchira, et al.
Pubblicazione: (2024)
di: Manke, Ruchira, et al.
Pubblicazione: (2024)
GenMorph: Automatically Generating Metamorphic Relations via Genetic Programming
di: Ayerdi, Jon, et al.
Pubblicazione: (2023)
di: Ayerdi, Jon, et al.
Pubblicazione: (2023)
Reinforcement Learning for Online Testing of Autonomous Driving Systems: a Replication and Extension Study
di: Giamattei, Luca, et al.
Pubblicazione: (2024)
di: Giamattei, Luca, et al.
Pubblicazione: (2024)
Two is Better Than One: Digital Siblings to Improve Autonomous Driving Testing
di: Biagiola, Matteo, et al.
Pubblicazione: (2023)
di: Biagiola, Matteo, et al.
Pubblicazione: (2023)
Neural Embeddings for Web Testing
di: Kanaththage, Kasun, et al.
Pubblicazione: (2023)
di: Kanaththage, Kasun, et al.
Pubblicazione: (2023)
Understanding LLM-Driven Test Oracle Generation
di: Bodicoat, Adam, et al.
Pubblicazione: (2026)
di: Bodicoat, Adam, et al.
Pubblicazione: (2026)
How Does Chunking Affect Retrieval-Augmented Code Completion? A Controlled Empirical Study
di: Wu, Xinjian, et al.
Pubblicazione: (2026)
di: Wu, Xinjian, et al.
Pubblicazione: (2026)
Inferring Data Preconditions from Deep Learning Models for Trustworthy Prediction in Deployment
di: Ahmed, Shibbir, et al.
Pubblicazione: (2024)
di: Ahmed, Shibbir, et al.
Pubblicazione: (2024)
Comparative Analysis of Carbon Footprint in Manual vs. LLM-Assisted Code Development
di: Cheung, Kuen Sum, et al.
Pubblicazione: (2025)
di: Cheung, Kuen Sum, et al.
Pubblicazione: (2025)
Benchmarking Generative AI Models for Deep Learning Test Input Generation
di: Maryam, et al.
Pubblicazione: (2024)
di: Maryam, et al.
Pubblicazione: (2024)
Automated Feature Extraction for Testing Deep Learning Systems Through Illumination Search
di: Tahereh Zohdinasab, et al.
Pubblicazione: (2026)
di: Tahereh Zohdinasab, et al.
Pubblicazione: (2026)
Simulator Ensembles for Trustworthy Autonomous Driving Testing
di: Sorokin, Lev, et al.
Pubblicazione: (2025)
di: Sorokin, Lev, et al.
Pubblicazione: (2025)
Towards Automated Page Object Generation for Web Testing using Large Language Models
di: Karagöz, Betül, et al.
Pubblicazione: (2026)
di: Karagöz, Betül, et al.
Pubblicazione: (2026)
When Agents Fail: A Comprehensive Study of Bugs in LLM Agents with Automated Labeling
di: Islam, Niful, et al.
Pubblicazione: (2026)
di: Islam, Niful, et al.
Pubblicazione: (2026)
LLMLOOP: Improving LLM-Generated Code and Tests through Automated Iterative Feedback Loops
di: Ravi, Ravin, et al.
Pubblicazione: (2026)
di: Ravi, Ravin, et al.
Pubblicazione: (2026)
IRepair: An Intent-Aware Approach to Repair Data-Driven Errors in Large Language Models
di: Imtiaz, Sayem Mohammad, et al.
Pubblicazione: (2025)
di: Imtiaz, Sayem Mohammad, et al.
Pubblicazione: (2025)
Hybrid Fault-Driven Mutation Testing for Python
di: Alimadadi, Saba, et al.
Pubblicazione: (2026)
di: Alimadadi, Saba, et al.
Pubblicazione: (2026)
XMutant: XAI-based Fuzzing for Deep Learning Systems
di: Chen, Xingcheng, et al.
Pubblicazione: (2025)
di: Chen, Xingcheng, et al.
Pubblicazione: (2025)
Data-Driven Evidence-Based Syntactic Sugar Design
di: OBrien, David, et al.
Pubblicazione: (2024)
di: OBrien, David, et al.
Pubblicazione: (2024)
PCLA: A Framework for Testing Autonomous Agents in the CARLA Simulator
di: Tehrani, Masoud Jamshidiyan, et al.
Pubblicazione: (2025)
di: Tehrani, Masoud Jamshidiyan, et al.
Pubblicazione: (2025)
LLMs-Powered Real-Time Fault Injection: An Approach Toward Intelligent Fault Test Cases Generation
di: Abboush, Mohammad, et al.
Pubblicazione: (2025)
di: Abboush, Mohammad, et al.
Pubblicazione: (2025)
SelfHeal: Empirical Fix Pattern Analysis and Bug Repair in LLM Agents
di: Islam, Niful, et al.
Pubblicazione: (2026)
di: Islam, Niful, et al.
Pubblicazione: (2026)
Testing for Fault Diversity in Reinforcement Learning
di: Mazouni, Quentin, et al.
Pubblicazione: (2024)
di: Mazouni, Quentin, et al.
Pubblicazione: (2024)
Bridging Research and Practice in Simulation-based Testing of Industrial Robot Navigation Systems
di: Khatiri, Sajad, et al.
Pubblicazione: (2025)
di: Khatiri, Sajad, et al.
Pubblicazione: (2025)
Benchmarking and Evaluating VLMs for Software Architecture Diagram Understanding
di: Ouyang, Shuyin, et al.
Pubblicazione: (2026)
di: Ouyang, Shuyin, et al.
Pubblicazione: (2026)
Predicting Safety Misbehaviours in Autonomous Driving Systems using Uncertainty Quantification
di: Grewal, Ruben, et al.
Pubblicazione: (2024)
di: Grewal, Ruben, et al.
Pubblicazione: (2024)
Detecting Trojaned DNNs via Spectral Regression Analysis
di: Pasini, Samuele, et al.
Pubblicazione: (2026)
di: Pasini, Samuele, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Real Faults in Deep Learning Fault Benchmarks: How Real Are They?
di: Jahangirova, Gunel, et al.
Pubblicazione: (2024) -
MuFF: Stable and Sensitive Post-training Mutation Testing for Deep Learning
di: Kim, Jinhan, et al.
Pubblicazione: (2025) -
An Empirical Study of Fault Localisation Techniques for Deep Learning
di: Humbatova, Nargiz, et al.
Pubblicazione: (2024) -
New Formulation of DNN Statistical Mutation Killing for Ensuring Monotonicity: A Technical Report
di: Kim, Jinhan, et al.
Pubblicazione: (2025) -
Fault Localisation and Repair for DL Systems: An Empirical Study with LLMs
di: Kim, Jinhan, et al.
Pubblicazione: (2025)