Testing of Deep Reinforcement Learning Agents with Surrogate Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Biagiola, Matteo, Tonella, Paolo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Two is Better Than One: Digital Siblings to Improve Autonomous Driving Testing
por: Biagiola, Matteo, et al.
Publicado: (2023)
por: Biagiola, Matteo, et al.
Publicado: (2023)
Benchmarking Generative AI Models for Deep Learning Test Input Generation
por: Maryam, et al.
Publicado: (2024)
por: Maryam, et al.
Publicado: (2024)
Enhancing Differential Testing With LLMs For Testing Deep Learning Libraries
por: Li, Meiziniu, et al.
Publicado: (2024)
por: Li, Meiziniu, et al.
Publicado: (2024)
COMET: Coverage-guided Model Generation For Deep Learning Library Testing
por: Li, Meiziniu, et al.
Publicado: (2022)
por: Li, Meiziniu, et al.
Publicado: (2022)
Reinforcement Learning for Online Testing of Autonomous Driving Systems: a Replication and Extension Study
por: Giamattei, Luca, et al.
Publicado: (2024)
por: Giamattei, Luca, et al.
Publicado: (2024)
DeepCodeProbe: Towards Understanding What Models Trained on Code Learn
por: Majdinasab, Vahid, et al.
Publicado: (2024)
por: Majdinasab, Vahid, et al.
Publicado: (2024)
The Explabox: Model-Agnostic Machine Learning Transparency & Analysis
por: Robeer, Marcel, et al.
Publicado: (2024)
por: Robeer, Marcel, et al.
Publicado: (2024)
Boundary State Generation for Testing and Improvement of Autonomous Driving Systems
por: Biagiola, Matteo, et al.
Publicado: (2023)
por: Biagiola, Matteo, et al.
Publicado: (2023)
On the Mistaken Assumption of Interchangeable Deep Reinforcement Learning Implementations
por: Hundal, Rajdeep Singh, et al.
Publicado: (2025)
por: Hundal, Rajdeep Singh, et al.
Publicado: (2025)
Toward Efficient Testing of Graph Neural Networks via Test Input Prioritization
por: Yang, Lichen, et al.
Publicado: (2025)
por: Yang, Lichen, et al.
Publicado: (2025)
LSPRAG: LSP-Guided RAG for Language-Agnostic Real-Time Unit Test Generation
por: Go, Gwihwan, et al.
Publicado: (2025)
por: Go, Gwihwan, et al.
Publicado: (2025)
LLMORPH: Automated Metamorphic Testing of Large Language Models
por: Cho, Steven, et al.
Publicado: (2026)
por: Cho, Steven, et al.
Publicado: (2026)
Autonomous QA Agent: A Retrieval-Augmented Framework for Reliable Selenium Script Generation
por: Vali, Dudekula Kasim
Publicado: (2025)
por: Vali, Dudekula Kasim
Publicado: (2025)
AgentAssay: Token-Efficient Regression Testing for Non-Deterministic AI Agent Workflows
por: Bhardwaj, Varun Pratap
Publicado: (2026)
por: Bhardwaj, Varun Pratap
Publicado: (2026)
Towards Explainable Test Case Prioritisation with Learning-to-Rank Models
por: Ramírez, Aurora, et al.
Publicado: (2024)
por: Ramírez, Aurora, et al.
Publicado: (2024)
Assessing Data Augmentation-Induced Bias in Training and Testing of Machine Learning Models
por: More, Riddhi, et al.
Publicado: (2025)
por: More, Riddhi, et al.
Publicado: (2025)
Demystifying the Silence of Correctness Bugs in PyTorch Compiler
por: Li, Meiziniu, et al.
Publicado: (2026)
por: Li, Meiziniu, et al.
Publicado: (2026)
CoverUp: Effective High Coverage Test Generation for Python
por: Pizzorno, Juan Altmayer, et al.
Publicado: (2024)
por: Pizzorno, Juan Altmayer, et al.
Publicado: (2024)
Navigating the growing field of research on AI for software testing -- the taxonomy for AI-augmented software testing and an ontology-driven literature survey
por: Schieferdecker, Ina K.
Publicado: (2025)
por: Schieferdecker, Ina K.
Publicado: (2025)
An Analysis of LLM Fine-Tuning and Few-Shot Learning for Flaky Test Detection and Classification
por: More, Riddhi, et al.
Publicado: (2025)
por: More, Riddhi, et al.
Publicado: (2025)
A Systematic Approach for Assessing Large Language Models' Test Case Generation Capability
por: Chang, Hung-Fu, et al.
Publicado: (2025)
por: Chang, Hung-Fu, et al.
Publicado: (2025)
Harnessing the Power of Large Language Models for Software Testing Education: A Focus on ISTQB Syllabus
por: Ngo, Tuan-Phong, et al.
Publicado: (2025)
por: Ngo, Tuan-Phong, et al.
Publicado: (2025)
Automated Vulnerability Detection Using Deep Learning Technique
por: Yang, Guan-Yan, et al.
Publicado: (2024)
por: Yang, Guan-Yan, et al.
Publicado: (2024)
DEFault++: Automated Fault Detection, Categorization, and Diagnosis for Transformer Architectures
por: Jahan, Sigma, et al.
Publicado: (2026)
por: Jahan, Sigma, et al.
Publicado: (2026)
From Untestable to Testable: Metamorphic Testing in the Age of LLMs
por: Terragni, Valerio
Publicado: (2026)
por: Terragni, Valerio
Publicado: (2026)
On the Soundness and Consistency of LLM Agents for Executing Test Cases Written in Natural Language
por: Salva, Sébastien, et al.
Publicado: (2025)
por: Salva, Sébastien, et al.
Publicado: (2025)
The Future of AI-Driven Software Engineering
por: Terragni, Valerio, et al.
Publicado: (2024)
por: Terragni, Valerio, et al.
Publicado: (2024)
A Comprehensive Study on Large Language Models for Mutation Testing
por: Wang, Bo, et al.
Publicado: (2024)
por: Wang, Bo, et al.
Publicado: (2024)
E-Test: E'er-Improving Test Suites
por: Qiu, Ketai, et al.
Publicado: (2025)
por: Qiu, Ketai, et al.
Publicado: (2025)
A semantic mutation metric for metamorphic relation adequacy in scientific computing programs
por: Li, Meng, et al.
Publicado: (2026)
por: Li, Meng, et al.
Publicado: (2026)
Analyzing Quantum Programs with LintQ: A Static Analysis Framework for Qiskit
por: Paltenghi, Matteo, et al.
Publicado: (2023)
por: Paltenghi, Matteo, et al.
Publicado: (2023)
Automated Generation of Issue-Reproducing Tests by Combining LLMs and Search-Based Testing
por: Kitsios, Konstantinos, et al.
Publicado: (2025)
por: Kitsios, Konstantinos, et al.
Publicado: (2025)
CodeTracer: Towards Traceable Agent States
por: Li, Han, et al.
Publicado: (2026)
por: Li, Han, et al.
Publicado: (2026)
A Match Made in Heaven? AI-driven Matching of Vulnerabilities and Security Unit Tests
por: Iannone, Emanuele, et al.
Publicado: (2025)
por: Iannone, Emanuele, et al.
Publicado: (2025)
VLM-Fuzz: Vision Language Model Assisted Recursive Depth-first Search Exploration for Effective UI Testing of Android Apps
por: Demissie, Biniam Fisseha, et al.
Publicado: (2025)
por: Demissie, Biniam Fisseha, et al.
Publicado: (2025)
Test-Driven AI Agent Definition (TDAD): Compiling Tool-Using Agents from Behavioral Specifications
por: Rehan, Tzafrir
Publicado: (2026)
por: Rehan, Tzafrir
Publicado: (2026)
RefactorBench: Evaluating Stateful Reasoning in Language Agents Through Code
por: Gautam, Dhruv, et al.
Publicado: (2025)
por: Gautam, Dhruv, et al.
Publicado: (2025)
Abstain and Validate: A Dual-LLM Policy for Reducing Noise in Agentic Program Repair
por: Cambronero, José, et al.
Publicado: (2025)
por: Cambronero, José, et al.
Publicado: (2025)
AcTracer: Active Testing of Large Language Model via Multi-Stage Sampling
por: Huang, Yuheng, et al.
Publicado: (2024)
por: Huang, Yuheng, et al.
Publicado: (2024)
QITE: Assembly-Level, Cross-Platform Testing of Quantum Computing Platforms
por: Paltenghi, Matteo, et al.
Publicado: (2025)
por: Paltenghi, Matteo, et al.
Publicado: (2025)
Ejemplares similares
-
Two is Better Than One: Digital Siblings to Improve Autonomous Driving Testing
por: Biagiola, Matteo, et al.
Publicado: (2023) -
Benchmarking Generative AI Models for Deep Learning Test Input Generation
por: Maryam, et al.
Publicado: (2024) -
Enhancing Differential Testing With LLMs For Testing Deep Learning Libraries
por: Li, Meiziniu, et al.
Publicado: (2024) -
COMET: Coverage-guided Model Generation For Deep Learning Library Testing
por: Li, Meiziniu, et al.
Publicado: (2022) -
Reinforcement Learning for Online Testing of Autonomous Driving Systems: a Replication and Extension Study
por: Giamattei, Luca, et al.
Publicado: (2024)