MutaBot: A Mutation Testing Approach for Chatbots
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Urrico, Michael Ferdinando, Clerissi, Diego, Mariani, Leonardo |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Towards Multi-Platform Mutation Testing of Task-based Chatbots
par: Clerissi, Diego, et autres
Publié: (2025)
par: Clerissi, Diego, et autres
Publié: (2025)
Automated Testing of Task-based Chatbots: How Far Are We?
par: Clerissi, Diego, et autres
Publié: (2026)
par: Clerissi, Diego, et autres
Publié: (2026)
Test Case Generation for Dialogflow Task-Based Chatbots
par: Rapisarda, Rocco Gianni, et autres
Publié: (2025)
par: Rapisarda, Rocco Gianni, et autres
Publié: (2025)
Assessing Task-based Chatbots: Snapshot and Curated Datasets for Dialogflow
par: Masserini, Elena, et autres
Publié: (2026)
par: Masserini, Elena, et autres
Publié: (2026)
Towards the Assessment of Task-based Chatbots: From the TOFU-R Snapshot to the BRASATO Curated Dataset
par: Masserini, Elena, et autres
Publié: (2025)
par: Masserini, Elena, et autres
Publié: (2025)
MutaGReP: Execution-Free Repository-Grounded Plan Search for Code-Use
par: Khan, Zaid, et autres
Publié: (2025)
par: Khan, Zaid, et autres
Publié: (2025)
QuanForge: A Mutation Testing Framework for Quantum Neural Networks
par: Shao, Minqi, et autres
Publié: (2026)
par: Shao, Minqi, et autres
Publié: (2026)
Mutation-based Consistency Testing for Evaluating the Code Understanding Capability of LLMs
par: Li, Ziyu, et autres
Publié: (2024)
par: Li, Ziyu, et autres
Publié: (2024)
Saving SWE-Bench: A Benchmark Mutation Approach for Realistic Agent Evaluation
par: Garg, Spandan, et autres
Publié: (2025)
par: Garg, Spandan, et autres
Publié: (2025)
Analyzing Prompt Influence on Automated Method Generation: An Empirical Study with Copilot
par: Fagadau, Ionut Daniel, et autres
Publié: (2024)
par: Fagadau, Ionut Daniel, et autres
Publié: (2024)
Understanding the Limits of Automated Evaluation for Code Review Bots in Practice
par: Karakaya, Veli, et autres
Publié: (2026)
par: Karakaya, Veli, et autres
Publié: (2026)
DroidBot-GPT: GPT-powered UI Automation for Android
par: Wen, Hao, et autres
Publié: (2023)
par: Wen, Hao, et autres
Publié: (2023)
Past, Present and Future: Exploring Adaptive AI in Software Development Bots
par: Elsisi, Omar, et autres
Publié: (2025)
par: Elsisi, Omar, et autres
Publié: (2025)
EduBot -- Can LLMs Solve Personalized Learning and Programming Assignments?
par: Wang, Yibin, et autres
Publié: (2025)
par: Wang, Yibin, et autres
Publié: (2025)
Mutation-Guided LLM-based Test Generation at Meta
par: Foster, Christopher, et autres
Publié: (2025)
par: Foster, Christopher, et autres
Publié: (2025)
Chatbot-Based Assessment of Code Understanding in Automated Programming Assessment Systems
par: Frankford, Eduard, et autres
Publié: (2026)
par: Frankford, Eduard, et autres
Publié: (2026)
Adaptive Testing for LLM-Based Applications: A Diversity-based Approach
par: Yoon, Juyeon, et autres
Publié: (2025)
par: Yoon, Juyeon, et autres
Publié: (2025)
A Metamorphic Testing Approach to Diagnosing Memorization in LLM-Based Program Repair
par: De Koning, Milan, et autres
Publié: (2026)
par: De Koning, Milan, et autres
Publié: (2026)
MIST-RL: Mutation-based Incremental Suite Testing via Reinforcement Learning
par: Zhu, Sicheng, et autres
Publié: (2026)
par: Zhu, Sicheng, et autres
Publié: (2026)
Re-Evaluating Code LLM Benchmarks Under Semantic Mutation
par: Pan, Zhiyuan, et autres
Publié: (2025)
par: Pan, Zhiyuan, et autres
Publié: (2025)
Understanding the Helpfulness of Stale Bot for Pull-based Development: An Empirical Study of 20 Large Open-Source Projects
par: Khatoonabadi, SayedHassan, et autres
Publié: (2023)
par: Khatoonabadi, SayedHassan, et autres
Publié: (2023)
Cooperative Multi-agent Approach for Automated Computer Game Testing
par: Shirzadeh-hajimahmood, Samira, et autres
Publié: (2024)
par: Shirzadeh-hajimahmood, Samira, et autres
Publié: (2024)
A Multi-Agent Approach for REST API Testing with Semantic Graphs and LLM-Driven Inputs
par: Kim, Myeongsoo, et autres
Publié: (2024)
par: Kim, Myeongsoo, et autres
Publié: (2024)
Are Large Language Models Robust in Understanding Code Against Semantics-Preserving Mutations?
par: Orvalho, Pedro, et autres
Publié: (2025)
par: Orvalho, Pedro, et autres
Publié: (2025)
RV4Chatbot: Are Chatbots Allowed to Dream of Electric Sheep?
par: Gatti, Andrea, et autres
Publié: (2024)
par: Gatti, Andrea, et autres
Publié: (2024)
TESTQUEST: A Web Gamification Tool to Improve Locators and Page Objects Quality
par: Olianas, Dario, et autres
Publié: (2025)
par: Olianas, Dario, et autres
Publié: (2025)
You Name It, I Run It: An LLM Agent to Execute Tests of Arbitrary Projects
par: Bouzenia, Islem, et autres
Publié: (2024)
par: Bouzenia, Islem, et autres
Publié: (2024)
The Present and Future of Bots in Software Engineering
par: Shihab, Emad, et autres
Publié: (2022)
par: Shihab, Emad, et autres
Publié: (2022)
An Approach for Auto Generation of Labeling Functions for Software Engineering Chatbots
par: Alor, Ebube, et autres
Publié: (2024)
par: Alor, Ebube, et autres
Publié: (2024)
Generating Minimalist Adversarial Perturbations to Test Object-Detection Models: An Adaptive Multi-Metric Evolutionary Search Approach
par: McIntyre-Garcia, Cristopher, et autres
Publié: (2024)
par: McIntyre-Garcia, Cristopher, et autres
Publié: (2024)
Tests as Prompt: A Test-Driven-Development Benchmark for LLM Code Generation
par: Cui, Yi
Publié: (2025)
par: Cui, Yi
Publié: (2025)
TREAT: A Code LLMs Trustworthiness / Reliability Evaluation and Testing Framework
par: Gao, Shuzheng, et autres
Publié: (2025)
par: Gao, Shuzheng, et autres
Publié: (2025)
AutoRestTest: A Tool for Automated REST API Testing Using LLMs and MARL
par: Stennett, Tyler, et autres
Publié: (2025)
par: Stennett, Tyler, et autres
Publié: (2025)
AI-Assisted Unit Test Writing and Test-Driven Code Refactoring: A Case Study
par: Smolic, Ema, et autres
Publié: (2026)
par: Smolic, Ema, et autres
Publié: (2026)
Fuzzy Inference System for Test Case Prioritization in Software Testing
par: Karatayev, Aron, et autres
Publié: (2024)
par: Karatayev, Aron, et autres
Publié: (2024)
AutoTest: Evolutionary Code Solution Selection with Test Cases
par: Duan, Zhihua, et autres
Publié: (2024)
par: Duan, Zhihua, et autres
Publié: (2024)
A System for Automated Unit Test Generation Using Large Language Models and Assessment of Generated Test Suites
par: Lops, Andrea, et autres
Publié: (2024)
par: Lops, Andrea, et autres
Publié: (2024)
Unit Testing in ASP Revisited: Language and Test-Driven Development Environment
par: Amendola, Giovanni, et autres
Publié: (2024)
par: Amendola, Giovanni, et autres
Publié: (2024)
QuanTest: Entanglement-Guided Testing of Quantum Neural Network Systems
par: Shi, Jinjing, et autres
Publié: (2024)
par: Shi, Jinjing, et autres
Publié: (2024)
The Future of Software Testing: AI-Powered Test Case Generation and Validation
par: Baqar, Mohammad, et autres
Publié: (2024)
par: Baqar, Mohammad, et autres
Publié: (2024)
Documents similaires
-
Towards Multi-Platform Mutation Testing of Task-based Chatbots
par: Clerissi, Diego, et autres
Publié: (2025) -
Automated Testing of Task-based Chatbots: How Far Are We?
par: Clerissi, Diego, et autres
Publié: (2026) -
Test Case Generation for Dialogflow Task-Based Chatbots
par: Rapisarda, Rocco Gianni, et autres
Publié: (2025) -
Assessing Task-based Chatbots: Snapshot and Curated Datasets for Dialogflow
par: Masserini, Elena, et autres
Publié: (2026) -
Towards the Assessment of Task-based Chatbots: From the TOFU-R Snapshot to the BRASATO Curated Dataset
par: Masserini, Elena, et autres
Publié: (2025)