Towards the Assessment of Task-based Chatbots: From the TOFU-R Snapshot to the BRASATO Curated Dataset
Fuente:
arXiv
Saved in:
| Main Authors: | Masserini, Elena, Clerissi, Diego, Micucci, Daniela, Campos, João R., Mariani, Leonardo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Assessing Task-based Chatbots: Snapshot and Curated Datasets for Dialogflow
by: Masserini, Elena, et al.
Published: (2026)
by: Masserini, Elena, et al.
Published: (2026)
Towards Multi-Platform Mutation Testing of Task-based Chatbots
by: Clerissi, Diego, et al.
Published: (2025)
by: Clerissi, Diego, et al.
Published: (2025)
Automated Testing of Task-based Chatbots: How Far Are We?
by: Clerissi, Diego, et al.
Published: (2026)
by: Clerissi, Diego, et al.
Published: (2026)
Bug Whispering: Towards Audio Bug Reporting
by: Masserini, Elena, et al.
Published: (2025)
by: Masserini, Elena, et al.
Published: (2025)
Anonymizing Test Data in Android: Does It Hurt?
by: Masserini, Elena, et al.
Published: (2024)
by: Masserini, Elena, et al.
Published: (2024)
Test Case Generation for Dialogflow Task-Based Chatbots
by: Rapisarda, Rocco Gianni, et al.
Published: (2025)
by: Rapisarda, Rocco Gianni, et al.
Published: (2025)
MutaBot: A Mutation Testing Approach for Chatbots
by: Urrico, Michael Ferdinando, et al.
Published: (2024)
by: Urrico, Michael Ferdinando, et al.
Published: (2024)
Assessing AI-Based Code Assistants in Method Generation Tasks
by: Corso, Vincenzo, et al.
Published: (2024)
by: Corso, Vincenzo, et al.
Published: (2024)
Generating Java Methods: An Empirical Assessment of Four AI-Based Code Assistants
by: Corso, Vincenzo, et al.
Published: (2024)
by: Corso, Vincenzo, et al.
Published: (2024)
MultiMind: A Plug-in for the Implementation of Development Tasks Aided by AI Assistants
by: Donato, Benedetta, et al.
Published: (2025)
by: Donato, Benedetta, et al.
Published: (2025)
On the Possibility of Breaking Copyleft Licenses When Reusing Code Generated by ChatGPT
by: Colombo, Gaia, et al.
Published: (2025)
by: Colombo, Gaia, et al.
Published: (2025)
Studying How Configurations Impact Code Generation in LLMs: the Case of ChatGPT
by: Donato, Benedetta, et al.
Published: (2025)
by: Donato, Benedetta, et al.
Published: (2025)
Multi-Level Testing of Conversational AI Systems
by: Masserini, Elena
Published: (2026)
by: Masserini, Elena
Published: (2026)
Analyzing Prompt Influence on Automated Method Generation: An Empirical Study with Copilot
by: Fagadau, Ionut Daniel, et al.
Published: (2024)
by: Fagadau, Ionut Daniel, et al.
Published: (2024)
TESTQUEST: A Web Gamification Tool to Improve Locators and Page Objects Quality
by: Olianas, Dario, et al.
Published: (2025)
by: Olianas, Dario, et al.
Published: (2025)
A Transformer-based Approach for Augmenting Software Engineering Chatbots Datasets
by: Abdellatif, Ahmad, et al.
Published: (2024)
by: Abdellatif, Ahmad, et al.
Published: (2024)
Towards Model-Driven Dashboard Generation for Systems-of-Systems
by: Rossi, Maria Teresa, et al.
Published: (2024)
by: Rossi, Maria Teresa, et al.
Published: (2024)
From PREVENTion to REACTion: Enhancing Failure Resolution in Naval Systems
by: Rossi, Maria Teresa, et al.
Published: (2025)
by: Rossi, Maria Teresa, et al.
Published: (2025)
Testing in the Evolving World of DL Systems:Insights from Python GitHub Projects
by: Ali, Qurban, et al.
Published: (2024)
by: Ali, Qurban, et al.
Published: (2024)
Chatbot-Based Assessment of Code Understanding in Automated Programming Assessment Systems
by: Frankford, Eduard, et al.
Published: (2026)
by: Frankford, Eduard, et al.
Published: (2026)
BugsRepo: A Comprehensive Curated Dataset of Bug Reports, Comments and Contributors Information from Bugzilla
by: Acharya, Jagrit, et al.
Published: (2025)
by: Acharya, Jagrit, et al.
Published: (2025)
OpenCat: Improving Interoperability of ADS Testing
by: Ali, Qurban, et al.
Published: (2025)
by: Ali, Qurban, et al.
Published: (2025)
An Evolutionary Approach to Adapt Tests Across Mobile Apps
by: Mariani, Leonardo, et al.
Published: (2021)
by: Mariani, Leonardo, et al.
Published: (2021)
Coverage-Guided Road Selection and Prioritization for Efficient Testing in Autonomous Driving Systems
by: Ali, Qurban, et al.
Published: (2026)
by: Ali, Qurban, et al.
Published: (2026)
Client--Library Compatibility Testing with API Interaction Snapshots
by: Monce, Gustave, et al.
Published: (2025)
by: Monce, Gustave, et al.
Published: (2025)
A Model-based Chatbot Generation Approach to Converse with Open Data Sources
by: Ed-douibi, Hamza, et al.
Published: (2020)
by: Ed-douibi, Hamza, et al.
Published: (2020)
PerfCurator: Curating a large-scale dataset of performance bug-related commits from public repositories
by: Azad, Md Abul Kalam, et al.
Published: (2024)
by: Azad, Md Abul Kalam, et al.
Published: (2024)
ReProbe: An Architecture for Reconfigurable and Adaptive Probes
by: Alessi, Federico, et al.
Published: (2024)
by: Alessi, Federico, et al.
Published: (2024)
CAM: A Collection of Snapshots of GitHub Java Repositories Together with Metrics
by: Bugayenko, Yegor
Published: (2024)
by: Bugayenko, Yegor
Published: (2024)
From Commits to Confidence: Towards Stability-Informed Risk Assessment in Open Source Software
by: Adejumo, Elijah Kayode, et al.
Published: (2025)
by: Adejumo, Elijah Kayode, et al.
Published: (2025)
CodeInsight: A Curated Dataset of Practical Coding Solutions from Stack Overflow
by: Beau, Nathanaël, et al.
Published: (2024)
by: Beau, Nathanaël, et al.
Published: (2024)
Harnessing Large Language Models for Curated Code Reviews
by: Sghaier, Oussama Ben, et al.
Published: (2025)
by: Sghaier, Oussama Ben, et al.
Published: (2025)
Reconsidering Conversational Norms in LLM Chatbots for Sustainable AI
by: Santos, Ronnie de Souza, et al.
Published: (2025)
by: Santos, Ronnie de Souza, et al.
Published: (2025)
"Where is My Troubleshooting Procedure?": Studying the Potential of RAG in Assisting Failure Resolution of Large Cyber-Physical System
by: Rossi, Maria Teresa, et al.
Published: (2026)
by: Rossi, Maria Teresa, et al.
Published: (2026)
Unveiling Assumptions: Exploring the Decisions of AI Chatbots and Human Testers
by: Neto, Francisco Gomes de Oliveira
Published: (2024)
by: Neto, Francisco Gomes de Oliveira
Published: (2024)
Using the SOCIO Chatbot for UML Modelling: A Family of Experiments
by: Ren, Ranci, et al.
Published: (2024)
by: Ren, Ranci, et al.
Published: (2024)
Improving Data Curation of Software Vulnerability Patches through Uncertainty Quantification
by: Chen, Hui, et al.
Published: (2024)
by: Chen, Hui, et al.
Published: (2024)
A Manually-Curated Dataset of Fixes to Vulnerabilities of Open-Source Software
by: Ponta, Serena E., et al.
Published: (2019)
by: Ponta, Serena E., et al.
Published: (2019)
SCOPE: A Dataset of Stereotyped Prompts for Counterfactual Fairness Assessment of LLMs
by: Parziale, Alessandra, et al.
Published: (2026)
by: Parziale, Alessandra, et al.
Published: (2026)
United We Stand: Towards End-to-End Log-based Fault Diagnosis via Interactive Multi-Task Learning
by: He, Minghua, et al.
Published: (2025)
by: He, Minghua, et al.
Published: (2025)
Similar Items
-
Assessing Task-based Chatbots: Snapshot and Curated Datasets for Dialogflow
by: Masserini, Elena, et al.
Published: (2026) -
Towards Multi-Platform Mutation Testing of Task-based Chatbots
by: Clerissi, Diego, et al.
Published: (2025) -
Automated Testing of Task-based Chatbots: How Far Are We?
by: Clerissi, Diego, et al.
Published: (2026) -
Bug Whispering: Towards Audio Bug Reporting
by: Masserini, Elena, et al.
Published: (2025) -
Anonymizing Test Data in Android: Does It Hurt?
by: Masserini, Elena, et al.
Published: (2024)