Assessing Task-based Chatbots: Snapshot and Curated Datasets for Dialogflow
Fuente:
arXiv
Salvato in:
| Autori principali: | Masserini, Elena, Clerissi, Diego, Micucci, Daniela, Mariani, Leonardo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards the Assessment of Task-based Chatbots: From the TOFU-R Snapshot to the BRASATO Curated Dataset
di: Masserini, Elena, et al.
Pubblicazione: (2025)
di: Masserini, Elena, et al.
Pubblicazione: (2025)
Automated Testing of Task-based Chatbots: How Far Are We?
di: Clerissi, Diego, et al.
Pubblicazione: (2026)
di: Clerissi, Diego, et al.
Pubblicazione: (2026)
Towards Multi-Platform Mutation Testing of Task-based Chatbots
di: Clerissi, Diego, et al.
Pubblicazione: (2025)
di: Clerissi, Diego, et al.
Pubblicazione: (2025)
Test Case Generation for Dialogflow Task-Based Chatbots
di: Rapisarda, Rocco Gianni, et al.
Pubblicazione: (2025)
di: Rapisarda, Rocco Gianni, et al.
Pubblicazione: (2025)
Bug Whispering: Towards Audio Bug Reporting
di: Masserini, Elena, et al.
Pubblicazione: (2025)
di: Masserini, Elena, et al.
Pubblicazione: (2025)
Anonymizing Test Data in Android: Does It Hurt?
di: Masserini, Elena, et al.
Pubblicazione: (2024)
di: Masserini, Elena, et al.
Pubblicazione: (2024)
MutaBot: A Mutation Testing Approach for Chatbots
di: Urrico, Michael Ferdinando, et al.
Pubblicazione: (2024)
di: Urrico, Michael Ferdinando, et al.
Pubblicazione: (2024)
Assessing AI-Based Code Assistants in Method Generation Tasks
di: Corso, Vincenzo, et al.
Pubblicazione: (2024)
di: Corso, Vincenzo, et al.
Pubblicazione: (2024)
MultiMind: A Plug-in for the Implementation of Development Tasks Aided by AI Assistants
di: Donato, Benedetta, et al.
Pubblicazione: (2025)
di: Donato, Benedetta, et al.
Pubblicazione: (2025)
Generating Java Methods: An Empirical Assessment of Four AI-Based Code Assistants
di: Corso, Vincenzo, et al.
Pubblicazione: (2024)
di: Corso, Vincenzo, et al.
Pubblicazione: (2024)
On the Possibility of Breaking Copyleft Licenses When Reusing Code Generated by ChatGPT
di: Colombo, Gaia, et al.
Pubblicazione: (2025)
di: Colombo, Gaia, et al.
Pubblicazione: (2025)
Studying How Configurations Impact Code Generation in LLMs: the Case of ChatGPT
di: Donato, Benedetta, et al.
Pubblicazione: (2025)
di: Donato, Benedetta, et al.
Pubblicazione: (2025)
Multi-Level Testing of Conversational AI Systems
di: Masserini, Elena
Pubblicazione: (2026)
di: Masserini, Elena
Pubblicazione: (2026)
Analyzing Prompt Influence on Automated Method Generation: An Empirical Study with Copilot
di: Fagadau, Ionut Daniel, et al.
Pubblicazione: (2024)
di: Fagadau, Ionut Daniel, et al.
Pubblicazione: (2024)
TESTQUEST: A Web Gamification Tool to Improve Locators and Page Objects Quality
di: Olianas, Dario, et al.
Pubblicazione: (2025)
di: Olianas, Dario, et al.
Pubblicazione: (2025)
A Transformer-based Approach for Augmenting Software Engineering Chatbots Datasets
di: Abdellatif, Ahmad, et al.
Pubblicazione: (2024)
di: Abdellatif, Ahmad, et al.
Pubblicazione: (2024)
Testing in the Evolving World of DL Systems:Insights from Python GitHub Projects
di: Ali, Qurban, et al.
Pubblicazione: (2024)
di: Ali, Qurban, et al.
Pubblicazione: (2024)
Towards Model-Driven Dashboard Generation for Systems-of-Systems
di: Rossi, Maria Teresa, et al.
Pubblicazione: (2024)
di: Rossi, Maria Teresa, et al.
Pubblicazione: (2024)
From PREVENTion to REACTion: Enhancing Failure Resolution in Naval Systems
di: Rossi, Maria Teresa, et al.
Pubblicazione: (2025)
di: Rossi, Maria Teresa, et al.
Pubblicazione: (2025)
BugsRepo: A Comprehensive Curated Dataset of Bug Reports, Comments and Contributors Information from Bugzilla
di: Acharya, Jagrit, et al.
Pubblicazione: (2025)
di: Acharya, Jagrit, et al.
Pubblicazione: (2025)
Coverage-Guided Road Selection and Prioritization for Efficient Testing in Autonomous Driving Systems
di: Ali, Qurban, et al.
Pubblicazione: (2026)
di: Ali, Qurban, et al.
Pubblicazione: (2026)
OpenCat: Improving Interoperability of ADS Testing
di: Ali, Qurban, et al.
Pubblicazione: (2025)
di: Ali, Qurban, et al.
Pubblicazione: (2025)
An Evolutionary Approach to Adapt Tests Across Mobile Apps
di: Mariani, Leonardo, et al.
Pubblicazione: (2021)
di: Mariani, Leonardo, et al.
Pubblicazione: (2021)
Client--Library Compatibility Testing with API Interaction Snapshots
di: Monce, Gustave, et al.
Pubblicazione: (2025)
di: Monce, Gustave, et al.
Pubblicazione: (2025)
A Model-based Chatbot Generation Approach to Converse with Open Data Sources
di: Ed-douibi, Hamza, et al.
Pubblicazione: (2020)
di: Ed-douibi, Hamza, et al.
Pubblicazione: (2020)
PerfCurator: Curating a large-scale dataset of performance bug-related commits from public repositories
di: Azad, Md Abul Kalam, et al.
Pubblicazione: (2024)
di: Azad, Md Abul Kalam, et al.
Pubblicazione: (2024)
ReProbe: An Architecture for Reconfigurable and Adaptive Probes
di: Alessi, Federico, et al.
Pubblicazione: (2024)
di: Alessi, Federico, et al.
Pubblicazione: (2024)
CAM: A Collection of Snapshots of GitHub Java Repositories Together with Metrics
di: Bugayenko, Yegor
Pubblicazione: (2024)
di: Bugayenko, Yegor
Pubblicazione: (2024)
CodeInsight: A Curated Dataset of Practical Coding Solutions from Stack Overflow
di: Beau, Nathanaël, et al.
Pubblicazione: (2024)
di: Beau, Nathanaël, et al.
Pubblicazione: (2024)
Harnessing Large Language Models for Curated Code Reviews
di: Sghaier, Oussama Ben, et al.
Pubblicazione: (2025)
di: Sghaier, Oussama Ben, et al.
Pubblicazione: (2025)
Reconsidering Conversational Norms in LLM Chatbots for Sustainable AI
di: Santos, Ronnie de Souza, et al.
Pubblicazione: (2025)
di: Santos, Ronnie de Souza, et al.
Pubblicazione: (2025)
Assessing and Advancing Benchmarks for Evaluating Large Language Models in Software Engineering Tasks
di: Hu, Xing, et al.
Pubblicazione: (2025)
di: Hu, Xing, et al.
Pubblicazione: (2025)
"Where is My Troubleshooting Procedure?": Studying the Potential of RAG in Assisting Failure Resolution of Large Cyber-Physical System
di: Rossi, Maria Teresa, et al.
Pubblicazione: (2026)
di: Rossi, Maria Teresa, et al.
Pubblicazione: (2026)
Unveiling Assumptions: Exploring the Decisions of AI Chatbots and Human Testers
di: Neto, Francisco Gomes de Oliveira
Pubblicazione: (2024)
di: Neto, Francisco Gomes de Oliveira
Pubblicazione: (2024)
Using the SOCIO Chatbot for UML Modelling: A Family of Experiments
di: Ren, Ranci, et al.
Pubblicazione: (2024)
di: Ren, Ranci, et al.
Pubblicazione: (2024)
Improving Data Curation of Software Vulnerability Patches through Uncertainty Quantification
di: Chen, Hui, et al.
Pubblicazione: (2024)
di: Chen, Hui, et al.
Pubblicazione: (2024)
A Manually-Curated Dataset of Fixes to Vulnerabilities of Open-Source Software
di: Ponta, Serena E., et al.
Pubblicazione: (2019)
di: Ponta, Serena E., et al.
Pubblicazione: (2019)
CodeMMLU: A Multi-Task Benchmark for Assessing Code Understanding & Reasoning Capabilities of CodeLLMs
di: Manh, Dung Nguyen, et al.
Pubblicazione: (2024)
di: Manh, Dung Nguyen, et al.
Pubblicazione: (2024)
TaskEval: Assessing Difficulty of Code Generation Tasks for Large Language Models
di: Tambon, Florian, et al.
Pubblicazione: (2024)
di: Tambon, Florian, et al.
Pubblicazione: (2024)
Code2Doc: A Quality-First Curated Dataset for Code Documentation
di: Karaman, Recep Kaan, et al.
Pubblicazione: (2025)
di: Karaman, Recep Kaan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Towards the Assessment of Task-based Chatbots: From the TOFU-R Snapshot to the BRASATO Curated Dataset
di: Masserini, Elena, et al.
Pubblicazione: (2025) -
Automated Testing of Task-based Chatbots: How Far Are We?
di: Clerissi, Diego, et al.
Pubblicazione: (2026) -
Towards Multi-Platform Mutation Testing of Task-based Chatbots
di: Clerissi, Diego, et al.
Pubblicazione: (2025) -
Test Case Generation for Dialogflow Task-Based Chatbots
di: Rapisarda, Rocco Gianni, et al.
Pubblicazione: (2025) -
Bug Whispering: Towards Audio Bug Reporting
di: Masserini, Elena, et al.
Pubblicazione: (2025)