A Typology of Synthetic Datasets for Dialogue Processing in Clinical Contexts
Fuente:
arXiv
Saved in:
| Main Authors: | Bedrick, Steven, Doğruöz, A. Seza, Nisioi, Sergiu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SG-UniBuc-NLP at SemEval-2026 Task 6: Multi-Head RoBERTa with Chunking for Long-Context Evasion Detection
by: Stefan, Gabriel, et al.
Published: (2026)
by: Stefan, Gabriel, et al.
Published: (2026)
Readability Measures and Automatic Text Simplification: In the Search of a Construct
by: Cardon, Rémi, et al.
Published: (2025)
by: Cardon, Rémi, et al.
Published: (2025)
Cheap Ways of Extracting Clinical Markers from Texts
by: Sandu, Anastasia, et al.
Published: (2024)
by: Sandu, Anastasia, et al.
Published: (2024)
Building Multilingual Datasets for Predicting Mental Health Severity through LLMs: Prospects and Challenges
by: Skianis, Konstantinos, et al.
Published: (2024)
by: Skianis, Konstantinos, et al.
Published: (2024)
Synthetic Dialogue Dataset Generation using LLM Agents
by: Abdullin, Yelaman, et al.
Published: (2024)
by: Abdullin, Yelaman, et al.
Published: (2024)
Clinical Document Metadata Extraction: A Scoping Review
by: Miller, Kurt, et al.
Published: (2025)
by: Miller, Kurt, et al.
Published: (2025)
Single- vs. Dual-Prompt Dialogue Generation with LLMs for Job Interviews in Human Resources
by: De Baer, Joachim, et al.
Published: (2025)
by: De Baer, Joachim, et al.
Published: (2025)
Automatic Correction of Writing Anomalies in Hausa Texts
by: Wali, Ahmad Mustapha, et al.
Published: (2025)
by: Wali, Ahmad Mustapha, et al.
Published: (2025)
Fine-Tuning Pre-Trained Code Models for AI-Generated Code Detection
by: Ispas, Jany-Gabriel, et al.
Published: (2026)
by: Ispas, Jany-Gabriel, et al.
Published: (2026)
Syn-TurnTurk: A Synthetic Dataset for Turn-Taking Prediction in Turkish Dialogues
by: Bayrak, Ahmet Tuğrul, et al.
Published: (2026)
by: Bayrak, Ahmet Tuğrul, et al.
Published: (2026)
Leveraging LLMs for Translating and Classifying Mental Health Data
by: Skianis, Konstantinos, et al.
Published: (2024)
by: Skianis, Konstantinos, et al.
Published: (2024)
Making Task-Oriented Dialogue Datasets More Natural by Synthetically Generating Indirect User Requests
by: Mannekote, Amogh, et al.
Published: (2024)
by: Mannekote, Amogh, et al.
Published: (2024)
From Medical Records to Diagnostic Dialogues: A Clinical-Grounded Approach and Dataset for Psychiatric Comorbidity
by: Wan, Tianxi, et al.
Published: (2025)
by: Wan, Tianxi, et al.
Published: (2025)
Dialectal and Low-Resource Machine Translation for Aromanian
by: Jerpelea, Alexandru-Iulius, et al.
Published: (2024)
by: Jerpelea, Alexandru-Iulius, et al.
Published: (2024)
MCPDial: A Minecraft Persona-driven Dialogue Dataset
by: Alavi, Seyed Hossein, et al.
Published: (2024)
by: Alavi, Seyed Hossein, et al.
Published: (2024)
Synthetic Patient-Physician Dialogue Generation from Clinical Notes Using LLM
by: Das, Trisha, et al.
Published: (2024)
by: Das, Trisha, et al.
Published: (2024)
MEDSAGE: Enhancing Robustness of Medical Dialogue Summarization to ASR Errors with LLM-generated Synthetic Dialogues
by: Binici, Kuluhan, et al.
Published: (2024)
by: Binici, Kuluhan, et al.
Published: (2024)
StreamingDialogue: Prolonged Dialogue Learning via Long Context Compression with Minimal Losses
by: Li, Jia-Nan, et al.
Published: (2024)
by: Li, Jia-Nan, et al.
Published: (2024)
Synthetic Dialogue Generation for Interactive Conversational Elicitation & Recommendation (ICER)
by: Ryu, Moonkyung, et al.
Published: (2025)
by: Ryu, Moonkyung, et al.
Published: (2025)
MedSynth: Realistic, Synthetic Medical Dialogue-Note Pairs
by: Mianroodi, Ahmad Rezaie, et al.
Published: (2025)
by: Mianroodi, Ahmad Rezaie, et al.
Published: (2025)
KMI: A Dataset of Korean Motivational Interviewing Dialogues for Psychotherapy
by: Kim, Hyunjong, et al.
Published: (2025)
by: Kim, Hyunjong, et al.
Published: (2025)
Measuring Diversity in Synthetic Datasets
by: Zhu, Yuchang, et al.
Published: (2025)
by: Zhu, Yuchang, et al.
Published: (2025)
ACR: Adaptive Context Refactoring via Context Refactoring Operators for Multi-Turn Dialogue
by: Shen, Jiawei, et al.
Published: (2026)
by: Shen, Jiawei, et al.
Published: (2026)
Domain-Adapted Retrieval for In-Context Annotation of Pedagogical Dialogue Acts
by: Lee, Jinsook, et al.
Published: (2026)
by: Lee, Jinsook, et al.
Published: (2026)
DYCP: Dynamic Context Pruning for Long-Form Dialogue with LLMs
by: Choi, Nayoung, et al.
Published: (2026)
by: Choi, Nayoung, et al.
Published: (2026)
Context-Agent: Dynamic Discourse Trees for Non-Linear Dialogue
by: Hu, Junan, et al.
Published: (2026)
by: Hu, Junan, et al.
Published: (2026)
The Convergence of Schema-Guided Dialogue Systems and the Model Context Protocol
by: Schlapbach, Andreas
Published: (2026)
by: Schlapbach, Andreas
Published: (2026)
Why Synthetic Isn't Real Yet: A Diagnostic Framework for Contact Center Dialogue Generation
by: Devanathan, Rishikesh, et al.
Published: (2025)
by: Devanathan, Rishikesh, et al.
Published: (2025)
Synthetic4Health: Generating Annotated Synthetic Clinical Letters
by: Ren, Libo, et al.
Published: (2024)
by: Ren, Libo, et al.
Published: (2024)
Recent Trends in Personalized Dialogue Generation: A Review of Datasets, Methodologies, and Evaluations
by: Chen, Yi-Pei, et al.
Published: (2024)
by: Chen, Yi-Pei, et al.
Published: (2024)
Red Teaming Language Models for Processing Contradictory Dialogues
by: Wen, Xiaofei, et al.
Published: (2024)
by: Wen, Xiaofei, et al.
Published: (2024)
URIEL+: Enhancing Linguistic Inclusion and Usability in a Typological and Multilingual Knowledge Base
by: Khan, Aditya, et al.
Published: (2024)
by: Khan, Aditya, et al.
Published: (2024)
Beyond Continuity: Challenges of Context Switching in Multi-Turn Dialogue with LLMs
by: Sinha, Aditya, et al.
Published: (2026)
by: Sinha, Aditya, et al.
Published: (2026)
Are they lovers or friends? Evaluating LLMs' Social Reasoning in English and Korean Dialogues
by: Kim, Eunsu, et al.
Published: (2025)
by: Kim, Eunsu, et al.
Published: (2025)
Generation of Synthetic Clinical Text: A Systematic Review
by: Alshaikhdeeb, Basel, et al.
Published: (2025)
by: Alshaikhdeeb, Basel, et al.
Published: (2025)
The Fragility of Chain-of-Thought Monitoring Across Typologically Diverse Languages
by: Onyame, Eric, et al.
Published: (2026)
by: Onyame, Eric, et al.
Published: (2026)
Leveraging Graph Structures and Large Language Models for End-to-End Synthetic Task-Oriented Dialogues
by: Medjad, Maya, et al.
Published: (2025)
by: Medjad, Maya, et al.
Published: (2025)
Speak or Stay Silent: Context-Aware Turn-Taking in Multi-Party Dialogue
by: Bhagtani, Kratika, et al.
Published: (2026)
by: Bhagtani, Kratika, et al.
Published: (2026)
Does Generative AI speak Nigerian-Pidgin?: Issues about Representativeness and Bias for Multilingualism in LLMs
by: Adelani, David Ifeoluwa, et al.
Published: (2024)
by: Adelani, David Ifeoluwa, et al.
Published: (2024)
MedKP: Medical Dialogue with Knowledge Enhancement and Clinical Pathway Encoding
by: Wu, Jiageng, et al.
Published: (2024)
by: Wu, Jiageng, et al.
Published: (2024)
Similar Items
-
SG-UniBuc-NLP at SemEval-2026 Task 6: Multi-Head RoBERTa with Chunking for Long-Context Evasion Detection
by: Stefan, Gabriel, et al.
Published: (2026) -
Readability Measures and Automatic Text Simplification: In the Search of a Construct
by: Cardon, Rémi, et al.
Published: (2025) -
Cheap Ways of Extracting Clinical Markers from Texts
by: Sandu, Anastasia, et al.
Published: (2024) -
Building Multilingual Datasets for Predicting Mental Health Severity through LLMs: Prospects and Challenges
by: Skianis, Konstantinos, et al.
Published: (2024) -
Synthetic Dialogue Dataset Generation using LLM Agents
by: Abdullin, Yelaman, et al.
Published: (2024)