Boosting Disfluency Detection with Large Language Model as Disfluency Generator
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cheng, Zhenrong, Guo, Jiayan, Sun, Hao, Zhang, Yan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Augmenting Automatic Speech Recognition Models with Disfluency Detection
von: Amann, Robin, et al.
Veröffentlicht: (2024)
von: Amann, Robin, et al.
Veröffentlicht: (2024)
Enhancing Naturalness in LLM-Generated Utterances through Disfluency Insertion
von: Hassan, Syed Zohaib, et al.
Veröffentlicht: (2024)
von: Hassan, Syed Zohaib, et al.
Veröffentlicht: (2024)
Missingness-resilient Video-enhanced Multimodal Disfluency Detection
von: Mohapatra, Payal, et al.
Veröffentlicht: (2024)
von: Mohapatra, Payal, et al.
Veröffentlicht: (2024)
Measuring the Effect of Disfluency in Multilingual Knowledge Probing Benchmarks
von: Semenov, Kirill, et al.
Veröffentlicht: (2025)
von: Semenov, Kirill, et al.
Veröffentlicht: (2025)
Humane Speech Synthesis through Zero-Shot Emotion and Disfluency Generation
von: Chaudhury, Rohan, et al.
Veröffentlicht: (2024)
von: Chaudhury, Rohan, et al.
Veröffentlicht: (2024)
Disfluencies We Live with in Japanese
Veröffentlicht: (2026)
Veröffentlicht: (2026)
Looks can be Deceptive: Distinguishing Repetition Disfluency from Reduplication
von: Ahmad, Arif, et al.
Veröffentlicht: (2024)
von: Ahmad, Arif, et al.
Veröffentlicht: (2024)
DRIVE: Disfluency-Rich Synthetic Dialog Data Generation Framework for Intelligent Vehicle Environments
von: Chavda, Anshul, et al.
Veröffentlicht: (2025)
von: Chavda, Anshul, et al.
Veröffentlicht: (2025)
DisfluencySpeech -- Single-Speaker Conversational Speech Dataset with Paralanguage
von: Wang, Kyra, et al.
Veröffentlicht: (2024)
von: Wang, Kyra, et al.
Veröffentlicht: (2024)
Z-Scores: A Metric for Linguistically Assessing Disfluency Removal
von: Teleki, Maria, et al.
Veröffentlicht: (2025)
von: Teleki, Maria, et al.
Veröffentlicht: (2025)
VocalBench-DF: A Benchmark for Evaluating Speech LLM Robustness to Disfluency
von: Liu, Hongcheng, et al.
Veröffentlicht: (2025)
von: Liu, Hongcheng, et al.
Veröffentlicht: (2025)
Automatic Speech Recognition for Non-Native English: Accuracy and Disfluency Handling
von: McGuire, Michael
Veröffentlicht: (2025)
von: McGuire, Michael
Veröffentlicht: (2025)
Mind the Pause: Disfluency-Aware Objective Tuning for Multilingual Speech Correction with LLMs
von: Kumar, Deepak, et al.
Veröffentlicht: (2026)
von: Kumar, Deepak, et al.
Veröffentlicht: (2026)
Smooth Operators: LLMs Translating Imperfect Hints into Disfluency-Rich Transcripts
von: Altinok, Duygu
Veröffentlicht: (2025)
von: Altinok, Duygu
Veröffentlicht: (2025)
Incorporating Co‐occurrence Into the Operationalization of Speech Disfluency for Second Language Pronunciation and Oral Proficiency Assessment
von: Xun Yan, et al.
Veröffentlicht: (2025)
von: Xun Yan, et al.
Veröffentlicht: (2025)
Distinguishing Repetition Disfluency from Morphological Reduplication in Bangla ASR Transcripts: A Novel Corpus and Benchmarking Analysis
von: Arpa, Zaara Zabeen, et al.
Veröffentlicht: (2025)
von: Arpa, Zaara Zabeen, et al.
Veröffentlicht: (2025)
Toward a Reinforcement-Learning-Based System for Adjusting Medication to Minimize Speech Disfluency
von: Constas, Pavlos, et al.
Veröffentlicht: (2023)
von: Constas, Pavlos, et al.
Veröffentlicht: (2023)
The Link Between Syntactic Complexity and Stuttering‐Like Disfluencies in French Speaking Adults
von: Alice Le Dévic, et al.
Veröffentlicht: (2026)
von: Alice Le Dévic, et al.
Veröffentlicht: (2026)
Adults with Autism Spectrum Disorder Demonstrate Increased Disfluency in Spontaneous Speech but Not in Reading
von: Marie Van Gaever, et al.
Veröffentlicht: (2025)
von: Marie Van Gaever, et al.
Veröffentlicht: (2025)
Full-Duplex-Bench-v3: Benchmarking Tool Use for Full-Duplex Voice Agents Under Real-World Disfluency
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2026)
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2026)
Partner‐Specific Adaptation in Disfluency Processing
von: Si On Yoon, et al.
Veröffentlicht: (2024)
von: Si On Yoon, et al.
Veröffentlicht: (2024)
Disfluencies and speech rate in spontaneous production and in oral reading in people who stutter and who do not stutter
von: Joana Cecilia Baptista Ramalho Pinto
Veröffentlicht: (2013)
von: Joana Cecilia Baptista Ramalho Pinto
Veröffentlicht: (2013)
A Novel Paradigm Boosting Translation Capabilities of Large Language Models
von: Guo, Jiaxin, et al.
Veröffentlicht: (2024)
von: Guo, Jiaxin, et al.
Veröffentlicht: (2024)
Assessing Others’ Knowledge Through Their Speech Disfluencies and Gestures
von: Can Avcı, et al.
Veröffentlicht: (2025)
von: Can Avcı, et al.
Veröffentlicht: (2025)
Iterative Structured Pruning for Large Language Models with Multi-Domain Calibration
von: Wu, Guangxin, et al.
Veröffentlicht: (2026)
von: Wu, Guangxin, et al.
Veröffentlicht: (2026)
CtrlDiff: Boosting Large Diffusion Language Models with Dynamic Block Prediction and Controllable Generation
von: Huang, Chihan, et al.
Veröffentlicht: (2025)
von: Huang, Chihan, et al.
Veröffentlicht: (2025)
Evaluating the Generation Capabilities of Large Chinese Language Models
von: Zeng, Hui, et al.
Veröffentlicht: (2023)
von: Zeng, Hui, et al.
Veröffentlicht: (2023)
FakeGPT: Fake News Generation, Explanation and Detection of Large Language Models
von: Huang, Yue, et al.
Veröffentlicht: (2023)
von: Huang, Yue, et al.
Veröffentlicht: (2023)
MI-PRUN: Optimize Large Language Model Pruning via Mutual Information
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
Boosting Large Language Models for Mental Manipulation Detection via Data Augmentation and Distillation
von: Gao, Yuansheng, et al.
Veröffentlicht: (2025)
von: Gao, Yuansheng, et al.
Veröffentlicht: (2025)
Correction to “Assessing Others’ Knowledge Through Their Speech Disfluencies and Gestures”
Veröffentlicht: (2026)
Veröffentlicht: (2026)
Unsupervised Distractor Generation via Large Language Model Distilling and Counterfactual Contrastive Decoding
von: Qu, Fanyi, et al.
Veröffentlicht: (2024)
von: Qu, Fanyi, et al.
Veröffentlicht: (2024)
Enhancing Large Language Models (LLMs) for Telecommunications using Knowledge Graphs and Retrieval-Augmented Generation
von: Yuan, Dun, et al.
Veröffentlicht: (2025)
von: Yuan, Dun, et al.
Veröffentlicht: (2025)
Unveil: Unified Visual-Textual Integration and Distillation for Multi-modal Document Retrieval
von: Sun, Hao, et al.
Veröffentlicht: (2026)
von: Sun, Hao, et al.
Veröffentlicht: (2026)
Re-Initialization Token Learning for Tool-Augmented Large Language Models
von: Li, Chenghao, et al.
Veröffentlicht: (2025)
von: Li, Chenghao, et al.
Veröffentlicht: (2025)
Say What You Mean: Natural Language Access Control with Large Language Models for Internet of Things
von: Cheng, Ye, et al.
Veröffentlicht: (2025)
von: Cheng, Ye, et al.
Veröffentlicht: (2025)
StableToolBench: Towards Stable Large-Scale Benchmarking on Tool Learning of Large Language Models
von: Guo, Zhicheng, et al.
Veröffentlicht: (2024)
von: Guo, Zhicheng, et al.
Veröffentlicht: (2024)
VaccineRAG: Boosting Multimodal Large Language Models' Immunity to Harmful RAG Samples
von: Sun, Qixin, et al.
Veröffentlicht: (2025)
von: Sun, Qixin, et al.
Veröffentlicht: (2025)
Pretraining Data Detection for Large Language Models: A Divergence-based Calibration Method
von: Zhang, Weichao, et al.
Veröffentlicht: (2024)
von: Zhang, Weichao, et al.
Veröffentlicht: (2024)
Generative Linguistics, Large Language Models, and the Social Nature of Scientific Success
von: Hao, Sophie
Veröffentlicht: (2025)
von: Hao, Sophie
Veröffentlicht: (2025)
Ähnliche Einträge
-
Augmenting Automatic Speech Recognition Models with Disfluency Detection
von: Amann, Robin, et al.
Veröffentlicht: (2024) -
Enhancing Naturalness in LLM-Generated Utterances through Disfluency Insertion
von: Hassan, Syed Zohaib, et al.
Veröffentlicht: (2024) -
Missingness-resilient Video-enhanced Multimodal Disfluency Detection
von: Mohapatra, Payal, et al.
Veröffentlicht: (2024) -
Measuring the Effect of Disfluency in Multilingual Knowledge Probing Benchmarks
von: Semenov, Kirill, et al.
Veröffentlicht: (2025) -
Humane Speech Synthesis through Zero-Shot Emotion and Disfluency Generation
von: Chaudhury, Rohan, et al.
Veröffentlicht: (2024)