Mind the Pause: Disfluency-Aware Objective Tuning for Multilingual Speech Correction with LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Kumar, Deepak, Gain, Baban, Ekbal, Asif |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
One Model to Translate Them All? A Journey to Mount Doom for Multilingual Model Merging
by: Gain, Baban, et al.
Published: (2026)
by: Gain, Baban, et al.
Published: (2026)
Beyond the Sentence: A Survey on Context-Aware Machine Translation with Large Language Models
by: Appicharla, Ramakrishna, et al.
Published: (2025)
by: Appicharla, Ramakrishna, et al.
Published: (2025)
Quality Estimation based Feedback Training for Improving Pronoun Translation
by: Dhankhar, Harshit, et al.
Published: (2025)
by: Dhankhar, Harshit, et al.
Published: (2025)
Bridging the Linguistic Divide: A Survey on Leveraging Large Language Models for Machine Translation
by: Gain, Baban, et al.
Published: (2025)
by: Gain, Baban, et al.
Published: (2025)
Impact of Visual Context on Noisy Multimodal NMT: An Empirical Study for English to Indian Languages
by: Gain, Baban, et al.
Published: (2023)
by: Gain, Baban, et al.
Published: (2023)
Indic-TunedLens: Interpreting Multilingual Models in Indian Languages
by: Panchal, Mihir, et al.
Published: (2026)
by: Panchal, Mihir, et al.
Published: (2026)
A Case Study on Context-Aware Neural Machine Translation with Multi-Task Learning
by: Appicharla, Ramakrishna, et al.
Published: (2024)
by: Appicharla, Ramakrishna, et al.
Published: (2024)
From Hallucinations to Facts: Enhancing Language Models with Curated Knowledge Graphs
by: Joshi, Ratnesh Kumar, et al.
Published: (2024)
by: Joshi, Ratnesh Kumar, et al.
Published: (2024)
Augmenting Automatic Speech Recognition Models with Disfluency Detection
by: Amann, Robin, et al.
Published: (2024)
by: Amann, Robin, et al.
Published: (2024)
Computational Politeness in Natural Language Processing: A Survey
by: Priya, Priyanshu, et al.
Published: (2024)
by: Priya, Priyanshu, et al.
Published: (2024)
Seeing Through VisualBERT: A Causal Adventure on Memetic Landscapes
by: Bandyopadhyay, Dibyanayan, et al.
Published: (2024)
by: Bandyopadhyay, Dibyanayan, et al.
Published: (2024)
CAuSE: Decoding Multimodal Classifiers using Faithful Natural Language Explanation
by: Bandyopadhyay, Dibyanayan, et al.
Published: (2025)
by: Bandyopadhyay, Dibyanayan, et al.
Published: (2025)
Real-Time Generation of Game Video Commentary with Multimodal LLMs: Pause-Aware Decoding Approaches
by: Afzal, Anum, et al.
Published: (2026)
by: Afzal, Anum, et al.
Published: (2026)
We Argue to Agree: Towards Personality-Driven Argumentation-Based Negotiation Dialogue Systems for Tourism
by: Priya, Priyanshu, et al.
Published: (2025)
by: Priya, Priyanshu, et al.
Published: (2025)
When Reviews Disagree: Fine-Grained Contradiction Analysis in Scientific Peer Reviews
by: Kumar, Sandeep, et al.
Published: (2026)
by: Kumar, Sandeep, et al.
Published: (2026)
Humane Speech Synthesis through Zero-Shot Emotion and Disfluency Generation
by: Chaudhury, Rohan, et al.
Published: (2024)
by: Chaudhury, Rohan, et al.
Published: (2024)
Can Large Language Models Unlock Novel Scientific Research Ideas?
by: Kumar, Sandeep, et al.
Published: (2024)
by: Kumar, Sandeep, et al.
Published: (2024)
Pause-Tuning for Long-Context Comprehension: A Lightweight Approach to LLM Attention Recalibration
by: Begin, James, et al.
Published: (2025)
by: Begin, James, et al.
Published: (2025)
'Quis custodiet ipsos custodes?' Who will watch the watchmen? On Detecting AI-generated peer-reviews
by: Kumar, Sandeep, et al.
Published: (2024)
by: Kumar, Sandeep, et al.
Published: (2024)
Looks can be Deceptive: Distinguishing Repetition Disfluency from Reduplication
by: Ahmad, Arif, et al.
Published: (2024)
by: Ahmad, Arif, et al.
Published: (2024)
Smooth Operators: LLMs Translating Imperfect Hints into Disfluency-Rich Transcripts
by: Altinok, Duygu
Published: (2025)
by: Altinok, Duygu
Published: (2025)
Strategic Prompting for Conversational Tasks: A Comparative Analysis of Large Language Models Across Diverse Conversational Tasks
by: Joshi, Ratnesh Kumar, et al.
Published: (2024)
by: Joshi, Ratnesh Kumar, et al.
Published: (2024)
Empowering Dysarthric Speech: Leveraging Advanced LLMs for Accurate Speech Correction and Multimodal Emotion Analysis
by: Attaluri, Kaushal, et al.
Published: (2024)
by: Attaluri, Kaushal, et al.
Published: (2024)
CorIL: Towards Enriching Indian Language to Indian Language Parallel Corpora and Machine Translation Systems
by: Bhattacharjee, Soham, et al.
Published: (2025)
by: Bhattacharjee, Soham, et al.
Published: (2025)
Data Quality Issues in Multilingual Speech Datasets: The Need for Sociolinguistic Awareness and Proactive Language Planning
by: Lau, Mingfei, et al.
Published: (2025)
by: Lau, Mingfei, et al.
Published: (2025)
Mind the Quote: Enabling Quotation-Aware Dialogue in LLMs via Plug-and-Play Modules
by: Zhang, Yueqi, et al.
Published: (2025)
by: Zhang, Yueqi, et al.
Published: (2025)
Analyzing and Fine-Tuning Whisper Models for Multilingual Pilot Speech Transcription in the Cockpit
by: Nareddy, Kartheek Kumar Reddy, et al.
Published: (2025)
by: Nareddy, Kartheek Kumar Reddy, et al.
Published: (2025)
The Riddle of Reflection: Evaluating Reasoning and Self-Awareness in Multilingual LLMs using Indian Riddles
by: M, Abhinav P, et al.
Published: (2025)
by: M, Abhinav P, et al.
Published: (2025)
Teaching LLMs How to Learn with Contextual Fine-Tuning
by: Choi, Younwoo, et al.
Published: (2025)
by: Choi, Younwoo, et al.
Published: (2025)
Think before you speak: Training Language Models With Pause Tokens
by: Goyal, Sachin, et al.
Published: (2023)
by: Goyal, Sachin, et al.
Published: (2023)
Sparse Semantic Dimension as a Generalization Certificate for LLMs
by: Bandyopadhyay, Dibyanayan, et al.
Published: (2026)
by: Bandyopadhyay, Dibyanayan, et al.
Published: (2026)
Enhancing Text Classification with a Novel Multi-Agent Collaboration Framework Leveraging BERT
by: Baban, Hediyeh, et al.
Published: (2025)
by: Baban, Hediyeh, et al.
Published: (2025)
Mind the Gap! Choice Independence in Using Multilingual LLMs for Persuasive Co-Writing Tasks in Different Languages
by: Biswas, Shreyan, et al.
Published: (2025)
by: Biswas, Shreyan, et al.
Published: (2025)
Building Accurate Translation-Tailored LLMs with Language Aware Instruction Tuning
by: Zan, Changtong, et al.
Published: (2024)
by: Zan, Changtong, et al.
Published: (2024)
Position: Pause Recycling LoRAs and Prioritize Mechanisms to Uncover Limits and Effectiveness
by: Chen, Mei-Yen, et al.
Published: (2025)
by: Chen, Mei-Yen, et al.
Published: (2025)
Multilingual LLMs Are Not Multilingual Thinkers: Evidence from Hindi Analogy Evaluation
by: Gupta, Ashray, et al.
Published: (2025)
by: Gupta, Ashray, et al.
Published: (2025)
L-TUNING: Synchronized Label Tuning for Prompt and Prefix in LLMs
by: Kowsher, Md., et al.
Published: (2023)
by: Kowsher, Md., et al.
Published: (2023)
Z-Scores: A Metric for Linguistically Assessing Disfluency Removal
by: Teleki, Maria, et al.
Published: (2025)
by: Teleki, Maria, et al.
Published: (2025)
Teaching According to Talents! Instruction Tuning LLMs with Competence-Aware Curriculum Learning
by: Li, Yangning, et al.
Published: (2025)
by: Li, Yangning, et al.
Published: (2025)
Multilingual Mathematical Reasoning: Advancing Open-Source LLMs in Hindi and English
by: Anand, Avinash, et al.
Published: (2024)
by: Anand, Avinash, et al.
Published: (2024)
Similar Items
-
One Model to Translate Them All? A Journey to Mount Doom for Multilingual Model Merging
by: Gain, Baban, et al.
Published: (2026) -
Beyond the Sentence: A Survey on Context-Aware Machine Translation with Large Language Models
by: Appicharla, Ramakrishna, et al.
Published: (2025) -
Quality Estimation based Feedback Training for Improving Pronoun Translation
by: Dhankhar, Harshit, et al.
Published: (2025) -
Bridging the Linguistic Divide: A Survey on Leveraging Large Language Models for Machine Translation
by: Gain, Baban, et al.
Published: (2025) -
Impact of Visual Context on Noisy Multimodal NMT: An Empirical Study for English to Indian Languages
by: Gain, Baban, et al.
Published: (2023)