Towards End-to-End Spoken Grammatical Error Correction
Fuente:
arXiv
Salvato in:
| Autori principali: | Bannò, Stefano, Ma, Rao, Qian, Mengjie, Knill, Kate M., Gales, Mark J. F. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
End-to-End Spoken Grammatical Error Correction
di: Qian, Mengjie, et al.
Pubblicazione: (2025)
di: Qian, Mengjie, et al.
Pubblicazione: (2025)
Scaling and Prompting for Improved End-to-End Spoken Grammatical Error Correction
di: Qian, Mengjie, et al.
Pubblicazione: (2025)
di: Qian, Mengjie, et al.
Pubblicazione: (2025)
Data Augmentation for Spoken Grammatical Error Correction
di: Karanasou, Penny, et al.
Pubblicazione: (2025)
di: Karanasou, Penny, et al.
Pubblicazione: (2025)
ASR Error Correction using Large Language Models
di: Ma, Rao, et al.
Pubblicazione: (2024)
di: Ma, Rao, et al.
Pubblicazione: (2024)
Natural Language-based Assessment of L2 Oral Proficiency using LLMs
di: Bannò, Stefano, et al.
Pubblicazione: (2025)
di: Bannò, Stefano, et al.
Pubblicazione: (2025)
Learn and Don't Forget: Adding a New Language to ASR Foundation Models
di: Qian, Mengjie, et al.
Pubblicazione: (2024)
di: Qian, Mengjie, et al.
Pubblicazione: (2024)
Assessment of L2 Oral Proficiency using Speech Large Language Models
di: Ma, Rao, et al.
Pubblicazione: (2025)
di: Ma, Rao, et al.
Pubblicazione: (2025)
Universal Acoustic Adversarial Attacks for Flexible Control of Speech-LLMs
di: Ma, Rao, et al.
Pubblicazione: (2025)
di: Ma, Rao, et al.
Pubblicazione: (2025)
Speaker Retrieval in the Wild: Challenges, Effectiveness and Robustness
di: Loweimi, Erfan, et al.
Pubblicazione: (2025)
di: Loweimi, Erfan, et al.
Pubblicazione: (2025)
Training Articulatory Inversion Models for Interspeaker Consistency
di: McGhee, Charles, et al.
Pubblicazione: (2025)
di: McGhee, Charles, et al.
Pubblicazione: (2025)
Zero-Shot End-To-End Spoken Question Answering In Medical Domain
di: Labrak, Yanis, et al.
Pubblicazione: (2024)
di: Labrak, Yanis, et al.
Pubblicazione: (2024)
Muting Whisper: A Universal Acoustic Adversarial Attack on Speech Foundation Models
di: Raina, Vyas, et al.
Pubblicazione: (2024)
di: Raina, Vyas, et al.
Pubblicazione: (2024)
URO-Bench: Towards Comprehensive Evaluation for End-to-End Spoken Dialogue Models
di: Yan, Ruiqi, et al.
Pubblicazione: (2025)
di: Yan, Ruiqi, et al.
Pubblicazione: (2025)
Privacy-Preserving End-to-End Spoken Language Understanding
di: Wang, Yinggui, et al.
Pubblicazione: (2024)
di: Wang, Yinggui, et al.
Pubblicazione: (2024)
The Speech-LLM Takes It All: A Truly Fully End-to-End Spoken Dialogue State Tracking Approach
di: Ghazal, Nizar El, et al.
Pubblicazione: (2025)
di: Ghazal, Nizar El, et al.
Pubblicazione: (2025)
Analyzing Mitigation Strategies for Catastrophic Forgetting in End-to-End Training of Spoken Language Models
di: Hsiao, Chi-Yuan, et al.
Pubblicazione: (2025)
di: Hsiao, Chi-Yuan, et al.
Pubblicazione: (2025)
Continual Learning for Monolingual End-to-End Automatic Speech Recognition
di: Eeckt, Steven Vander, et al.
Pubblicazione: (2021)
di: Eeckt, Steven Vander, et al.
Pubblicazione: (2021)
Finetuning End-to-End Models for Estonian Conversational Spoken Language Translation
di: Sildam, Tiia, et al.
Pubblicazione: (2024)
di: Sildam, Tiia, et al.
Pubblicazione: (2024)
GLM-4-Voice: Towards Intelligent and Human-Like End-to-End Spoken Chatbot
di: Zeng, Aohan, et al.
Pubblicazione: (2024)
di: Zeng, Aohan, et al.
Pubblicazione: (2024)
PRoDeliberation: Parallel Robust Deliberation for End-to-End Spoken Language Understanding
di: Le, Trang, et al.
Pubblicazione: (2024)
di: Le, Trang, et al.
Pubblicazione: (2024)
Towards End-to-End Training of Automatic Speech Recognition for Nigerian Pidgin
di: Rufai, Amina Mardiyyah, et al.
Pubblicazione: (2020)
di: Rufai, Amina Mardiyyah, et al.
Pubblicazione: (2020)
End-to-end streaming model for low-latency speech anonymization
di: Quamer, Waris, et al.
Pubblicazione: (2024)
di: Quamer, Waris, et al.
Pubblicazione: (2024)
Predictive Speech Recognition and End-of-Utterance Detection Towards Spoken Dialog Systems
di: Zink, Oswald, et al.
Pubblicazione: (2024)
di: Zink, Oswald, et al.
Pubblicazione: (2024)
Chain-of-Thought Reasoning in Streaming Full-Duplex End-to-End Spoken Dialogue Systems
di: Arora, Siddhant, et al.
Pubblicazione: (2025)
di: Arora, Siddhant, et al.
Pubblicazione: (2025)
VoxEval: Benchmarking the Knowledge Understanding Capabilities of End-to-End Spoken Language Models
di: Cui, Wenqian, et al.
Pubblicazione: (2025)
di: Cui, Wenqian, et al.
Pubblicazione: (2025)
Joint Speech and Text Training for LLM-Based End-to-End Spoken Dialogue State Tracking
di: Vendrame, Katia, et al.
Pubblicazione: (2025)
di: Vendrame, Katia, et al.
Pubblicazione: (2025)
Retrieval Augmented End-to-End Spoken Dialog Models
di: Wang, Mingqiu, et al.
Pubblicazione: (2024)
di: Wang, Mingqiu, et al.
Pubblicazione: (2024)
Integrating Pre-Trained Speech and Language Models for End-to-End Speech Recognition
di: Hono, Yukiya, et al.
Pubblicazione: (2023)
di: Hono, Yukiya, et al.
Pubblicazione: (2023)
SpeechDPR: End-to-End Spoken Passage Retrieval for Open-Domain Spoken Question Answering
di: Lin, Chyi-Jiunn, et al.
Pubblicazione: (2024)
di: Lin, Chyi-Jiunn, et al.
Pubblicazione: (2024)
Spoken Language Understanding on Unseen Tasks With In-Context Learning
di: Agrawal, Neeraj, et al.
Pubblicazione: (2025)
di: Agrawal, Neeraj, et al.
Pubblicazione: (2025)
End-to-End Integration of Speech Separation and Voice Activity Detection for Low-Latency Diarization of Telephone Conversations
di: Morrone, Giovanni, et al.
Pubblicazione: (2023)
di: Morrone, Giovanni, et al.
Pubblicazione: (2023)
An Automated End-to-End Open-Source Software for High-Quality Text-to-Speech Dataset Generation
di: Gunduz, Ahmet, et al.
Pubblicazione: (2024)
di: Gunduz, Ahmet, et al.
Pubblicazione: (2024)
SDiaReward: Modeling and Benchmarking Spoken Dialogue Rewards with Modality and Colloquialness
di: Lu, Jingyu, et al.
Pubblicazione: (2026)
di: Lu, Jingyu, et al.
Pubblicazione: (2026)
Revisiting ASR Error Correction with Specialized Models
di: Gu, Zijin, et al.
Pubblicazione: (2024)
di: Gu, Zijin, et al.
Pubblicazione: (2024)
Medical Spoken Named Entity Recognition
di: Le-Duc, Khai, et al.
Pubblicazione: (2024)
di: Le-Duc, Khai, et al.
Pubblicazione: (2024)
Two Heads Are Better Than One: Audio-Visual Speech Error Correction with Dual Hypotheses
di: Kim, Sungnyun, et al.
Pubblicazione: (2025)
di: Kim, Sungnyun, et al.
Pubblicazione: (2025)
BanglaDialecto: An End-to-End AI-Powered Regional Speech Standardization
di: Samin, Md. Nazmus Sadat, et al.
Pubblicazione: (2024)
di: Samin, Md. Nazmus Sadat, et al.
Pubblicazione: (2024)
"KAN you hear me?" Exploring Kolmogorov-Arnold Networks for Spoken Language Understanding
di: Koudounas, Alkis, et al.
Pubblicazione: (2025)
di: Koudounas, Alkis, et al.
Pubblicazione: (2025)
Enhanced ASR Robustness to Packet Loss with a Front-End Adaptation Network
di: Dissen, Yehoshua, et al.
Pubblicazione: (2024)
di: Dissen, Yehoshua, et al.
Pubblicazione: (2024)
FlashLabs Chroma 1.0: A Real-Time End-to-End Spoken Dialogue Model with Personalized Voice Cloning
di: Chen, Tanyu, et al.
Pubblicazione: (2026)
di: Chen, Tanyu, et al.
Pubblicazione: (2026)
Documenti analoghi
-
End-to-End Spoken Grammatical Error Correction
di: Qian, Mengjie, et al.
Pubblicazione: (2025) -
Scaling and Prompting for Improved End-to-End Spoken Grammatical Error Correction
di: Qian, Mengjie, et al.
Pubblicazione: (2025) -
Data Augmentation for Spoken Grammatical Error Correction
di: Karanasou, Penny, et al.
Pubblicazione: (2025) -
ASR Error Correction using Large Language Models
di: Ma, Rao, et al.
Pubblicazione: (2024) -
Natural Language-based Assessment of L2 Oral Proficiency using LLMs
di: Bannò, Stefano, et al.
Pubblicazione: (2025)