Incorporating Error Level Noise Embedding for Improving LLM-Assisted Robustness in Persian Speech Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Rahmani, Zahra, Sameti, Hossein |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Direct Persian-English Speech-to-Speech Translation with Discrete Units and Synthetic Parallel Data
by: Rashidi, Sina, et al.
Published: (2025)
by: Rashidi, Sina, et al.
Published: (2025)
MegaChat: A Synthetic Persian Q&A Dataset for High-Quality Sales Chatbot Evaluation
by: Rahmani, Mahdi, et al.
Published: (2025)
by: Rahmani, Mahdi, et al.
Published: (2025)
Improving Speech Recognition Error Prediction for Modern and Off-the-shelf Speech Recognizers
by: Serai, Prashant, et al.
Published: (2024)
by: Serai, Prashant, et al.
Published: (2024)
PersianRAG: A Retrieval-Augmented Generation System for Persian Language
by: Hosseini, Hossein, et al.
Published: (2024)
by: Hosseini, Hossein, et al.
Published: (2024)
Accent-Invariant Automatic Speech Recognition via Saliency-Driven Spectrogram Masking
by: Sameti, Mohammad Hossein, et al.
Published: (2025)
by: Sameti, Mohammad Hossein, et al.
Published: (2025)
Improving Speech Recognition of Named Entities in Classroom Speech with LLM Revision and Phonetic-Semantic Context
by: Trinh, Viet Anh, et al.
Published: (2025)
by: Trinh, Viet Anh, et al.
Published: (2025)
Khayyam Challenge (PersianMMLU): Is Your LLM Truly Wise to The Persian Language?
by: Ghahroodi, Omid, et al.
Published: (2024)
by: Ghahroodi, Omid, et al.
Published: (2024)
Towards Unsupervised Speech Recognition at the Syllable-Level
by: Wang, Liming, et al.
Published: (2025)
by: Wang, Liming, et al.
Published: (2025)
Error-preserving Automatic Speech Recognition of Young English Learners' Language
by: Michot, Janick, et al.
Published: (2024)
by: Michot, Janick, et al.
Published: (2024)
OPSD: an Offensive Persian Social media Dataset and its baseline evaluations
by: Safayani, Mehran, et al.
Published: (2024)
by: Safayani, Mehran, et al.
Published: (2024)
Sharif-STR at SemEval-2024 Task 1: Transformer as a Regression Model for Fine-Grained Scoring of Textual Semantic Relations
by: Ebrahimi, Seyedeh Fatemeh, et al.
Published: (2024)
by: Ebrahimi, Seyedeh Fatemeh, et al.
Published: (2024)
LipGER: Visually-Conditioned Generative Error Correction for Robust Automatic Speech Recognition
by: Ghosh, Sreyan, et al.
Published: (2024)
by: Ghosh, Sreyan, et al.
Published: (2024)
LLM-based Agentic Reasoning Frameworks: A Survey from Methods to Scenarios
by: Zhao, Bingxi, et al.
Published: (2025)
by: Zhao, Bingxi, et al.
Published: (2025)
findsylls: A Language-Agnostic Toolkit for Syllable-Level Speech Tokenization and Embedding
by: Martínez, Héctor Javier Vázquez
Published: (2026)
by: Martínez, Héctor Javier Vázquez
Published: (2026)
SENS-ASR: Semantic Embedding injection in Neural-transducer for Streaming Automatic Speech Recognition
by: Dkhissi, Youness, et al.
Published: (2026)
by: Dkhissi, Youness, et al.
Published: (2026)
SLPL SHROOM at SemEval2024 Task 06: A comprehensive study on models ability to detect hallucination
by: Fallah, Pouya, et al.
Published: (2024)
by: Fallah, Pouya, et al.
Published: (2024)
Sharif-MGTD at SemEval-2024 Task 8: A Transformer-Based Approach to Detect Machine Generated Text
by: Ebrahimi, Seyedeh Fatemeh, et al.
Published: (2024)
by: Ebrahimi, Seyedeh Fatemeh, et al.
Published: (2024)
Mitigating Out-of-Entity Errors in Named Entity Recognition: A Sentence-Level Strategy
by: Jiang, Guochao, et al.
Published: (2024)
by: Jiang, Guochao, et al.
Published: (2024)
PARSI: Persian Authorship Recognition via Stylometric Integration
by: Shahnazari, Kourosh, et al.
Published: (2025)
by: Shahnazari, Kourosh, et al.
Published: (2025)
Large Language Models are Efficient Learners of Noise-Robust Speech Recognition
by: Hu, Yuchen, et al.
Published: (2024)
by: Hu, Yuchen, et al.
Published: (2024)
Automated Rubrics for Reliable Evaluation of Medical Dialogue Systems
by: Chen, Yinzhu, et al.
Published: (2026)
by: Chen, Yinzhu, et al.
Published: (2026)
Embedding Ontologies via Incorporating Extensional and Intensional Knowledge
by: Wang, Keyu, et al.
Published: (2024)
by: Wang, Keyu, et al.
Published: (2024)
Benchmarking Rotary Position Embeddings for Automatic Speech Recognition
by: Zhang, Shucong, et al.
Published: (2025)
by: Zhang, Shucong, et al.
Published: (2025)
Benchmarking and Improving LLM Robustness for Personalized Generation
by: Okite, Chimaobi, et al.
Published: (2025)
by: Okite, Chimaobi, et al.
Published: (2025)
Adversarial Question Answering Robustness: A Multi-Level Error Analysis and Mitigation Study
by: Choudhury, Agniv Roy, et al.
Published: (2026)
by: Choudhury, Agniv Roy, et al.
Published: (2026)
Empowering Persian LLMs for Instruction Following: A Novel Dataset and Training Approach
by: Mokhtarabadi, Hojjat, et al.
Published: (2024)
by: Mokhtarabadi, Hojjat, et al.
Published: (2024)
MooER: LLM-based Speech Recognition and Translation Models from Moore Threads
by: Xu, Junhao, et al.
Published: (2024)
by: Xu, Junhao, et al.
Published: (2024)
Not All Errors Are Equal: Investigation of Speech Recognition Errors in Alzheimer's Disease Detection
by: Kang, Jiawen, et al.
Published: (2024)
by: Kang, Jiawen, et al.
Published: (2024)
Benchmarking Commercial ASR Systems on Code-Switching Speech: Arabic, Persian, and German
by: Abdoli, Sajjad, et al.
Published: (2026)
by: Abdoli, Sajjad, et al.
Published: (2026)
Vividh-ASR: A Complexity-Tiered Benchmark and Optimization Dynamics for Robust Indic Speech Recognition
by: Juvekar, Kush, et al.
Published: (2026)
by: Juvekar, Kush, et al.
Published: (2026)
Efficient Streaming LLM for Speech Recognition
by: Jia, Junteng, et al.
Published: (2024)
by: Jia, Junteng, et al.
Published: (2024)
PerMedCQA: Benchmarking Large Language Models on Medical Consumer Question Answering in Persian Language
by: Jamali, Naghmeh, et al.
Published: (2025)
by: Jamali, Naghmeh, et al.
Published: (2025)
Persian Typographical Error Type Detection Using Deep Neural Networks on Algorithmically-Generated Misspellings
by: Dehghani, Mohammad, et al.
Published: (2023)
by: Dehghani, Mohammad, et al.
Published: (2023)
Self-Correcting Large Language Models: Generation vs. Multiple Choice
by: Rahmani, Hossein A., et al.
Published: (2025)
by: Rahmani, Hossein A., et al.
Published: (2025)
Leveraging Online Data to Enhance Medical Knowledge in a Small Persian Language Model
by: Ghassabi, Mehrdad, et al.
Published: (2025)
by: Ghassabi, Mehrdad, et al.
Published: (2025)
Speak & Spell: LLM-Driven Controllable Phonetic Error Augmentation for Robust Dialogue State Tracking
by: Lee, Jihyun, et al.
Published: (2024)
by: Lee, Jihyun, et al.
Published: (2024)
MEDSAGE: Enhancing Robustness of Medical Dialogue Summarization to ASR Errors with LLM-generated Synthetic Dialogues
by: Binici, Kuluhan, et al.
Published: (2024)
by: Binici, Kuluhan, et al.
Published: (2024)
LLM-Assisted Content Conditional Debiasing for Fair Text Embedding
by: Deng, Wenlong, et al.
Published: (2024)
by: Deng, Wenlong, et al.
Published: (2024)
LLM NL2SQL Robustness: Surface Noise vs. Linguistic Variation in Traditional and Agentic Settings
by: Tu, Lifu, et al.
Published: (2026)
by: Tu, Lifu, et al.
Published: (2026)
Automatic Speech Recognition for Sanskrit with Transfer Learning
by: Sadhukhan, Bidit, et al.
Published: (2025)
by: Sadhukhan, Bidit, et al.
Published: (2025)
Similar Items
-
Improving Direct Persian-English Speech-to-Speech Translation with Discrete Units and Synthetic Parallel Data
by: Rashidi, Sina, et al.
Published: (2025) -
MegaChat: A Synthetic Persian Q&A Dataset for High-Quality Sales Chatbot Evaluation
by: Rahmani, Mahdi, et al.
Published: (2025) -
Improving Speech Recognition Error Prediction for Modern and Off-the-shelf Speech Recognizers
by: Serai, Prashant, et al.
Published: (2024) -
PersianRAG: A Retrieval-Augmented Generation System for Persian Language
by: Hosseini, Hossein, et al.
Published: (2024) -
Accent-Invariant Automatic Speech Recognition via Saliency-Driven Spectrogram Masking
by: Sameti, Mohammad Hossein, et al.
Published: (2025)