ASR Under Noise: Exploring Robustness for Sundanese and Javanese
Fuente:
arXiv
Guardado en:
| Autores principales: | Pranida, Salsabila Zahirah, Airlangga, Muhammad Cendekia, Genadi, Rifo Ahmad, Shehata, Shady |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Culturally-Nuanced Story Generation for Reasoning in Low-Resource Languages: The Case of Javanese and Sundanese
por: Pranida, Salsabila Zahirah, et al.
Publicado: (2025)
por: Pranida, Salsabila Zahirah, et al.
Publicado: (2025)
Multilingual Idioms in Sentences and Conversations Across High-, Medium-, and Low-Resource Languages
por: Almheiri, Saeed, et al.
Publicado: (2026)
por: Almheiri, Saeed, et al.
Publicado: (2026)
Emergence of Primacy and Recency Effect in Mamba: A Mechanistic Point of View
por: Airlangga, Muhammad Cendekia, et al.
Publicado: (2025)
por: Airlangga, Muhammad Cendekia, et al.
Publicado: (2025)
XLS-R Deep Learning Model for Multilingual ASR on Low- Resource Languages: Indonesian, Javanese, and Sundanese
por: Arisaputra, Panji, et al.
Publicado: (2024)
por: Arisaputra, Panji, et al.
Publicado: (2024)
Surfacing Subtle Stereotypes: A Multilingual, Debate-Oriented Evaluation of Modern LLMs
por: Saeed, Muhammed, et al.
Publicado: (2025)
por: Saeed, Muhammed, et al.
Publicado: (2025)
Measuring AI Reasoning: A Guide for Researchers
por: Nwadike, Munachiso Samuel, et al.
Publicado: (2026)
por: Nwadike, Munachiso Samuel, et al.
Publicado: (2026)
Sycophancy Hides Linearly in the Attention Heads
por: Genadi, Rifo, et al.
Publicado: (2026)
por: Genadi, Rifo, et al.
Publicado: (2026)
Detecting Propaganda Techniques in Code-Switched Social Media Text
por: Salman, Muhammad Umar, et al.
Publicado: (2023)
por: Salman, Muhammad Umar, et al.
Publicado: (2023)
ArabEmoNet: A Lightweight Hybrid 2D CNN-BiLSTM Model with Attention for Robust Arabic Speech Emotion Recognition
por: Abouzeid, Ali, et al.
Publicado: (2025)
por: Abouzeid, Ali, et al.
Publicado: (2025)
Exploring the Limitations of Detecting Machine-Generated Text
por: Doughman, Jad, et al.
Publicado: (2024)
por: Doughman, Jad, et al.
Publicado: (2024)
NileChat: Towards Linguistically Diverse and Culturally Aware LLMs for Local Communities
por: Mekki, Abdellah El, et al.
Publicado: (2025)
por: Mekki, Abdellah El, et al.
Publicado: (2025)
NurseLLM: The First Specialized Language Model for Nursing
por: Khondaker, Md Tawkat Islam, et al.
Publicado: (2025)
por: Khondaker, Md Tawkat Islam, et al.
Publicado: (2025)
Beyond Content: How Grammatical Gender Shapes Visual Representation in Text-to-Image Models
por: Saeed, Muhammed, et al.
Publicado: (2025)
por: Saeed, Muhammed, et al.
Publicado: (2025)
Desert Camels and Oil Sheikhs: Arab-Centric Red Teaming of Frontier LLMs
por: Saeed, Muhammed, et al.
Publicado: (2024)
por: Saeed, Muhammed, et al.
Publicado: (2024)
ArFake: A Multi-Dialect Benchmark and Baselines for Arabic Spoof-Speech Detection
por: Maged, Mohamed, et al.
Publicado: (2025)
por: Maged, Mohamed, et al.
Publicado: (2025)
Explainable Disentanglement on Discrete Speech Representations for Noise-Robust ASR
por: Gopal, Shreyas, et al.
Publicado: (2025)
por: Gopal, Shreyas, et al.
Publicado: (2025)
Can LLM Generate Culturally Relevant Commonsense QA Data? Case Study in Indonesian and Sundanese
por: Putri, Rifki Afina, et al.
Publicado: (2024)
por: Putri, Rifki Afina, et al.
Publicado: (2024)
Revisiting Acoustic Features for Robust ASR
por: Shah, Muhammad A., et al.
Publicado: (2024)
por: Shah, Muhammad A., et al.
Publicado: (2024)
Do Language Models Understand Honorific Systems in Javanese?
por: Farhansyah, Mohammad Rifqi, et al.
Publicado: (2025)
por: Farhansyah, Mohammad Rifqi, et al.
Publicado: (2025)
Cross-lingual Transfer Learning for Javanese Dependency Parsing
por: Ghiffari, Fadli Aulawi Al, et al.
Publicado: (2024)
por: Ghiffari, Fadli Aulawi Al, et al.
Publicado: (2024)
Jawaher: A Multidialectal Dataset of Arabic Proverbs for LLM Benchmarking
por: Magdy, Samar M., et al.
Publicado: (2025)
por: Magdy, Samar M., et al.
Publicado: (2025)
On the Robust Approximation of ASR Metrics
por: Waheed, Abdul, et al.
Publicado: (2025)
por: Waheed, Abdul, et al.
Publicado: (2025)
A Comprehensive Study on the Effectiveness of ASR Representations for Noise-Robust Speech Emotion Recognition
por: Shi, Xiaohan, et al.
Publicado: (2023)
por: Shi, Xiaohan, et al.
Publicado: (2023)
Open Universal Arabic ASR Leaderboard
por: Wang, Yingzhi, et al.
Publicado: (2024)
por: Wang, Yingzhi, et al.
Publicado: (2024)
Noise-Robust AV-ASR Using Visual Features Both in the Whisper Encoder and Decoder
por: Li, Zhengyang, et al.
Publicado: (2026)
por: Li, Zhengyang, et al.
Publicado: (2026)
Exploring the Potential of Multimodal LLM with Knowledge-Intensive Multimodal ASR
por: Wang, Minghan, et al.
Publicado: (2024)
por: Wang, Minghan, et al.
Publicado: (2024)
Zero-Shot Context-Aware ASR for Diverse Arabic Varieties
por: Talafha, Bashar, et al.
Publicado: (2025)
por: Talafha, Bashar, et al.
Publicado: (2025)
Automatic Speech Recognition (ASR) for African Low-Resource Languages: A Systematic Literature Review
por: Imam, Sukairaj Hafiz, et al.
Publicado: (2025)
por: Imam, Sukairaj Hafiz, et al.
Publicado: (2025)
Building Robust and Scalable Multilingual ASR for Indian Languages
por: Gangwar, Arjun, et al.
Publicado: (2025)
por: Gangwar, Arjun, et al.
Publicado: (2025)
Exploring SSL Discrete Tokens for Multilingual ASR
por: Cui, Mingyu, et al.
Publicado: (2024)
por: Cui, Mingyu, et al.
Publicado: (2024)
LAHAJA: A Robust Multi-accent Benchmark for Evaluating Hindi ASR Systems
por: Javed, Tahir, et al.
Publicado: (2024)
por: Javed, Tahir, et al.
Publicado: (2024)
Interventional Speech Noise Injection for ASR Generalizable Spoken Language Understanding
por: Jung, Yeonjoon, et al.
Publicado: (2024)
por: Jung, Yeonjoon, et al.
Publicado: (2024)
NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR
por: Xie, Yuan, et al.
Publicado: (2026)
por: Xie, Yuan, et al.
Publicado: (2026)
Robust ASR Error Correction with Conservative Data Filtering
por: Udagawa, Takuma, et al.
Publicado: (2024)
por: Udagawa, Takuma, et al.
Publicado: (2024)
Bridging ASR and LLMs for Dysarthric Speech Recognition: Benchmarking Self-Supervised and Generative Approaches
por: Aboeitta, Ahmed, et al.
Publicado: (2025)
por: Aboeitta, Ahmed, et al.
Publicado: (2025)
Towards Pretraining Robust ASR Foundation Model with Acoustic-Aware Data Augmentation
por: Liu, Dancheng, et al.
Publicado: (2025)
por: Liu, Dancheng, et al.
Publicado: (2025)
Noise-Robust De-Duplication at Scale
por: Silcock, Emily, et al.
Publicado: (2022)
por: Silcock, Emily, et al.
Publicado: (2022)
Can We Trust LLMs for Mental Health Screening? Consistency, ASR Robustness, and Evidence Faithfulness
por: Loweimi, Erfan, et al.
Publicado: (2026)
por: Loweimi, Erfan, et al.
Publicado: (2026)
Exploring the potential and limitations of Model Merging for Multi-Domain Adaptation in ASR
por: Carvalho, Carlos, et al.
Publicado: (2026)
por: Carvalho, Carlos, et al.
Publicado: (2026)
Universal-2-TF: Robust All-Neural Text Formatting for ASR
por: Khare, Yash, et al.
Publicado: (2025)
por: Khare, Yash, et al.
Publicado: (2025)
Ejemplares similares
-
Culturally-Nuanced Story Generation for Reasoning in Low-Resource Languages: The Case of Javanese and Sundanese
por: Pranida, Salsabila Zahirah, et al.
Publicado: (2025) -
Multilingual Idioms in Sentences and Conversations Across High-, Medium-, and Low-Resource Languages
por: Almheiri, Saeed, et al.
Publicado: (2026) -
Emergence of Primacy and Recency Effect in Mamba: A Mechanistic Point of View
por: Airlangga, Muhammad Cendekia, et al.
Publicado: (2025) -
XLS-R Deep Learning Model for Multilingual ASR on Low- Resource Languages: Indonesian, Javanese, and Sundanese
por: Arisaputra, Panji, et al.
Publicado: (2024) -
Surfacing Subtle Stereotypes: A Multilingual, Debate-Oriented Evaluation of Modern LLMs
por: Saeed, Muhammed, et al.
Publicado: (2025)