Guardado en:
| Autores principales: | Wahida, Farah, Chamikara, M. A. P., Shanmugarasa, Yashothara, Chhetri, Mohan Baruwal, Ranbaduge, Thilina, Khalil, Ibrahim |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2508.05409 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Hidden in the Noise: Unveiling Backdoors in Audio LLMs Alignment through Latent Acoustic Pattern Triggers
por: Lin, Liang, et al.
Publicado: (2025)
por: Lin, Liang, et al.
Publicado: (2025)
Multichannel Voice Trigger Detection Based on Transform-average-concatenate
por: Higuchi, Takuya, et al.
Publicado: (2023)
por: Higuchi, Takuya, et al.
Publicado: (2023)
Two-pass Endpoint Detection for Speech Recognition
por: Raju, Anirudh, et al.
Publicado: (2024)
por: Raju, Anirudh, et al.
Publicado: (2024)
Dual-Branch Knowledge Distillation for Noise-Robust Synthetic Speech Detection
por: Fan, Cunhang, et al.
Publicado: (2023)
por: Fan, Cunhang, et al.
Publicado: (2023)
Suppressing Noise Disparity in Training Data for Automatic Pathological Speech Detection
por: Amiri, Mahdi, et al.
Publicado: (2024)
por: Amiri, Mahdi, et al.
Publicado: (2024)
CEC: A Noisy Label Detection Method for Speaker Recognition
por: Shen, Yao, et al.
Publicado: (2024)
por: Shen, Yao, et al.
Publicado: (2024)
Leveraging LLM and Text-Queried Separation for Noise-Robust Sound Event Detection
por: Yin, Han, et al.
Publicado: (2024)
por: Yin, Han, et al.
Publicado: (2024)
Noise-Robust Contrastive Learning with an MFCC-Conformer For Coronary Artery Disease Detection
por: Marocchi, Milan, et al.
Publicado: (2026)
por: Marocchi, Milan, et al.
Publicado: (2026)
CTC Blank Triggered Dynamic Layer-Skipping for Efficient CTC-based Speech Recognition
por: Hou, Junfeng, et al.
Publicado: (2024)
por: Hou, Junfeng, et al.
Publicado: (2024)
Noise-Robust Sound Event Detection and Counting via Language-Queried Sound Separation
por: Chen, Yuanjian, et al.
Publicado: (2025)
por: Chen, Yuanjian, et al.
Publicado: (2025)
Noisy Disentanglement with Tri-stage Training for Noise-Robust Speech Recognition
por: Chen, Shuangyuan, et al.
Publicado: (2025)
por: Chen, Shuangyuan, et al.
Publicado: (2025)
Leveraging Mamba with Full-Face Vision for Audio-Visual Speech Enhancement
por: Chao, Rong, et al.
Publicado: (2025)
por: Chao, Rong, et al.
Publicado: (2025)
Face-Voice Association for Audiovisual Active Speaker Detection in Egocentric Recordings
por: Clarke, Jason, et al.
Publicado: (2025)
por: Clarke, Jason, et al.
Publicado: (2025)
Findings of the 2024 Mandarin Stuttering Event Detection and Automatic Speech Recognition Challenge
por: Xue, Hongfei, et al.
Publicado: (2024)
por: Xue, Hongfei, et al.
Publicado: (2024)
Leveraging LLM for Stuttering Speech: A Unified Architecture Bridging Recognition and Event Detection
por: Huang, Shangkun, et al.
Publicado: (2025)
por: Huang, Shangkun, et al.
Publicado: (2025)
Testing Correctness, Fairness, and Robustness of Speech Emotion Recognition Models
por: Derington, Anna, et al.
Publicado: (2023)
por: Derington, Anna, et al.
Publicado: (2023)
Retrieval Augmented Correction of Named Entity Speech Recognition Errors
por: Pusateri, Ernest, et al.
Publicado: (2024)
por: Pusateri, Ernest, et al.
Publicado: (2024)
Descriptor:: Extended-Length Audio Dataset for Synthetic Voice Detection and Speaker Recognition (ELAD-SVDSR)
por: Vijaykumar, Rahul, et al.
Publicado: (2025)
por: Vijaykumar, Rahul, et al.
Publicado: (2025)
Enhancing Automatic Speech Recognition Through Integrated Noise Detection Architecture
por: Singh, Karamvir
Publicado: (2025)
por: Singh, Karamvir
Publicado: (2025)
Misophonia Trigger Sound Detection on Synthetic Soundscapes Using a Hybrid Model with a Frozen Pre-Trained CNN and a Time-Series Module
por: Sashida, Kurumi, et al.
Publicado: (2026)
por: Sashida, Kurumi, et al.
Publicado: (2026)
Description and Discussion on DCASE 2026 Challenge Task 2: Noise-aware Unsupervised Anomalous Sound Detection for Machine Condition Monitoring
por: Nishida, Tomoya, et al.
Publicado: (2026)
por: Nishida, Tomoya, et al.
Publicado: (2026)
End-to-End Integration of Speech Emotion Recognition with Voice Activity Detection using Self-Supervised Learning Features
por: Yamashita, Natsuo, et al.
Publicado: (2024)
por: Yamashita, Natsuo, et al.
Publicado: (2024)
Gradient Norm-based Fine-Tuning for Backdoor Defense in Automatic Speech Recognition
por: Zhou, Nanjun, et al.
Publicado: (2025)
por: Zhou, Nanjun, et al.
Publicado: (2025)
Towards Robust Dysarthric Speech Recognition: LLM-Agent Post-ASR Correction Beyond WER
por: Zheng, Xiuwen, et al.
Publicado: (2026)
por: Zheng, Xiuwen, et al.
Publicado: (2026)
MMGER: Multi-modal and Multi-granularity Generative Error Correction with LLM for Joint Accent and Speech Recognition
por: Mu, Bingshen, et al.
Publicado: (2024)
por: Mu, Bingshen, et al.
Publicado: (2024)
Adaptive Noise Resilient Keyword Spotting Using One-Shot Learning
por: Martinez-Rau, Luciano Sebastian, et al.
Publicado: (2025)
por: Martinez-Rau, Luciano Sebastian, et al.
Publicado: (2025)
Imperceptible Rhythm Backdoor Attacks: Exploring Rhythm Transformation for Embedding Undetectable Vulnerabilities on Speech Recognition
por: Yao, Wenhan, et al.
Publicado: (2024)
por: Yao, Wenhan, et al.
Publicado: (2024)
Exploration of Adapter for Noise Robust Automatic Speech Recognition
por: Shi, Hao, et al.
Publicado: (2024)
por: Shi, Hao, et al.
Publicado: (2024)
Mixture of LoRA Experts with Multi-Modal and Multi-Granularity LLM Generative Error Correction for Accented Speech Recognition
por: Mu, Bingshen, et al.
Publicado: (2025)
por: Mu, Bingshen, et al.
Publicado: (2025)
From Audio Deepfake Detection to AI-Generated Music Detection -- A Pathway and Overview
por: Li, Yupei, et al.
Publicado: (2024)
por: Li, Yupei, et al.
Publicado: (2024)
Pitch Accent Detection improves Pretrained Automatic Speech Recognition
por: Sasu, David, et al.
Publicado: (2025)
por: Sasu, David, et al.
Publicado: (2025)
Unified Audio Event Detection
por: Jiang, Yidi, et al.
Publicado: (2024)
por: Jiang, Yidi, et al.
Publicado: (2024)
Generalizable Detection of Audio Deepfakes
por: Lopez, Jose A., et al.
Publicado: (2025)
por: Lopez, Jose A., et al.
Publicado: (2025)
Detection of Deepfake Environmental Audio
por: Ouajdi, Hafsa, et al.
Publicado: (2024)
por: Ouajdi, Hafsa, et al.
Publicado: (2024)
Automatic Pronunciation Error Detection and Correction of the Holy Quran's Learners Using Deep Learning
por: Abdelfattah, Abdullah, et al.
Publicado: (2025)
por: Abdelfattah, Abdullah, et al.
Publicado: (2025)
Personalized Speech Emotion Recognition in Human-Robot Interaction using Vision Transformers
por: Mishra, Ruchik, et al.
Publicado: (2024)
por: Mishra, Ruchik, et al.
Publicado: (2024)
Water Flow Detection Device Based on Sound Data Analysis and Machine Learning to Detect Water Leakage
por: Pourmehrani, Hossein, et al.
Publicado: (2025)
por: Pourmehrani, Hossein, et al.
Publicado: (2025)
Speech as a Biomarker for Disease Detection
por: Botelho, Catarina, et al.
Publicado: (2024)
por: Botelho, Catarina, et al.
Publicado: (2024)
ASPED: An Audio Dataset for Detecting Pedestrians
por: Seshadri, Pavan, et al.
Publicado: (2023)
por: Seshadri, Pavan, et al.
Publicado: (2023)
Mind the Gap: Detecting Cluster Exits for Robust Local Density-Based Score Normalization in Anomalous Sound Detection
por: Wilkinghoff, Kevin, et al.
Publicado: (2026)
por: Wilkinghoff, Kevin, et al.
Publicado: (2026)
Ejemplares similares
-
Hidden in the Noise: Unveiling Backdoors in Audio LLMs Alignment through Latent Acoustic Pattern Triggers
por: Lin, Liang, et al.
Publicado: (2025) -
Multichannel Voice Trigger Detection Based on Transform-average-concatenate
por: Higuchi, Takuya, et al.
Publicado: (2023) -
Two-pass Endpoint Detection for Speech Recognition
por: Raju, Anirudh, et al.
Publicado: (2024) -
Dual-Branch Knowledge Distillation for Noise-Robust Synthetic Speech Detection
por: Fan, Cunhang, et al.
Publicado: (2023) -
Suppressing Noise Disparity in Training Data for Automatic Pathological Speech Detection
por: Amiri, Mahdi, et al.
Publicado: (2024)