Fairness of Automatic Speech Recognition in Cleft Lip and Palate Speech
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bhattacharjee, Susmita, Mishra, Jagabandhu, Shekhawat, H. S., Prasanna, S. R. Mahadeva |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Parameter-Efficient Fine-Tuning of Foundation Models for CLP Speech Classification
von: Bhattacharjee, Susmita, et al.
Veröffentlicht: (2025)
von: Bhattacharjee, Susmita, et al.
Veröffentlicht: (2025)
Spoken language change detection inspired by speaker change detection
von: Mishra, Jagabandhu, et al.
Veröffentlicht: (2023)
von: Mishra, Jagabandhu, et al.
Veröffentlicht: (2023)
Fusion of Modulation Spectrogram and SSL with Multi-head Attention for Fake Speech Detection
von: N, Rishith Sadashiv T, et al.
Veröffentlicht: (2025)
von: N, Rishith Sadashiv T, et al.
Veröffentlicht: (2025)
Implicit Self-supervised Language Representation for Spoken Language Diarization
von: Mishra, Jagabandhu, et al.
Veröffentlicht: (2023)
von: Mishra, Jagabandhu, et al.
Veröffentlicht: (2023)
Timbre-Aware LLM-based Direct Speech-to-Speech Translation Extendable to Multiple Language Pairs
von: Arya, Lalaram, et al.
Veröffentlicht: (2026)
von: Arya, Lalaram, et al.
Veröffentlicht: (2026)
Advancing Zero-Shot Open-Set Speech Deepfake Source Tracing
von: Chhibber, Manasi, et al.
Veröffentlicht: (2025)
von: Chhibber, Manasi, et al.
Veröffentlicht: (2025)
Towards Explainable Spoofed Speech Attribution and Detection:a Probabilistic Approach for Characterizing Speech Synthesizer Components
von: Mishra, Jagabandhu, et al.
Veröffentlicht: (2025)
von: Mishra, Jagabandhu, et al.
Veröffentlicht: (2025)
FairASR: Fair Audio Contrastive Learning for Automatic Speech Recognition
von: Kim, Jongsuk, et al.
Veröffentlicht: (2025)
von: Kim, Jongsuk, et al.
Veröffentlicht: (2025)
An Explainable Probabilistic Attribute Embedding Approach for Spoofed Speech Characterization
von: Chhibber, Manasi, et al.
Veröffentlicht: (2024)
von: Chhibber, Manasi, et al.
Veröffentlicht: (2024)
Joint Optimization of Speaker and Spoof Detectors for Spoofing-Robust Automatic Speaker Verification
von: Kurnaz, Oğuzhan, et al.
Veröffentlicht: (2025)
von: Kurnaz, Oğuzhan, et al.
Veröffentlicht: (2025)
Kinship Verification Using Voice
von: Mishra, Jagabandhu, et al.
Veröffentlicht: (2026)
von: Mishra, Jagabandhu, et al.
Veröffentlicht: (2026)
Speaker Attributed Automatic Speech Recognition Using Speech Aware LLMS
von: Aronowitz, Hagai, et al.
Veröffentlicht: (2026)
von: Aronowitz, Hagai, et al.
Veröffentlicht: (2026)
LipDiffuser: Lip-to-Speech Generation with Conditional Diffusion Models
von: Richter, Julius, et al.
Veröffentlicht: (2025)
von: Richter, Julius, et al.
Veröffentlicht: (2025)
Unsupervised Online Continual Learning for Automatic Speech Recognition
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2024)
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2024)
Using Songs to Improve Kazakh Automatic Speech Recognition
von: Yeshpanov, Rustem
Veröffentlicht: (2026)
von: Yeshpanov, Rustem
Veröffentlicht: (2026)
The RoyalFlush Automatic Speech Diarization and Recognition System for In-Car Multi-Channel Automatic Speech Recognition Challenge
von: Tian, Jingguang, et al.
Veröffentlicht: (2024)
von: Tian, Jingguang, et al.
Veröffentlicht: (2024)
Spoofing-Robust Speaker Verification Using Parallel Embedding Fusion: BTU Speech Group's Approach for ASVspoof5 Challenge
von: Kurnaz, Oğuzhan, et al.
Veröffentlicht: (2024)
von: Kurnaz, Oğuzhan, et al.
Veröffentlicht: (2024)
Too Good to Be True: A Study on Modern Automatic Speech Recognition for the Evaluation of Speech Enhancement
von: de Oliveira, Danilo, et al.
Veröffentlicht: (2026)
von: de Oliveira, Danilo, et al.
Veröffentlicht: (2026)
LipVoicer: Generating Speech from Silent Videos Guided by Lip Reading
von: Yemini, Yochai, et al.
Veröffentlicht: (2023)
von: Yemini, Yochai, et al.
Veröffentlicht: (2023)
LipGER: Visually-Conditioned Generative Error Correction for Robust Automatic Speech Recognition
von: Ghosh, Sreyan, et al.
Veröffentlicht: (2024)
von: Ghosh, Sreyan, et al.
Veröffentlicht: (2024)
Non-Intrusive Automatic Speech Recognition Refinement: A Survey
von: Peyghan, Mohammad Reza, et al.
Veröffentlicht: (2025)
von: Peyghan, Mohammad Reza, et al.
Veröffentlicht: (2025)
Disentangled-Transformer: An Explainable End-to-End Automatic Speech Recognition Model with Speech Content-Context Separation
von: Wang, Pu, et al.
Veröffentlicht: (2024)
von: Wang, Pu, et al.
Veröffentlicht: (2024)
Zero Shot Text to Speech Augmentation for Automatic Speech Recognition on Low-Resource Accented Speech Corpora
von: Nespoli, Francesco, et al.
Veröffentlicht: (2024)
von: Nespoli, Francesco, et al.
Veröffentlicht: (2024)
Robust Nasality Representation Learning for Cleft Palate-Related Velopharyngeal Dysfunction Screening in Real-World Settings
von: Liu, Weixin, et al.
Veröffentlicht: (2026)
von: Liu, Weixin, et al.
Veröffentlicht: (2026)
Group-Aware Partial Model Merging for Children's Automatic Speech Recognition
von: Rolland, Thomas, et al.
Veröffentlicht: (2025)
von: Rolland, Thomas, et al.
Veröffentlicht: (2025)
UME: Upcycling Mixture-of-Experts for Scalable and Efficient Automatic Speech Recognition
von: Fu, Li, et al.
Veröffentlicht: (2024)
von: Fu, Li, et al.
Veröffentlicht: (2024)
Training Data Augmentation for Dysarthric Automatic Speech Recognition by Text-to-Dysarthric-Speech Synthesis
von: Leung, Wing-Zin, et al.
Veröffentlicht: (2024)
von: Leung, Wing-Zin, et al.
Veröffentlicht: (2024)
Zero-Shot Recognition of Dysarthric Speech Using Commercial Automatic Speech Recognition and Multimodal Large Language Models
von: Alsayegh, Ali, et al.
Veröffentlicht: (2025)
von: Alsayegh, Ali, et al.
Veröffentlicht: (2025)
Detecting and Defending Against Adversarial Attacks on Automatic Speech Recognition via Diffusion Models
von: Kühne, Nikolai L., et al.
Veröffentlicht: (2024)
von: Kühne, Nikolai L., et al.
Veröffentlicht: (2024)
Aligning Speech to Languages to Enhance Code-switching Speech Recognition
von: Liu, Hexin, et al.
Veröffentlicht: (2024)
von: Liu, Hexin, et al.
Veröffentlicht: (2024)
Streaming Decoder-Only Automatic Speech Recognition with Discrete Speech Units: A Pilot Study
von: Chen, Peikun, et al.
Veröffentlicht: (2024)
von: Chen, Peikun, et al.
Veröffentlicht: (2024)
Automatic Speech Recognition for Hindi
von: Saha, Anish, et al.
Veröffentlicht: (2024)
von: Saha, Anish, et al.
Veröffentlicht: (2024)
Improving Automatic Speech Recognition with Decoder-Centric Regularisation in Encoder-Decoder Models
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
von: Polok, Alexander, et al.
Veröffentlicht: (2024)
Using Adapters to Overcome Catastrophic Forgetting in End-to-End Automatic Speech Recognition
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2022)
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2022)
Testing Correctness, Fairness, and Robustness of Speech Emotion Recognition Models
von: Derington, Anna, et al.
Veröffentlicht: (2023)
von: Derington, Anna, et al.
Veröffentlicht: (2023)
A Semi-spontaneous Dutch Speech Dataset for Speech Enhancement and Speech Recognition
von: de Groot, Dimme, et al.
Veröffentlicht: (2026)
von: de Groot, Dimme, et al.
Veröffentlicht: (2026)
Optimizing a-DCF for Spoofing-Robust Speaker Verification
von: Kurnaz, Oğuzhan, et al.
Veröffentlicht: (2024)
von: Kurnaz, Oğuzhan, et al.
Veröffentlicht: (2024)
Augmenting Polish Automatic Speech Recognition System With Synthetic Data
von: Bondaruk, Łukasz, et al.
Veröffentlicht: (2024)
von: Bondaruk, Łukasz, et al.
Veröffentlicht: (2024)
Leveraging Self-Supervised Models for Automatic Whispered Speech Recognition
von: Farhadipour, Aref, et al.
Veröffentlicht: (2024)
von: Farhadipour, Aref, et al.
Veröffentlicht: (2024)
Rehearsal-Free Online Continual Learning for Automatic Speech Recognition
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2023)
von: Eeckt, Steven Vander, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Parameter-Efficient Fine-Tuning of Foundation Models for CLP Speech Classification
von: Bhattacharjee, Susmita, et al.
Veröffentlicht: (2025) -
Spoken language change detection inspired by speaker change detection
von: Mishra, Jagabandhu, et al.
Veröffentlicht: (2023) -
Fusion of Modulation Spectrogram and SSL with Multi-head Attention for Fake Speech Detection
von: N, Rishith Sadashiv T, et al.
Veröffentlicht: (2025) -
Implicit Self-supervised Language Representation for Spoken Language Diarization
von: Mishra, Jagabandhu, et al.
Veröffentlicht: (2023) -
Timbre-Aware LLM-based Direct Speech-to-Speech Translation Extendable to Multiple Language Pairs
von: Arya, Lalaram, et al.
Veröffentlicht: (2026)