Helsinki Speech Challenge 2024
Fuente:
arXiv
Salvato in:
| Autori principali: | Ludvigsen, Martin, Karvonen, Elli, Juvonen, Markus, Siltanen, Samuli |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The Interspeech 2024 Challenge on Speech Processing Using Discrete Units
di: Chang, Xuankai, et al.
Pubblicazione: (2024)
di: Chang, Xuankai, et al.
Pubblicazione: (2024)
The VoiceMOS Challenge 2024: Beyond Speech Quality Prediction
di: Huang, Wen-Chin, et al.
Pubblicazione: (2024)
di: Huang, Wen-Chin, et al.
Pubblicazione: (2024)
Recomposed realities: animating still images via patch clustering and randomness
di: Juvonen, Markus, et al.
Pubblicazione: (2025)
di: Juvonen, Markus, et al.
Pubblicazione: (2025)
UTDUSS: UTokyo-SaruLab System for Interspeech2024 Speech Processing Using Discrete Speech Unit Challenge
di: Nakata, Wataru, et al.
Pubblicazione: (2024)
di: Nakata, Wataru, et al.
Pubblicazione: (2024)
Findings of the 2024 Mandarin Stuttering Event Detection and Automatic Speech Recognition Challenge
di: Xue, Hongfei, et al.
Pubblicazione: (2024)
di: Xue, Hongfei, et al.
Pubblicazione: (2024)
Lessons Learned from the URGENT 2024 Speech Enhancement Challenge
di: Zhang, Wangyou, et al.
Pubblicazione: (2025)
di: Zhang, Wangyou, et al.
Pubblicazione: (2025)
Interspeech 2025 URGENT Speech Enhancement Challenge
di: Saijo, Kohei, et al.
Pubblicazione: (2025)
di: Saijo, Kohei, et al.
Pubblicazione: (2025)
Double Multi-Head Attention Multimodal System for Odyssey 2024 Speech Emotion Recognition Challenge
di: Costa, Federico, et al.
Pubblicazione: (2024)
di: Costa, Federico, et al.
Pubblicazione: (2024)
Benchmarking Speech Systems for Frontline Health Conversations: The DISPLACE-M Challenge
di: E, Dhanya, et al.
Pubblicazione: (2026)
di: E, Dhanya, et al.
Pubblicazione: (2026)
The X-LANCE Technical Report for Interspeech 2024 Speech Processing Using Discrete Speech Unit Challenge
di: Guo, Yiwei, et al.
Pubblicazione: (2024)
di: Guo, Yiwei, et al.
Pubblicazione: (2024)
NPU-NTU System for Voice Privacy 2024 Challenge
di: Yao, Jixun, et al.
Pubblicazione: (2024)
di: Yao, Jixun, et al.
Pubblicazione: (2024)
The Database and Benchmark for the Source Speaker Tracing Challenge 2024
di: Li, Ze, et al.
Pubblicazione: (2024)
di: Li, Ze, et al.
Pubblicazione: (2024)
Acoustic modeling for Overlapping Speech Recognition: JHU Chime-5 Challenge System
di: Manohar, Vimal, et al.
Pubblicazione: (2024)
di: Manohar, Vimal, et al.
Pubblicazione: (2024)
P.808 Multilingual Speech Enhancement Testing: Approach and Results of URGENT 2025 Challenge
di: Sach, Marvin, et al.
Pubblicazione: (2025)
di: Sach, Marvin, et al.
Pubblicazione: (2025)
ICASSP 2026 URGENT Speech Enhancement Challenge
di: Li, Chenda, et al.
Pubblicazione: (2026)
di: Li, Chenda, et al.
Pubblicazione: (2026)
Collecting, Curating, and Annotating Good Quality Speech deepfake dataset for Famous Figures: Process and Challenges
di: Ali, Hashim, et al.
Pubblicazione: (2025)
di: Ali, Hashim, et al.
Pubblicazione: (2025)
ASASVIcomtech: The Vicomtech-UGR Speech Deepfake Detection and SASV Systems for the ASVspoof5 Challenge
di: Martín-Doñas, Juan M., et al.
Pubblicazione: (2024)
di: Martín-Doñas, Juan M., et al.
Pubblicazione: (2024)
URGENT Challenge: Universality, Robustness, and Generalizability For Speech Enhancement
di: Zhang, Wangyou, et al.
Pubblicazione: (2024)
di: Zhang, Wangyou, et al.
Pubblicazione: (2024)
Advances in Speech Separation: Techniques, Challenges, and Future Trends
di: Li, Kai, et al.
Pubblicazione: (2025)
di: Li, Kai, et al.
Pubblicazione: (2025)
The TEA-ASLP System for Multilingual Conversational Speech Recognition and Speech Diarization in MLC-SLM 2025 Challenge
di: Xue, Hongfei, et al.
Pubblicazione: (2025)
di: Xue, Hongfei, et al.
Pubblicazione: (2025)
Textless Streaming Speech-to-Speech Translation using Semantic Speech Tokens
di: Zhao, Jinzheng, et al.
Pubblicazione: (2024)
di: Zhao, Jinzheng, et al.
Pubblicazione: (2024)
The CCF AATC 2025 Speech Restoration Challenge: A Retrospective
di: Zhang, Junan, et al.
Pubblicazione: (2025)
di: Zhang, Junan, et al.
Pubblicazione: (2025)
Context-Driven Dynamic Pruning for Large Speech Foundation Models
di: Someki, Masao, et al.
Pubblicazione: (2025)
di: Someki, Masao, et al.
Pubblicazione: (2025)
AS-Speech: Adaptive Style For Speech Synthesis
di: Li, Zhipeng, et al.
Pubblicazione: (2024)
di: Li, Zhipeng, et al.
Pubblicazione: (2024)
Data-Efficient Low-Complexity Acoustic Scene Classification in the DCASE 2024 Challenge
di: Schmid, Florian, et al.
Pubblicazione: (2024)
di: Schmid, Florian, et al.
Pubblicazione: (2024)
A Semi-spontaneous Dutch Speech Dataset for Speech Enhancement and Speech Recognition
di: de Groot, Dimme, et al.
Pubblicazione: (2026)
di: de Groot, Dimme, et al.
Pubblicazione: (2026)
The RoyalFlush Automatic Speech Diarization and Recognition System for In-Car Multi-Channel Automatic Speech Recognition Challenge
di: Tian, Jingguang, et al.
Pubblicazione: (2024)
di: Tian, Jingguang, et al.
Pubblicazione: (2024)
Robust Speech Activity Detection in the Presence of Singing Voice
di: Grundhuber, Philipp, et al.
Pubblicazione: (2025)
di: Grundhuber, Philipp, et al.
Pubblicazione: (2025)
Speech Quality-Based Localization of Low-Quality Speech and Text-to-Speech Synthesis Artefacts
di: Kuhlmann, Michael, et al.
Pubblicazione: (2026)
di: Kuhlmann, Michael, et al.
Pubblicazione: (2026)
Influence of Clean Speech Characteristics on Speech Enhancement Performance
di: Hou, Mingchi, et al.
Pubblicazione: (2025)
di: Hou, Mingchi, et al.
Pubblicazione: (2025)
Aligning Speech to Languages to Enhance Code-switching Speech Recognition
di: Liu, Hexin, et al.
Pubblicazione: (2024)
di: Liu, Hexin, et al.
Pubblicazione: (2024)
Fairness of Automatic Speech Recognition in Cleft Lip and Palate Speech
di: Bhattacharjee, Susmita, et al.
Pubblicazione: (2025)
di: Bhattacharjee, Susmita, et al.
Pubblicazione: (2025)
Assessing the Impact of Noise and Speech Enhancement on the Intelligibility of Speech Codecs
di: Behringer, Lyonel, et al.
Pubblicazione: (2026)
di: Behringer, Lyonel, et al.
Pubblicazione: (2026)
FlexSpeech: Towards Stable, Controllable and Expressive Text-to-Speech
di: Ma, Linhan, et al.
Pubblicazione: (2025)
di: Ma, Linhan, et al.
Pubblicazione: (2025)
Squeeze-and-Excite ResNet-Conformers for Sound Event Localization, Detection, and Distance Estimation for DCASE 2024 Challenge
di: Yeow, Jun Wei, et al.
Pubblicazione: (2024)
di: Yeow, Jun Wei, et al.
Pubblicazione: (2024)
Tackling Cognitive Impairment Detection from Speech: A submission to the PROCESS Challenge
di: Botelho, Catarina, et al.
Pubblicazione: (2024)
di: Botelho, Catarina, et al.
Pubblicazione: (2024)
The DKU System for Multi-Speaker Automatic Speech Recognition in MLC-SLM Challenge
di: Lin, Yuke, et al.
Pubblicazione: (2025)
di: Lin, Yuke, et al.
Pubblicazione: (2025)
Advancing Speech Quality Assessment Through Scientific Challenges and Open-source Activities
di: Huang, Wen-Chin
Pubblicazione: (2025)
di: Huang, Wen-Chin
Pubblicazione: (2025)
The 1st SpeechWellness Challenge: Detecting Suicide Risk Among Adolescents
di: Wu, Wen, et al.
Pubblicazione: (2025)
di: Wu, Wen, et al.
Pubblicazione: (2025)
Speech Quality Embeddings for Improved Detection and Classification of Degradations in Speech Signals
di: Kuhlmann, Michael, et al.
Pubblicazione: (2026)
di: Kuhlmann, Michael, et al.
Pubblicazione: (2026)
Documenti analoghi
-
The Interspeech 2024 Challenge on Speech Processing Using Discrete Units
di: Chang, Xuankai, et al.
Pubblicazione: (2024) -
The VoiceMOS Challenge 2024: Beyond Speech Quality Prediction
di: Huang, Wen-Chin, et al.
Pubblicazione: (2024) -
Recomposed realities: animating still images via patch clustering and randomness
di: Juvonen, Markus, et al.
Pubblicazione: (2025) -
UTDUSS: UTokyo-SaruLab System for Interspeech2024 Speech Processing Using Discrete Speech Unit Challenge
di: Nakata, Wataru, et al.
Pubblicazione: (2024) -
Findings of the 2024 Mandarin Stuttering Event Detection and Automatic Speech Recognition Challenge
di: Xue, Hongfei, et al.
Pubblicazione: (2024)