MADUV: The 1st INTERSPEECH Mice Autism Detection via Ultrasound Vocalization Challenge
Fuente:
arXiv
Guardado en:
| Autores principales: | Yang, Zijiang, Song, Meishu, Jing, Xin, Zhang, Haojie, Qian, Kun, Hu, Bin, Tamada, Kota, Takumi, Toru, Schuller, Björn W., Yamamoto, Yoshiharu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Enhancing Emotional Text-to-Speech Controllability with Natural Language Guidance through Contrastive Learning and Diffusion Models
por: Jing, Xin, et al.
Publicado: (2024)
por: Jing, Xin, et al.
Publicado: (2024)
Enhancing Child Vocalization Classification with Phonetically-Tuned Embeddings for Assisting Autism Diagnosis
por: Li, Jialu, et al.
Publicado: (2023)
por: Li, Jialu, et al.
Publicado: (2023)
RMVPE: A Robust Model for Vocal Pitch Estimation in Polyphonic Music
por: Wei, Haojie, et al.
Publicado: (2023)
por: Wei, Haojie, et al.
Publicado: (2023)
ParaCLAP -- Towards a general language-audio model for computational paralinguistic tasks
por: Jing, Xin, et al.
Publicado: (2024)
por: Jing, Xin, et al.
Publicado: (2024)
Charting 15 years of progress in deep learning for speech emotion recognition: A replication study
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2025)
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2025)
DJCM: A Deep Joint Cascade Model for Singing Voice Separation and Vocal Pitch Estimation
por: Wei, Haojie, et al.
Publicado: (2024)
por: Wei, Haojie, et al.
Publicado: (2024)
Audio-based Step-count Estimation for Running -- Windowing and Neural Network Baselines
por: Wagner, Philipp, et al.
Publicado: (2024)
por: Wagner, Philipp, et al.
Publicado: (2024)
Quantifying Dimensional Independence in Speech: An Information-Theoretic Framework for Disentangled Representation Learning
por: Kashyap, Bipasha, et al.
Publicado: (2026)
por: Kashyap, Bipasha, et al.
Publicado: (2026)
From Audio Deepfake Detection to AI-Generated Music Detection -- A Pathway and Overview
por: Li, Yupei, et al.
Publicado: (2024)
por: Li, Yupei, et al.
Publicado: (2024)
Intelligent Cardiac Auscultation for Murmur Detection via Parallel-Attentive Models with Uncertainty Estimation
por: Zhang, Zixing, et al.
Publicado: (2024)
por: Zhang, Zixing, et al.
Publicado: (2024)
Cross-Dialect Bird Species Recognition with Dialect-Calibrated Augmentation
por: Ding, Jiani, et al.
Publicado: (2025)
por: Ding, Jiani, et al.
Publicado: (2025)
Abusive Speech Detection in Indic Languages Using Acoustic Features
por: Spiesberger, Anika A., et al.
Publicado: (2024)
por: Spiesberger, Anika A., et al.
Publicado: (2024)
Can Large Language Models Aid in Annotating Speech Emotional Data? Uncovering New Frontiers
por: Latif, Siddique, et al.
Publicado: (2023)
por: Latif, Siddique, et al.
Publicado: (2023)
An automatic analysis of ultrasound vocalisations for the prediction of interaction context in captive Egyptian fruit bats
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024)
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024)
Mel-RoFormer for Vocal Separation and Vocal Melody Transcription
por: Wang, Ju-Chiang, et al.
Publicado: (2024)
por: Wang, Ju-Chiang, et al.
Publicado: (2024)
Explainable Detection of Machine Generated Music and Early Systematic Evaluation
por: Li, Yupei, et al.
Publicado: (2024)
por: Li, Yupei, et al.
Publicado: (2024)
A Comprehensive Survey on Heart Sound Analysis in the Deep Learning Era
por: Ren, Zhao, et al.
Publicado: (2023)
por: Ren, Zhao, et al.
Publicado: (2023)
Computer Audition: From Task-Specific Machine Learning to Foundation Models
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024)
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2024)
M6: Multi-generator, Multi-domain, Multi-lingual and cultural, Multi-genres, Multi-instrument Machine-Generated Music Detection Databases
por: Li, Yupei, et al.
Publicado: (2024)
por: Li, Yupei, et al.
Publicado: (2024)
DOTA-ME-CS: Daily Oriented Text Audio-Mandarin English-Code Switching Dataset
por: Li, Yupei, et al.
Publicado: (2025)
por: Li, Yupei, et al.
Publicado: (2025)
Testing Correctness, Fairness, and Robustness of Speech Emotion Recognition Models
por: Derington, Anna, et al.
Publicado: (2023)
por: Derington, Anna, et al.
Publicado: (2023)
Using voice analysis as an early indicator of risk for depression in young adults
por: Scherer, Klaus R., et al.
Publicado: (2024)
por: Scherer, Klaus R., et al.
Publicado: (2024)
CVSM: Contrastive Vocal Similarity Modeling
por: Garoufis, Christos, et al.
Publicado: (2025)
por: Garoufis, Christos, et al.
Publicado: (2025)
The 1st SpeechWellness Challenge: Detecting Suicide Risk Among Adolescents
por: Wu, Wen, et al.
Publicado: (2025)
por: Wu, Wen, et al.
Publicado: (2025)
Breaking Resource Barriers in Speech Emotion Recognition via Data Distillation
por: Chang, Yi, et al.
Publicado: (2024)
por: Chang, Yi, et al.
Publicado: (2024)
Emotion-Aware Contrastive Adaptation Network for Source-Free Cross-Corpus Speech Emotion Recognition
por: Zhao, Yan, et al.
Publicado: (2024)
por: Zhao, Yan, et al.
Publicado: (2024)
SmoothCLAP: Soft-Target Enhanced Contrastive Language\--Audio Pretraining for Affective Computing
por: Jing, Xin, et al.
Publicado: (2026)
por: Jing, Xin, et al.
Publicado: (2026)
Audio Explanation Synthesis with Generative Foundation Models
por: Akman, Alican, et al.
Publicado: (2024)
por: Akman, Alican, et al.
Publicado: (2024)
Wav2Small: Distilling Wav2Vec2 to 72K parameters for Low-Resource Speech emotion recognition
por: Kounadis-Bastian, Dionyssos, et al.
Publicado: (2024)
por: Kounadis-Bastian, Dionyssos, et al.
Publicado: (2024)
Melodic and Metrical Elements of Expressiveness in Hindustani Vocal Music
por: Bhake, Yash, et al.
Publicado: (2025)
por: Bhake, Yash, et al.
Publicado: (2025)
Auditory Representation Effective for Estimating Vocal Tract Information
por: Irino, Toshio, et al.
Publicado: (2023)
por: Irino, Toshio, et al.
Publicado: (2023)
A Reliable and Efficient Detection Pipeline for Rodent Ultrasonic Vocalizations
por: Anis, Sabah Shahnoor, et al.
Publicado: (2025)
por: Anis, Sabah Shahnoor, et al.
Publicado: (2025)
Biodenoising: Animal Vocalization Denoising without Access to Clean Data
por: Miron, Marius, et al.
Publicado: (2024)
por: Miron, Marius, et al.
Publicado: (2024)
Drum-to-Vocal Percussion Sound Conversion and Its Evaluation Methodology
por: Nobukawa, Rinka, et al.
Publicado: (2025)
por: Nobukawa, Rinka, et al.
Publicado: (2025)
1st Place Solution to Odyssey Emotion Recognition Challenge Task1: Tackling Class Imbalance Problem
por: Chen, Mingjie, et al.
Publicado: (2024)
por: Chen, Mingjie, et al.
Publicado: (2024)
ProsodyFM: Unsupervised Phrasing and Intonation Control for Intelligible Speech Synthesis
por: He, Xiangheng, et al.
Publicado: (2024)
por: He, Xiangheng, et al.
Publicado: (2024)
Domain Adapting Deep Reinforcement Learning for Real-world Speech Emotion Recognition
por: Rajapakshe, Thejan, et al.
Publicado: (2022)
por: Rajapakshe, Thejan, et al.
Publicado: (2022)
The T12 System for AudioMOS Challenge 2025: Audio Aesthetics Score Prediction System Using KAN- and VERSA-based Models
por: Yamamoto, Katsuhiko, et al.
Publicado: (2025)
por: Yamamoto, Katsuhiko, et al.
Publicado: (2025)
Analysis of Self-Supervised Speech Models on Children's Speech and Infant Vocalizations
por: Li, Jialu, et al.
Publicado: (2024)
por: Li, Jialu, et al.
Publicado: (2024)
voc2vec: A Foundation Model for Non-Verbal Vocalization
por: Koudounas, Alkis, et al.
Publicado: (2025)
por: Koudounas, Alkis, et al.
Publicado: (2025)
Ejemplares similares
-
Enhancing Emotional Text-to-Speech Controllability with Natural Language Guidance through Contrastive Learning and Diffusion Models
por: Jing, Xin, et al.
Publicado: (2024) -
Enhancing Child Vocalization Classification with Phonetically-Tuned Embeddings for Assisting Autism Diagnosis
por: Li, Jialu, et al.
Publicado: (2023) -
RMVPE: A Robust Model for Vocal Pitch Estimation in Polyphonic Music
por: Wei, Haojie, et al.
Publicado: (2023) -
ParaCLAP -- Towards a general language-audio model for computational paralinguistic tasks
por: Jing, Xin, et al.
Publicado: (2024) -
Charting 15 years of progress in deep learning for speech emotion recognition: A replication study
por: Triantafyllopoulos, Andreas, et al.
Publicado: (2025)