CryCeleb: A Speaker Verification Dataset Based on Infant Cry Sounds
Fuente:
arXiv
Salvato in:
| Autori principali: | Budaghyan, David, Onu, Charles C., Gorin, Arsenii, Subakan, Cem, Precup, Doina |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ICSD: An Open-source Dataset for Infant Cry and Snoring Detection
di: Liu, Qingyu, et al.
Pubblicazione: (2024)
di: Liu, Qingyu, et al.
Pubblicazione: (2024)
Enhancing Infant Crying Detection with Gradient Boosting for Improved Emotional and Mental Health Diagnostics
di: Lee, Kyunghun, et al.
Pubblicazione: (2024)
di: Lee, Kyunghun, et al.
Pubblicazione: (2024)
ALAS: Measuring Latent Speech-Text Alignment For Spoken Language Understanding In Multimodal LLMs
di: Mousavi, Pooneh, et al.
Pubblicazione: (2025)
di: Mousavi, Pooneh, et al.
Pubblicazione: (2025)
LMU-Based Sequential Learning and Posterior Ensemble Fusion for Cross-Domain Infant Cry Classification
di: Jazaeri, Niloofar, et al.
Pubblicazione: (2026)
di: Jazaeri, Niloofar, et al.
Pubblicazione: (2026)
Focal Modulation Networks for Interpretable Sound Classification
di: Della Libera, Luca, et al.
Pubblicazione: (2024)
di: Della Libera, Luca, et al.
Pubblicazione: (2024)
Making deep neural networks work for medical audio: representation, compression and domain adaptation
di: Onu, Charles C
Pubblicazione: (2025)
di: Onu, Charles C
Pubblicazione: (2025)
InfantCryNet: A Data-driven Framework for Intelligent Analysis of Infant Cries
di: Hong, Mengze, et al.
Pubblicazione: (2024)
di: Hong, Mengze, et al.
Pubblicazione: (2024)
Listen First, Then Answer: Timestamp-Grounded Speech Reasoning
di: Jeong, Jihoon, et al.
Pubblicazione: (2026)
di: Jeong, Jihoon, et al.
Pubblicazione: (2026)
LL-SDR: Low-Latency Speech enhancement through Discrete Representations
di: Li, Jingyi, et al.
Pubblicazione: (2026)
di: Li, Jingyi, et al.
Pubblicazione: (2026)
HeightCeleb - an enrichment of VoxCeleb dataset with speaker height information
di: Kacprzak, Stanisław, et al.
Pubblicazione: (2024)
di: Kacprzak, Stanisław, et al.
Pubblicazione: (2024)
The VoxCeleb Speaker Recognition Challenge: A Retrospective
di: Huh, Jaesung, et al.
Pubblicazione: (2024)
di: Huh, Jaesung, et al.
Pubblicazione: (2024)
Phoneme Discretized Saliency Maps for Explainable Detection of AI-Generated Voice
di: Gupta, Shubham, et al.
Pubblicazione: (2024)
di: Gupta, Shubham, et al.
Pubblicazione: (2024)
Diffusion-Based Adversarial Purification for Speaker Verification
di: Bai, Yibo, et al.
Pubblicazione: (2023)
di: Bai, Yibo, et al.
Pubblicazione: (2023)
Analysis of Speech Temporal Dynamics in the Context of Speaker Verification and Voice Anonymization
di: Tomashenko, Natalia, et al.
Pubblicazione: (2024)
di: Tomashenko, Natalia, et al.
Pubblicazione: (2024)
AdvSV: An Over-the-Air Adversarial Attack Dataset for Speaker Verification
di: Wang, Li, et al.
Pubblicazione: (2023)
di: Wang, Li, et al.
Pubblicazione: (2023)
Learning Emotion-Invariant Speaker Representations for Speaker Verification
di: Tian, Jingguang, et al.
Pubblicazione: (2025)
di: Tian, Jingguang, et al.
Pubblicazione: (2025)
Autoregressive Speech Enhancement via Acoustic Tokens
di: Della Libera, Luca, et al.
Pubblicazione: (2025)
di: Della Libera, Luca, et al.
Pubblicazione: (2025)
Disentangling Speaker Traits for Deepfake Source Verification via Chebyshev Polynomial and Riemannian Metric Learning
di: Xuan, Xi, et al.
Pubblicazione: (2026)
di: Xuan, Xi, et al.
Pubblicazione: (2026)
Listenable Maps for Audio Classifiers
di: Paissan, Francesco, et al.
Pubblicazione: (2024)
di: Paissan, Francesco, et al.
Pubblicazione: (2024)
Adversarial Reweighting for Speaker Verification Fairness
di: Jin, Minho, et al.
Pubblicazione: (2022)
di: Jin, Minho, et al.
Pubblicazione: (2022)
TidyVoice: A Curated Multilingual Dataset for Speaker Verification Derived from Common Voice
di: Farhadipour, Aref, et al.
Pubblicazione: (2026)
di: Farhadipour, Aref, et al.
Pubblicazione: (2026)
LightCAM: A Fast and Light Implementation of Context-Aware Masking based D-TDNN for Speaker Verification
di: Cao, Di, et al.
Pubblicazione: (2024)
di: Cao, Di, et al.
Pubblicazione: (2024)
Emphasized Non-Target Speaker Knowledge in Knowledge Distillation for Automatic Speaker Verification
di: Truong, Duc-Tuan, et al.
Pubblicazione: (2023)
di: Truong, Duc-Tuan, et al.
Pubblicazione: (2023)
PhiNet: Speaker Verification with Phonetic Interpretability
di: Ma, Yi, et al.
Pubblicazione: (2026)
di: Ma, Yi, et al.
Pubblicazione: (2026)
Phonetic Richness for Improved Automatic Speaker Verification
di: Klein, Nicholas, et al.
Pubblicazione: (2024)
di: Klein, Nicholas, et al.
Pubblicazione: (2024)
VoxVietnam: a Large-Scale Multi-Genre Dataset for Vietnamese Speaker Recognition
di: Vu, Hoang Long, et al.
Pubblicazione: (2024)
di: Vu, Hoang Long, et al.
Pubblicazione: (2024)
Investigation of Speaker Representation for Target-Speaker Speech Processing
di: Ashihara, Takanori, et al.
Pubblicazione: (2024)
di: Ashihara, Takanori, et al.
Pubblicazione: (2024)
Robust Training for Speaker Verification against Noisy Labels
di: Fang, Zhihua, et al.
Pubblicazione: (2022)
di: Fang, Zhihua, et al.
Pubblicazione: (2022)
Enhancing Age-Related Robustness in Children Speaker Verification
di: Shetty, Vishwas M., et al.
Pubblicazione: (2025)
di: Shetty, Vishwas M., et al.
Pubblicazione: (2025)
MASV: Speaker Verification with Global and Local Context Mamba
di: Liu, Yang, et al.
Pubblicazione: (2024)
di: Liu, Yang, et al.
Pubblicazione: (2024)
Spectral-Aware Low-Rank Adaptation for Speaker Verification
di: Li, Zhe, et al.
Pubblicazione: (2025)
di: Li, Zhe, et al.
Pubblicazione: (2025)
How Should We Extract Discrete Audio Tokens from Self-Supervised Models?
di: Mousavi, Pooneh, et al.
Pubblicazione: (2024)
di: Mousavi, Pooneh, et al.
Pubblicazione: (2024)
A Probabilistic Fusion Framework for Spoofing Aware Speaker Verification
di: Zhang, You, et al.
Pubblicazione: (2022)
di: Zhang, You, et al.
Pubblicazione: (2022)
An Investigation of Reprogramming for Cross-Language Adaptation in Speaker Verification Systems
di: Li, Jingyu, et al.
Pubblicazione: (2024)
di: Li, Jingyu, et al.
Pubblicazione: (2024)
Neural Codec-based Adversarial Sample Detection for Speaker Verification
di: Chen, Xuanjun, et al.
Pubblicazione: (2024)
di: Chen, Xuanjun, et al.
Pubblicazione: (2024)
Study on the Fairness of Speaker Verification Systems on Underrepresented Accents in English
di: Estevez, Mariel, et al.
Pubblicazione: (2022)
di: Estevez, Mariel, et al.
Pubblicazione: (2022)
Integrated Multi-Level Knowledge Distillation for Enhanced Speaker Verification
di: Yang, Wenhao, et al.
Pubblicazione: (2024)
di: Yang, Wenhao, et al.
Pubblicazione: (2024)
Neighborhood Attention Transformer with Progressive Channel Fusion for Speaker Verification
di: Li, Nian, et al.
Pubblicazione: (2024)
di: Li, Nian, et al.
Pubblicazione: (2024)
Self-Supervised Syllable Discovery Based on Speaker-Disentangled HuBERT
di: Komatsu, Ryota, et al.
Pubblicazione: (2024)
di: Komatsu, Ryota, et al.
Pubblicazione: (2024)
Speaker-Reasoner: Scaling Interaction Turns and Reasoning Patterns for Timestamped Speaker-Attributed ASR
di: Lin, Zhennan, et al.
Pubblicazione: (2026)
di: Lin, Zhennan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
ICSD: An Open-source Dataset for Infant Cry and Snoring Detection
di: Liu, Qingyu, et al.
Pubblicazione: (2024) -
Enhancing Infant Crying Detection with Gradient Boosting for Improved Emotional and Mental Health Diagnostics
di: Lee, Kyunghun, et al.
Pubblicazione: (2024) -
ALAS: Measuring Latent Speech-Text Alignment For Spoken Language Understanding In Multimodal LLMs
di: Mousavi, Pooneh, et al.
Pubblicazione: (2025) -
LMU-Based Sequential Learning and Posterior Ensemble Fusion for Cross-Domain Infant Cry Classification
di: Jazaeri, Niloofar, et al.
Pubblicazione: (2026) -
Focal Modulation Networks for Interpretable Sound Classification
di: Della Libera, Luca, et al.
Pubblicazione: (2024)