GIRAFE: Glottal Imaging Dataset for Advanced Segmentation, Analysis, and Facilitative Playbacks Evaluation
Fuente:
arXiv
Salvato in:
| Autori principali: | Andrade-Miranda, G., Chatzipapas, K., Arias-Londoño, J. D., Godino-Llorente, J. I. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Singing Voice Synthesis Using Differentiable LPC and Glottal-Flow-Inspired Wavetables
di: Yu, Chin-Yun, et al.
Pubblicazione: (2023)
di: Yu, Chin-Yun, et al.
Pubblicazione: (2023)
SNC: A Stem-Native Codec for Efficient Lossless Audio Storage with Adaptive Playback Capabilities
di: Sufi, Shaad
Pubblicazione: (2026)
di: Sufi, Shaad
Pubblicazione: (2026)
Modeling and Estimation of Vocal Tract and Glottal Source Parameters Using ARMAX-LF Model
di: Lia, Kai, et al.
Pubblicazione: (2024)
di: Lia, Kai, et al.
Pubblicazione: (2024)
Del Visual al Auditivo: Sonorización de Escenas Guiada por Imagen
di: Sánchez, María, et al.
Pubblicazione: (2024)
di: Sánchez, María, et al.
Pubblicazione: (2024)
NeuroVoz: a Castillian Spanish corpus of parkinsonian speech
di: Mendes-Laureano, Janaína, et al.
Pubblicazione: (2024)
di: Mendes-Laureano, Janaína, et al.
Pubblicazione: (2024)
Construction and Analysis of Impression Caption Dataset for Environmental Sounds
di: Okamoto, Yuki, et al.
Pubblicazione: (2024)
di: Okamoto, Yuki, et al.
Pubblicazione: (2024)
Baseline Systems and Evaluation Metrics for Spatial Semantic Segmentation of Sound Scenes
di: Nguyen, Binh Thien, et al.
Pubblicazione: (2025)
di: Nguyen, Binh Thien, et al.
Pubblicazione: (2025)
Can Audio Reveal Music Performance Difficulty? Insights from the Piano Syllabus Dataset
di: Ramoneda, Pedro, et al.
Pubblicazione: (2024)
di: Ramoneda, Pedro, et al.
Pubblicazione: (2024)
Unseen but not Unknown: Using Dataset Concealment to Robustly Evaluate Speech Quality Estimation Models
di: Pieper, Jaden, et al.
Pubblicazione: (2026)
di: Pieper, Jaden, et al.
Pubblicazione: (2026)
DFADD: The Diffusion and Flow-Matching Based Audio Deepfake Dataset
di: Du, Jiawei, et al.
Pubblicazione: (2024)
di: Du, Jiawei, et al.
Pubblicazione: (2024)
MusicEval: A Generative Music Dataset with Expert Ratings for Automatic Text-to-Music Evaluation
di: Liu, Cheng, et al.
Pubblicazione: (2025)
di: Liu, Cheng, et al.
Pubblicazione: (2025)
SimuSOE: A Simulated Snoring Dataset for Obstructive Sleep Apnea-Hypopnea Syndrome Evaluation during Wakefulness
di: Lin, Jie, et al.
Pubblicazione: (2024)
di: Lin, Jie, et al.
Pubblicazione: (2024)
CodecFake+: A Large-Scale Neural Audio Codec-Based Deepfake Speech Dataset
di: Chen, Xuanjun, et al.
Pubblicazione: (2025)
di: Chen, Xuanjun, et al.
Pubblicazione: (2025)
Toward Multimodal Industrial Fault Analysis: A Single-Speed Chain Conveyor Dataset with Audio and Vibration Signals
di: Chen, Zhang, et al.
Pubblicazione: (2026)
di: Chen, Zhang, et al.
Pubblicazione: (2026)
An Extensive Analysis of the Singing Voice Conversion Challenge 2025 Evaluation Results
di: Violeta, Lester Phillip, et al.
Pubblicazione: (2025)
di: Violeta, Lester Phillip, et al.
Pubblicazione: (2025)
Comparative Evaluation of Acoustic Feature Extraction Tools for Clinical Speech Analysis
di: Choi, Anna Seo Gyeong, et al.
Pubblicazione: (2025)
di: Choi, Anna Seo Gyeong, et al.
Pubblicazione: (2025)
Facilitating deep acoustic phenotyping: A basic coding scheme of infant vocalisations preluding computational analysis, machine learning and clinical reasoning
di: Kulvicius, Tomas, et al.
Pubblicazione: (2023)
di: Kulvicius, Tomas, et al.
Pubblicazione: (2023)
Evaluating Parkinson's Disease Detection in Anonymized Speech: A Performance and Acoustic Analysis
di: Franzreb, Carlos, et al.
Pubblicazione: (2026)
di: Franzreb, Carlos, et al.
Pubblicazione: (2026)
ASPED: An Audio Dataset for Detecting Pedestrians
di: Seshadri, Pavan, et al.
Pubblicazione: (2023)
di: Seshadri, Pavan, et al.
Pubblicazione: (2023)
STraDa: A Singer Traits Dataset
di: Kong, Yuexuan, et al.
Pubblicazione: (2024)
di: Kong, Yuexuan, et al.
Pubblicazione: (2024)
Mamba-based Segmentation Model for Speaker Diarization
di: Plaquet, Alexis, et al.
Pubblicazione: (2024)
di: Plaquet, Alexis, et al.
Pubblicazione: (2024)
A Dataset for Automatic Assessment of TTS Quality in Spanish
di: Welford, Alejandro Sosa, et al.
Pubblicazione: (2025)
di: Welford, Alejandro Sosa, et al.
Pubblicazione: (2025)
Audio-Language Datasets of Scenes and Events: A Survey
di: Wijngaard, Gijs, et al.
Pubblicazione: (2024)
di: Wijngaard, Gijs, et al.
Pubblicazione: (2024)
CUEMPATHY: A Counseling Speech Dataset for Psychotherapy Research
di: Tao, Dehua, et al.
Pubblicazione: (2024)
di: Tao, Dehua, et al.
Pubblicazione: (2024)
Dataset-Distillation Generative Model for Speech Emotion Recognition
di: Ritter-Gutierrez, Fabian, et al.
Pubblicazione: (2024)
di: Ritter-Gutierrez, Fabian, et al.
Pubblicazione: (2024)
MLAAD: The Multi-Language Audio Anti-Spoofing Dataset
di: Müller, Nicolas M., et al.
Pubblicazione: (2024)
di: Müller, Nicolas M., et al.
Pubblicazione: (2024)
Vision Transformer Segmentation for Visual Bird Sound Denoising
di: Kumar, Sahil, et al.
Pubblicazione: (2024)
di: Kumar, Sahil, et al.
Pubblicazione: (2024)
SoundCollage: Automated Discovery of New Classes in Audio Datasets
di: Choi, Ryuhaerang, et al.
Pubblicazione: (2024)
di: Choi, Ryuhaerang, et al.
Pubblicazione: (2024)
UrBAN: Urban Beehive Acoustics and PheNotyping Dataset
di: Abdollahi, Mahsa, et al.
Pubblicazione: (2024)
di: Abdollahi, Mahsa, et al.
Pubblicazione: (2024)
EmoFake: An Initial Dataset for Emotion Fake Audio Detection
di: Zhao, Yan, et al.
Pubblicazione: (2022)
di: Zhao, Yan, et al.
Pubblicazione: (2022)
The Florence Price Art Song Dataset and Piano Accompaniment Generator
di: He, Tao-Tao, et al.
Pubblicazione: (2025)
di: He, Tao-Tao, et al.
Pubblicazione: (2025)
Binamix -- A Python Library for Generating Binaural Audio Datasets
di: Barry, Dan, et al.
Pubblicazione: (2025)
di: Barry, Dan, et al.
Pubblicazione: (2025)
ICSD: An Open-source Dataset for Infant Cry and Snoring Detection
di: Liu, Qingyu, et al.
Pubblicazione: (2024)
di: Liu, Qingyu, et al.
Pubblicazione: (2024)
The Extended SONICOM HRTF Dataset and Spatial Audio Metrics Toolbox
di: Poole, Katarina C., et al.
Pubblicazione: (2025)
di: Poole, Katarina C., et al.
Pubblicazione: (2025)
Multi-Utterance Speech Separation and Association Trained on Short Segments
di: Wang, Yuzhu, et al.
Pubblicazione: (2025)
di: Wang, Yuzhu, et al.
Pubblicazione: (2025)
Dissecting the Segmentation Model of End-to-End Diarization with Vector Clustering
di: Plaquet, Alexis, et al.
Pubblicazione: (2025)
di: Plaquet, Alexis, et al.
Pubblicazione: (2025)
Advances in Speech Separation: Techniques, Challenges, and Future Trends
di: Li, Kai, et al.
Pubblicazione: (2025)
di: Li, Kai, et al.
Pubblicazione: (2025)
ASAudio: A Survey of Advanced Spatial Audio Research
di: Zhu, Zhiyuan, et al.
Pubblicazione: (2025)
di: Zhu, Zhiyuan, et al.
Pubblicazione: (2025)
Advancing Continual Learning for Robust Deepfake Audio Classification
di: Dong, Feiyi, et al.
Pubblicazione: (2024)
di: Dong, Feiyi, et al.
Pubblicazione: (2024)
Comparative Analysis of ASR Methods for Speech Deepfake Detection
di: Salvi, Davide, et al.
Pubblicazione: (2024)
di: Salvi, Davide, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Singing Voice Synthesis Using Differentiable LPC and Glottal-Flow-Inspired Wavetables
di: Yu, Chin-Yun, et al.
Pubblicazione: (2023) -
SNC: A Stem-Native Codec for Efficient Lossless Audio Storage with Adaptive Playback Capabilities
di: Sufi, Shaad
Pubblicazione: (2026) -
Modeling and Estimation of Vocal Tract and Glottal Source Parameters Using ARMAX-LF Model
di: Lia, Kai, et al.
Pubblicazione: (2024) -
Del Visual al Auditivo: Sonorización de Escenas Guiada por Imagen
di: Sánchez, María, et al.
Pubblicazione: (2024) -
NeuroVoz: a Castillian Spanish corpus of parkinsonian speech
di: Mendes-Laureano, Janaína, et al.
Pubblicazione: (2024)