STraDa: A Singer Traits Dataset
Fuente:
arXiv
Guardado en:
| Autores principales: | Kong, Yuexuan, Tran, Viet-Anh, Hennequin, Romain |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
S-KEY: Self-supervised Learning of Major and Minor Keys from Audio
por: Kong, Yuexuan, et al.
Publicado: (2025)
por: Kong, Yuexuan, et al.
Publicado: (2025)
STONE: Self-supervised Tonality Estimator
por: Kong, Yuexuan, et al.
Publicado: (2024)
por: Kong, Yuexuan, et al.
Publicado: (2024)
From Real to Cloned Singer Identification
por: Desblancs, Dorian, et al.
Publicado: (2024)
por: Desblancs, Dorian, et al.
Publicado: (2024)
Emergent musical properties of a transformer under contrastive self-supervised learning
por: Kong, Yuexuan, et al.
Publicado: (2025)
por: Kong, Yuexuan, et al.
Publicado: (2025)
AI-Generated Music Detection and its Challenges
por: Afchar, Darius, et al.
Publicado: (2025)
por: Afchar, Darius, et al.
Publicado: (2025)
SingIt! Singer Voice Transformation
por: Eliav, Amit, et al.
Publicado: (2024)
por: Eliav, Amit, et al.
Publicado: (2024)
Synthetic Singers: A Review of Deep-Learning-based Singing Voice Synthesis Approaches
por: Pan, Changhao, et al.
Publicado: (2026)
por: Pan, Changhao, et al.
Publicado: (2026)
Singer separation for karaoke content generation
por: Lin, Hsuan-Yu, et al.
Publicado: (2021)
por: Lin, Hsuan-Yu, et al.
Publicado: (2021)
An Experimental Comparison Of Multi-view Self-supervised Methods For Music Tagging
por: Meseguer-Brocal, Gabriel, et al.
Publicado: (2024)
por: Meseguer-Brocal, Gabriel, et al.
Publicado: (2024)
Detecting music deepfakes is easy but actually hard
por: Afchar, Darius, et al.
Publicado: (2024)
por: Afchar, Darius, et al.
Publicado: (2024)
VS-Singer: Vision-Guided Stereo Singing Voice Synthesis with Consistency Schrödinger Bridge
por: Zhao, Zijing, et al.
Publicado: (2025)
por: Zhao, Zijing, et al.
Publicado: (2025)
IdolSongsJp Corpus: A Multi-Singer Song Corpus in the Style of Japanese Idol Groups
por: Suda, Hitoshi, et al.
Publicado: (2025)
por: Suda, Hitoshi, et al.
Publicado: (2025)
MuSE-SVS: Multi-Singer Emotional Singing Voice Synthesizer that Controls Emotional Intensity
por: Kim, Sungjae, et al.
Publicado: (2022)
por: Kim, Sungjae, et al.
Publicado: (2022)
Reducing Geographic Disparities in Automatic Speech Recognition via Elastic Weight Consolidation
por: Trinh, Viet Anh, et al.
Publicado: (2022)
por: Trinh, Viet Anh, et al.
Publicado: (2022)
Period Singer: Integrating Periodic and Aperiodic Variational Autoencoders for Natural-Sounding End-to-End Singing Voice Synthesis
por: Kim, Taewoo, et al.
Publicado: (2024)
por: Kim, Taewoo, et al.
Publicado: (2024)
LDM-SVC: Latent Diffusion Model Based Zero-Shot Any-to-Any Singing Voice Conversion with Singer Guidance
por: Chen, Shihao, et al.
Publicado: (2024)
por: Chen, Shihao, et al.
Publicado: (2024)
YingMusic-Singer-Plus: Controllable Singing Voice Synthesis with Flexible Lyric Manipulation and Annotation-free Melody Guidance
por: Hao, Chunbo, et al.
Publicado: (2026)
por: Hao, Chunbo, et al.
Publicado: (2026)
Tracking of Intermittent and Moving Speakers : Dataset and Metrics
por: Iatariene, Taous, et al.
Publicado: (2025)
por: Iatariene, Taous, et al.
Publicado: (2025)
Towards High-Fidelity and Controllable Bioacoustic Generation via Enhanced Diffusion Learning
por: Song, Tianyu, et al.
Publicado: (2025)
por: Song, Tianyu, et al.
Publicado: (2025)
BiSinger: Bilingual Singing Voice Synthesis
por: Zhou, Huali, et al.
Publicado: (2023)
por: Zhou, Huali, et al.
Publicado: (2023)
Performance and energy balance: a comprehensive study of state-of-the-art sound event detection systems
por: Ronchini, Francesca, et al.
Publicado: (2023)
por: Ronchini, Francesca, et al.
Publicado: (2023)
LlamaPartialSpoof: An LLM-Driven Fake Speech Dataset Simulating Disinformation Generation
por: Luong, Hieu-Thi, et al.
Publicado: (2024)
por: Luong, Hieu-Thi, et al.
Publicado: (2024)
Vox-Profile: A Speech Foundation Model Benchmark for Characterizing Diverse Speaker and Speech Traits
por: Feng, Tiantian, et al.
Publicado: (2025)
por: Feng, Tiantian, et al.
Publicado: (2025)
METEOR: Melody-aware Texture-controllable Symbolic Orchestral Music Generation via Transformer VAE
por: Le, Dinh-Viet-Toan, et al.
Publicado: (2024)
por: Le, Dinh-Viet-Toan, et al.
Publicado: (2024)
A Phoneme-Scale Assessment of Multichannel Speech Enhancement Algorithms
por: Monir, Nasser-Eddine, et al.
Publicado: (2024)
por: Monir, Nasser-Eddine, et al.
Publicado: (2024)
PerformSinger: Multimodal Singing Voice Synthesis Leveraging Synchronized Lip Cues from Singing Performance Videos
por: Gu, Ke, et al.
Publicado: (2025)
por: Gu, Ke, et al.
Publicado: (2025)
Singer Identity Representation Learning using Self-Supervised Techniques
por: Torres, Bernardo, et al.
Publicado: (2024)
por: Torres, Bernardo, et al.
Publicado: (2024)
ChildMandarin: A Comprehensive Mandarin Speech Dataset for Young Children Aged 3-5
por: Zhou, Jiaming, et al.
Publicado: (2024)
por: Zhou, Jiaming, et al.
Publicado: (2024)
Angular Distance Distribution Loss for Audio Classification
por: Almudévar, Antonio, et al.
Publicado: (2024)
por: Almudévar, Antonio, et al.
Publicado: (2024)
Two-pass Endpoint Detection for Speech Recognition
por: Raju, Anirudh, et al.
Publicado: (2024)
por: Raju, Anirudh, et al.
Publicado: (2024)
StyleSinger: Style Transfer for Out-of-Domain Singing Voice Synthesis
por: Zhang, Yu, et al.
Publicado: (2023)
por: Zhang, Yu, et al.
Publicado: (2023)
Evaluating Multichannel Speech Enhancement Algorithms at the Phoneme Scale Across Genders
por: Monir, Nasser-Eddine, et al.
Publicado: (2025)
por: Monir, Nasser-Eddine, et al.
Publicado: (2025)
Echoes of Ideology: Toward an Audio Analysis Pipeline to Unveil Character Traits in Historical Nazi Propaganda Films
por: Ruth, Nicolas, et al.
Publicado: (2026)
por: Ruth, Nicolas, et al.
Publicado: (2026)
Domain-Invariant Representation Learning of Bird Sounds
por: Moummad, Ilyass, et al.
Publicado: (2024)
por: Moummad, Ilyass, et al.
Publicado: (2024)
Production and Manufacturing of 3D Printed Acoustic Guitars
por: Tran, Timothy, et al.
Publicado: (2025)
por: Tran, Timothy, et al.
Publicado: (2025)
A decade of DCASE: Achievements, practices, evaluations and future challenges
por: Mesaros, Annamaria, et al.
Publicado: (2024)
por: Mesaros, Annamaria, et al.
Publicado: (2024)
ConSinger: Efficient High-Fidelity Singing Voice Generation with Minimal Steps
por: Song, Yulin, et al.
Publicado: (2024)
por: Song, Yulin, et al.
Publicado: (2024)
Optimal Scalogram for Computational Complexity Reduction in Acoustic Recognition Using Deep Learning
por: Phan, Dang Thoai, et al.
Publicado: (2025)
por: Phan, Dang Thoai, et al.
Publicado: (2025)
Audio-Language Datasets of Scenes and Events: A Survey
por: Wijngaard, Gijs, et al.
Publicado: (2024)
por: Wijngaard, Gijs, et al.
Publicado: (2024)
CUEMPATHY: A Counseling Speech Dataset for Psychotherapy Research
por: Tao, Dehua, et al.
Publicado: (2024)
por: Tao, Dehua, et al.
Publicado: (2024)
Ejemplares similares
-
S-KEY: Self-supervised Learning of Major and Minor Keys from Audio
por: Kong, Yuexuan, et al.
Publicado: (2025) -
STONE: Self-supervised Tonality Estimator
por: Kong, Yuexuan, et al.
Publicado: (2024) -
From Real to Cloned Singer Identification
por: Desblancs, Dorian, et al.
Publicado: (2024) -
Emergent musical properties of a transformer under contrastive self-supervised learning
por: Kong, Yuexuan, et al.
Publicado: (2025) -
AI-Generated Music Detection and its Challenges
por: Afchar, Darius, et al.
Publicado: (2025)