Adversarial Multi-Task Learning for Disentangling Timbre and Pitch in Singing Voice Synthesis
Fuente:
arXiv
Salvato in:
| Autori principali: | Kim, Tae-Woo, Kang, Min-Su, Lee, Gyeong-Hoon |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2022
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CONTUNER: Singing Voice Beautifying with Pitch and Expressiveness Condition
di: Wang, Jianzong, et al.
Pubblicazione: (2024)
di: Wang, Jianzong, et al.
Pubblicazione: (2024)
UNMIXX: Untangling Highly Correlated Singing Voices Mixtures
di: Jung, Jihoo, et al.
Pubblicazione: (2026)
di: Jung, Jihoo, et al.
Pubblicazione: (2026)
TokSing: Singing Voice Synthesis based on Discrete Tokens
di: Wu, Yuning, et al.
Pubblicazione: (2024)
di: Wu, Yuning, et al.
Pubblicazione: (2024)
DJCM: A Deep Joint Cascade Model for Singing Voice Separation and Vocal Pitch Estimation
di: Wei, Haojie, et al.
Pubblicazione: (2024)
di: Wei, Haojie, et al.
Pubblicazione: (2024)
Binaural Sound Event Localization and Detection Neural Network based on HRTF Localization Cues for Humanoid Robots
di: Lee, Gyeong-Tae
Pubblicazione: (2025)
di: Lee, Gyeong-Tae
Pubblicazione: (2025)
Robust Singing Voice Transcription Serves Synthesis
di: Li, Ruiqi, et al.
Pubblicazione: (2024)
di: Li, Ruiqi, et al.
Pubblicazione: (2024)
Everyone-Can-Sing: Zero-Shot Singing Voice Synthesis and Conversion with Speech Reference
di: Dai, Shuqi, et al.
Pubblicazione: (2025)
di: Dai, Shuqi, et al.
Pubblicazione: (2025)
Period Singer: Integrating Periodic and Aperiodic Variational Autoencoders for Natural-Sounding End-to-End Singing Voice Synthesis
di: Kim, Taewoo, et al.
Pubblicazione: (2024)
di: Kim, Taewoo, et al.
Pubblicazione: (2024)
SPA-SVC: Self-supervised Pitch Augmentation for Singing Voice Conversion
di: Bai, Bingsong, et al.
Pubblicazione: (2024)
di: Bai, Bingsong, et al.
Pubblicazione: (2024)
Jointly Recognizing Speech and Singing Voices Based on Multi-Task Audio Source Separation
di: Bai, Ye, et al.
Pubblicazione: (2024)
di: Bai, Ye, et al.
Pubblicazione: (2024)
MuSE-SVS: Multi-Singer Emotional Singing Voice Synthesizer that Controls Emotional Intensity
di: Kim, Sungjae, et al.
Pubblicazione: (2022)
di: Kim, Sungjae, et al.
Pubblicazione: (2022)
QvTAD: Differential Relative Attribute Learning for Voice Timbre Attribute Detection
di: Wu, Zhiyu, et al.
Pubblicazione: (2025)
di: Wu, Zhiyu, et al.
Pubblicazione: (2025)
Singing Voice Graph Modeling for SingFake Detection
di: Chen, Xuanjun, et al.
Pubblicazione: (2024)
di: Chen, Xuanjun, et al.
Pubblicazione: (2024)
SingIt! Singer Voice Transformation
di: Eliav, Amit, et al.
Pubblicazione: (2024)
di: Eliav, Amit, et al.
Pubblicazione: (2024)
Synthetic Singers: A Review of Deep-Learning-based Singing Voice Synthesis Approaches
di: Pan, Changhao, et al.
Pubblicazione: (2026)
di: Pan, Changhao, et al.
Pubblicazione: (2026)
RobustSVC: HuBERT-based Melody Extractor and Adversarial Learning for Robust Singing Voice Conversion
di: Chen, Wei, et al.
Pubblicazione: (2024)
di: Chen, Wei, et al.
Pubblicazione: (2024)
Mitigating Latent Mismatch in cVAE-Based Singing Voice Synthesis via Flow Matching
di: Yun, Minhyeok, et al.
Pubblicazione: (2026)
di: Yun, Minhyeok, et al.
Pubblicazione: (2026)
VISinger2+: End-to-End Singing Voice Synthesis Augmented by Self-Supervised Learning Representation
di: Yu, Yifeng, et al.
Pubblicazione: (2024)
di: Yu, Yifeng, et al.
Pubblicazione: (2024)
Pitch-and-Spectrum-Aware Singing Quality Assessment with Bias Correction and Model Fusion
di: Shi, Yu-Fei, et al.
Pubblicazione: (2024)
di: Shi, Yu-Fei, et al.
Pubblicazione: (2024)
Neural Concatenative Singing Voice Conversion: Rethinking Concatenation-Based Approach for One-Shot Singing Voice Conversion
di: Sha, Binzhu, et al.
Pubblicazione: (2023)
di: Sha, Binzhu, et al.
Pubblicazione: (2023)
Muskits-ESPnet: A Comprehensive Toolkit for Singing Voice Synthesis in New Paradigm
di: Wu, Yuning, et al.
Pubblicazione: (2024)
di: Wu, Yuning, et al.
Pubblicazione: (2024)
Singing Voice Synthesis Using Differentiable LPC and Glottal-Flow-Inspired Wavetables
di: Yu, Chin-Yun, et al.
Pubblicazione: (2023)
di: Yu, Chin-Yun, et al.
Pubblicazione: (2023)
SingMOS: An extensive Open-Source Singing Voice Dataset for MOS Prediction
di: Tang, Yuxun, et al.
Pubblicazione: (2024)
di: Tang, Yuxun, et al.
Pubblicazione: (2024)
SingVERSE: A Diverse, Real-World Benchmark for Singing Voice Enhancement
di: Jiang, Shaohan, et al.
Pubblicazione: (2025)
di: Jiang, Shaohan, et al.
Pubblicazione: (2025)
Self-Supervised Singing Voice Pre-Training towards Speech-to-Singing Conversion
di: Li, Ruiqi, et al.
Pubblicazione: (2024)
di: Li, Ruiqi, et al.
Pubblicazione: (2024)
InstructSing: High-Fidelity Singing Voice Generation via Instructing Yourself
di: Zeng, Chang, et al.
Pubblicazione: (2024)
di: Zeng, Chang, et al.
Pubblicazione: (2024)
DisMix: Disentangling Mixtures of Musical Instruments for Source-level Pitch and Timbre Manipulation
di: Luo, Yin-Jyun, et al.
Pubblicazione: (2024)
di: Luo, Yin-Jyun, et al.
Pubblicazione: (2024)
Singing Voice Data Scaling-up: An Introduction to ACE-Opencpop and ACE-KiSing
di: Shi, Jiatong, et al.
Pubblicazione: (2024)
di: Shi, Jiatong, et al.
Pubblicazione: (2024)
SingNet: Towards a Large-Scale, Diverse, and In-the-Wild Singing Voice Dataset
di: Gu, Yicheng, et al.
Pubblicazione: (2025)
di: Gu, Yicheng, et al.
Pubblicazione: (2025)
VS-Singer: Vision-Guided Stereo Singing Voice Synthesis with Consistency Schrödinger Bridge
di: Zhao, Zijing, et al.
Pubblicazione: (2025)
di: Zhao, Zijing, et al.
Pubblicazione: (2025)
RAF: Relativistic Adversarial Feedback For Universal Speech Synthesis
di: Lee, Yongjoon, et al.
Pubblicazione: (2026)
di: Lee, Yongjoon, et al.
Pubblicazione: (2026)
Enhancing Spectrogram Realism in Singing Voice Synthesis via Explicit Bandwidth Extension Prior to Vocoder
di: Yang, Runxuan, et al.
Pubblicazione: (2025)
di: Yang, Runxuan, et al.
Pubblicazione: (2025)
A Preliminary Investigation on Flexible Singing Voice Synthesis Through Decomposed Framework with Inferrable Features
di: Violeta, Lester Phillip, et al.
Pubblicazione: (2024)
di: Violeta, Lester Phillip, et al.
Pubblicazione: (2024)
Controllable Singing Voice Synthesis using Phoneme-Level Energy Sequence
di: Ryu, Yerin, et al.
Pubblicazione: (2025)
di: Ryu, Yerin, et al.
Pubblicazione: (2025)
BiSinger: Bilingual Singing Voice Synthesis
di: Zhou, Huali, et al.
Pubblicazione: (2023)
di: Zhou, Huali, et al.
Pubblicazione: (2023)
DiffAttack: Diffusion-based Timbre-reserved Adversarial Attack in Speaker Identification
di: Wang, Qing, et al.
Pubblicazione: (2025)
di: Wang, Qing, et al.
Pubblicazione: (2025)
PerformSinger: Multimodal Singing Voice Synthesis Leveraging Synchronized Lip Cues from Singing Performance Videos
di: Gu, Ke, et al.
Pubblicazione: (2025)
di: Gu, Ke, et al.
Pubblicazione: (2025)
The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan
di: Sheng, Zhengyan, et al.
Pubblicazione: (2025)
di: Sheng, Zhengyan, et al.
Pubblicazione: (2025)
Stepback: Enhanced Disentanglement for Voice Conversion via Multi-Task Learning
di: Yang, Qian, et al.
Pubblicazione: (2025)
di: Yang, Qian, et al.
Pubblicazione: (2025)
Binaural Sound Event Localization and Detection based on HRTF Cues for Humanoid Robots
di: Lee, Gyeong-Tae, et al.
Pubblicazione: (2025)
di: Lee, Gyeong-Tae, et al.
Pubblicazione: (2025)
Documenti analoghi
-
CONTUNER: Singing Voice Beautifying with Pitch and Expressiveness Condition
di: Wang, Jianzong, et al.
Pubblicazione: (2024) -
UNMIXX: Untangling Highly Correlated Singing Voices Mixtures
di: Jung, Jihoo, et al.
Pubblicazione: (2026) -
TokSing: Singing Voice Synthesis based on Discrete Tokens
di: Wu, Yuning, et al.
Pubblicazione: (2024) -
DJCM: A Deep Joint Cascade Model for Singing Voice Separation and Vocal Pitch Estimation
di: Wei, Haojie, et al.
Pubblicazione: (2024) -
Binaural Sound Event Localization and Detection Neural Network based on HRTF Localization Cues for Humanoid Robots
di: Lee, Gyeong-Tae
Pubblicazione: (2025)