Low-Resource Cross-Domain Singing Voice Synthesis via Reduced Self-Supervised Speech Representations
Fuente:
arXiv
Saved in:
| Main Authors: | Kakoulidis, Panos, Ellinas, Nikolaos, Vamvoukakis, Georgios, Christidou, Myrsini, Vioni, Alexandra, Maniati, Georgia, Oh, Junkwang, Jho, Gunu, Hwang, Inchul, Tsiakoulis, Pirros, Chalamandaris, Aimilios |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Pseudo-Cepstrum: Pitch Modification for Mel-Based Neural Vocoders
by: Ellinas, Nikolaos, et al.
Published: (2025)
by: Ellinas, Nikolaos, et al.
Published: (2025)
Improved Text Emotion Prediction Using Combined Valence and Arousal Ordinal Classification
by: Mitsios, Michael, et al.
Published: (2024)
by: Mitsios, Michael, et al.
Published: (2024)
MambaRate: Speech Quality Assessment Across Different Sampling Rates
by: Kakoulidis, Panos, et al.
Published: (2025)
by: Kakoulidis, Panos, et al.
Published: (2025)
Cross-lingual Text-To-Speech with Flow-based Voice Conversion for Improved Pronunciation
by: Ellinas, Nikolaos, et al.
Published: (2022)
by: Ellinas, Nikolaos, et al.
Published: (2022)
Investigating Disentanglement in a Phoneme-level Speech Codec for Prosody Modeling
by: Karapiperis, Sotirios, et al.
Published: (2024)
by: Karapiperis, Sotirios, et al.
Published: (2024)
Counterfactual Activation Editing for Post-hoc Prosody and Mispronunciation Correction in TTS Models
by: Lee, Kyowoon, et al.
Published: (2025)
by: Lee, Kyowoon, et al.
Published: (2025)
Self-Supervised Singing Voice Pre-Training towards Speech-to-Singing Conversion
by: Li, Ruiqi, et al.
Published: (2024)
by: Li, Ruiqi, et al.
Published: (2024)
Protecting Cryptographic Libraries against Side-Channel and Code-Reuse Attacks
by: Tsoupidi, Rodothea Myrsini, et al.
Published: (2024)
by: Tsoupidi, Rodothea Myrsini, et al.
Published: (2024)
SingFake: Singing Voice Deepfake Detection
by: Zang, Yongyi, et al.
Published: (2023)
by: Zang, Yongyi, et al.
Published: (2023)
PolySinger: Singing-Voice to Singing-Voice Translation from English to Japanese
by: Antonisen, Silas, et al.
Published: (2024)
by: Antonisen, Silas, et al.
Published: (2024)
Singing Voice Graph Modeling for SingFake Detection
by: Chen, Xuanjun, et al.
Published: (2024)
by: Chen, Xuanjun, et al.
Published: (2024)
Periodic and Event-Triggering for Joint Capacity Maximization and Safe Intersection Crossing
by: Vitale, Christian, et al.
Published: (2022)
by: Vitale, Christian, et al.
Published: (2022)
Chapter Multi-Period Attack-Aware Optical Network Planning under Demand Uncertainty
by: Manousakis, Konstantinos, et al.
Published: (2021)
by: Manousakis, Konstantinos, et al.
Published: (2021)
Scheduling Aerial Vehicles in an Urban Air Mobility Scheme
by: Rigas, Emmanouil S., et al.
Published: (2021)
by: Rigas, Emmanouil S., et al.
Published: (2021)
Multi-Step Traffic Prediction for Multi-Period Planning in Optical Networks
by: Maryam, Hafsa, et al.
Published: (2024)
by: Maryam, Hafsa, et al.
Published: (2024)
Distributed Estimation and Control for Jamming an Aerial Target With Multiple Agents
by: Papaioannou, Savvas, et al.
Published: (2023)
by: Papaioannou, Savvas, et al.
Published: (2023)
Issues on Modeling the Singing Voice
by: Loscos, Alex
Published: (2003)
by: Loscos, Alex
Published: (2003)
SingIt! Singer Voice Transformation
by: Eliav, Amit, et al.
Published: (2024)
by: Eliav, Amit, et al.
Published: (2024)
Singing Voice Therapy Revisited
by: Ilter Denizoglu
Published: (2021)
by: Ilter Denizoglu
Published: (2021)
TokSing: Singing Voice Synthesis based on Discrete Tokens
by: Wu, Yuning, et al.
Published: (2024)
by: Wu, Yuning, et al.
Published: (2024)
Singing Voice Conversion with Accompaniment Using Self-Supervised Representation-Based Melody Features
by: Chen, Wei, et al.
Published: (2025)
by: Chen, Wei, et al.
Published: (2025)
SingVisio: Visual Analytics of Diffusion Model for Singing Voice Conversion
by: Xue, Liumeng, et al.
Published: (2024)
by: Xue, Liumeng, et al.
Published: (2024)
Descriptive Singing Voice Manipulation Dataset
by: Gohari, Mahyar
Published: (2026)
by: Gohari, Mahyar
Published: (2026)
Neural Concatenative Singing Voice Conversion: Rethinking Concatenation-Based Approach for One-Shot Singing Voice Conversion
by: Sha, Binzhu, et al.
Published: (2023)
by: Sha, Binzhu, et al.
Published: (2023)
SingMOS: An extensive Open-Source Singing Voice Dataset for MOS Prediction
by: Tang, Yuxun, et al.
Published: (2024)
by: Tang, Yuxun, et al.
Published: (2024)
SingVERSE: A Diverse, Real-World Benchmark for Singing Voice Enhancement
by: Jiang, Shaohan, et al.
Published: (2025)
by: Jiang, Shaohan, et al.
Published: (2025)
InstructSing: High-Fidelity Singing Voice Generation via Instructing Yourself
by: Zeng, Chang, et al.
Published: (2024)
by: Zeng, Chang, et al.
Published: (2024)
VISinger2+: End-to-End Singing Voice Synthesis Augmented by Self-Supervised Learning Representation
by: Yu, Yifeng, et al.
Published: (2024)
by: Yu, Yifeng, et al.
Published: (2024)
Activity delay patterns in project networks
by: Vazquez, Alexei, et al.
Published: (2023)
by: Vazquez, Alexei, et al.
Published: (2023)
Probabilistically Robust Trajectory Planning of Multiple Aerial Agents
by: Vitale, Christian, et al.
Published: (2024)
by: Vitale, Christian, et al.
Published: (2024)
Cooperative Search and Track of Rogue Drones using Multiagent Reinforcement Learning
by: Valianti, Panayiota, et al.
Published: (2025)
by: Valianti, Panayiota, et al.
Published: (2025)
Deepfake Detection of Singing Voices With Whisper Encodings
by: Sharma, Falguni, et al.
Published: (2025)
by: Sharma, Falguni, et al.
Published: (2025)
BiSinger: Bilingual Singing Voice Synthesis
by: Zhou, Huali, et al.
Published: (2023)
by: Zhou, Huali, et al.
Published: (2023)
Automatic Estimation of Singing Voice Musical Dynamics
by: Narang, Jyoti, et al.
Published: (2024)
by: Narang, Jyoti, et al.
Published: (2024)
Perceived Femininity in Singing Voice: Analysis and Prediction
by: Kong, Yuexuan, et al.
Published: (2025)
by: Kong, Yuexuan, et al.
Published: (2025)
Robust Singing Voice Transcription Serves Synthesis
by: Li, Ruiqi, et al.
Published: (2024)
by: Li, Ruiqi, et al.
Published: (2024)
Singing Voice Data Scaling-up: An Introduction to ACE-Opencpop and ACE-KiSing
by: Shi, Jiatong, et al.
Published: (2024)
by: Shi, Jiatong, et al.
Published: (2024)
SingNet: Towards a Large-Scale, Diverse, and In-the-Wild Singing Voice Dataset
by: Gu, Yicheng, et al.
Published: (2025)
by: Gu, Yicheng, et al.
Published: (2025)
Everyone-Can-Sing: Zero-Shot Singing Voice Synthesis and Conversion with Speech Reference
by: Dai, Shuqi, et al.
Published: (2025)
by: Dai, Shuqi, et al.
Published: (2025)
VibE-SVC: Vibrato Extraction with High-frequency F0 Contour for Singing Voice Conversion
by: Choi, Joon-Seung, et al.
Published: (2025)
by: Choi, Joon-Seung, et al.
Published: (2025)
Similar Items
-
Pseudo-Cepstrum: Pitch Modification for Mel-Based Neural Vocoders
by: Ellinas, Nikolaos, et al.
Published: (2025) -
Improved Text Emotion Prediction Using Combined Valence and Arousal Ordinal Classification
by: Mitsios, Michael, et al.
Published: (2024) -
MambaRate: Speech Quality Assessment Across Different Sampling Rates
by: Kakoulidis, Panos, et al.
Published: (2025) -
Cross-lingual Text-To-Speech with Flow-based Voice Conversion for Improved Pronunciation
by: Ellinas, Nikolaos, et al.
Published: (2022) -
Investigating Disentanglement in a Phoneme-level Speech Codec for Prosody Modeling
by: Karapiperis, Sotirios, et al.
Published: (2024)