Spectral Mapping of Singing Voices: U-Net-Assisted Vocal Segmentation
Fuente:
arXiv
Guardado en:
| Autor principal: | Sorrenti, Adam |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SingFake: Singing Voice Deepfake Detection
por: Zang, Yongyi, et al.
Publicado: (2023)
por: Zang, Yongyi, et al.
Publicado: (2023)
Deepfake Detection of Singing Voices With Whisper Encodings
por: Sharma, Falguni, et al.
Publicado: (2025)
por: Sharma, Falguni, et al.
Publicado: (2025)
RDSinger: Reference-based Diffusion Network for Singing Voice Synthesis
por: Sui, Kehan, et al.
Publicado: (2024)
por: Sui, Kehan, et al.
Publicado: (2024)
SPA-SVC: Self-supervised Pitch Augmentation for Singing Voice Conversion
por: Bai, Bingsong, et al.
Publicado: (2024)
por: Bai, Bingsong, et al.
Publicado: (2024)
LHQ-SVC: Lightweight and High Quality Singing Voice Conversion Modeling
por: Huang, Yubo, et al.
Publicado: (2024)
por: Huang, Yubo, et al.
Publicado: (2024)
Controllable Singing Voice Synthesis using Phoneme-Level Energy Sequence
por: Ryu, Yerin, et al.
Publicado: (2025)
por: Ryu, Yerin, et al.
Publicado: (2025)
DiTSinger: Scaling Singing Voice Synthesis with Diffusion Transformer and Implicit Alignment
por: Du, Zongcai, et al.
Publicado: (2025)
por: Du, Zongcai, et al.
Publicado: (2025)
SoulX-Singer: Towards High-Quality Zero-Shot Singing Voice Synthesis
por: Qian, Jiale, et al.
Publicado: (2026)
por: Qian, Jiale, et al.
Publicado: (2026)
Mitigating Latent Mismatch in cVAE-Based Singing Voice Synthesis via Flow Matching
por: Yun, Minhyeok, et al.
Publicado: (2026)
por: Yun, Minhyeok, et al.
Publicado: (2026)
Speech Foundation Model Ensembles for the Controlled Singing Voice Deepfake Detection (CtrSVDD) Challenge 2024
por: Guragain, Anmol, et al.
Publicado: (2024)
por: Guragain, Anmol, et al.
Publicado: (2024)
R2-SVC: Towards Real-World Robust and Expressive Zero-shot Singing Voice Conversion
por: Zheng, Junjie, et al.
Publicado: (2025)
por: Zheng, Junjie, et al.
Publicado: (2025)
VibE-SVC: Vibrato Extraction with High-frequency F0 Contour for Singing Voice Conversion
por: Choi, Joon-Seung, et al.
Publicado: (2025)
por: Choi, Joon-Seung, et al.
Publicado: (2025)
HQ-SVC: Towards High-Quality Zero-Shot Singing Voice Conversion in Low-Resource Scenarios
por: Bai, Bingsong, et al.
Publicado: (2025)
por: Bai, Bingsong, et al.
Publicado: (2025)
Multimodal Laryngoscopic Video Analysis for Assisted Diagnosis of Vocal Fold Paralysis
por: Zhang, Yucong, et al.
Publicado: (2024)
por: Zhang, Yucong, et al.
Publicado: (2024)
SingMOS-Pro: An Comprehensive Benchmark for Singing Quality Assessment
por: Tang, Yuxun, et al.
Publicado: (2025)
por: Tang, Yuxun, et al.
Publicado: (2025)
SingNet: Towards a Large-Scale, Diverse, and In-the-Wild Singing Voice Dataset
por: Gu, Yicheng, et al.
Publicado: (2025)
por: Gu, Yicheng, et al.
Publicado: (2025)
LAPS-Diff: A Diffusion-Based Framework for Singing Voice Synthesis With Language Aware Prosody-Style Guided Learning
por: Dhar, Sandipan, et al.
Publicado: (2025)
por: Dhar, Sandipan, et al.
Publicado: (2025)
DJCM: A Deep Joint Cascade Model for Singing Voice Separation and Vocal Pitch Estimation
por: Wei, Haojie, et al.
Publicado: (2024)
por: Wei, Haojie, et al.
Publicado: (2024)
SVDD Challenge 2024: A Singing Voice Deepfake Detection Challenge Evaluation Plan
por: Zhang, You, et al.
Publicado: (2024)
por: Zhang, You, et al.
Publicado: (2024)
VocalAgent: Large Language Models for Vocal Health Diagnostics with Safety-Aware Evaluation
por: Kim, Yubin, et al.
Publicado: (2025)
por: Kim, Yubin, et al.
Publicado: (2025)
Learning Marmoset Vocal Patterns with a Masked Autoencoder for Robust Call Segmentation, Classification, and Caller Identification
por: Wu, Bin, et al.
Publicado: (2024)
por: Wu, Bin, et al.
Publicado: (2024)
CoMoSVC: Consistency Model-based Singing Voice Conversion
por: Lu, Yiwen, et al.
Publicado: (2024)
por: Lu, Yiwen, et al.
Publicado: (2024)
Prompt-Singer: Controllable Singing-Voice-Synthesis with Natural Language Prompt
por: Wang, Yongqi, et al.
Publicado: (2024)
por: Wang, Yongqi, et al.
Publicado: (2024)
GeHirNet: A Gender-Aware Hierarchical Model for Voice Pathology Classification
por: Wu, Fan, et al.
Publicado: (2025)
por: Wu, Fan, et al.
Publicado: (2025)
SingIt! Singer Voice Transformation
por: Eliav, Amit, et al.
Publicado: (2024)
por: Eliav, Amit, et al.
Publicado: (2024)
TokSing: Singing Voice Synthesis based on Discrete Tokens
por: Wu, Yuning, et al.
Publicado: (2024)
por: Wu, Yuning, et al.
Publicado: (2024)
Facing the Music: Tackling Singing Voice Separation in Cinematic Audio Source Separation
por: Watcharasupat, Karn N., et al.
Publicado: (2024)
por: Watcharasupat, Karn N., et al.
Publicado: (2024)
Sing-On-Your-Beat: Simple Text-Controllable Accompaniment Generations
por: Trinh, Quoc-Huy, et al.
Publicado: (2024)
por: Trinh, Quoc-Huy, et al.
Publicado: (2024)
USM-VC: Mitigating Timbre Leakage with Universal Semantic Mapping Residual Block for Voice Conversion
por: Li, Na, et al.
Publicado: (2025)
por: Li, Na, et al.
Publicado: (2025)
Neural Concatenative Singing Voice Conversion: Rethinking Concatenation-Based Approach for One-Shot Singing Voice Conversion
por: Sha, Binzhu, et al.
Publicado: (2023)
por: Sha, Binzhu, et al.
Publicado: (2023)
Acoustic Imaging for UAV Detection: Dense Beamformed Energy Maps and U-Net SELD
por: Rodriguez, Belman Jahir, et al.
Publicado: (2025)
por: Rodriguez, Belman Jahir, et al.
Publicado: (2025)
SingMOS: An extensive Open-Source Singing Voice Dataset for MOS Prediction
por: Tang, Yuxun, et al.
Publicado: (2024)
por: Tang, Yuxun, et al.
Publicado: (2024)
Self-Supervised Singing Voice Pre-Training towards Speech-to-Singing Conversion
por: Li, Ruiqi, et al.
Publicado: (2024)
por: Li, Ruiqi, et al.
Publicado: (2024)
InstructSing: High-Fidelity Singing Voice Generation via Instructing Yourself
por: Zeng, Chang, et al.
Publicado: (2024)
por: Zeng, Chang, et al.
Publicado: (2024)
SingVERSE: A Diverse, Real-World Benchmark for Singing Voice Enhancement
por: Jiang, Shaohan, et al.
Publicado: (2025)
por: Jiang, Shaohan, et al.
Publicado: (2025)
Learning Physiology-Informed Vocal Spectrotemporal Representations for Speech Emotion Recognition
por: Zhang, Xu, et al.
Publicado: (2026)
por: Zhang, Xu, et al.
Publicado: (2026)
Mamba2 Meets Silence: Robust Vocal Source Separation for Sparse Regions
por: Kim, Euiyeon, et al.
Publicado: (2025)
por: Kim, Euiyeon, et al.
Publicado: (2025)
NV-Bench: Benchmark of Nonverbal Vocalization Synthesis for Expressive Text-to-Speech Generation
por: Ni, Qinke, et al.
Publicado: (2026)
por: Ni, Qinke, et al.
Publicado: (2026)
Robust Singing Voice Transcription Serves Synthesis
por: Li, Ruiqi, et al.
Publicado: (2024)
por: Li, Ruiqi, et al.
Publicado: (2024)
Singing Voice Data Scaling-up: An Introduction to ACE-Opencpop and ACE-KiSing
por: Shi, Jiatong, et al.
Publicado: (2024)
por: Shi, Jiatong, et al.
Publicado: (2024)
Ejemplares similares
-
SingFake: Singing Voice Deepfake Detection
por: Zang, Yongyi, et al.
Publicado: (2023) -
Deepfake Detection of Singing Voices With Whisper Encodings
por: Sharma, Falguni, et al.
Publicado: (2025) -
RDSinger: Reference-based Diffusion Network for Singing Voice Synthesis
por: Sui, Kehan, et al.
Publicado: (2024) -
SPA-SVC: Self-supervised Pitch Augmentation for Singing Voice Conversion
por: Bai, Bingsong, et al.
Publicado: (2024) -
LHQ-SVC: Lightweight and High Quality Singing Voice Conversion Modeling
por: Huang, Yubo, et al.
Publicado: (2024)