A Comparative Analysis of Poetry Reading Audio: Singing, Narrating, or Somewhere In Between?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Choi, Kahyun, Kim, Minje |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Hyperbolic Distance-Based Speech Separation
von: Petermann, Darius, et al.
Veröffentlicht: (2024)
von: Petermann, Darius, et al.
Veröffentlicht: (2024)
Neural Speech and Audio Coding: Modern AI Technology Meets Traditional Codecs
von: Kim, Minje, et al.
Veröffentlicht: (2024)
von: Kim, Minje, et al.
Veröffentlicht: (2024)
Serenade: A Singing Style Conversion Framework Based On Audio Infilling
von: Violeta, Lester Phillip, et al.
Veröffentlicht: (2025)
von: Violeta, Lester Phillip, et al.
Veröffentlicht: (2025)
Leveraging Diverse Semantic-based Audio Pretrained Models for Singing Voice Conversion
von: Zhang, Xueyao, et al.
Veröffentlicht: (2023)
von: Zhang, Xueyao, et al.
Veröffentlicht: (2023)
Jointly Recognizing Speech and Singing Voices Based on Multi-Task Audio Source Separation
von: Bai, Ye, et al.
Veröffentlicht: (2024)
von: Bai, Ye, et al.
Veröffentlicht: (2024)
PromptSep: Generative Audio Separation via Multimodal Prompting
von: Wen, Yutong, et al.
Veröffentlicht: (2025)
von: Wen, Yutong, et al.
Veröffentlicht: (2025)
Task-Specific Audio Coding for Machines: Machine-Learned Latent Features Are Codes for That Machine
von: Kuznetsova, Anastasia, et al.
Veröffentlicht: (2025)
von: Kuznetsova, Anastasia, et al.
Veröffentlicht: (2025)
Low-Resource Audio Codec (LRAC): 2025 Challenge Description
von: Wojcicki, Kamil, et al.
Veröffentlicht: (2025)
von: Wojcicki, Kamil, et al.
Veröffentlicht: (2025)
A Survey on 30+ Years of Automatic Singing Assessment and Singing Information Processing
von: Santos, Arthur N. dos, et al.
Veröffentlicht: (2026)
von: Santos, Arthur N. dos, et al.
Veröffentlicht: (2026)
SingVERSE: A Diverse, Real-World Benchmark for Singing Voice Enhancement
von: Jiang, Shaohan, et al.
Veröffentlicht: (2025)
von: Jiang, Shaohan, et al.
Veröffentlicht: (2025)
TokSing: Singing Voice Synthesis based on Discrete Tokens
von: Wu, Yuning, et al.
Veröffentlicht: (2024)
von: Wu, Yuning, et al.
Veröffentlicht: (2024)
SingMOS: An extensive Open-Source Singing Voice Dataset for MOS Prediction
von: Tang, Yuxun, et al.
Veröffentlicht: (2024)
von: Tang, Yuxun, et al.
Veröffentlicht: (2024)
Self-Supervised Singing Voice Pre-Training towards Speech-to-Singing Conversion
von: Li, Ruiqi, et al.
Veröffentlicht: (2024)
von: Li, Ruiqi, et al.
Veröffentlicht: (2024)
InstructSing: High-Fidelity Singing Voice Generation via Instructing Yourself
von: Zeng, Chang, et al.
Veröffentlicht: (2024)
von: Zeng, Chang, et al.
Veröffentlicht: (2024)
MuSE-SVS: Multi-Singer Emotional Singing Voice Synthesizer that Controls Emotional Intensity
von: Kim, Sungjae, et al.
Veröffentlicht: (2022)
von: Kim, Sungjae, et al.
Veröffentlicht: (2022)
UNMIXX: Untangling Highly Correlated Singing Voices Mixtures
von: Jung, Jihoo, et al.
Veröffentlicht: (2026)
von: Jung, Jihoo, et al.
Veröffentlicht: (2026)
Personalized Neural Speech Codec
von: Jang, Inseon, et al.
Veröffentlicht: (2024)
von: Jang, Inseon, et al.
Veröffentlicht: (2024)
Singing Voice Data Scaling-up: An Introduction to ACE-Opencpop and ACE-KiSing
von: Shi, Jiatong, et al.
Veröffentlicht: (2024)
von: Shi, Jiatong, et al.
Veröffentlicht: (2024)
SingNet: Towards a Large-Scale, Diverse, and In-the-Wild Singing Voice Dataset
von: Gu, Yicheng, et al.
Veröffentlicht: (2025)
von: Gu, Yicheng, et al.
Veröffentlicht: (2025)
Everyone-Can-Sing: Zero-Shot Singing Voice Synthesis and Conversion with Speech Reference
von: Dai, Shuqi, et al.
Veröffentlicht: (2025)
von: Dai, Shuqi, et al.
Veröffentlicht: (2025)
An Extensive Analysis of the Singing Voice Conversion Challenge 2025 Evaluation Results
von: Violeta, Lester Phillip, et al.
Veröffentlicht: (2025)
von: Violeta, Lester Phillip, et al.
Veröffentlicht: (2025)
SingIt! Singer Voice Transformation
von: Eliav, Amit, et al.
Veröffentlicht: (2024)
von: Eliav, Amit, et al.
Veröffentlicht: (2024)
SingOMD: Singing Oriented Multi-resolution Discrete Representation Construction from Speech Models
von: Tang, Yuxun, et al.
Veröffentlicht: (2024)
von: Tang, Yuxun, et al.
Veröffentlicht: (2024)
Adversarial Multi-Task Learning for Disentangling Timbre and Pitch in Singing Voice Synthesis
von: Kim, Tae-Woo, et al.
Veröffentlicht: (2022)
von: Kim, Tae-Woo, et al.
Veröffentlicht: (2022)
Comparative Analysis of Finite Difference and Finite Element Method for Audio Waveform Simulation
von: Florin, Juliette
Veröffentlicht: (2025)
von: Florin, Juliette
Veröffentlicht: (2025)
Neural Concatenative Singing Voice Conversion: Rethinking Concatenation-Based Approach for One-Shot Singing Voice Conversion
von: Sha, Binzhu, et al.
Veröffentlicht: (2023)
von: Sha, Binzhu, et al.
Veröffentlicht: (2023)
Gencho: Room Impulse Response Generation from Reverberant Speech and Text via Diffusion Transformers
von: Lin, Jackie, et al.
Veröffentlicht: (2026)
von: Lin, Jackie, et al.
Veröffentlicht: (2026)
Singing Voice Graph Modeling for SingFake Detection
von: Chen, Xuanjun, et al.
Veröffentlicht: (2024)
von: Chen, Xuanjun, et al.
Veröffentlicht: (2024)
Robust Singing Voice Transcription Serves Synthesis
von: Li, Ruiqi, et al.
Veröffentlicht: (2024)
von: Li, Ruiqi, et al.
Veröffentlicht: (2024)
Multimodal Representation Loss Between Timed Text and Audio for Regularized Speech Separation
von: Hsieh, Tsun-An, et al.
Veröffentlicht: (2024)
von: Hsieh, Tsun-An, et al.
Veröffentlicht: (2024)
Period Singer: Integrating Periodic and Aperiodic Variational Autoencoders for Natural-Sounding End-to-End Singing Voice Synthesis
von: Kim, Taewoo, et al.
Veröffentlicht: (2024)
von: Kim, Taewoo, et al.
Veröffentlicht: (2024)
CONTUNER: Singing Voice Beautifying with Pitch and Expressiveness Condition
von: Wang, Jianzong, et al.
Veröffentlicht: (2024)
von: Wang, Jianzong, et al.
Veröffentlicht: (2024)
Combolutional Neural Networks
von: Churchwell, Cameron, et al.
Veröffentlicht: (2025)
von: Churchwell, Cameron, et al.
Veröffentlicht: (2025)
User-guided Generative Source Separation
von: Wen, Yutong, et al.
Veröffentlicht: (2025)
von: Wen, Yutong, et al.
Veröffentlicht: (2025)
Comparative Evaluation of Text and Audio Simplification: A Methodological Replication Study
von: Barai, Prosanta, et al.
Veröffentlicht: (2025)
von: Barai, Prosanta, et al.
Veröffentlicht: (2025)
Zero-Shot Duet Singing Voices Separation with Diffusion Models
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2023)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2023)
Muskits-ESPnet: A Comprehensive Toolkit for Singing Voice Synthesis in New Paradigm
von: Wu, Yuning, et al.
Veröffentlicht: (2024)
von: Wu, Yuning, et al.
Veröffentlicht: (2024)
STARS: A Unified Framework for Singing Transcription, Alignment, and Refined Style Annotation
von: Guo, Wenxiang, et al.
Veröffentlicht: (2025)
von: Guo, Wenxiang, et al.
Veröffentlicht: (2025)
Correlation of Fréchet Audio Distance With Human Perception of Environmental Audio Is Embedding Dependant
von: Tailleur, Modan, et al.
Veröffentlicht: (2024)
von: Tailleur, Modan, et al.
Veröffentlicht: (2024)
DeFT-Mamba: Universal Multichannel Sound Separation and Polyphonic Audio Classification
von: Lee, Dongheon, et al.
Veröffentlicht: (2024)
von: Lee, Dongheon, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Hyperbolic Distance-Based Speech Separation
von: Petermann, Darius, et al.
Veröffentlicht: (2024) -
Neural Speech and Audio Coding: Modern AI Technology Meets Traditional Codecs
von: Kim, Minje, et al.
Veröffentlicht: (2024) -
Serenade: A Singing Style Conversion Framework Based On Audio Infilling
von: Violeta, Lester Phillip, et al.
Veröffentlicht: (2025) -
Leveraging Diverse Semantic-based Audio Pretrained Models for Singing Voice Conversion
von: Zhang, Xueyao, et al.
Veröffentlicht: (2023) -
Jointly Recognizing Speech and Singing Voices Based on Multi-Task Audio Source Separation
von: Bai, Ye, et al.
Veröffentlicht: (2024)