Wavespace: A Highly Explorable Wavetable Generator
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lee, Hazounne, Kim, Kihong, Lee, Sungho, Lee, Kyogu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Rethinking Speech Representation Aggregation in Speech Enhancement: A Phonetic Mutual Information Perspective
von: Han, Seungu, et al.
Veröffentlicht: (2026)
von: Han, Seungu, et al.
Veröffentlicht: (2026)
Few-step Adversarial Schrödinger Bridge for Generative Speech Enhancement
von: Han, Seungu, et al.
Veröffentlicht: (2025)
von: Han, Seungu, et al.
Veröffentlicht: (2025)
Vo-Ve: An Explainable Voice-Vector for Speaker Identity Evaluation
von: Lee, Jaejun, et al.
Veröffentlicht: (2025)
von: Lee, Jaejun, et al.
Veröffentlicht: (2025)
Towards Bitrate-Efficient and Noise-Robust Speech Coding with Variable Bitrate RVQ
von: Chae, Yunkee, et al.
Veröffentlicht: (2025)
von: Chae, Yunkee, et al.
Veröffentlicht: (2025)
Music De-limiter Networks via Sample-wise Gain Inversion
von: Jeon, Chang-Bin, et al.
Veröffentlicht: (2023)
von: Jeon, Chang-Bin, et al.
Veröffentlicht: (2023)
GRAFX: An Open-Source Library for Audio Processing Graphs in PyTorch
von: Lee, Sungho, et al.
Veröffentlicht: (2024)
von: Lee, Sungho, et al.
Veröffentlicht: (2024)
Do Captioning Metrics Reflect Music Semantic Alignment?
von: Lee, Jinwoo, et al.
Veröffentlicht: (2024)
von: Lee, Jinwoo, et al.
Veröffentlicht: (2024)
MGE-LDM: Joint Latent Diffusion for Simultaneous Music Generation and Source Extraction
von: Chae, Yunkee, et al.
Veröffentlicht: (2025)
von: Chae, Yunkee, et al.
Veröffentlicht: (2025)
Differentiable Acoustic Radiance Transfer
von: Lee, Sungho, et al.
Veröffentlicht: (2025)
von: Lee, Sungho, et al.
Veröffentlicht: (2025)
Inverse Nonlinearity Compensation of Hyperelastic Deformation in Dielectric Elastomer for Acoustic Actuation
von: Lee, Jin Woo, et al.
Veröffentlicht: (2024)
von: Lee, Jin Woo, et al.
Veröffentlicht: (2024)
DDD: A Perceptually Superior Low-Response-Time DNN-based Declipper
von: Yi, Jayeon, et al.
Veröffentlicht: (2024)
von: Yi, Jayeon, et al.
Veröffentlicht: (2024)
Removing Speaker Information from Speech Representation using Variable-Length Soft Pooling
von: Hwang, Injune, et al.
Veröffentlicht: (2024)
von: Hwang, Injune, et al.
Veröffentlicht: (2024)
Reverse Engineering of Music Mixing Graphs with Differentiable Processors and Iterative Pruning
von: Lee, Sungho, et al.
Veröffentlicht: (2025)
von: Lee, Sungho, et al.
Veröffentlicht: (2025)
DOSE : Drum One-Shot Extraction from Music Mixture
von: Hwang, Suntae, et al.
Veröffentlicht: (2025)
von: Hwang, Suntae, et al.
Veröffentlicht: (2025)
String Sound Synthesizer on GPU-accelerated Finite Difference Scheme
von: Lee, Jin Woo, et al.
Veröffentlicht: (2023)
von: Lee, Jin Woo, et al.
Veröffentlicht: (2023)
Learning Semantic Information from Raw Audio Signal Using Both Contextual and Phonetic Representations
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2024)
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2024)
Singing Voice Synthesis Using Differentiable LPC and Glottal-Flow-Inspired Wavetables
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2023)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2023)
Practical and Reproducible Symbolic Music Generation by Large Language Models with Structural Embeddings
von: Rhyu, Seungyeon, et al.
Veröffentlicht: (2024)
von: Rhyu, Seungyeon, et al.
Veröffentlicht: (2024)
Guiding Frame-Level CTC Alignments Using Self-knowledge Distillation
von: Kim, Eungbeom, et al.
Veröffentlicht: (2024)
von: Kim, Eungbeom, et al.
Veröffentlicht: (2024)
Music Auto-Tagging with Robust Music Representation Learned via Domain Adversarial Training
von: Joung, Haesun, et al.
Veröffentlicht: (2024)
von: Joung, Haesun, et al.
Veröffentlicht: (2024)
Differentiable Modal Synthesis for Physical Modeling of Planar String Sound and Motion Simulation
von: Lee, Jin Woo, et al.
Veröffentlicht: (2024)
von: Lee, Jin Woo, et al.
Veröffentlicht: (2024)
Hear Your Face: Face-based voice conversion with F0 estimation
von: Lee, Jaejun, et al.
Veröffentlicht: (2024)
von: Lee, Jaejun, et al.
Veröffentlicht: (2024)
Distance Sampling-based Paraphraser Leveraging ChatGPT for Text Data Manipulation
von: Oh, Yoori, et al.
Veröffentlicht: (2024)
von: Oh, Yoori, et al.
Veröffentlicht: (2024)
Wavetable Synthesis Using CVAE for Timbre Control Based on Semantic Label
von: Yutani, Tsugumasa, et al.
Veröffentlicht: (2024)
von: Yutani, Tsugumasa, et al.
Veröffentlicht: (2024)
SAM: A Mamba-2 State-Space Audio-Language Model
von: Lee, Taehan, et al.
Veröffentlicht: (2025)
von: Lee, Taehan, et al.
Veröffentlicht: (2025)
Improving Test-Time Performance of RVQ-based Neural Codecs
von: Kim, Hyeongju, et al.
Veröffentlicht: (2025)
von: Kim, Hyeongju, et al.
Veröffentlicht: (2025)
SEMamba++: A General Speech Restoration Framework Leveraging Global, Local, and Periodic Spectral Patterns
von: Lee, Yongjoon, et al.
Veröffentlicht: (2026)
von: Lee, Yongjoon, et al.
Veröffentlicht: (2026)
VoiceTailor: Lightweight Plug-In Adapter for Diffusion-Based Personalized Text-to-Speech
von: Kim, Heeseung, et al.
Veröffentlicht: (2024)
von: Kim, Heeseung, et al.
Veröffentlicht: (2024)
Erasing Your Voice Before It's Heard: Training-free Speaker Unlearning for Zero-shot Text-to-Speech
von: Lee, Myungjin, et al.
Veröffentlicht: (2026)
von: Lee, Myungjin, et al.
Veröffentlicht: (2026)
Ambisonizer: Neural Upmixing as Spherical Harmonics Generation
von: Zang, Yongyi, et al.
Veröffentlicht: (2024)
von: Zang, Yongyi, et al.
Veröffentlicht: (2024)
DualSpeech: Enhancing Speaker-Fidelity and Text-Intelligibility Through Dual Classifier-Free Guidance
von: Yang, Jinhyeok, et al.
Veröffentlicht: (2024)
von: Yang, Jinhyeok, et al.
Veröffentlicht: (2024)
Instance-Specific Test-Time Training for Speech Editing in the Wild
von: Kim, Taewoo, et al.
Veröffentlicht: (2025)
von: Kim, Taewoo, et al.
Veröffentlicht: (2025)
Audio-Based Linguistic Feature Extraction for Enhancing Multi-lingual and Low-Resource Text-to-Speech
von: Kim, Youngjae, et al.
Veröffentlicht: (2024)
von: Kim, Youngjae, et al.
Veröffentlicht: (2024)
Enhancing Zero-Shot Multi-Speaker TTS with Negated Speaker Representations
von: Jeon, Yejin, et al.
Veröffentlicht: (2024)
von: Jeon, Yejin, et al.
Veröffentlicht: (2024)
Period Singer: Integrating Periodic and Aperiodic Variational Autoencoders for Natural-Sounding End-to-End Singing Voice Synthesis
von: Kim, Taewoo, et al.
Veröffentlicht: (2024)
von: Kim, Taewoo, et al.
Veröffentlicht: (2024)
UNMIXX: Untangling Highly Correlated Singing Voices Mixtures
von: Jung, Jihoo, et al.
Veröffentlicht: (2026)
von: Jung, Jihoo, et al.
Veröffentlicht: (2026)
Inter-channel Conv-TasNet for multichannel speech enhancement
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
Naturalness-Aware Curriculum Learning with Dynamic Temperature for Speech Deepfake Detection
von: Kim, Taewoo, et al.
Veröffentlicht: (2025)
von: Kim, Taewoo, et al.
Veröffentlicht: (2025)
Binaural Sound Event Localization and Detection Neural Network based on HRTF Localization Cues for Humanoid Robots
von: Lee, Gyeong-Tae
Veröffentlicht: (2025)
von: Lee, Gyeong-Tae
Veröffentlicht: (2025)
Wave-U-Mamba: An End-To-End Framework For High-Quality And Efficient Speech Super Resolution
von: Lee, Yongjoon, et al.
Veröffentlicht: (2024)
von: Lee, Yongjoon, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Rethinking Speech Representation Aggregation in Speech Enhancement: A Phonetic Mutual Information Perspective
von: Han, Seungu, et al.
Veröffentlicht: (2026) -
Few-step Adversarial Schrödinger Bridge for Generative Speech Enhancement
von: Han, Seungu, et al.
Veröffentlicht: (2025) -
Vo-Ve: An Explainable Voice-Vector for Speaker Identity Evaluation
von: Lee, Jaejun, et al.
Veröffentlicht: (2025) -
Towards Bitrate-Efficient and Noise-Robust Speech Coding with Variable Bitrate RVQ
von: Chae, Yunkee, et al.
Veröffentlicht: (2025) -
Music De-limiter Networks via Sample-wise Gain Inversion
von: Jeon, Chang-Bin, et al.
Veröffentlicht: (2023)