Mix2Morph: Learning Sound Morphing from Noisy Mixes
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chu, Annie, García, Hugo Flores, Nieto, Oriol, Salamon, Justin, Pardo, Bryan, Seetharaman, Prem |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Generative Audio Extension and Morphing
von: Seetharaman, Prem, et al.
Veröffentlicht: (2026)
von: Seetharaman, Prem, et al.
Veröffentlicht: (2026)
Sketch2Sound: Controllable Audio Generation via Time-Varying Signals and Sonic Imitations
von: García, Hugo Flores, et al.
Veröffentlicht: (2024)
von: García, Hugo Flores, et al.
Veröffentlicht: (2024)
Audiocards: Structured Metadata Improves Audio Language Models For Sound Design
von: Sridhar, Sripathi, et al.
Veröffentlicht: (2026)
von: Sridhar, Sripathi, et al.
Veröffentlicht: (2026)
SILA: Signal-to-Language Augmentation for Enhanced Control in Text-to-Audio Generation
von: Kumar, Sonal, et al.
Veröffentlicht: (2024)
von: Kumar, Sonal, et al.
Veröffentlicht: (2024)
Video-Guided Foley Sound Generation with Multimodal Controls
von: Chen, Ziyang, et al.
Veröffentlicht: (2024)
von: Chen, Ziyang, et al.
Veröffentlicht: (2024)
AudioChat: Unified Audio Storytelling, Editing, and Understanding with Transfusion Forcing
von: Chen, William, et al.
Veröffentlicht: (2026)
von: Chen, William, et al.
Veröffentlicht: (2026)
The Rhythm In Anything: Audio-Prompted Drums Generation with Masked Language Modeling
von: O'Reilly, Patrick, et al.
Veröffentlicht: (2025)
von: O'Reilly, Patrick, et al.
Veröffentlicht: (2025)
TAC: Timestamped Audio Captioning
von: Kumar, Sonal, et al.
Veröffentlicht: (2026)
von: Kumar, Sonal, et al.
Veröffentlicht: (2026)
FLAM: Frame-Wise Language-Audio Modeling
von: Wu, Yusong, et al.
Veröffentlicht: (2025)
von: Wu, Yusong, et al.
Veröffentlicht: (2025)
Augment, Drop & Swap: Improving Diversity in LLM Captions for Efficient Music-Text Representation Learning
von: Manco, Ilaria, et al.
Veröffentlicht: (2024)
von: Manco, Ilaria, et al.
Veröffentlicht: (2024)
PromptSep: Generative Audio Separation via Multimodal Prompting
von: Wen, Yutong, et al.
Veröffentlicht: (2025)
von: Wen, Yutong, et al.
Veröffentlicht: (2025)
Code Drift: Towards Idempotent Neural Audio Codecs
von: O'Reilly, Patrick, et al.
Veröffentlicht: (2024)
von: O'Reilly, Patrick, et al.
Veröffentlicht: (2024)
SoundMorpher: Perceptually-Uniform Sound Morphing with Diffusion Model
von: Niu, Xinlei, et al.
Veröffentlicht: (2024)
von: Niu, Xinlei, et al.
Veröffentlicht: (2024)
Sound Effects and the Morphing Façade: Emotion in Sound of Projection Mapping
von: Roger Pastó Cortina
Veröffentlicht: (2019)
von: Roger Pastó Cortina
Veröffentlicht: (2019)
Time-Domain Voice Identity Morphing (TD-VIM): A Signal-Level Approach to Morphing Attacks on Speaker Verification Systems
von: PN, Aravinda Reddy, et al.
Veröffentlicht: (2026)
von: PN, Aravinda Reddy, et al.
Veröffentlicht: (2026)
Learning Perceptually Relevant Temporal Envelope Morphing
von: Dixit, Satvik, et al.
Veröffentlicht: (2025)
von: Dixit, Satvik, et al.
Veröffentlicht: (2025)
Text2FX: Harnessing CLAP Embeddings for Text-Guided Audio Effects
von: Chu, Annie, et al.
Veröffentlicht: (2024)
von: Chu, Annie, et al.
Veröffentlicht: (2024)
VoxMorph: Scalable Zero-shot Voice Identity Morphing via Disentangled Embeddings
von: Krishnamurthy, Bharath, et al.
Veröffentlicht: (2026)
von: Krishnamurthy, Bharath, et al.
Veröffentlicht: (2026)
Towards Controllable Audio Texture Morphing
von: Gupta, Chitralekha, et al.
Veröffentlicht: (2023)
von: Gupta, Chitralekha, et al.
Veröffentlicht: (2023)
Taming Audio VAEs via Target-KL Regularization
von: Seetharaman, Prem, et al.
Veröffentlicht: (2026)
von: Seetharaman, Prem, et al.
Veröffentlicht: (2026)
Exploring Musical Roots: Applying Audio Embeddings to Empower Influence Attribution for a Generative Music Model
von: Barnett, Julia, et al.
Veröffentlicht: (2024)
von: Barnett, Julia, et al.
Veröffentlicht: (2024)
Ethics Statements in AI Music Papers: The Effective and the Ineffective
von: Barnett, Julia, et al.
Veröffentlicht: (2025)
von: Barnett, Julia, et al.
Veröffentlicht: (2025)
Audio Hallucination Attacks: Probing the Reliability of Large Audio Language Models
von: Seth, Ashish, et al.
Veröffentlicht: (2026)
von: Seth, Ashish, et al.
Veröffentlicht: (2026)
Patch-Mix Contrastive Learning with Audio Spectrogram Transformer on Respiratory Sound Classification
von: Bae, Sangmin, et al.
Veröffentlicht: (2023)
von: Bae, Sangmin, et al.
Veröffentlicht: (2023)
SoundShift: Exploring Sound Manipulations for Accessible Mixed-Reality Awareness
von: Chang, Ruei-Che, et al.
Veröffentlicht: (2024)
von: Chang, Ruei-Che, et al.
Veröffentlicht: (2024)
Sound Field Translation and Mixed Source Model for Virtual Applications with Perceptual Validation
von: Birnie, Lachlan, et al.
Veröffentlicht: (2020)
von: Birnie, Lachlan, et al.
Veröffentlicht: (2020)
ReCLAP: Improving Zero Shot Audio Classification by Describing Sounds
von: Ghosh, Sreyan, et al.
Veröffentlicht: (2024)
von: Ghosh, Sreyan, et al.
Veröffentlicht: (2024)
MT-HuBERT: Self-Supervised Mix-Training for Few-Shot Keyword Spotting in Mixed Speech
von: Yuan, Junming, et al.
Veröffentlicht: (2025)
von: Yuan, Junming, et al.
Veröffentlicht: (2025)
Few-Shot Keyword Spotting from Mixed Speech
von: Yuan, Junming, et al.
Veröffentlicht: (2024)
von: Yuan, Junming, et al.
Veröffentlicht: (2024)
MixFake: Benchmarking and Enhancing Audio Deepfake Detection in Diverse Real-world Mixed Audio
von: Li, Qingcao, et al.
Veröffentlicht: (2026)
von: Li, Qingcao, et al.
Veröffentlicht: (2026)
WhAM: Towards A Translative Model of Sperm Whale Vocalization
von: Paradise, Orr, et al.
Veröffentlicht: (2025)
von: Paradise, Orr, et al.
Veröffentlicht: (2025)
Searching For Music Mixing Graphs: A Pruning Approach
von: Lee, Sungho, et al.
Veröffentlicht: (2024)
von: Lee, Sungho, et al.
Veröffentlicht: (2024)
High-Fidelity Neural Phonetic Posteriorgrams
von: Churchwell, Cameron, et al.
Veröffentlicht: (2024)
von: Churchwell, Cameron, et al.
Veröffentlicht: (2024)
Maximum Likelihood Estimation of the Direction of Sound In A Reverberant Noisy Environment
von: Mansour, Mohamed F.
Veröffentlicht: (2024)
von: Mansour, Mohamed F.
Veröffentlicht: (2024)
Audio Flamingo Sound-CoT Technical Report: Improving Chain-of-Thought Reasoning in Sound Understanding
von: Kong, Zhifeng, et al.
Veröffentlicht: (2025)
von: Kong, Zhifeng, et al.
Veröffentlicht: (2025)
Sound Masking Strategies for Interference with Mosquito Hearing
von: Faber, Justin, et al.
Veröffentlicht: (2025)
von: Faber, Justin, et al.
Veröffentlicht: (2025)
Sound Check: Auditing Audio Datasets
von: Agnew, William, et al.
Veröffentlicht: (2024)
von: Agnew, William, et al.
Veröffentlicht: (2024)
Diff-MST: Differentiable Mixing Style Transfer
von: Vanka, Soumya Sai, et al.
Veröffentlicht: (2024)
von: Vanka, Soumya Sai, et al.
Veröffentlicht: (2024)
MixAssist: An Audio-Language Dataset for Co-Creative AI Assistance in Music Mixing
von: Clemens, Michael, et al.
Veröffentlicht: (2025)
von: Clemens, Michael, et al.
Veröffentlicht: (2025)
Passive Acoustic Monitoring of Noisy Coral Reefs
von: Vishnu, Hari, et al.
Veröffentlicht: (2025)
von: Vishnu, Hari, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Generative Audio Extension and Morphing
von: Seetharaman, Prem, et al.
Veröffentlicht: (2026) -
Sketch2Sound: Controllable Audio Generation via Time-Varying Signals and Sonic Imitations
von: García, Hugo Flores, et al.
Veröffentlicht: (2024) -
Audiocards: Structured Metadata Improves Audio Language Models For Sound Design
von: Sridhar, Sripathi, et al.
Veröffentlicht: (2026) -
SILA: Signal-to-Language Augmentation for Enhanced Control in Text-to-Audio Generation
von: Kumar, Sonal, et al.
Veröffentlicht: (2024) -
Video-Guided Foley Sound Generation with Multimodal Controls
von: Chen, Ziyang, et al.
Veröffentlicht: (2024)