Blind Audio Bandwidth Extension: A Diffusion-Based Zero-Shot Approach
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Moliner, Eloi, Elvander, Filip, Välimäki, Vesa |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Diffusion-Based Generative Equalizer for Music Restoration
von: Moliner, Eloi, et al.
Veröffentlicht: (2024)
von: Moliner, Eloi, et al.
Veröffentlicht: (2024)
Diffusion-Based Audio Inpainting
von: Moliner, Eloi, et al.
Veröffentlicht: (2023)
von: Moliner, Eloi, et al.
Veröffentlicht: (2023)
Similarity-Guided Diffusion for Long-Gap Music Inpainting
von: Turland, Sean, et al.
Veröffentlicht: (2025)
von: Turland, Sean, et al.
Veröffentlicht: (2025)
BUDDy: Single-Channel Blind Unsupervised Dereverberation with Diffusion Models
von: Moliner, Eloi, et al.
Veröffentlicht: (2024)
von: Moliner, Eloi, et al.
Veröffentlicht: (2024)
Unsupervised Blind Joint Dereverberation and Room Acoustics Estimation with Diffusion Models
von: Lemercier, Jean-Marie, et al.
Veröffentlicht: (2024)
von: Lemercier, Jean-Marie, et al.
Veröffentlicht: (2024)
Diffusion Models for Audio Restoration
von: Lemercier, Jean-Marie, et al.
Veröffentlicht: (2024)
von: Lemercier, Jean-Marie, et al.
Veröffentlicht: (2024)
HRTF Estimation using a Score-based Prior
von: Thuillier, Etienne, et al.
Veröffentlicht: (2024)
von: Thuillier, Etienne, et al.
Veröffentlicht: (2024)
Gaussian Flow Bridges for Audio Domain Transfer with Unpaired Data
von: Moliner, Eloi, et al.
Veröffentlicht: (2024)
von: Moliner, Eloi, et al.
Veröffentlicht: (2024)
Resampling Filter Design for Multirate Neural Audio Effect Processing
von: Carson, Alistair, et al.
Veröffentlicht: (2025)
von: Carson, Alistair, et al.
Veröffentlicht: (2025)
Fast and Flexible Audio Bandwidth Extension via Vocos
von: Sharma, Yatharth
Veröffentlicht: (2026)
von: Sharma, Yatharth
Veröffentlicht: (2026)
An Octave-based Multi-Resolution CQT Architecture for Diffusion-based Audio Generation
von: da Costa, Maurício do V. M., et al.
Veröffentlicht: (2025)
von: da Costa, Maurício do V. M., et al.
Veröffentlicht: (2025)
Zero Shot Audio to Audio Emotion Transfer With Speaker Disentanglement
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
A Method for Capturing and Reproducing Directional Reverberation in Six Degrees of Freedom
von: Alary, Benoit, et al.
Veröffentlicht: (2021)
von: Alary, Benoit, et al.
Veröffentlicht: (2021)
Zero-Shot Unsupervised and Text-Based Audio Editing Using DDPM Inversion
von: Manor, Hila, et al.
Veröffentlicht: (2024)
von: Manor, Hila, et al.
Veröffentlicht: (2024)
Automatic Music Mixing using a Generative Model of Effect Embeddings
von: Moliner, Eloi, et al.
Veröffentlicht: (2025)
von: Moliner, Eloi, et al.
Veröffentlicht: (2025)
Multi-label Zero-Shot Audio Classification with Temporal Attention
von: Dogan, Duygu, et al.
Veröffentlicht: (2024)
von: Dogan, Duygu, et al.
Veröffentlicht: (2024)
On Class Separability Pitfalls In Audio-Text Contrastive Zero-Shot Learning
von: Tavares, Tiago, et al.
Veröffentlicht: (2024)
von: Tavares, Tiago, et al.
Veröffentlicht: (2024)
Listenable Maps for Zero-Shot Audio Classifiers
von: Paissan, Francesco, et al.
Veröffentlicht: (2024)
von: Paissan, Francesco, et al.
Veröffentlicht: (2024)
Embedding-Space Diffusion for Zero-Shot Environmental Sound Classification
von: Sims, Ysobel, et al.
Veröffentlicht: (2024)
von: Sims, Ysobel, et al.
Veröffentlicht: (2024)
Sample Rate Independent Recurrent Neural Networks for Audio Effects Processing
von: Carson, Alistair, et al.
Veröffentlicht: (2024)
von: Carson, Alistair, et al.
Veröffentlicht: (2024)
Estimation and Restoration of Unknown Nonlinear Distortion using Diffusion
von: Švento, Michal, et al.
Veröffentlicht: (2025)
von: Švento, Michal, et al.
Veröffentlicht: (2025)
UBGAN: Enhancing Coded Speech with Blind and Guided Bandwidth Extension
von: Gupta, Kishan, et al.
Veröffentlicht: (2025)
von: Gupta, Kishan, et al.
Veröffentlicht: (2025)
Unsupervised Estimation of Nonlinear Audio Effects: Comparing Diffusion-Based and Adversarial approaches
von: Moliner, Eloi, et al.
Veröffentlicht: (2025)
von: Moliner, Eloi, et al.
Veröffentlicht: (2025)
Boundary-Informed Sound Field Reconstruction
von: Sundström, David, et al.
Veröffentlicht: (2025)
von: Sundström, David, et al.
Veröffentlicht: (2025)
Vector Quantized Diffusion Model Based Speech Bandwidth Extension
von: Fang, Yuan, et al.
Veröffentlicht: (2024)
von: Fang, Yuan, et al.
Veröffentlicht: (2024)
Audio Flamingo: A Novel Audio Language Model with Few-Shot Learning and Dialogue Abilities
von: Kong, Zhifeng, et al.
Veröffentlicht: (2024)
von: Kong, Zhifeng, et al.
Veröffentlicht: (2024)
MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
von: Jiang, Ziyue, et al.
Veröffentlicht: (2025)
von: Jiang, Ziyue, et al.
Veröffentlicht: (2025)
Zero-Shot Mono-to-Binaural Speech Synthesis
von: Levkovitch, Alon, et al.
Veröffentlicht: (2024)
von: Levkovitch, Alon, et al.
Veröffentlicht: (2024)
Optimizing tiny colorless feedback delay networks
von: Santo, Gloria Dal, et al.
Veröffentlicht: (2024)
von: Santo, Gloria Dal, et al.
Veröffentlicht: (2024)
On the Transferability of Large-Scale Self-Supervision to Few-Shot Audio Classification
von: Heggan, Calum, et al.
Veröffentlicht: (2024)
von: Heggan, Calum, et al.
Veröffentlicht: (2024)
Zero-Shot Multi-Lingual Speaker Verification in Clinical Trials
von: Akram, Ali, et al.
Veröffentlicht: (2024)
von: Akram, Ali, et al.
Veröffentlicht: (2024)
Multi-modal Adversarial Training for Zero-Shot Voice Cloning
von: Janiczek, John, et al.
Veröffentlicht: (2024)
von: Janiczek, John, et al.
Veröffentlicht: (2024)
Exploring Meta Information for Audio-based Zero-shot Bird Classification
von: Gebhard, Alexander, et al.
Veröffentlicht: (2023)
von: Gebhard, Alexander, et al.
Veröffentlicht: (2023)
Fast Timing-Conditioned Latent Audio Diffusion
von: Evans, Zach, et al.
Veröffentlicht: (2024)
von: Evans, Zach, et al.
Veröffentlicht: (2024)
LC-Protonets: Multi-Label Few-Shot Learning for World Music Audio Tagging
von: Papaioannou, Charilaos, et al.
Veröffentlicht: (2024)
von: Papaioannou, Charilaos, et al.
Veröffentlicht: (2024)
Estimated Audio-Caption Correspondences Improve Language-Based Audio Retrieval
von: Primus, Paul, et al.
Veröffentlicht: (2024)
von: Primus, Paul, et al.
Veröffentlicht: (2024)
Multi-Source Localization and Data Association for Time-Difference of Arrival Measurements
von: Flood, Gabrielle, et al.
Veröffentlicht: (2024)
von: Flood, Gabrielle, et al.
Veröffentlicht: (2024)
Zero-shot Voice Conversion with Diffusion Transformers
von: Liu, Songting
Veröffentlicht: (2024)
von: Liu, Songting
Veröffentlicht: (2024)
Music Boomerang: Reusing Diffusion Models for Data Augmentation and Audio Manipulation
von: Fichtinger, Alexander, et al.
Veröffentlicht: (2025)
von: Fichtinger, Alexander, et al.
Veröffentlicht: (2025)
Audio Texture Manipulation by Exemplar-Based Analogy
von: Cheng, Kan Jen, et al.
Veröffentlicht: (2025)
von: Cheng, Kan Jen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Diffusion-Based Generative Equalizer for Music Restoration
von: Moliner, Eloi, et al.
Veröffentlicht: (2024) -
Diffusion-Based Audio Inpainting
von: Moliner, Eloi, et al.
Veröffentlicht: (2023) -
Similarity-Guided Diffusion for Long-Gap Music Inpainting
von: Turland, Sean, et al.
Veröffentlicht: (2025) -
BUDDy: Single-Channel Blind Unsupervised Dereverberation with Diffusion Models
von: Moliner, Eloi, et al.
Veröffentlicht: (2024) -
Unsupervised Blind Joint Dereverberation and Room Acoustics Estimation with Diffusion Models
von: Lemercier, Jean-Marie, et al.
Veröffentlicht: (2024)