Audio synthesizer inversion in symmetric parameter spaces with approximately equivariant flow matching
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hayes, Ben, Saitis, Charalampos, Fazekas, György |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mamba-Diffusion Model with Learnable Wavelet for Controllable Symbolic Music Generation
von: Zhang, Jincheng, et al.
Veröffentlicht: (2025)
von: Zhang, Jincheng, et al.
Veröffentlicht: (2025)
Composer Style-specific Symbolic Music Generation Using Vector Quantized Discrete Diffusion Models
von: Zhang, Jincheng, et al.
Veröffentlicht: (2023)
von: Zhang, Jincheng, et al.
Veröffentlicht: (2023)
Real-time Timbre Remapping with Differentiable DSP
von: Shier, Jordie, et al.
Veröffentlicht: (2024)
von: Shier, Jordie, et al.
Veröffentlicht: (2024)
Time-of-arrival Estimation and Phase Unwrapping of Head-related Transfer Functions With Integer Linear Programming
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024)
Assessing the Alignment of Audio Representations with Timbre Similarity Ratings
von: Tian, Haokun, et al.
Veröffentlicht: (2025)
von: Tian, Haokun, et al.
Veröffentlicht: (2025)
Conditioning and Sampling in Variational Diffusion Models for Speech Super-Resolution
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2022)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2022)
Listenable Maps for Audio Classifiers
von: Paissan, Francesco, et al.
Veröffentlicht: (2024)
von: Paissan, Francesco, et al.
Veröffentlicht: (2024)
Exploring Transformer-Based Music Overpainting for Jazz Piano Variations
von: Row, Eleanor, et al.
Veröffentlicht: (2024)
von: Row, Eleanor, et al.
Veröffentlicht: (2024)
Listenable Maps for Zero-Shot Audio Classifiers
von: Paissan, Francesco, et al.
Veröffentlicht: (2024)
von: Paissan, Francesco, et al.
Veröffentlicht: (2024)
Latent Granular Resynthesis using Neural Audio Codecs
von: Tokui, Nao, et al.
Veröffentlicht: (2025)
von: Tokui, Nao, et al.
Veröffentlicht: (2025)
Differentiable All-pole Filters for Time-varying Audio Systems
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024)
Resampling Filter Design for Multirate Neural Audio Effect Processing
von: Carson, Alistair, et al.
Veröffentlicht: (2025)
von: Carson, Alistair, et al.
Veröffentlicht: (2025)
A Generalized Bandsplit Neural Network for Cinematic Audio Source Separation
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2023)
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2023)
EmotionCaps: Enhancing Audio Captioning Through Emotion-Augmented Data Generation
von: Manivannan, Mithun, et al.
Veröffentlicht: (2024)
von: Manivannan, Mithun, et al.
Veröffentlicht: (2024)
FlowDec: A flow-based full-band general audio codec with high perceptual quality
von: Welker, Simon, et al.
Veröffentlicht: (2025)
von: Welker, Simon, et al.
Veröffentlicht: (2025)
Music2Latent: Consistency Autoencoders for Latent Audio Compression
von: Pasini, Marco, et al.
Veröffentlicht: (2024)
von: Pasini, Marco, et al.
Veröffentlicht: (2024)
WavTokenizer: an Efficient Acoustic Discrete Codec Tokenizer for Audio Language Modeling
von: Ji, Shengpeng, et al.
Veröffentlicht: (2024)
von: Ji, Shengpeng, et al.
Veröffentlicht: (2024)
Hybrid Losses for Hierarchical Embedding Learning
von: Tian, Haokun, et al.
Veröffentlicht: (2025)
von: Tian, Haokun, et al.
Veröffentlicht: (2025)
Bayesian Restoration of Audio Degraded by Low-Frequency Pulses Modeled via Gaussian Process
von: de Carvalho, Hugo Tremonte, et al.
Veröffentlicht: (2020)
von: de Carvalho, Hugo Tremonte, et al.
Veröffentlicht: (2020)
SLAP: Siamese Language-Audio Pretraining Without Negative Samples for Music Understanding
von: Guinot, Julien, et al.
Veröffentlicht: (2025)
von: Guinot, Julien, et al.
Veröffentlicht: (2025)
Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning
von: Tuncay, Ludovic, et al.
Veröffentlicht: (2025)
von: Tuncay, Ludovic, et al.
Veröffentlicht: (2025)
An Explainable Proxy Model for Multiabel Audio Segmentation
von: Mariotte, Théo, et al.
Veröffentlicht: (2024)
von: Mariotte, Théo, et al.
Veröffentlicht: (2024)
Gull: A Generative Multifunctional Audio Codec
von: Luo, Yi, et al.
Veröffentlicht: (2024)
von: Luo, Yi, et al.
Veröffentlicht: (2024)
Singing Voice Synthesis Using Differentiable LPC and Glottal-Flow-Inspired Wavetables
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2023)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2023)
Differentiable Time-Varying Linear Prediction in the Context of End-to-End Analysis-by-Synthesis
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024)
Designing Neural Synthesizers for Low-Latency Interaction
von: Caspe, Franco, et al.
Veröffentlicht: (2025)
von: Caspe, Franco, et al.
Veröffentlicht: (2025)
LMAC-TD: Producing Time Domain Explanations for Audio Classifiers
von: Mancini, Eleonora, et al.
Veröffentlicht: (2024)
von: Mancini, Eleonora, et al.
Veröffentlicht: (2024)
Joint Source-Environment Adaptation for Deep Learning-Based Underwater Acoustic Source Ranging
von: Kari, Dariush, et al.
Veröffentlicht: (2025)
von: Kari, Dariush, et al.
Veröffentlicht: (2025)
Adaptive Control Attention Network for Underwater Acoustic Localization and Domain Adaptation
von: Vo, Quoc Thinh, et al.
Veröffentlicht: (2025)
von: Vo, Quoc Thinh, et al.
Veröffentlicht: (2025)
Diff-TONE: Timestep Optimization for iNstrument Editing in Text-to-Music Diffusion Models
von: Baoueb, Teysir, et al.
Veröffentlicht: (2025)
von: Baoueb, Teysir, et al.
Veröffentlicht: (2025)
Mismatch-Robust Underwater Acoustic Localization Using A Differentiable Modular Forward Model
von: Kari, Dariush, et al.
Veröffentlicht: (2025)
von: Kari, Dariush, et al.
Veröffentlicht: (2025)
Lightweight DNN for Full-Band Speech Denoising on Mobile Devices: Exploiting Long and Short Temporal Patterns
von: Drossos, Konstantinos, et al.
Veröffentlicht: (2025)
von: Drossos, Konstantinos, et al.
Veröffentlicht: (2025)
The Inverse Drum Machine: Source Separation Through Joint Transcription and Analysis-by-Synthesis
von: Torres, Bernardo, et al.
Veröffentlicht: (2025)
von: Torres, Bernardo, et al.
Veröffentlicht: (2025)
XAI-Driven Spectral Analysis of Cough Sounds for Respiratory Disease Characterization
von: Amado-Caballero, Patricia, et al.
Veröffentlicht: (2025)
von: Amado-Caballero, Patricia, et al.
Veröffentlicht: (2025)
GLA-Grad++: An Improved Griffin-Lim Guided Diffusion Model for Speech Synthesis
von: Baoueb, Teysir, et al.
Veröffentlicht: (2025)
von: Baoueb, Teysir, et al.
Veröffentlicht: (2025)
Learnable Adaptive Time-Frequency Representation via Differentiable Short-Time Fourier Transform
von: Leiber, Maxime, et al.
Veröffentlicht: (2025)
von: Leiber, Maxime, et al.
Veröffentlicht: (2025)
AI-Assisted Music Production: A User Study on Text-to-Music Models
von: Ronchini, Francesca, et al.
Veröffentlicht: (2025)
von: Ronchini, Francesca, et al.
Veröffentlicht: (2025)
A Convolutional Framework for Mapping Imagined Auditory MEG into Listened Brain Responses
von: Maghsoudi, Maryam, et al.
Veröffentlicht: (2025)
von: Maghsoudi, Maryam, et al.
Veröffentlicht: (2025)
Real-Time Streaming Mel Vocoding with Generative Flow Matching
von: Welker, Simon, et al.
Veröffentlicht: (2025)
von: Welker, Simon, et al.
Veröffentlicht: (2025)
Joint Source-Environment Adaptation of Data-Driven Underwater Acoustic Source Ranging Based on Model Uncertainty
von: Kari, Dariush, et al.
Veröffentlicht: (2025)
von: Kari, Dariush, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Mamba-Diffusion Model with Learnable Wavelet for Controllable Symbolic Music Generation
von: Zhang, Jincheng, et al.
Veröffentlicht: (2025) -
Composer Style-specific Symbolic Music Generation Using Vector Quantized Discrete Diffusion Models
von: Zhang, Jincheng, et al.
Veröffentlicht: (2023) -
Real-time Timbre Remapping with Differentiable DSP
von: Shier, Jordie, et al.
Veröffentlicht: (2024) -
Time-of-arrival Estimation and Phase Unwrapping of Head-related Transfer Functions With Integer Linear Programming
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024) -
Assessing the Alignment of Audio Representations with Timbre Similarity Ratings
von: Tian, Haokun, et al.
Veröffentlicht: (2025)