Open-Amp: Synthetic Data Framework for Audio Effect Foundation Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wright, Alec, Carson, Alistair, Juvela, Lauri |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Collaborative Watermarking for Adversarial Speech Synthesis
von: Juvela, Lauri, et al.
Veröffentlicht: (2023)
von: Juvela, Lauri, et al.
Veröffentlicht: (2023)
Gradient-based Optimisation of Modulation Effects
von: Carson, Alistair, et al.
Veröffentlicht: (2026)
von: Carson, Alistair, et al.
Veröffentlicht: (2026)
Resampling Filter Design for Multirate Neural Audio Effect Processing
von: Carson, Alistair, et al.
Veröffentlicht: (2025)
von: Carson, Alistair, et al.
Veröffentlicht: (2025)
Audio Codec Augmentation for Robust Collaborative Watermarking of Speech Synthesis
von: Juvela, Lauri, et al.
Veröffentlicht: (2024)
von: Juvela, Lauri, et al.
Veröffentlicht: (2024)
Interpolation Filter Design for Sample Rate Independent Audio Effect RNNs
von: Carson, Alistair, et al.
Veröffentlicht: (2024)
von: Carson, Alistair, et al.
Veröffentlicht: (2024)
End-to-End Amp Modeling: From Data to Controllable Guitar Amplifier Models
von: Juvela, Lauri, et al.
Veröffentlicht: (2024)
von: Juvela, Lauri, et al.
Veröffentlicht: (2024)
Sample Rate Independent Recurrent Neural Networks for Audio Effects Processing
von: Carson, Alistair, et al.
Veröffentlicht: (2024)
von: Carson, Alistair, et al.
Veröffentlicht: (2024)
NatureLM-audio: an Audio-Language Foundation Model for Bioacoustics
von: Robinson, David, et al.
Veröffentlicht: (2024)
von: Robinson, David, et al.
Veröffentlicht: (2024)
Sparse Autoencoders Make Audio Foundation Models more Explainable
von: Mariotte, Théo, et al.
Veröffentlicht: (2025)
von: Mariotte, Théo, et al.
Veröffentlicht: (2025)
Towards Open Respiratory Acoustic Foundation Models: Pretraining and Benchmarking
von: Zhang, Yuwei, et al.
Veröffentlicht: (2024)
von: Zhang, Yuwei, et al.
Veröffentlicht: (2024)
Unsupervised Estimation of Nonlinear Audio Effects: Comparing Diffusion-Based and Adversarial approaches
von: Moliner, Eloi, et al.
Veröffentlicht: (2025)
von: Moliner, Eloi, et al.
Veröffentlicht: (2025)
Guiding Audio Editing with Audio Language Model
von: Lan, Zitong, et al.
Veröffentlicht: (2025)
von: Lan, Zitong, et al.
Veröffentlicht: (2025)
AU-Harness: An Open-Source Toolkit for Holistic Evaluation of Audio LLMs
von: Nguyen, Hoang, et al.
Veröffentlicht: (2025)
von: Nguyen, Hoang, et al.
Veröffentlicht: (2025)
Parametric Neural Amp Modeling with Active Learning
von: Grötschla, Florian, et al.
Veröffentlicht: (2025)
von: Grötschla, Florian, et al.
Veröffentlicht: (2025)
AudioGenX: Explainability on Text-to-Audio Generative Models
von: Kang, Hyunju, et al.
Veröffentlicht: (2025)
von: Kang, Hyunju, et al.
Veröffentlicht: (2025)
Can Synthetic Audio From Generative Foundation Models Assist Audio Recognition and Speech Modeling?
von: Feng, Tiantian, et al.
Veröffentlicht: (2024)
von: Feng, Tiantian, et al.
Veröffentlicht: (2024)
Differentiable All-pole Filters for Time-varying Audio Systems
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2024)
Tuning In: Analysis of Audio Classifier Performance in Clinical Settings with Limited Data
von: Mahdi, Hamza, et al.
Veröffentlicht: (2024)
von: Mahdi, Hamza, et al.
Veröffentlicht: (2024)
UltraEval-Audio: A Unified Framework for Comprehensive Evaluation of Audio Foundation Models
von: Shi, Qundong, et al.
Veröffentlicht: (2026)
von: Shi, Qundong, et al.
Veröffentlicht: (2026)
From Alignment to Advancement: Bootstrapping Audio-Language Alignment with Synthetic Data
von: Kuan, Chun-Yi, et al.
Veröffentlicht: (2025)
von: Kuan, Chun-Yi, et al.
Veröffentlicht: (2025)
De-crackling Virtual Analog Controls with Asymptotically Stable Recurrent Neural Networks
von: Kallinen, Valtteri, et al.
Veröffentlicht: (2025)
von: Kallinen, Valtteri, et al.
Veröffentlicht: (2025)
Prompt-guided Precise Audio Editing with Diffusion Models
von: Xu, Manjie, et al.
Veröffentlicht: (2024)
von: Xu, Manjie, et al.
Veröffentlicht: (2024)
Benchmarking Language Modeling for Lossless Compression of Full-Fidelity Audio
von: Long, Phillip, et al.
Veröffentlicht: (2026)
von: Long, Phillip, et al.
Veröffentlicht: (2026)
Text-Queried Audio Source Separation via Hierarchical Modeling
von: Yin, Xinlei, et al.
Veröffentlicht: (2025)
von: Yin, Xinlei, et al.
Veröffentlicht: (2025)
PoDAR: Power-Disentangled Audio Representation for Generative Modeling
von: Luebs, Alejandro, et al.
Veröffentlicht: (2026)
von: Luebs, Alejandro, et al.
Veröffentlicht: (2026)
Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework
von: Zhang, Kuiyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Kuiyuan, et al.
Veröffentlicht: (2025)
A Framework for Synthetic Audio Conversations Generation using Large Language Models
von: Kyaw, Kaung Myat, et al.
Veröffentlicht: (2024)
von: Kyaw, Kaung Myat, et al.
Veröffentlicht: (2024)
SSLAM: Enhancing Self-Supervised Models with Audio Mixtures for Polyphonic Soundscapes
von: Alex, Tony, et al.
Veröffentlicht: (2025)
von: Alex, Tony, et al.
Veröffentlicht: (2025)
Diffused Responsibility: Analyzing the Energy Consumption of Generative Text-to-Audio Diffusion Models
von: Passoni, Riccardo, et al.
Veröffentlicht: (2025)
von: Passoni, Riccardo, et al.
Veröffentlicht: (2025)
$\texttt{AVROBUSTBENCH}$: Benchmarking the Robustness of Audio-Visual Recognition Models at Test-Time
von: Maharana, Sarthak Kumar, et al.
Veröffentlicht: (2025)
von: Maharana, Sarthak Kumar, et al.
Veröffentlicht: (2025)
Scaling Ambiguity: Augmenting Human Annotation in Speech Emotion Recognition with Audio-Language Models
von: Zhang, Wenda, et al.
Veröffentlicht: (2026)
von: Zhang, Wenda, et al.
Veröffentlicht: (2026)
Tuberculosis Screening from Cough Audio: Baseline Models, Clinical Variables, and Uncertainty Quantification
von: Kafentzis, George P., et al.
Veröffentlicht: (2026)
von: Kafentzis, George P., et al.
Veröffentlicht: (2026)
HEAR: Holistic Evaluation of Audio Representations
von: Turian, Joseph, et al.
Veröffentlicht: (2022)
von: Turian, Joseph, et al.
Veröffentlicht: (2022)
Enhancing Synthetic Training Data for Speech Commands: From ASR-Based Filtering to Domain Adaptation in SSL Latent Space
von: Quintas, Sebastião, et al.
Veröffentlicht: (2024)
von: Quintas, Sebastião, et al.
Veröffentlicht: (2024)
Learning Source Disentanglement in Neural Audio Codec
von: Bie, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Bie, Xiaoyu, et al.
Veröffentlicht: (2024)
TACNET: Temporal Audio Source Counting Network
von: Ahmadnejad, Amirreza, et al.
Veröffentlicht: (2023)
von: Ahmadnejad, Amirreza, et al.
Veröffentlicht: (2023)
Exploring and Applying Audio-Based Sentiment Analysis in Music
von: Jhanji, Etash
Veröffentlicht: (2024)
von: Jhanji, Etash
Veröffentlicht: (2024)
A Survey of Deep Learning Audio Generation Methods
von: Božić, Matej, et al.
Veröffentlicht: (2024)
von: Božić, Matej, et al.
Veröffentlicht: (2024)
Training-Free Multi-Step Audio Source Separation
von: Zang, Yongyi, et al.
Veröffentlicht: (2025)
von: Zang, Yongyi, et al.
Veröffentlicht: (2025)
On Temporal Guidance and Iterative Refinement in Audio Source Separation
von: Morocutti, Tobias, et al.
Veröffentlicht: (2025)
von: Morocutti, Tobias, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Collaborative Watermarking for Adversarial Speech Synthesis
von: Juvela, Lauri, et al.
Veröffentlicht: (2023) -
Gradient-based Optimisation of Modulation Effects
von: Carson, Alistair, et al.
Veröffentlicht: (2026) -
Resampling Filter Design for Multirate Neural Audio Effect Processing
von: Carson, Alistair, et al.
Veröffentlicht: (2025) -
Audio Codec Augmentation for Robust Collaborative Watermarking of Speech Synthesis
von: Juvela, Lauri, et al.
Veröffentlicht: (2024) -
Interpolation Filter Design for Sample Rate Independent Audio Effect RNNs
von: Carson, Alistair, et al.
Veröffentlicht: (2024)