GRAFX: An Open-Source Library for Audio Processing Graphs in PyTorch
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lee, Sungho, Martínez-Ramírez, Marco, Liao, Wei-Hsiang, Uhlich, Stefan, Fabbro, Giorgio, Lee, Kyogu, Mitsufuji, Yuki |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reverse Engineering of Music Mixing Graphs with Differentiable Processors and Iterative Pruning
von: Lee, Sungho, et al.
Veröffentlicht: (2025)
von: Lee, Sungho, et al.
Veröffentlicht: (2025)
ITO-Master: Inference-Time Optimization for Audio Effects Modeling of Music Mastering Processors
von: Koo, Junghyun, et al.
Veröffentlicht: (2025)
von: Koo, Junghyun, et al.
Veröffentlicht: (2025)
The Whole Is Greater than the Sum of Its Parts: Improving Music Source Separation by Bridging Network
von: Sawata, Ryosuke, et al.
Veröffentlicht: (2023)
von: Sawata, Ryosuke, et al.
Veröffentlicht: (2023)
Rethinking Speech Representation Aggregation in Speech Enhancement: A Phonetic Mutual Information Perspective
von: Han, Seungu, et al.
Veröffentlicht: (2026)
von: Han, Seungu, et al.
Veröffentlicht: (2026)
Latent Diffusion Bridges for Unsupervised Musical Audio Timbre Transfer
von: Mancusi, Michele, et al.
Veröffentlicht: (2024)
von: Mancusi, Michele, et al.
Veröffentlicht: (2024)
Few-step Adversarial Schrödinger Bridge for Generative Speech Enhancement
von: Han, Seungu, et al.
Veröffentlicht: (2025)
von: Han, Seungu, et al.
Veröffentlicht: (2025)
Wavespace: A Highly Explorable Wavetable Generator
von: Lee, Hazounne, et al.
Veröffentlicht: (2024)
von: Lee, Hazounne, et al.
Veröffentlicht: (2024)
Searching For Music Mixing Graphs: A Pruning Approach
von: Lee, Sungho, et al.
Veröffentlicht: (2024)
von: Lee, Sungho, et al.
Veröffentlicht: (2024)
Variable Bitrate Residual Vector Quantization for Audio Coding
von: Chae, Yunkee, et al.
Veröffentlicht: (2024)
von: Chae, Yunkee, et al.
Veröffentlicht: (2024)
Fx-Encoder++: Extracting Instrument-Wise Audio Effects Representations from Mixtures
von: Yeh, Yen-Tung, et al.
Veröffentlicht: (2025)
von: Yeh, Yen-Tung, et al.
Veröffentlicht: (2025)
Can Large Language Models Predict Audio Effects Parameters from Natural Language?
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
Vo-Ve: An Explainable Voice-Vector for Speaker Identity Evaluation
von: Lee, Jaejun, et al.
Veröffentlicht: (2025)
von: Lee, Jaejun, et al.
Veröffentlicht: (2025)
TorchFX: A modern approach to Audio DSP with PyTorch and GPU acceleration
von: Spanio, Matteo, et al.
Veröffentlicht: (2025)
von: Spanio, Matteo, et al.
Veröffentlicht: (2025)
MGE-LDM: Joint Latent Diffusion for Simultaneous Music Generation and Source Extraction
von: Chae, Yunkee, et al.
Veröffentlicht: (2025)
von: Chae, Yunkee, et al.
Veröffentlicht: (2025)
Towards Bitrate-Efficient and Noise-Robust Speech Coding with Variable Bitrate RVQ
von: Chae, Yunkee, et al.
Veröffentlicht: (2025)
von: Chae, Yunkee, et al.
Veröffentlicht: (2025)
Music De-limiter Networks via Sample-wise Gain Inversion
von: Jeon, Chang-Bin, et al.
Veröffentlicht: (2023)
von: Jeon, Chang-Bin, et al.
Veröffentlicht: (2023)
SilentCipher: Deep Audio Watermarking
von: Singh, Mayank Kumar, et al.
Veröffentlicht: (2024)
von: Singh, Mayank Kumar, et al.
Veröffentlicht: (2024)
Learning Semantic Information from Raw Audio Signal Using Both Contextual and Phonetic Representations
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2024)
von: Kim, Jaeyeon, et al.
Veröffentlicht: (2024)
DiffVox: A Differentiable Model for Capturing and Analysing Vocal Effects Distributions
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2025)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2025)
Towards Assessing Data Replication in Music Generation with Music Similarity Metrics on Raw Audio
von: Batlle-Roca, Roser, et al.
Veröffentlicht: (2024)
von: Batlle-Roca, Roser, et al.
Veröffentlicht: (2024)
The Sound Demixing Challenge 2023 $\unicode{x2013}$ Cinematic Demixing Track
von: Uhlich, Stefan, et al.
Veröffentlicht: (2023)
von: Uhlich, Stefan, et al.
Veröffentlicht: (2023)
Do Captioning Metrics Reflect Music Semantic Alignment?
von: Lee, Jinwoo, et al.
Veröffentlicht: (2024)
von: Lee, Jinwoo, et al.
Veröffentlicht: (2024)
SteerMusic: Enhanced Musical Consistency for Zero-shot Text-guided and Personalized Music Editing
von: Niu, Xinlei, et al.
Veröffentlicht: (2025)
von: Niu, Xinlei, et al.
Veröffentlicht: (2025)
Automatic Music Mixing using a Generative Model of Effect Embeddings
von: Moliner, Eloi, et al.
Veröffentlicht: (2025)
von: Moliner, Eloi, et al.
Veröffentlicht: (2025)
Enhancing Neural Audio Fingerprint Robustness to Audio Degradation for Music Identification
von: Araz, R. Oguz, et al.
Veröffentlicht: (2025)
von: Araz, R. Oguz, et al.
Veröffentlicht: (2025)
Removing Speaker Information from Speech Representation using Variable-Length Soft Pooling
von: Hwang, Injune, et al.
Veröffentlicht: (2024)
von: Hwang, Injune, et al.
Veröffentlicht: (2024)
SpecMaskGIT: Masked Generative Modeling of Audio Spectrograms for Efficient Audio Synthesis and Beyond
von: Comunità, Marco, et al.
Veröffentlicht: (2024)
von: Comunità, Marco, et al.
Veröffentlicht: (2024)
Music Foundation Model as Generic Booster for Music Downstream Tasks
von: Liao, WeiHsiang, et al.
Veröffentlicht: (2024)
von: Liao, WeiHsiang, et al.
Veröffentlicht: (2024)
The Sound Demixing Challenge 2023 $\unicode{x2013}$ Music Demixing Track
von: Fabbro, Giorgio, et al.
Veröffentlicht: (2023)
von: Fabbro, Giorgio, et al.
Veröffentlicht: (2023)
Inverse Nonlinearity Compensation of Hyperelastic Deformation in Dielectric Elastomer for Acoustic Actuation
von: Lee, Jin Woo, et al.
Veröffentlicht: (2024)
von: Lee, Jin Woo, et al.
Veröffentlicht: (2024)
Improving Inference-Time Optimisation for Vocal Effects Style Transfer with a Gaussian Prior
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2025)
von: Yu, Chin-Yun, et al.
Veröffentlicht: (2025)
Differentiable Acoustic Radiance Transfer
von: Lee, Sungho, et al.
Veröffentlicht: (2025)
von: Lee, Sungho, et al.
Veröffentlicht: (2025)
SAVGBench: Benchmarking Spatially Aligned Audio-Video Generation
von: Shimada, Kazuki, et al.
Veröffentlicht: (2024)
von: Shimada, Kazuki, et al.
Veröffentlicht: (2024)
Improving Unsupervised Clean-to-Rendered Guitar Tone Transformation Using GANs and Integrated Unaligned Clean Data
von: Chen, Yu-Hua, et al.
Veröffentlicht: (2024)
von: Chen, Yu-Hua, et al.
Veröffentlicht: (2024)
Large-Scale Training Data Attribution for Music Generative Models via Unlearning
von: Choi, Woosung, et al.
Veröffentlicht: (2025)
von: Choi, Woosung, et al.
Veröffentlicht: (2025)
COCOLA: Coherence-Oriented Contrastive Learning of Musical Audio Representations
von: Ciranni, Ruben, et al.
Veröffentlicht: (2024)
von: Ciranni, Ruben, et al.
Veröffentlicht: (2024)
Timbre-Trap: A Low-Resource Framework for Instrument-Agnostic Music Transcription
von: Cwitkowitz, Frank, et al.
Veröffentlicht: (2023)
von: Cwitkowitz, Frank, et al.
Veröffentlicht: (2023)
DOSE : Drum One-Shot Extraction from Music Mixture
von: Hwang, Suntae, et al.
Veröffentlicht: (2025)
von: Hwang, Suntae, et al.
Veröffentlicht: (2025)
Open-Set Source Tracing of Audio Deepfake Systems
von: Klein, Nicholas, et al.
Veröffentlicht: (2025)
von: Klein, Nicholas, et al.
Veröffentlicht: (2025)
FoleyBench: A Benchmark For Video-to-Audio Models
von: Dixit, Satvik, et al.
Veröffentlicht: (2025)
von: Dixit, Satvik, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Reverse Engineering of Music Mixing Graphs with Differentiable Processors and Iterative Pruning
von: Lee, Sungho, et al.
Veröffentlicht: (2025) -
ITO-Master: Inference-Time Optimization for Audio Effects Modeling of Music Mastering Processors
von: Koo, Junghyun, et al.
Veröffentlicht: (2025) -
The Whole Is Greater than the Sum of Its Parts: Improving Music Source Separation by Bridging Network
von: Sawata, Ryosuke, et al.
Veröffentlicht: (2023) -
Rethinking Speech Representation Aggregation in Speech Enhancement: A Phonetic Mutual Information Perspective
von: Han, Seungu, et al.
Veröffentlicht: (2026) -
Latent Diffusion Bridges for Unsupervised Musical Audio Timbre Transfer
von: Mancusi, Michele, et al.
Veröffentlicht: (2024)