Rethinking Non-Negative Matrix Factorization with Implicit Neural Representations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Subramani, Krishna, Smaragdis, Paris, Higuchi, Takuya, Souden, Mehrez |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Combolutional Neural Networks
von: Churchwell, Cameron, et al.
Veröffentlicht: (2025)
von: Churchwell, Cameron, et al.
Veröffentlicht: (2025)
Re-Bottleneck: Latent Re-Structuring for Neural Audio Autoencoders
von: Bralios, Dimitrios, et al.
Veröffentlicht: (2025)
von: Bralios, Dimitrios, et al.
Veröffentlicht: (2025)
Noise-Robust DSP-Assisted Neural Pitch Estimation with Very Low Complexity
von: Subramani, Krishna, et al.
Veröffentlicht: (2023)
von: Subramani, Krishna, et al.
Veröffentlicht: (2023)
ImmerseDiffusion: A Generative Spatial Audio Latent Diffusion Model
von: Heydari, Mojtaba, et al.
Veröffentlicht: (2024)
von: Heydari, Mojtaba, et al.
Veröffentlicht: (2024)
Learning to Upsample and Upmix Audio in the Latent Domain
von: Bralios, Dimitrios, et al.
Veröffentlicht: (2025)
von: Bralios, Dimitrios, et al.
Veröffentlicht: (2025)
Resource-constrained stereo singing voice cancellation
von: Borrelli, Clara, et al.
Veröffentlicht: (2024)
von: Borrelli, Clara, et al.
Veröffentlicht: (2024)
On Class Separability Pitfalls In Audio-Text Contrastive Zero-Shot Learning
von: Tavares, Tiago, et al.
Veröffentlicht: (2024)
von: Tavares, Tiago, et al.
Veröffentlicht: (2024)
Audio Editing with Non-Rigid Text Prompts
von: Paissan, Francesco, et al.
Veröffentlicht: (2023)
von: Paissan, Francesco, et al.
Veröffentlicht: (2023)
Sound Source Separation Using Latent Variational Block-Wise Disentanglement
von: Helwani, Karim, et al.
Veröffentlicht: (2024)
von: Helwani, Karim, et al.
Veröffentlicht: (2024)
Adaptive Slimming for Scalable and Efficient Speech Enhancement
von: Miccini, Riccardo, et al.
Veröffentlicht: (2025)
von: Miccini, Riccardo, et al.
Veröffentlicht: (2025)
Ambisonics Super-Resolution Using A Waveform-Domain Neural Network
von: Nawfal, Ismael, et al.
Veröffentlicht: (2025)
von: Nawfal, Ismael, et al.
Veröffentlicht: (2025)
Scaling Up Adaptive Filter Optimizers
von: Casebeer, Jonah, et al.
Veröffentlicht: (2024)
von: Casebeer, Jonah, et al.
Veröffentlicht: (2024)
StereoFoley: Object-Aware Stereo Audio Generation from Video
von: Karchkhadze, Tornike, et al.
Veröffentlicht: (2025)
von: Karchkhadze, Tornike, et al.
Veröffentlicht: (2025)
User-guided Generative Source Separation
von: Wen, Yutong, et al.
Veröffentlicht: (2025)
von: Wen, Yutong, et al.
Veröffentlicht: (2025)
Large Language Models and Non-Negative Matrix Factorization for Bioacoustic Signal Decomposition
von: Torabi, Yasaman, et al.
Veröffentlicht: (2025)
von: Torabi, Yasaman, et al.
Veröffentlicht: (2025)
Gencho: Room Impulse Response Generation from Reverberant Speech and Text via Diffusion Transformers
von: Lin, Jackie, et al.
Veröffentlicht: (2026)
von: Lin, Jackie, et al.
Veröffentlicht: (2026)
Bayesian Negative Binomial Regression of Afrobeats Chart Persistence
von: Cabansag, Ian Jacob, et al.
Veröffentlicht: (2026)
von: Cabansag, Ian Jacob, et al.
Veröffentlicht: (2026)
HAAQI-Net: A Non-intrusive Neural Music Audio Quality Assessment Model for Hearing Aids
von: Wisnu, Dyah A. M. G., et al.
Veröffentlicht: (2024)
von: Wisnu, Dyah A. M. G., et al.
Veröffentlicht: (2024)
Contextual Speech Extraction: Leveraging Textual History as an Implicit Cue for Target Speech Extraction
von: Kim, Minsu, et al.
Veröffentlicht: (2025)
von: Kim, Minsu, et al.
Veröffentlicht: (2025)
Unsupervised Composable Representations for Audio
von: Bindi, Giovanni, et al.
Veröffentlicht: (2024)
von: Bindi, Giovanni, et al.
Veröffentlicht: (2024)
Learning Disentangled Speech Representations
von: Brima, Yusuf, et al.
Veröffentlicht: (2023)
von: Brima, Yusuf, et al.
Veröffentlicht: (2023)
Feature Representations for Automatic Meerkat Vocalization Classification
von: Mahmoud, Imen Ben, et al.
Veröffentlicht: (2024)
von: Mahmoud, Imen Ben, et al.
Veröffentlicht: (2024)
Benchmarking Representations for Speech, Music, and Acoustic Events
von: La Quatra, Moreno, et al.
Veröffentlicht: (2024)
von: La Quatra, Moreno, et al.
Veröffentlicht: (2024)
Evaluating Disentangled Representations for Controllable Music Generation
von: Ibáñez-Martínez, Laura, et al.
Veröffentlicht: (2026)
von: Ibáñez-Martínez, Laura, et al.
Veröffentlicht: (2026)
Learning Music Audio Representations With Limited Data
von: Plachouras, Christos, et al.
Veröffentlicht: (2025)
von: Plachouras, Christos, et al.
Veröffentlicht: (2025)
Acoustic-to-articulatory inversion for dysarthric speech: Are pre-trained self-supervised representations favorable?
von: Maharana, Sarthak Kumar, et al.
Veröffentlicht: (2023)
von: Maharana, Sarthak Kumar, et al.
Veröffentlicht: (2023)
Multichannel Voice Trigger Detection Based on Transform-average-concatenate
von: Higuchi, Takuya, et al.
Veröffentlicht: (2023)
von: Higuchi, Takuya, et al.
Veröffentlicht: (2023)
Learning Disentangled Audio Representations through Controlled Synthesis
von: Brima, Yusuf, et al.
Veröffentlicht: (2024)
von: Brima, Yusuf, et al.
Veröffentlicht: (2024)
ASTRA: Aligning Speech and Text Representations for Asr without Sampling
von: Gaur, Neeraj, et al.
Veröffentlicht: (2024)
von: Gaur, Neeraj, et al.
Veröffentlicht: (2024)
Singer Identity Representation Learning using Self-Supervised Techniques
von: Torres, Bernardo, et al.
Veröffentlicht: (2024)
von: Torres, Bernardo, et al.
Veröffentlicht: (2024)
COCOLA: Coherence-Oriented Contrastive Learning of Musical Audio Representations
von: Ciranni, Ruben, et al.
Veröffentlicht: (2024)
von: Ciranni, Ruben, et al.
Veröffentlicht: (2024)
Motif Mining and Unsupervised Representation Learning for BirdCLEF 2022
von: Miyaguchi, Anthony, et al.
Veröffentlicht: (2022)
von: Miyaguchi, Anthony, et al.
Veröffentlicht: (2022)
RepCodec: A Speech Representation Codec for Speech Tokenization
von: Huang, Zhichao, et al.
Veröffentlicht: (2023)
von: Huang, Zhichao, et al.
Veröffentlicht: (2023)
PromptSep: Generative Audio Separation via Multimodal Prompting
von: Wen, Yutong, et al.
Veröffentlicht: (2025)
von: Wen, Yutong, et al.
Veröffentlicht: (2025)
Towards the Synthesis of Non-speech Vocalizations
von: Hoq, Enjamamul, et al.
Veröffentlicht: (2024)
von: Hoq, Enjamamul, et al.
Veröffentlicht: (2024)
Speech After Gender: A Trans-Feminine Perspective on Next Steps for Speech Science and Technology
von: Netzorg, Robin, et al.
Veröffentlicht: (2024)
von: Netzorg, Robin, et al.
Veröffentlicht: (2024)
TheGlueNote: Learned Representations for Robust and Flexible Note Alignment
von: Peter, Silvan David, et al.
Veröffentlicht: (2024)
von: Peter, Silvan David, et al.
Veröffentlicht: (2024)
The Effect of Batch Size on Contrastive Self-Supervised Speech Representation Learning
von: Vaessen, Nik, et al.
Veröffentlicht: (2024)
von: Vaessen, Nik, et al.
Veröffentlicht: (2024)
An LSTM-Based Chord Generation System Using Chroma Histogram Representations
von: Hardwick, Jack
Veröffentlicht: (2024)
von: Hardwick, Jack
Veröffentlicht: (2024)
Knowledge boosting during low-latency inference
von: Srinivas, Vidya, et al.
Veröffentlicht: (2024)
von: Srinivas, Vidya, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Combolutional Neural Networks
von: Churchwell, Cameron, et al.
Veröffentlicht: (2025) -
Re-Bottleneck: Latent Re-Structuring for Neural Audio Autoencoders
von: Bralios, Dimitrios, et al.
Veröffentlicht: (2025) -
Noise-Robust DSP-Assisted Neural Pitch Estimation with Very Low Complexity
von: Subramani, Krishna, et al.
Veröffentlicht: (2023) -
ImmerseDiffusion: A Generative Spatial Audio Latent Diffusion Model
von: Heydari, Mojtaba, et al.
Veröffentlicht: (2024) -
Learning to Upsample and Upmix Audio in the Latent Domain
von: Bralios, Dimitrios, et al.
Veröffentlicht: (2025)