AEROMamba: An efficient architecture for audio super-resolution using generative adversarial networks and state space models
Fuente:
arXiv
Salvato in:
| Autori principali: | Abreu, Wallace, Biscainho, Luiz Wagner Pereira |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AudioRepInceptionNeXt: A lightweight single-stream architecture for efficient audio recognition
di: Lau, Kin Wai, et al.
Pubblicazione: (2024)
di: Lau, Kin Wai, et al.
Pubblicazione: (2024)
FiPA-SR -- FiLM-Conditioned Perceptually Informed Audio Super-Resolution
di: Abreu, Wallace, et al.
Pubblicazione: (2026)
di: Abreu, Wallace, et al.
Pubblicazione: (2026)
Deep learning-based filtering of cross-spectral matrices using generative adversarial networks
di: Puhle, Christof
Pubblicazione: (2025)
di: Puhle, Christof
Pubblicazione: (2025)
Bayesian Restoration of Audio Degraded by Low-Frequency Pulses Modeled via Gaussian Process
di: de Carvalho, Hugo Tremonte, et al.
Pubblicazione: (2020)
di: de Carvalho, Hugo Tremonte, et al.
Pubblicazione: (2020)
Speaker anonymization using neural audio codec language models
di: Panariello, Michele, et al.
Pubblicazione: (2023)
di: Panariello, Michele, et al.
Pubblicazione: (2023)
Synthetic training set generation using text-to-audio models for environmental sound classification
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
di: Ronchini, Francesca, et al.
Pubblicazione: (2024)
Towards audio language modeling -- an overview
di: Wu, Haibin, et al.
Pubblicazione: (2024)
di: Wu, Haibin, et al.
Pubblicazione: (2024)
STASE: A spatialized text-to-audio synthesis engine for music generation
di: Chi, Tutti, et al.
Pubblicazione: (2025)
di: Chi, Tutti, et al.
Pubblicazione: (2025)
DashengTokenizer: One layer is enough for unified audio understanding and generation
di: Dinkel, Heinrich, et al.
Pubblicazione: (2026)
di: Dinkel, Heinrich, et al.
Pubblicazione: (2026)
Are audio DeepFake detection models polyglots?
di: Marek, Bartłomiej, et al.
Pubblicazione: (2024)
di: Marek, Bartłomiej, et al.
Pubblicazione: (2024)
Enhancement by postfiltering for speech and audio coding in ad-hoc sensor networks
di: Das, Sneha, et al.
Pubblicazione: (2020)
di: Das, Sneha, et al.
Pubblicazione: (2020)
Modeling strategies for speech enhancement in the latent space of a neural audio codec
di: Kammoun, Sofiene, et al.
Pubblicazione: (2025)
di: Kammoun, Sofiene, et al.
Pubblicazione: (2025)
Scaling up masked audio encoder learning for general audio classification
di: Dinkel, Heinrich, et al.
Pubblicazione: (2024)
di: Dinkel, Heinrich, et al.
Pubblicazione: (2024)
Regularized autoregressive modeling and its application to audio signal reconstruction
di: Mokrý, Ondřej, et al.
Pubblicazione: (2024)
di: Mokrý, Ondřej, et al.
Pubblicazione: (2024)
Visual-based spatial audio generation system for multi-speaker environments
di: Liu, Xiaojing, et al.
Pubblicazione: (2025)
di: Liu, Xiaojing, et al.
Pubblicazione: (2025)
Reconstructing the Charlie Parker Omnibook using an audio-to-score automatic transcription pipeline
di: Riley, Xavier, et al.
Pubblicazione: (2024)
di: Riley, Xavier, et al.
Pubblicazione: (2024)
Discriminating real and synthetic super-resolved audio samples using embedding-based classifiers
di: Silaev, Mikhail, et al.
Pubblicazione: (2026)
di: Silaev, Mikhail, et al.
Pubblicazione: (2026)
Codec-SUPERB @ SLT 2024: A lightweight benchmark for neural audio codec models
di: Wu, Haibin, et al.
Pubblicazione: (2024)
di: Wu, Haibin, et al.
Pubblicazione: (2024)
ParaCLAP -- Towards a general language-audio model for computational paralinguistic tasks
di: Jing, Xin, et al.
Pubblicazione: (2024)
di: Jing, Xin, et al.
Pubblicazione: (2024)
WavJEPA: Semantic learning unlocks robust audio foundation models for raw waveforms
di: Yuksel, Goksenin, et al.
Pubblicazione: (2025)
di: Yuksel, Goksenin, et al.
Pubblicazione: (2025)
Leveraging Real Electric Guitar Tones and Effects to Improve Robustness in Guitar Tablature Transcription Modeling
di: Pedroza, Hegel, et al.
Pubblicazione: (2024)
di: Pedroza, Hegel, et al.
Pubblicazione: (2024)
Self-supervised learning method using multiple sampling strategies for general-purpose audio representation
di: Kuroyanagi, Ibuki, et al.
Pubblicazione: (2025)
di: Kuroyanagi, Ibuki, et al.
Pubblicazione: (2025)
Blind estimation of audio effects using an auto-encoder approach and differentiable digital signal processing
di: Peladeau, Côme, et al.
Pubblicazione: (2023)
di: Peladeau, Côme, et al.
Pubblicazione: (2023)
An overview of neural architectures for self-supervised audio representation learning from masked spectrograms
di: Yadav, Sarthak, et al.
Pubblicazione: (2025)
di: Yadav, Sarthak, et al.
Pubblicazione: (2025)
Tweaking autoregressive methods for inpainting of gaps in audio signals
di: Mokrý, Ondřej, et al.
Pubblicazione: (2024)
di: Mokrý, Ondřej, et al.
Pubblicazione: (2024)
MBCodec:Thorough disentangle for high-fidelity audio compression
di: Zhang, Ruonan, et al.
Pubblicazione: (2025)
di: Zhang, Ruonan, et al.
Pubblicazione: (2025)
Real-time implementation of vibrato transfer as an audio effect
di: Hyrkas, Jeremy
Pubblicazione: (2025)
di: Hyrkas, Jeremy
Pubblicazione: (2025)
A SOUND APPROACH: Using Large Language Models to generate audio descriptions for egocentric text-audio retrieval
di: Oncescu, Andreea-Maria, et al.
Pubblicazione: (2024)
di: Oncescu, Andreea-Maria, et al.
Pubblicazione: (2024)
EDTC: enhance depth of text comprehension in automated audio captioning
di: Tan, Liwen, et al.
Pubblicazione: (2024)
di: Tan, Liwen, et al.
Pubblicazione: (2024)
FxSearcher: gradient-free text-driven audio transformation
di: Ki, Hojoon, et al.
Pubblicazione: (2025)
di: Ki, Hojoon, et al.
Pubblicazione: (2025)
Where are we in audio deepfake detection? A systematic analysis over generative and detection models
di: Li, Xiang, et al.
Pubblicazione: (2024)
di: Li, Xiang, et al.
Pubblicazione: (2024)
Guitar-TECHS: An Electric Guitar Dataset Covering Techniques, Musical Excerpts, Chords and Scales Using a Diverse Array of Hardware
di: Pedroza, Hegel, et al.
Pubblicazione: (2025)
di: Pedroza, Hegel, et al.
Pubblicazione: (2025)
A state-space representation of the boundary integral equation for room acoustic modelling
di: Ali, Randall, et al.
Pubblicazione: (2026)
di: Ali, Randall, et al.
Pubblicazione: (2026)
Testing chatbots on the creation of encoders for audio conditioned image generation
di: León, Jorge E., et al.
Pubblicazione: (2025)
di: León, Jorge E., et al.
Pubblicazione: (2025)
Towards generalizing deep-audio fake detection networks
di: Gasenzer, Konstantin, et al.
Pubblicazione: (2023)
di: Gasenzer, Konstantin, et al.
Pubblicazione: (2023)
AxLSTMs: learning self-supervised audio representations with xLSTMs
di: Yadav, Sarthak, et al.
Pubblicazione: (2024)
di: Yadav, Sarthak, et al.
Pubblicazione: (2024)
Deep learning based spatial aliasing reduction in beamforming for audio capture
di: Guzik, Mateusz, et al.
Pubblicazione: (2025)
di: Guzik, Mateusz, et al.
Pubblicazione: (2025)
Human-CLAP: Human-perception-based contrastive language-audio pretraining
di: Takano, Taisei, et al.
Pubblicazione: (2025)
di: Takano, Taisei, et al.
Pubblicazione: (2025)
ICGAN: An implicit conditioning method for interpretable feature control of neural audio synthesis
di: Liu, Yunyi, et al.
Pubblicazione: (2024)
di: Liu, Yunyi, et al.
Pubblicazione: (2024)
A robust audio deepfake detection system via multi-view feature
di: Yang, Yujie, et al.
Pubblicazione: (2024)
di: Yang, Yujie, et al.
Pubblicazione: (2024)
Documenti analoghi
-
AudioRepInceptionNeXt: A lightweight single-stream architecture for efficient audio recognition
di: Lau, Kin Wai, et al.
Pubblicazione: (2024) -
FiPA-SR -- FiLM-Conditioned Perceptually Informed Audio Super-Resolution
di: Abreu, Wallace, et al.
Pubblicazione: (2026) -
Deep learning-based filtering of cross-spectral matrices using generative adversarial networks
di: Puhle, Christof
Pubblicazione: (2025) -
Bayesian Restoration of Audio Degraded by Low-Frequency Pulses Modeled via Gaussian Process
di: de Carvalho, Hugo Tremonte, et al.
Pubblicazione: (2020) -
Speaker anonymization using neural audio codec language models
di: Panariello, Michele, et al.
Pubblicazione: (2023)