QINCODEC: Neural Audio Compression with Implicit Neural Codebooks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lahrichi, Zineb, Hadjeres, Gaëtan, Richard, Gael, Peeters, Geoffroy |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
S-PRESSO: Ultra Low Bitrate Sound Effect Compression With Diffusion Autoencoders And Offline Quantization
von: Lahrichi, Zineb, et al.
Veröffentlicht: (2026)
von: Lahrichi, Zineb, et al.
Veröffentlicht: (2026)
The Inverse Drum Machine: Source Separation Through Joint Transcription and Analysis-by-Synthesis
von: Torres, Bernardo, et al.
Veröffentlicht: (2025)
von: Torres, Bernardo, et al.
Veröffentlicht: (2025)
Unsupervised Harmonic Parameter Estimation Using Differentiable DSP and Spectral Optimal Transport
von: Torres, Bernardo, et al.
Veröffentlicht: (2023)
von: Torres, Bernardo, et al.
Veröffentlicht: (2023)
Episodic fine-tuning prototypical networks for optimization-based few-shot learning: Application to audio classification
von: Zhuang, Xuanyu, et al.
Veröffentlicht: (2024)
von: Zhuang, Xuanyu, et al.
Veröffentlicht: (2024)
Investigating Design Choices in Joint-Embedding Predictive Architectures for General Audio Representation Learning
von: Riou, Alain, et al.
Veröffentlicht: (2024)
von: Riou, Alain, et al.
Veröffentlicht: (2024)
PESTO: Pitch Estimation with Self-supervised Transposition-equivariant Objective
von: Riou, Alain, et al.
Veröffentlicht: (2023)
von: Riou, Alain, et al.
Veröffentlicht: (2023)
PESTO: Real-Time Pitch Estimation with Self-supervised Transposition-equivariant Objective
von: Riou, Alain, et al.
Veröffentlicht: (2025)
von: Riou, Alain, et al.
Veröffentlicht: (2025)
Zero-shot Musical Stem Retrieval with Joint-Embedding Predictive Architectures
von: Riou, Alain, et al.
Veröffentlicht: (2024)
von: Riou, Alain, et al.
Veröffentlicht: (2024)
Stem-JEPA: A Joint-Embedding Predictive Architecture for Musical Stem Compatibility Estimation
von: Riou, Alain, et al.
Veröffentlicht: (2024)
von: Riou, Alain, et al.
Veröffentlicht: (2024)
Using Random Codebooks for Audio Neural AutoEncoders
von: Giniès, Benoît, et al.
Veröffentlicht: (2024)
von: Giniès, Benoît, et al.
Veröffentlicht: (2024)
Compressing Quaternion Convolutional Neural Networks for Audio Classification
von: Singh, Arshdeep, et al.
Veröffentlicht: (2025)
von: Singh, Arshdeep, et al.
Veröffentlicht: (2025)
Aliasing-Free Neural Audio Synthesis
von: Gu, Yicheng, et al.
Veröffentlicht: (2025)
von: Gu, Yicheng, et al.
Veröffentlicht: (2025)
Low-Complexity Neural Wind Noise Reduction for Audio Recordings
von: Eftekhari, Hesam, et al.
Veröffentlicht: (2025)
von: Eftekhari, Hesam, et al.
Veröffentlicht: (2025)
Sample Rate Independent Recurrent Neural Networks for Audio Effects Processing
von: Carson, Alistair, et al.
Veröffentlicht: (2024)
von: Carson, Alistair, et al.
Veröffentlicht: (2024)
Latent Granular Resynthesis using Neural Audio Codecs
von: Tokui, Nao, et al.
Veröffentlicht: (2025)
von: Tokui, Nao, et al.
Veröffentlicht: (2025)
Resampling Filter Design for Multirate Neural Audio Effect Processing
von: Carson, Alistair, et al.
Veröffentlicht: (2025)
von: Carson, Alistair, et al.
Veröffentlicht: (2025)
Audio Compression using Periodic Gabor with Biorthogonal Exchange: Implementation Using the Zak Transform
von: Alimi, Roger, et al.
Veröffentlicht: (2025)
von: Alimi, Roger, et al.
Veröffentlicht: (2025)
Speech dereverberation constrained on room impulse response characteristics
von: Bahrman, Louis, et al.
Veröffentlicht: (2024)
von: Bahrman, Louis, et al.
Veröffentlicht: (2024)
A Generalized Bandsplit Neural Network for Cinematic Audio Source Separation
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2023)
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2023)
A Multi-decoder Neural Tracking Method for Accurately Predicting Speech Intelligibility
von: Sonck, Rien, et al.
Veröffentlicht: (2026)
von: Sonck, Rien, et al.
Veröffentlicht: (2026)
Neural Speech and Audio Coding: Modern AI Technology Meets Traditional Codecs
von: Kim, Minje, et al.
Veröffentlicht: (2024)
von: Kim, Minje, et al.
Veröffentlicht: (2024)
A Hybrid Model for Weakly-Supervised Speech Dereverberation
von: Bahrman, Louis, et al.
Veröffentlicht: (2025)
von: Bahrman, Louis, et al.
Veröffentlicht: (2025)
Audio-Visual Speech Enhancement: Architectural Design and Deployment Strategies
von: Hamadouche, Anis, et al.
Veröffentlicht: (2025)
von: Hamadouche, Anis, et al.
Veröffentlicht: (2025)
Diff-TONE: Timestep Optimization for iNstrument Editing in Text-to-Music Diffusion Models
von: Baoueb, Teysir, et al.
Veröffentlicht: (2025)
von: Baoueb, Teysir, et al.
Veröffentlicht: (2025)
U-DREAM: Unsupervised Dereverberation guided by a Reverberation Model
von: Bahrman, Louis, et al.
Veröffentlicht: (2025)
von: Bahrman, Louis, et al.
Veröffentlicht: (2025)
GLA-Grad++: An Improved Griffin-Lim Guided Diffusion Model for Speech Synthesis
von: Baoueb, Teysir, et al.
Veröffentlicht: (2025)
von: Baoueb, Teysir, et al.
Veröffentlicht: (2025)
Physics-Informed Direction-Aware Neural Acoustic Fields
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2025)
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2025)
Aliasing Reduction in Neural Amp Modeling by Smoothing Activations
von: Sato, Ryota, et al.
Veröffentlicht: (2025)
von: Sato, Ryota, et al.
Veröffentlicht: (2025)
Relating the Neural Representations of Vocalized, Mimed, and Imagined Speech
von: Maghsoudi, Maryam, et al.
Veröffentlicht: (2026)
von: Maghsoudi, Maryam, et al.
Veröffentlicht: (2026)
ANIRA: An Architecture for Neural Network Inference in Real-Time Audio Applications
von: Ackva, Valentin, et al.
Veröffentlicht: (2025)
von: Ackva, Valentin, et al.
Veröffentlicht: (2025)
On Improving Error Resilience of Neural End-to-End Speech Coders
von: Gupta, Kishan, et al.
Veröffentlicht: (2024)
von: Gupta, Kishan, et al.
Veröffentlicht: (2024)
Translation-Equivariant Self-Supervised Learning for Pitch Estimation with Optimal Transport
von: Torres, Bernardo, et al.
Veröffentlicht: (2025)
von: Torres, Bernardo, et al.
Veröffentlicht: (2025)
Velocity Potential Neural Field for Efficient Ambisonics Impulse Response Modeling
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2026)
von: Masuyama, Yoshiki, et al.
Veröffentlicht: (2026)
Multimodal Self-Attention Network with Temporal Alignment for Audio-Visual Emotion Recognition
von: Koo, Inyong, et al.
Veröffentlicht: (2026)
von: Koo, Inyong, et al.
Veröffentlicht: (2026)
GLA-Grad: A Griffin-Lim Extended Waveform Generation Diffusion Model
von: Liu, Haocheng, et al.
Veröffentlicht: (2024)
von: Liu, Haocheng, et al.
Veröffentlicht: (2024)
SpecDiff-GAN: A Spectrally-Shaped Noise Diffusion GAN for Speech and Music Synthesis
von: Baoueb, Teysir, et al.
Veröffentlicht: (2024)
von: Baoueb, Teysir, et al.
Veröffentlicht: (2024)
Audio signal interpolation using optimal transportation of spectrograms
von: Valdivia, David, et al.
Veröffentlicht: (2025)
von: Valdivia, David, et al.
Veröffentlicht: (2025)
STNet: Prediction of Underwater Sound Speed Profiles with An Advanced Semi-Transformer Neural Network
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
QHARMA-GAN: Quasi-Harmonic Neural Vocoder based on Autoregressive Moving Average Model
von: Chen, Shaowen, et al.
Veröffentlicht: (2025)
von: Chen, Shaowen, et al.
Veröffentlicht: (2025)
Beyond Identity: A Generalizable Approach for Deepfake Audio Detection
von: Ahmadiadli, Yasaman, et al.
Veröffentlicht: (2025)
von: Ahmadiadli, Yasaman, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
S-PRESSO: Ultra Low Bitrate Sound Effect Compression With Diffusion Autoencoders And Offline Quantization
von: Lahrichi, Zineb, et al.
Veröffentlicht: (2026) -
The Inverse Drum Machine: Source Separation Through Joint Transcription and Analysis-by-Synthesis
von: Torres, Bernardo, et al.
Veröffentlicht: (2025) -
Unsupervised Harmonic Parameter Estimation Using Differentiable DSP and Spectral Optimal Transport
von: Torres, Bernardo, et al.
Veröffentlicht: (2023) -
Episodic fine-tuning prototypical networks for optimization-based few-shot learning: Application to audio classification
von: Zhuang, Xuanyu, et al.
Veröffentlicht: (2024) -
Investigating Design Choices in Joint-Embedding Predictive Architectures for General Audio Representation Learning
von: Riou, Alain, et al.
Veröffentlicht: (2024)