TQCodec: Towards neural audio codec for high-fidelity music streaming
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | He, Lixing, Chen, Zhouxuan, Liu, Mingshuai, Sun, Xinran, Wang, Wucheng, Li, Minfu, Kong, Lingcheng, Zhao, Weifeng, Zhou, Wenjiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Speaker anonymization using neural audio codec language models
von: Panariello, Michele, et al.
Veröffentlicht: (2023)
von: Panariello, Michele, et al.
Veröffentlicht: (2023)
Modeling strategies for speech enhancement in the latent space of a neural audio codec
von: Kammoun, Sofiene, et al.
Veröffentlicht: (2025)
von: Kammoun, Sofiene, et al.
Veröffentlicht: (2025)
Codec-SUPERB @ SLT 2024: A lightweight benchmark for neural audio codec models
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
EnCodecMAE: Leveraging neural codecs for universal audio representation learning
von: Pepino, Leonardo, et al.
Veröffentlicht: (2023)
von: Pepino, Leonardo, et al.
Veröffentlicht: (2023)
Probing neural audio codecs for distinctions among English nuclear tunes
von: Vigneaux, Juan Pablo, et al.
Veröffentlicht: (2026)
von: Vigneaux, Juan Pablo, et al.
Veröffentlicht: (2026)
Switchcodec: Adaptive residual-expert sparse quantization for high-fidelity neural audio coding
von: Wang, Xiangbo, et al.
Veröffentlicht: (2026)
von: Wang, Xiangbo, et al.
Veröffentlicht: (2026)
MBCodec:Thorough disentangle for high-fidelity audio compression
von: Zhang, Ruonan, et al.
Veröffentlicht: (2025)
von: Zhang, Ruonan, et al.
Veröffentlicht: (2025)
FlowDec: A flow-based full-band general audio codec with high perceptual quality
von: Welker, Simon, et al.
Veröffentlicht: (2025)
von: Welker, Simon, et al.
Veröffentlicht: (2025)
FreeCodec: A disentangled neural speech codec with fewer tokens
von: Zheng, Youqiang, et al.
Veröffentlicht: (2024)
von: Zheng, Youqiang, et al.
Veröffentlicht: (2024)
LaDA-Band: Language Diffusion Models for Vocal-to-Accompaniment Generation
von: Wang, Qi, et al.
Veröffentlicht: (2026)
von: Wang, Qi, et al.
Veröffentlicht: (2026)
Audio Dialogues: Dialogues dataset for audio and music understanding
von: Goel, Arushi, et al.
Veröffentlicht: (2024)
von: Goel, Arushi, et al.
Veröffentlicht: (2024)
SMART: Tuning a symbolic music generation system with an audio domain aesthetic reward
von: Jonason, Nicolas, et al.
Veröffentlicht: (2025)
von: Jonason, Nicolas, et al.
Veröffentlicht: (2025)
VibOmni: Towards Scalable Bone-conduction Speech Enhancement on Earables
von: He, Lixing, et al.
Veröffentlicht: (2025)
von: He, Lixing, et al.
Veröffentlicht: (2025)
STASE: A spatialized text-to-audio synthesis engine for music generation
von: Chi, Tutti, et al.
Veröffentlicht: (2025)
von: Chi, Tutti, et al.
Veröffentlicht: (2025)
Towards audio language modeling -- an overview
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
Joint sentiment analysis of lyrics and audio in music
von: Schaab, Lea, et al.
Veröffentlicht: (2024)
von: Schaab, Lea, et al.
Veröffentlicht: (2024)
Poly-SVC: Polyphony-Aware Singing Voice Conversion with Harmonic Modeling
von: Geng, Chen, et al.
Veröffentlicht: (2026)
von: Geng, Chen, et al.
Veröffentlicht: (2026)
AudioRepInceptionNeXt: A lightweight single-stream architecture for efficient audio recognition
von: Lau, Kin Wai, et al.
Veröffentlicht: (2024)
von: Lau, Kin Wai, et al.
Veröffentlicht: (2024)
CoHear: Conversation Enhancement via Multi-Earphone Collaboration
von: He, Lixing, et al.
Veröffentlicht: (2025)
von: He, Lixing, et al.
Veröffentlicht: (2025)
ICGAN: An implicit conditioning method for interpretable feature control of neural audio synthesis
von: Liu, Yunyi, et al.
Veröffentlicht: (2024)
von: Liu, Yunyi, et al.
Veröffentlicht: (2024)
Vocos: Closing the gap between time-domain and Fourier-based neural vocoders for high-quality audio synthesis
von: Siuzdak, Hubert
Veröffentlicht: (2023)
von: Siuzdak, Hubert
Veröffentlicht: (2023)
Bird detection in audio: a survey and a challenge
von: Stowell, Dan, et al.
Veröffentlicht: (2016)
von: Stowell, Dan, et al.
Veröffentlicht: (2016)
Scaling up masked audio encoder learning for general audio classification
von: Dinkel, Heinrich, et al.
Veröffentlicht: (2024)
von: Dinkel, Heinrich, et al.
Veröffentlicht: (2024)
Towards generalizing deep-audio fake detection networks
von: Gasenzer, Konstantin, et al.
Veröffentlicht: (2023)
von: Gasenzer, Konstantin, et al.
Veröffentlicht: (2023)
Supervised contrastive learning from weakly-labeled audio segments for musical version matching
von: Serrà, Joan, et al.
Veröffentlicht: (2025)
von: Serrà, Joan, et al.
Veröffentlicht: (2025)
Emoanti: audio anti-deepfake with refined emotion-guided representations
von: Li, Xiaokang, et al.
Veröffentlicht: (2025)
von: Li, Xiaokang, et al.
Veröffentlicht: (2025)
Making deep neural networks work for medical audio: representation, compression and domain adaptation
von: Onu, Charles C
Veröffentlicht: (2025)
von: Onu, Charles C
Veröffentlicht: (2025)
An overview of neural architectures for self-supervised audio representation learning from masked spectrograms
von: Yadav, Sarthak, et al.
Veröffentlicht: (2025)
von: Yadav, Sarthak, et al.
Veröffentlicht: (2025)
EDTC: enhance depth of text comprehension in automated audio captioning
von: Tan, Liwen, et al.
Veröffentlicht: (2024)
von: Tan, Liwen, et al.
Veröffentlicht: (2024)
Mixer Metaphors: audio interfaces for non-musical applications
von: McNamara, Tace, et al.
Veröffentlicht: (2025)
von: McNamara, Tace, et al.
Veröffentlicht: (2025)
Stage-adaptive audio diffusion modeling
von: Zhang, Xuanhao, et al.
Veröffentlicht: (2026)
von: Zhang, Xuanhao, et al.
Veröffentlicht: (2026)
GRAM: Spatial general-purpose audio representations for real-world environments
von: Yuksel, Goksenin, et al.
Veröffentlicht: (2026)
von: Yuksel, Goksenin, et al.
Veröffentlicht: (2026)
DashengTokenizer: One layer is enough for unified audio understanding and generation
von: Dinkel, Heinrich, et al.
Veröffentlicht: (2026)
von: Dinkel, Heinrich, et al.
Veröffentlicht: (2026)
ACAVCaps: Enabling large-scale training for fine-grained and diverse audio understanding
von: Niu, Yadong, et al.
Veröffentlicht: (2026)
von: Niu, Yadong, et al.
Veröffentlicht: (2026)
ParaCLAP -- Towards a general language-audio model for computational paralinguistic tasks
von: Jing, Xin, et al.
Veröffentlicht: (2024)
von: Jing, Xin, et al.
Veröffentlicht: (2024)
LDCodec: A high quality neural audio codec with low-complexity decoder
von: Jiang, Jiawei, et al.
Veröffentlicht: (2025)
von: Jiang, Jiawei, et al.
Veröffentlicht: (2025)
An audio-quality-based multi-strategy approach for target speaker extraction in the MISP 2023 Challenge
von: Han, Runduo, et al.
Veröffentlicht: (2024)
von: Han, Runduo, et al.
Veröffentlicht: (2024)
Quantized Approximate Signal Processing (QASP): Towards Homomorphic Encryption for audio
von: Nguyen, Tu Duyen, et al.
Veröffentlicht: (2025)
von: Nguyen, Tu Duyen, et al.
Veröffentlicht: (2025)
Speech-XL: Towards Long-Form Speech Understanding in Large Speech Language Models
von: Sun, Haoqin, et al.
Veröffentlicht: (2026)
von: Sun, Haoqin, et al.
Veröffentlicht: (2026)
WikiMuTe: A web-sourced dataset of semantic descriptions for music audio
von: Weck, Benno, et al.
Veröffentlicht: (2023)
von: Weck, Benno, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Speaker anonymization using neural audio codec language models
von: Panariello, Michele, et al.
Veröffentlicht: (2023) -
Modeling strategies for speech enhancement in the latent space of a neural audio codec
von: Kammoun, Sofiene, et al.
Veröffentlicht: (2025) -
Codec-SUPERB @ SLT 2024: A lightweight benchmark for neural audio codec models
von: Wu, Haibin, et al.
Veröffentlicht: (2024) -
EnCodecMAE: Leveraging neural codecs for universal audio representation learning
von: Pepino, Leonardo, et al.
Veröffentlicht: (2023) -
Probing neural audio codecs for distinctions among English nuclear tunes
von: Vigneaux, Juan Pablo, et al.
Veröffentlicht: (2026)