MSR-Codec: A Low-Bitrate Multi-Stream Residual Codec for High-Fidelity Speech Generation with Information Disentanglement
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Li, Jingyu, Zhang, Guangyan, Ye, Zhen, Guo, Yiwen |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
VoCodec: An Efficient Lightweight Low-Bitrate Speech Codec
par: Yang, Leyan, et autres
Publié: (2026)
par: Yang, Leyan, et autres
Publié: (2026)
BigCodec: Pushing the Limits of Low-Bitrate Neural Speech Codec
par: Xin, Detai, et autres
Publié: (2024)
par: Xin, Detai, et autres
Publié: (2024)
MuCodec: Ultra Low-Bitrate Music Codec
par: Xu, Yaoxun, et autres
Publié: (2024)
par: Xu, Yaoxun, et autres
Publié: (2024)
CodeSep: Low-Bitrate Codec-Driven Speech Separation with Base-Token Disentanglement and Auxiliary-Token Serial Prediction
par: Du, Hui-Peng, et autres
Publié: (2026)
par: Du, Hui-Peng, et autres
Publié: (2026)
Entropy-Guided GRVQ for Ultra-Low Bitrate Neural Speech Codec
par: Ren, Yanzhou, et autres
Publié: (2026)
par: Ren, Yanzhou, et autres
Publié: (2026)
LSCodec: Low-Bitrate and Speaker-Decoupled Discrete Speech Codec
par: Guo, Yiwei, et autres
Publié: (2024)
par: Guo, Yiwei, et autres
Publié: (2024)
Single-Codec: Single-Codebook Speech Codec towards High-Performance Speech Generation
par: Li, Hanzhao, et autres
Publié: (2024)
par: Li, Hanzhao, et autres
Publié: (2024)
Optimizing Neural Speech Codec for Low-Bitrate Compression via Multi-Scale Encoding
par: Yang, Peiji, et autres
Publié: (2024)
par: Yang, Peiji, et autres
Publié: (2024)
OmniCodec: Low Frame Rate Universal Audio Codec with Semantic-Acoustic Disentanglement
par: Hu, Jingbin, et autres
Publié: (2026)
par: Hu, Jingbin, et autres
Publié: (2026)
PURE Codec: Progressive Unfolding of Residual Entropy for Speech Codec Learning
par: Shi, Jiatong, et autres
Publié: (2025)
par: Shi, Jiatong, et autres
Publié: (2025)
FocalCodec-Stream: Streaming Low-Bitrate Speech Coding via Causal Distillation
par: Della Libera, Luca, et autres
Publié: (2025)
par: Della Libera, Luca, et autres
Publié: (2025)
SoCodec: A Semantic-Ordered Multi-Stream Speech Codec for Efficient Language Model Based Text-to-Speech Synthesis
par: Guo, Haohan, et autres
Publié: (2024)
par: Guo, Haohan, et autres
Publié: (2024)
SemantiCodec: An Ultra Low Bitrate Semantic Audio Codec for General Sound
par: Liu, Haohe, et autres
Publié: (2024)
par: Liu, Haohe, et autres
Publié: (2024)
XY-Tokenizer: Mitigating the Semantic-Acoustic Conflict in Low-Bitrate Speech Codecs
par: Gong, Yitian, et autres
Publié: (2025)
par: Gong, Yitian, et autres
Publié: (2025)
CFMDCTCodec: A Low-Bitrate Neural Speech Codec with Noise-Prior-aware Conditional Flow Matching for MDCT-Spectral Enhancement
par: Jiang, Xiao-Hang, et autres
Publié: (2026)
par: Jiang, Xiao-Hang, et autres
Publié: (2026)
SPG-Codec: Exploring the Role and Boundaries of Semantic Priors in Ultra-Low-Bitrate Neural Speech Coding
par: Zhao, Mingyu, et autres
Publié: (2026)
par: Zhao, Mingyu, et autres
Publié: (2026)
MDCTCodec: A Lightweight MDCT-based Neural Audio Codec towards High Sampling Rate and Low Bitrate Scenarios
par: Jiang, Xiao-Hang, et autres
Publié: (2024)
par: Jiang, Xiao-Hang, et autres
Publié: (2024)
Universal Speech Token Learning via Low-Bitrate Neural Codec and Pretrained Representations
par: Jiang, Xue, et autres
Publié: (2025)
par: Jiang, Xue, et autres
Publié: (2025)
SwitchCodec: A High-Fidelity Nerual Audio Codec With Sparse Quantization
par: Wang, Jin, et autres
Publié: (2025)
par: Wang, Jin, et autres
Publié: (2025)
A High-Quality and Low-Complexity Streamable Neural Speech Codec with Knowledge Distillation
par: Zhang, En-Wei, et autres
Publié: (2025)
par: Zhang, En-Wei, et autres
Publié: (2025)
TS3-Codec: Transformer-Based Simple Streaming Single Codec
par: Wu, Haibin, et autres
Publié: (2024)
par: Wu, Haibin, et autres
Publié: (2024)
DualCodec: A Low-Frame-Rate, Semantically-Enhanced Neural Audio Codec for Speech Generation
par: Li, Jiaqi, et autres
Publié: (2025)
par: Li, Jiaqi, et autres
Publié: (2025)
Spectral Codecs: Improving Non-Autoregressive Speech Synthesis with Spectrogram-Based Audio Codecs
par: Langman, Ryan, et autres
Publié: (2024)
par: Langman, Ryan, et autres
Publié: (2024)
FocalCodec: Low-Bitrate Speech Coding via Focal Modulation Networks
par: Della Libera, Luca, et autres
Publié: (2025)
par: Della Libera, Luca, et autres
Publié: (2025)
SuperCodec: A Neural Speech Codec with Selective Back-Projection Network
par: Zheng, Youqiang, et autres
Publié: (2024)
par: Zheng, Youqiang, et autres
Publié: (2024)
CodecSlime: Temporal Redundancy Compression of Neural Speech Codec via Dynamic Frame Rate
par: Wang, Hankun, et autres
Publié: (2025)
par: Wang, Hankun, et autres
Publié: (2025)
Exploring Disentangled Neural Speech Codecs from Self-Supervised Representations
par: Aihara, Ryo, et autres
Publié: (2025)
par: Aihara, Ryo, et autres
Publié: (2025)
HILCodec: High-Fidelity and Lightweight Neural Audio Codec
par: Ahn, Sunghwan, et autres
Publié: (2024)
par: Ahn, Sunghwan, et autres
Publié: (2024)
Language-Codec: Bridging Discrete Codec Representations and Speech Language Models
par: Ji, Shengpeng, et autres
Publié: (2024)
par: Ji, Shengpeng, et autres
Publié: (2024)
A Low-Complexity Speech Codec Using Parametric Dithering for ASR
par: Murray, Ellison, et autres
Publié: (2025)
par: Murray, Ellison, et autres
Publié: (2025)
SAC: Neural Speech Codec with Semantic-Acoustic Dual-Stream Quantization
par: Chen, Wenxi, et autres
Publié: (2025)
par: Chen, Wenxi, et autres
Publié: (2025)
RepCodec: A Speech Representation Codec for Speech Tokenization
par: Huang, Zhichao, et autres
Publié: (2023)
par: Huang, Zhichao, et autres
Publié: (2023)
ESPnet-Codec: Comprehensive Training and Evaluation of Neural Codecs for Audio, Music, and Speech
par: Shi, Jiatong, et autres
Publié: (2024)
par: Shi, Jiatong, et autres
Publié: (2024)
Low Frame-rate Speech Codec: a Codec Designed for Fast High-quality Speech LLM Training and Inference
par: Casanova, Edresson, et autres
Publié: (2024)
par: Casanova, Edresson, et autres
Publié: (2024)
Personalized Neural Speech Codec
par: Jang, Inseon, et autres
Publié: (2024)
par: Jang, Inseon, et autres
Publié: (2024)
Investigating Neural Audio Codecs for Speech Language Model-Based Speech Generation
par: Li, Jiaqi, et autres
Publié: (2024)
par: Li, Jiaqi, et autres
Publié: (2024)
CodecFake+: A Large-Scale Neural Audio Codec-Based Deepfake Speech Dataset
par: Chen, Xuanjun, et autres
Publié: (2025)
par: Chen, Xuanjun, et autres
Publié: (2025)
Benchmarking Neural Speech Codec Intelligibility with SITool
par: Leschanowsky, Anna, et autres
Publié: (2025)
par: Leschanowsky, Anna, et autres
Publié: (2025)
An Ultra-Low Latency, End-to-End Streaming Speech Synthesis Architecture via Block-Wise Generation and Depth-Wise Codec Decoding
par: Su, Tianhui, et autres
Publié: (2026)
par: Su, Tianhui, et autres
Publié: (2026)
SecoustiCodec: Cross-Modal Aligned Streaming Single-Codecbook Speech Codec
par: Qiang, Chunyu, et autres
Publié: (2025)
par: Qiang, Chunyu, et autres
Publié: (2025)
Documents similaires
-
VoCodec: An Efficient Lightweight Low-Bitrate Speech Codec
par: Yang, Leyan, et autres
Publié: (2026) -
BigCodec: Pushing the Limits of Low-Bitrate Neural Speech Codec
par: Xin, Detai, et autres
Publié: (2024) -
MuCodec: Ultra Low-Bitrate Music Codec
par: Xu, Yaoxun, et autres
Publié: (2024) -
CodeSep: Low-Bitrate Codec-Driven Speech Separation with Base-Token Disentanglement and Auxiliary-Token Serial Prediction
par: Du, Hui-Peng, et autres
Publié: (2026) -
Entropy-Guided GRVQ for Ultra-Low Bitrate Neural Speech Codec
par: Ren, Yanzhou, et autres
Publié: (2026)