ComplexDec: A Domain-robust High-fidelity Neural Audio Codec with Complex Spectrum Modeling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Yi-Chiao, Marković, Dejan, Krenn, Steven, Gebru, Israel D., Richard, Alexander |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ScoreDec: A Phase-preserving High-Fidelity Audio Codec with A Generalized Score-based Diffusion Post-filter
von: Wu, Yi-Chiao, et al.
Veröffentlicht: (2024)
von: Wu, Yi-Chiao, et al.
Veröffentlicht: (2024)
BinauralFlow: A Causal and Streamable Approach for High-Quality Binaural Speech Synthesis with Flow Matching Models
von: Liang, Susan, et al.
Veröffentlicht: (2025)
von: Liang, Susan, et al.
Veröffentlicht: (2025)
EMO-Codec: An In-Depth Look at Emotion Preservation capacity of Legacy and Neural Codec Models With Subjective and Objective Evaluations
von: Ren, Wenze, et al.
Veröffentlicht: (2024)
von: Ren, Wenze, et al.
Veröffentlicht: (2024)
ESPnet-Codec: Comprehensive Training and Evaluation of Neural Codecs for Audio, Music, and Speech
von: Shi, Jiatong, et al.
Veröffentlicht: (2024)
von: Shi, Jiatong, et al.
Veröffentlicht: (2024)
HILCodec: High-Fidelity and Lightweight Neural Audio Codec
von: Ahn, Sunghwan, et al.
Veröffentlicht: (2024)
von: Ahn, Sunghwan, et al.
Veröffentlicht: (2024)
UniCodec: Unified Audio Codec with Single Domain-Adaptive Codebook
von: Jiang, Yidi, et al.
Veröffentlicht: (2025)
von: Jiang, Yidi, et al.
Veröffentlicht: (2025)
APCodec: A Neural Audio Codec with Parallel Amplitude and Phase Spectrum Encoding and Decoding
von: Ai, Yang, et al.
Veröffentlicht: (2024)
von: Ai, Yang, et al.
Veröffentlicht: (2024)
Codec-Based Deepfake Source Tracing via Neural Audio Codec Taxonomy
von: Chen, Xuanjun, et al.
Veröffentlicht: (2025)
von: Chen, Xuanjun, et al.
Veröffentlicht: (2025)
APCodec+: A Spectrum-Coding-Based High-Fidelity and High-Compression-Rate Neural Audio Codec with Staged Training Paradigm
von: Du, Hui-Peng, et al.
Veröffentlicht: (2024)
von: Du, Hui-Peng, et al.
Veröffentlicht: (2024)
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling
von: Xue, Rongkun, et al.
Veröffentlicht: (2025)
von: Xue, Rongkun, et al.
Veröffentlicht: (2025)
Towards Neural Audio Codec Source Parsing
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2025)
von: Phukan, Orchid Chetia, et al.
Veröffentlicht: (2025)
CodecFake+: A Large-Scale Neural Audio Codec-Based Deepfake Speech Dataset
von: Chen, Xuanjun, et al.
Veröffentlicht: (2025)
von: Chen, Xuanjun, et al.
Veröffentlicht: (2025)
DualCodec: A Low-Frame-Rate, Semantically-Enhanced Neural Audio Codec for Speech Generation
von: Li, Jiaqi, et al.
Veröffentlicht: (2025)
von: Li, Jiaqi, et al.
Veröffentlicht: (2025)
SwitchCodec: A High-Fidelity Nerual Audio Codec With Sparse Quantization
von: Wang, Jin, et al.
Veröffentlicht: (2025)
von: Wang, Jin, et al.
Veröffentlicht: (2025)
Code Drift: Towards Idempotent Neural Audio Codecs
von: O'Reilly, Patrick, et al.
Veröffentlicht: (2024)
von: O'Reilly, Patrick, et al.
Veröffentlicht: (2024)
SpecTokenizer: A Lightweight Streaming Codec in the Compressed Spectrum Domain
von: Wan, Zixiang, et al.
Veröffentlicht: (2025)
von: Wan, Zixiang, et al.
Veröffentlicht: (2025)
Investigating Neural Audio Codecs for Speech Language Model-Based Speech Generation
von: Li, Jiaqi, et al.
Veröffentlicht: (2024)
von: Li, Jiaqi, et al.
Veröffentlicht: (2024)
BANC: Towards Efficient Binaural Audio Neural Codec for Overlapping Speech
von: Ratnarajah, Anton, et al.
Veröffentlicht: (2023)
von: Ratnarajah, Anton, et al.
Veröffentlicht: (2023)
On the Language and Gender Biases in PSTN, VoIP and Neural Audio Codecs
von: Altwlkany, Kemal, et al.
Veröffentlicht: (2025)
von: Altwlkany, Kemal, et al.
Veröffentlicht: (2025)
EARS: An Anechoic Fullband Speech Dataset Benchmarked for Speech Enhancement and Dereverberation
von: Richter, Julius, et al.
Veröffentlicht: (2024)
von: Richter, Julius, et al.
Veröffentlicht: (2024)
An Intra-BRNN and GB-RVQ Based END-TO-END Neural Audio Codec
von: Xu, Linping, et al.
Veröffentlicht: (2024)
von: Xu, Linping, et al.
Veröffentlicht: (2024)
NDVQ: Robust Neural Audio Codec with Normal Distribution-Based Vector Quantization
von: Niu, Zhikang, et al.
Veröffentlicht: (2024)
von: Niu, Zhikang, et al.
Veröffentlicht: (2024)
Analysis of Speaker Verification Performance Trade-offs with Neural Audio Codec Transmission
von: Thakur, Nirmalya Mallick, et al.
Veröffentlicht: (2025)
von: Thakur, Nirmalya Mallick, et al.
Veröffentlicht: (2025)
Analyzing and Mitigating Inconsistency in Discrete Audio Tokens for Neural Codec Language Models
von: Liu, Wenrui, et al.
Veröffentlicht: (2024)
von: Liu, Wenrui, et al.
Veröffentlicht: (2024)
CodecFake: Enhancing Anti-Spoofing Models Against Deepfake Audios from Codec-Based Speech Synthesis Systems
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
Codec-SUPERB: An In-Depth Analysis of Sound Codec Models
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
EuleroDec: A Complex-Valued RVQ-VAE for Efficient and Robust Audio Coding
von: Cerovaz, Luca, et al.
Veröffentlicht: (2026)
von: Cerovaz, Luca, et al.
Veröffentlicht: (2026)
Toward a Sparse and Interpretable Audio Codec
von: Vinyard, John
Veröffentlicht: (2025)
von: Vinyard, John
Veröffentlicht: (2025)
Streaming Endpointer for Spoken Dialogue using Neural Audio Codecs and Label-Delayed Training
von: Udupa, Sathvik, et al.
Veröffentlicht: (2025)
von: Udupa, Sathvik, et al.
Veröffentlicht: (2025)
VCNAC: A Variable-Channel Neural Audio Codec for Mono, Stereo, and Surround Sound
von: Grötschla, Florian, et al.
Veröffentlicht: (2026)
von: Grötschla, Florian, et al.
Veröffentlicht: (2026)
MDCTCodec: A Lightweight MDCT-based Neural Audio Codec towards High Sampling Rate and Low Bitrate Scenarios
von: Jiang, Xiao-Hang, et al.
Veröffentlicht: (2024)
von: Jiang, Xiao-Hang, et al.
Veröffentlicht: (2024)
SNAC: Multi-Scale Neural Audio Codec
von: Siuzdak, Hubert, et al.
Veröffentlicht: (2024)
von: Siuzdak, Hubert, et al.
Veröffentlicht: (2024)
Trade-offs Between Capacity and Robustness in Neural Audio Codecs for Adversarially Robust Speech Recognition
von: Prescott, Jordan, et al.
Veröffentlicht: (2026)
von: Prescott, Jordan, et al.
Veröffentlicht: (2026)
Codec-SUPERB @ SLT 2024: A lightweight benchmark for neural audio codec models
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
von: Wu, Haibin, et al.
Veröffentlicht: (2024)
UniAudio 1.5: Large Language Model-driven Audio Codec is A Few-shot Audio Task Learner
von: Yang, Dongchao, et al.
Veröffentlicht: (2024)
von: Yang, Dongchao, et al.
Veröffentlicht: (2024)
Learning Source Disentanglement in Neural Audio Codec
von: Bie, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Bie, Xiaoyu, et al.
Veröffentlicht: (2024)
BigCodec: Pushing the Limits of Low-Bitrate Neural Speech Codec
von: Xin, Detai, et al.
Veröffentlicht: (2024)
von: Xin, Detai, et al.
Veröffentlicht: (2024)
Personalized Neural Speech Codec
von: Jang, Inseon, et al.
Veröffentlicht: (2024)
von: Jang, Inseon, et al.
Veröffentlicht: (2024)
Low-Resource Audio Codec (LRAC): 2025 Challenge Description
von: Wojcicki, Kamil, et al.
Veröffentlicht: (2025)
von: Wojcicki, Kamil, et al.
Veröffentlicht: (2025)
DAC-JAX: A JAX Implementation of the Descript Audio Codec
von: Braun, David
Veröffentlicht: (2024)
von: Braun, David
Veröffentlicht: (2024)
Ähnliche Einträge
-
ScoreDec: A Phase-preserving High-Fidelity Audio Codec with A Generalized Score-based Diffusion Post-filter
von: Wu, Yi-Chiao, et al.
Veröffentlicht: (2024) -
BinauralFlow: A Causal and Streamable Approach for High-Quality Binaural Speech Synthesis with Flow Matching Models
von: Liang, Susan, et al.
Veröffentlicht: (2025) -
EMO-Codec: An In-Depth Look at Emotion Preservation capacity of Legacy and Neural Codec Models With Subjective and Objective Evaluations
von: Ren, Wenze, et al.
Veröffentlicht: (2024) -
ESPnet-Codec: Comprehensive Training and Evaluation of Neural Codecs for Audio, Music, and Speech
von: Shi, Jiatong, et al.
Veröffentlicht: (2024) -
HILCodec: High-Fidelity and Lightweight Neural Audio Codec
von: Ahn, Sunghwan, et al.
Veröffentlicht: (2024)