Solid State Bus-Comp: A Large-Scale and Diverse Dataset for Dynamic Range Compressor Virtual Analog Modeling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gu, Yicheng, Zhang, Runsong, Juvela, Lauri, Wu, Zhizheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Aliasing-Free Neural Audio Synthesis
von: Gu, Yicheng, et al.
Veröffentlicht: (2025)
von: Gu, Yicheng, et al.
Veröffentlicht: (2025)
Neurodyne: Neural Pitch Manipulation with Representation Learning and Cycle-Consistency GAN
von: Gu, Yicheng, et al.
Veröffentlicht: (2025)
von: Gu, Yicheng, et al.
Veröffentlicht: (2025)
De-crackling Virtual Analog Controls with Asymptotically Stable Recurrent Neural Networks
von: Kallinen, Valtteri, et al.
Veröffentlicht: (2025)
von: Kallinen, Valtteri, et al.
Veröffentlicht: (2025)
An Investigation of Time-Frequency Representation Discriminators for High-Fidelity Vocoder
von: Gu, Yicheng, et al.
Veröffentlicht: (2024)
von: Gu, Yicheng, et al.
Veröffentlicht: (2024)
SingNet: Towards a Large-Scale, Diverse, and In-the-Wild Singing Voice Dataset
von: Gu, Yicheng, et al.
Veröffentlicht: (2025)
von: Gu, Yicheng, et al.
Veröffentlicht: (2025)
HiFi-Glot: High-Fidelity Neural Formant Synthesis with Differentiable Resonant Filters
von: Gu, Yicheng, et al.
Veröffentlicht: (2024)
von: Gu, Yicheng, et al.
Veröffentlicht: (2024)
AffectSpeech: A Large-Scale Emotional Speech Dataset with Fine-Grained Textual Descriptions for Speech Emotion Captioning and Synthesis
von: Qi, Tianhua, et al.
Veröffentlicht: (2026)
von: Qi, Tianhua, et al.
Veröffentlicht: (2026)
Modeling Analog Dynamic Range Compressors using Deep Learning and State-space Models
von: Yin, Hanzhi, et al.
Veröffentlicht: (2024)
von: Yin, Hanzhi, et al.
Veröffentlicht: (2024)
Toward Universal Speech Enhancement for Diverse Input Conditions
von: Zhang, Wangyou, et al.
Veröffentlicht: (2023)
von: Zhang, Wangyou, et al.
Veröffentlicht: (2023)
Tracking of Intermittent and Moving Speakers : Dataset and Metrics
von: Iatariene, Taous, et al.
Veröffentlicht: (2025)
von: Iatariene, Taous, et al.
Veröffentlicht: (2025)
Leveraging Diverse Semantic-based Audio Pretrained Models for Singing Voice Conversion
von: Zhang, Xueyao, et al.
Veröffentlicht: (2023)
von: Zhang, Xueyao, et al.
Veröffentlicht: (2023)
Audio Codec Augmentation for Robust Collaborative Watermarking of Speech Synthesis
von: Juvela, Lauri, et al.
Veröffentlicht: (2024)
von: Juvela, Lauri, et al.
Veröffentlicht: (2024)
Emilia: A Large-Scale, Extensive, Multilingual, and Diverse Dataset for Speech Generation
von: He, Haorui, et al.
Veröffentlicht: (2025)
von: He, Haorui, et al.
Veröffentlicht: (2025)
A Data-Driven Exploration of Elevation Cues in HRTFs: An Explainable AI Perspective Across Multiple Datasets
von: De Rus, Juan Antonio, et al.
Veröffentlicht: (2025)
von: De Rus, Juan Antonio, et al.
Veröffentlicht: (2025)
Mel-McNet: A Mel-Scale Framework for Online Multichannel Speech Enhancement
von: Yang, Yujie, et al.
Veröffentlicht: (2025)
von: Yang, Yujie, et al.
Veröffentlicht: (2025)
SSM2Mel: State Space Model to Reconstruct Mel Spectrogram from the EEG
von: Fan, Cunhang, et al.
Veröffentlicht: (2025)
von: Fan, Cunhang, et al.
Veröffentlicht: (2025)
Dynamic Prediction of Full-Ocean Depth SSP by Hierarchical LSTM: An Experimental Result
von: Lu, Jiajun, et al.
Veröffentlicht: (2023)
von: Lu, Jiajun, et al.
Veröffentlicht: (2023)
Soundscape Captioning using Sound Affective Quality Network and Large Language Model
von: Hou, Yuanbo, et al.
Veröffentlicht: (2024)
von: Hou, Yuanbo, et al.
Veröffentlicht: (2024)
Large Language Model-based Nonnegative Matrix Factorization For Cardiorespiratory Sound Separation
von: Torabi, Yasaman, et al.
Veröffentlicht: (2025)
von: Torabi, Yasaman, et al.
Veröffentlicht: (2025)
SpeakerBeam-SS: Real-time Target Speaker Extraction with Lightweight Conv-TasNet and State Space Modeling
von: Sato, Hiroshi, et al.
Veröffentlicht: (2024)
von: Sato, Hiroshi, et al.
Veröffentlicht: (2024)
Frequency-Based Alignment of EEG and Audio Signals Using Contrastive Learning and SincNet for Auditory Attention Detection
von: Liao, Yuan, et al.
Veröffentlicht: (2025)
von: Liao, Yuan, et al.
Veröffentlicht: (2025)
Future Full-Ocean Deep SSPs Prediction based on Hierarchical Long Short-Term Memory Neural Networks
von: Lu, Jiajun, et al.
Veröffentlicht: (2023)
von: Lu, Jiajun, et al.
Veröffentlicht: (2023)
Singing Voice Graph Modeling for SingFake Detection
von: Chen, Xuanjun, et al.
Veröffentlicht: (2024)
von: Chen, Xuanjun, et al.
Veröffentlicht: (2024)
Using Ear-EEG to Decode Auditory Attention in Multiple-speaker Environment
von: Zhu, Haolin, et al.
Veröffentlicht: (2024)
von: Zhu, Haolin, et al.
Veröffentlicht: (2024)
PromptEVC: Controllable Emotional Voice Conversion with Natural Language Prompts
von: Qi, Tianhua, et al.
Veröffentlicht: (2025)
von: Qi, Tianhua, et al.
Veröffentlicht: (2025)
Self-supervised speech representation and contextual text embedding for match-mismatch classification with EEG recording
von: Wang, Bo, et al.
Veröffentlicht: (2024)
von: Wang, Bo, et al.
Veröffentlicht: (2024)
Joint Fullband-Subband Modeling for High-Resolution SingFake Detection
von: Chen, Xuanjun, et al.
Veröffentlicht: (2026)
von: Chen, Xuanjun, et al.
Veröffentlicht: (2026)
How Does Instrumental Music Help SingFake Detection?
von: Chen, Xuanjun, et al.
Veröffentlicht: (2025)
von: Chen, Xuanjun, et al.
Veröffentlicht: (2025)
Array-Aware Ambisonics and HRTF Encoding for Binaural Reproduction With Wearable Arrays
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2025)
von: Gayer, Yhonatan, et al.
Veröffentlicht: (2025)
SIRUP: A diffusion-based virtual upmixer of steering vectors for highly-directive spatialization with first-order ambisonics
von: Picard, Emilio, et al.
Veröffentlicht: (2026)
von: Picard, Emilio, et al.
Veröffentlicht: (2026)
Real time fault detection in 3D printers using Convolutional Neural Networks and acoustic signals
von: Waheed, Muhammad Fasih, et al.
Veröffentlicht: (2026)
von: Waheed, Muhammad Fasih, et al.
Veröffentlicht: (2026)
AutoMashup: Automatic Music Mashups Creation
von: Delabaere, Marine, et al.
Veröffentlicht: (2025)
von: Delabaere, Marine, et al.
Veröffentlicht: (2025)
Neural Tracking of Sustained Attention, Attention Switching, and Natural Conversation in Audiovisual Environments using Mobile EEG
von: Wilroth, Johanna, et al.
Veröffentlicht: (2026)
von: Wilroth, Johanna, et al.
Veröffentlicht: (2026)
Beamforming in the Reproducing Kernel Domain Based on Spatial Differentiation
von: Iwami, Takahiro, et al.
Veröffentlicht: (2025)
von: Iwami, Takahiro, et al.
Veröffentlicht: (2025)
A Survey of Advancing Audio Super-Resolution and Bandwidth Extension from Discriminative to Generative Models
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
DDD: A Perceptually Superior Low-Response-Time DNN-based Declipper
von: Yi, Jayeon, et al.
Veröffentlicht: (2024)
von: Yi, Jayeon, et al.
Veröffentlicht: (2024)
Towards Realistic Emotional Voice Conversion using Controllable Emotional Intensity
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
von: Qi, Tianhua, et al.
Veröffentlicht: (2024)
Steered Response Power-Based Direction-of-Arrival Estimation Exploiting an Auxiliary Microphone
von: Brümann, Klaus, et al.
Veröffentlicht: (2024)
von: Brümann, Klaus, et al.
Veröffentlicht: (2024)
Significance of Chirp MFCC as a Feature in Speech and Audio Applications
von: Joysingh, S. Johanan, et al.
Veröffentlicht: (2024)
von: Joysingh, S. Johanan, et al.
Veröffentlicht: (2024)
Cross-Talk Reduction
von: Wang, Zhong-Qiu, et al.
Veröffentlicht: (2024)
von: Wang, Zhong-Qiu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Aliasing-Free Neural Audio Synthesis
von: Gu, Yicheng, et al.
Veröffentlicht: (2025) -
Neurodyne: Neural Pitch Manipulation with Representation Learning and Cycle-Consistency GAN
von: Gu, Yicheng, et al.
Veröffentlicht: (2025) -
De-crackling Virtual Analog Controls with Asymptotically Stable Recurrent Neural Networks
von: Kallinen, Valtteri, et al.
Veröffentlicht: (2025) -
An Investigation of Time-Frequency Representation Discriminators for High-Fidelity Vocoder
von: Gu, Yicheng, et al.
Veröffentlicht: (2024) -
SingNet: Towards a Large-Scale, Diverse, and In-the-Wild Singing Voice Dataset
von: Gu, Yicheng, et al.
Veröffentlicht: (2025)