End-to-end transfer learning for speaker-independent cross-language and cross-corpus speech emotion recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Tang, Duowei, Kuppens, Peter, Geurts, Lucca, van Waterschoot, Toon |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Deep, data-driven modeling of room acoustics: literature review and research perspectives
by: van Waterschoot, Toon
Published: (2025)
by: van Waterschoot, Toon
Published: (2025)
Sound Field Reconstruction Using Physics-Informed Boundary Integral Networks
by: Damiano, Stefano, et al.
Published: (2025)
by: Damiano, Stefano, et al.
Published: (2025)
On Time Delay Interpolation for Improved Acoustic Reflector Localization
by: Rosseel, Hannes, et al.
Published: (2025)
by: Rosseel, Hannes, et al.
Published: (2025)
Accelerated Interactive Auralization of Highly Reverberant Spaces using Graphics Hardware
by: Rosseel, Hannes, et al.
Published: (2025)
by: Rosseel, Hannes, et al.
Published: (2025)
End-to-end multi-channel speaker extraction and binaural speech synthesis
by: Chi, Cheng, et al.
Published: (2024)
by: Chi, Cheng, et al.
Published: (2024)
Cascaded noise reduction and acoustic echo cancellation based on an extended noise reduction
by: Roebben, Arnout, et al.
Published: (2024)
by: Roebben, Arnout, et al.
Published: (2024)
Frequency Tracking Features for Data-Efficient Deep Siren Identification
by: Damiano, Stefano, et al.
Published: (2024)
by: Damiano, Stefano, et al.
Published: (2024)
Topology-Independent GEVD-Based Distributed Adaptive Node-Specific Signal Estimation in Ad-Hoc Wireless Acoustic Sensor Networks
by: Didier, Paul, et al.
Published: (2024)
by: Didier, Paul, et al.
Published: (2024)
A Comparative Analysis of Generalised Echo and Interference Cancelling and Extended Multichannel Wiener Filtering for Combined Noise Reduction and Acoustic Echo Cancellation
by: Roebben, Arnout, et al.
Published: (2025)
by: Roebben, Arnout, et al.
Published: (2025)
Thinking in cocktail party: Chain-of-Thought and reinforcement learning for target speaker automatic speech recognition
by: Zhang, Yiru, et al.
Published: (2025)
by: Zhang, Yiru, et al.
Published: (2025)
Dereverberation in Acoustic Sensor Networks Using Weighted Prediction Error With Microphone-dependent Prediction Delays
by: Lohmann, Anselm, et al.
Published: (2023)
by: Lohmann, Anselm, et al.
Published: (2023)
On the Use of Dereverberation for Acoustic Feedback Cancellation
by: Liekens, Basil, et al.
Published: (2026)
by: Liekens, Basil, et al.
Published: (2026)
Reference Microphone Selection for the Weighted Prediction Error Algorithm using the Normalized L-p Norm
by: Lohmann, Anselm, et al.
Published: (2024)
by: Lohmann, Anselm, et al.
Published: (2024)
Microphone Subset Selection for the Weighted Prediction Error Algorithm using a Group Sparsity Penalty
by: Lohmann, Anselm, et al.
Published: (2024)
by: Lohmann, Anselm, et al.
Published: (2024)
Integrated Minimum Mean Squared Error Algorithms for Combined Acoustic Echo Cancellation and Noise Reduction
by: Roebben, Arnout, et al.
Published: (2024)
by: Roebben, Arnout, et al.
Published: (2024)
Identifiability Conditions for Acoustic Feedback Cancellation with the Two-Channel Adaptive Feedback Canceller Algorithm
by: Roebben, Arnout, et al.
Published: (2025)
by: Roebben, Arnout, et al.
Published: (2025)
Scalable-Complexity Steered Response Power Mapping based on Low-Rank and Sparse Interpolation
by: Dietzen, Thomas, et al.
Published: (2023)
by: Dietzen, Thomas, et al.
Published: (2023)
Tracking of Spatially Dynamic Room Impulse Responses Along Locally Linearized Trajectories
by: MacWilliam, Kathleen, et al.
Published: (2025)
by: MacWilliam, Kathleen, et al.
Published: (2025)
Exploring synthetic data for cross-speaker style transfer in style representation based TTS
by: Ueda, Lucas H., et al.
Published: (2024)
by: Ueda, Lucas H., et al.
Published: (2024)
One-Shot Distributed Node-Specific Signal Estimation with Non-Overlapping Latent Subspaces in Acoustic Sensor Networks
by: Didier, Paul, et al.
Published: (2024)
by: Didier, Paul, et al.
Published: (2024)
Improved Topology-Independent Distributed Adaptive Node-Specific Signal Estimation for Wireless Acoustic Sensor Networks
by: Didier, Paul, et al.
Published: (2025)
by: Didier, Paul, et al.
Published: (2025)
Charting 15 years of progress in deep learning for speech emotion recognition: A replication study
by: Triantafyllopoulos, Andreas, et al.
Published: (2025)
by: Triantafyllopoulos, Andreas, et al.
Published: (2025)
State-Space Estimation of Spatially Dynamic Room Impulse Responses using a Room Acoustic Model-based Prior
by: MacWilliam, Kathleen, et al.
Published: (2024)
by: MacWilliam, Kathleen, et al.
Published: (2024)
The trajectoRIR Database: Room Acoustic Recordings Along a Trajectory of Moving Microphones
by: Damiano, Stefano, et al.
Published: (2025)
by: Damiano, Stefano, et al.
Published: (2025)
The Neural-SRP method for positional sound source localization
by: Grinstein, Eric, et al.
Published: (2024)
by: Grinstein, Eric, et al.
Published: (2024)
Graph-based multi-Feature fusion method for speech emotion recognition
by: Liu, Xueyu, et al.
Published: (2024)
by: Liu, Xueyu, et al.
Published: (2024)
A State-of-the-Art Review on Acoustic Preservation of Historical Worship Spaces through Auralization
by: Rosseel, Hannes, et al.
Published: (2025)
by: Rosseel, Hannes, et al.
Published: (2025)
Lightweight speech enhancement guided target speech extraction in noisy multi-speaker scenarios
by: Huang, Ziling, et al.
Published: (2025)
by: Huang, Ziling, et al.
Published: (2025)
Fast-Converging Distributed Signal Estimation in Topology-Unconstrained Wireless Acoustic Sensor Networks
by: Didier, Paul, et al.
Published: (2025)
by: Didier, Paul, et al.
Published: (2025)
Can Synthetic Data Boost the Training of Deep Acoustic Vehicle Counting Networks?
by: Damiano, Stefano, et al.
Published: (2024)
by: Damiano, Stefano, et al.
Published: (2024)
SelfTTS: cross-speaker style transfer through explicit embedding disentanglement and self-refinement using self-augmentation
by: Ueda, Lucas H., et al.
Published: (2026)
by: Ueda, Lucas H., et al.
Published: (2026)
Heterogeneous bimodal attention fusion for speech emotion recognition
by: Luo, Jiachen, et al.
Published: (2025)
by: Luo, Jiachen, et al.
Published: (2025)
learning discriminative features from spectrograms using center loss for speech emotion recognition
by: Dai, Dongyang, et al.
Published: (2025)
by: Dai, Dongyang, et al.
Published: (2025)
A state-space representation of the boundary integral equation for room acoustic modelling
by: Ali, Randall, et al.
Published: (2026)
by: Ali, Randall, et al.
Published: (2026)
Advancing automatic speech recognition using feature fusion with self-supervised learning features: A case study on Fearless Steps Apollo corpus
by: Chen, Szu-Jui, et al.
Published: (2026)
by: Chen, Szu-Jui, et al.
Published: (2026)
Phoneme-based speech recognition driven by large language models and sampling marginalization
by: Ma, Te, et al.
Published: (2025)
by: Ma, Te, et al.
Published: (2025)
Explainable speech emotion recognition through attentive pooling: insights from attention-based temporal localization
by: Leygue, Tahitoa, et al.
Published: (2025)
by: Leygue, Tahitoa, et al.
Published: (2025)
Sound field estimation with moving microphones using kernel ridge regression
by: Brunnström, Jesper, et al.
Published: (2025)
by: Brunnström, Jesper, et al.
Published: (2025)
Analyzing the relationships between pretraining language, phonetic, tonal, and speaker information in self-supervised speech models
by: Gubian, Michele, et al.
Published: (2025)
by: Gubian, Michele, et al.
Published: (2025)
SCDiar: a streaming diarization system based on speaker change detection and speech recognition
by: Zheng, Naijun, et al.
Published: (2025)
by: Zheng, Naijun, et al.
Published: (2025)
Similar Items
-
Deep, data-driven modeling of room acoustics: literature review and research perspectives
by: van Waterschoot, Toon
Published: (2025) -
Sound Field Reconstruction Using Physics-Informed Boundary Integral Networks
by: Damiano, Stefano, et al.
Published: (2025) -
On Time Delay Interpolation for Improved Acoustic Reflector Localization
by: Rosseel, Hannes, et al.
Published: (2025) -
Accelerated Interactive Auralization of Highly Reverberant Spaces using Graphics Hardware
by: Rosseel, Hannes, et al.
Published: (2025) -
End-to-end multi-channel speaker extraction and binaural speech synthesis
by: Chi, Cheng, et al.
Published: (2024)