Understanding Self-Supervised Learning of Speech Representation via Invariance and Redundancy Reduction
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Brima, Yusuf, Krumnack, Ulf, Pika, Simone, Heidemann, Gunther |
|---|---|
| Format: | Preprint |
| Publié: |
2023
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Learning Disentangled Speech Representations
par: Brima, Yusuf, et autres
Publié: (2023)
par: Brima, Yusuf, et autres
Publié: (2023)
Learning Disentangled Audio Representations through Controlled Synthesis
par: Brima, Yusuf, et autres
Publié: (2024)
par: Brima, Yusuf, et autres
Publié: (2024)
The Effect of Batch Size on Contrastive Self-Supervised Speech Representation Learning
par: Vaessen, Nik, et autres
Publié: (2024)
par: Vaessen, Nik, et autres
Publié: (2024)
Low-Resource Cross-Domain Singing Voice Synthesis via Reduced Self-Supervised Speech Representations
par: Kakoulidis, Panos, et autres
Publié: (2024)
par: Kakoulidis, Panos, et autres
Publié: (2024)
Losses Can Be Blessings: Routing Self-Supervised Speech Representations Towards Efficient Multilingual and Multitask Speech Processing
par: Fu, Yonggan, et autres
Publié: (2022)
par: Fu, Yonggan, et autres
Publié: (2022)
Singer Identity Representation Learning using Self-Supervised Techniques
par: Torres, Bernardo, et autres
Publié: (2024)
par: Torres, Bernardo, et autres
Publié: (2024)
Self-Supervised Disentangled Representation Learning for Robust Target Speech Extraction
par: Mu, Zhaoxi, et autres
Publié: (2023)
par: Mu, Zhaoxi, et autres
Publié: (2023)
Speech Self-Supervised Representations Benchmarking: a Case for Larger Probing Heads
par: Zaiem, Salah, et autres
Publié: (2023)
par: Zaiem, Salah, et autres
Publié: (2023)
MT-SLVR: Multi-Task Self-Supervised Learning for Transformation In(Variant) Representations
par: Heggan, Calum, et autres
Publié: (2023)
par: Heggan, Calum, et autres
Publié: (2023)
Additive Margin in Contrastive Self-Supervised Frameworks to Learn Discriminative Speaker Representations
par: Lepage, Theo, et autres
Publié: (2024)
par: Lepage, Theo, et autres
Publié: (2024)
How Redundant Is the Transformer Stack in Speech Representation Models?
par: Dorszewski, Teresa, et autres
Publié: (2024)
par: Dorszewski, Teresa, et autres
Publié: (2024)
Windowed SummaryMixing: An Efficient Fine-Tuning of Self-Supervised Learning Models for Low-resource Speech Recognition
par: Menon, Aditya Srinivas, et autres
Publié: (2026)
par: Menon, Aditya Srinivas, et autres
Publié: (2026)
CA-SSLR: Condition-Aware Self-Supervised Learning Representation for Generalized Speech Processing
par: Lu, Yen-Ju, et autres
Publié: (2024)
par: Lu, Yen-Ju, et autres
Publié: (2024)
Soft Clustering Anchors for Self-Supervised Speech Representation Learning in Joint Embedding Prediction Architectures
par: Ioannides, Georgios, et autres
Publié: (2026)
par: Ioannides, Georgios, et autres
Publié: (2026)
Singing Voice Conversion with Accompaniment Using Self-Supervised Representation-Based Melody Features
par: Chen, Wei, et autres
Publié: (2025)
par: Chen, Wei, et autres
Publié: (2025)
DAISY: Data Adaptive Self-Supervised Early Exit for Speech Representation Models
par: Lin, Tzu-Quan, et autres
Publié: (2024)
par: Lin, Tzu-Quan, et autres
Publié: (2024)
RepCodec: A Speech Representation Codec for Speech Tokenization
par: Huang, Zhichao, et autres
Publié: (2023)
par: Huang, Zhichao, et autres
Publié: (2023)
Representation Learning with Parameterised Quantum Circuits for Advancing Speech Emotion Recognition
par: Rajapakshe, Thejan, et autres
Publié: (2025)
par: Rajapakshe, Thejan, et autres
Publié: (2025)
Self-Supervised Learning for Few-Shot Bird Sound Classification
par: Moummad, Ilyass, et autres
Publié: (2023)
par: Moummad, Ilyass, et autres
Publié: (2023)
Improving the Adversarial Robustness for Speaker Verification by Self-Supervised Learning
par: Wu, Haibin, et autres
Publié: (2021)
par: Wu, Haibin, et autres
Publié: (2021)
Self-Supervised Learning for Speaker Recognition: A study and review
par: Lepage, Theo, et autres
Publié: (2026)
par: Lepage, Theo, et autres
Publié: (2026)
Towards Supervised Performance on Speaker Verification with Self-Supervised Learning by Leveraging Large-Scale ASR Models
par: Miara, Victor, et autres
Publié: (2024)
par: Miara, Victor, et autres
Publié: (2024)
Self-Supervised Frameworks for Speaker Verification via Bootstrapped Positive Sampling
par: Lepage, Theo, et autres
Publié: (2025)
par: Lepage, Theo, et autres
Publié: (2025)
Emotion-Aware Speech Self-Supervised Representation Learning with Intensity Knowledge
par: Liu, Rui, et autres
Publié: (2024)
par: Liu, Rui, et autres
Publié: (2024)
Benchmarking Representations for Speech, Music, and Acoustic Events
par: La Quatra, Moreno, et autres
Publié: (2024)
par: La Quatra, Moreno, et autres
Publié: (2024)
ParaMETA: Towards Learning Disentangled Paralinguistic Speaking Styles Representations from Speech
par: Lou, Haowei, et autres
Publié: (2026)
par: Lou, Haowei, et autres
Publié: (2026)
Label-Efficient Self-Supervised Speaker Verification With Information Maximization and Contrastive Learning
par: Lepage, Théo, et autres
Publié: (2022)
par: Lepage, Théo, et autres
Publié: (2022)
Property Neurons in Self-Supervised Speech Transformers
par: Lin, Tzu-Quan, et autres
Publié: (2024)
par: Lin, Tzu-Quan, et autres
Publié: (2024)
Improving Perceptual Audio Aesthetic Assessment via Triplet Loss and Self-Supervised Embeddings
par: Wisnu, Dyah A. M. G., et autres
Publié: (2025)
par: Wisnu, Dyah A. M. G., et autres
Publié: (2025)
Scaling Self-Supervised Representation Learning for Symbolic Piano Performance
par: Bradshaw, Louis, et autres
Publié: (2025)
par: Bradshaw, Louis, et autres
Publié: (2025)
Africa-Centric Self-Supervised Pre-Training for Multilingual Speech Representation in a Sub-Saharan Context
par: Caubrière, Antoine, et autres
Publié: (2024)
par: Caubrière, Antoine, et autres
Publié: (2024)
Wav2code: Restore Clean Speech Representations via Codebook Lookup for Noise-Robust ASR
par: Hu, Yuchen, et autres
Publié: (2023)
par: Hu, Yuchen, et autres
Publié: (2023)
SKILL: Similarity-aware Knowledge distILLation for Speech Self-Supervised Learning
par: Zampierin, Luca, et autres
Publié: (2024)
par: Zampierin, Luca, et autres
Publié: (2024)
Safeguarding Privacy in Edge Speech Understanding with Tiny Foundation Models
par: Benazir, Afsara, et autres
Publié: (2025)
par: Benazir, Afsara, et autres
Publié: (2025)
ASTRA: Aligning Speech and Text Representations for Asr without Sampling
par: Gaur, Neeraj, et autres
Publié: (2024)
par: Gaur, Neeraj, et autres
Publié: (2024)
RobustSpeechFlow: Learning Robust Text-to-Speech Trajectories via Augmentation-based Contrastive Flow Matching
par: Yang, Jinhyeok, et autres
Publié: (2026)
par: Yang, Jinhyeok, et autres
Publié: (2026)
Self-Supervised Speech Quality Estimation and Enhancement Using Only Clean Speech
par: Fu, Szu-Wei, et autres
Publié: (2024)
par: Fu, Szu-Wei, et autres
Publié: (2024)
Enhancing Lung Disease Diagnosis via Semi-Supervised Machine Learning
par: Xu, Xiaoran, et autres
Publié: (2025)
par: Xu, Xiaoran, et autres
Publié: (2025)
Toward Fully Self-Supervised Multi-Pitch Estimation
par: Cwitkowitz, Frank, et autres
Publié: (2024)
par: Cwitkowitz, Frank, et autres
Publié: (2024)
Self-Supervised Embeddings for Detecting Individual Symptoms of Depression
par: Dumpala, Sri Harsha, et autres
Publié: (2024)
par: Dumpala, Sri Harsha, et autres
Publié: (2024)
Documents similaires
-
Learning Disentangled Speech Representations
par: Brima, Yusuf, et autres
Publié: (2023) -
Learning Disentangled Audio Representations through Controlled Synthesis
par: Brima, Yusuf, et autres
Publié: (2024) -
The Effect of Batch Size on Contrastive Self-Supervised Speech Representation Learning
par: Vaessen, Nik, et autres
Publié: (2024) -
Low-Resource Cross-Domain Singing Voice Synthesis via Reduced Self-Supervised Speech Representations
par: Kakoulidis, Panos, et autres
Publié: (2024) -
Losses Can Be Blessings: Routing Self-Supervised Speech Representations Towards Efficient Multilingual and Multitask Speech Processing
par: Fu, Yonggan, et autres
Publié: (2022)