Controlling Contrastive Self-Supervised Learning with Knowledge-Driven Multiple Hypothesis: Application to Beat Tracking
Fuente:
arXiv
Salvato in:
| Autori principali: | Gagnere, Antonin, Essid, Slim, Peeters, Geoffroy |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Contrastive Self-Supervised Learning scheme for beat tracking amenable to few-shot learning
di: Gagnere, Antonin, et al.
Pubblicazione: (2024)
di: Gagnere, Antonin, et al.
Pubblicazione: (2024)
Zero-shot Musical Stem Retrieval with Joint-Embedding Predictive Architectures
di: Riou, Alain, et al.
Pubblicazione: (2024)
di: Riou, Alain, et al.
Pubblicazione: (2024)
Contrastive Knowledge Distillation for Embedding Refinement in Personalized Speech Enhancement
di: Serre, Thomas, et al.
Pubblicazione: (2026)
di: Serre, Thomas, et al.
Pubblicazione: (2026)
A lightweight dual-stage framework for personalized speech enhancement based on DeepFilterNet2
di: Serre, Thomas, et al.
Pubblicazione: (2024)
di: Serre, Thomas, et al.
Pubblicazione: (2024)
Less Forgetting for Better Generalization: Exploring Continual-learning Fine-tuning Methods for Speech Self-supervised Representations
di: Zaiem, Salah, et al.
Pubblicazione: (2024)
di: Zaiem, Salah, et al.
Pubblicazione: (2024)
SALT: Standardized Audio event Label Taxonomy
di: Stamatiadis, Paraskevas, et al.
Pubblicazione: (2024)
di: Stamatiadis, Paraskevas, et al.
Pubblicazione: (2024)
Blind estimation of audio effects using an auto-encoder approach and differentiable digital signal processing
di: Peladeau, Côme, et al.
Pubblicazione: (2023)
di: Peladeau, Côme, et al.
Pubblicazione: (2023)
Multiple Choice Learning for Efficient Speech Separation with Many Speakers
di: Perera, David, et al.
Pubblicazione: (2024)
di: Perera, David, et al.
Pubblicazione: (2024)
Speech Self-Supervised Representations Benchmarking: a Case for Larger Probing Heads
di: Zaiem, Salah, et al.
Pubblicazione: (2023)
di: Zaiem, Salah, et al.
Pubblicazione: (2023)
PESTO: Pitch Estimation with Self-supervised Transposition-equivariant Objective
di: Riou, Alain, et al.
Pubblicazione: (2023)
di: Riou, Alain, et al.
Pubblicazione: (2023)
A sound description: Exploring prompt templates and class descriptions to enhance zero-shot audio classification
di: Olvera, Michel, et al.
Pubblicazione: (2024)
di: Olvera, Michel, et al.
Pubblicazione: (2024)
Translation-Equivariant Self-Supervised Learning for Pitch Estimation with Optimal Transport
di: Torres, Bernardo, et al.
Pubblicazione: (2025)
di: Torres, Bernardo, et al.
Pubblicazione: (2025)
Perceptual Noise-Masking with Music through Deep Spectral Envelope Shaping
di: Berger, Clémentine, et al.
Pubblicazione: (2025)
di: Berger, Clémentine, et al.
Pubblicazione: (2025)
Online speaker diarization of meetings guided by speech separation
di: Gruttadauria, Elio, et al.
Pubblicazione: (2024)
di: Gruttadauria, Elio, et al.
Pubblicazione: (2024)
Annealed Multiple Choice Learning: Overcoming limitations of Winner-takes-all with annealing
di: Perera, David, et al.
Pubblicazione: (2024)
di: Perera, David, et al.
Pubblicazione: (2024)
Efficient Adapter Tuning for Joint Singing Voice Beat and Downbeat Tracking with Self-supervised Learning Features
di: Deng, Jiajun, et al.
Pubblicazione: (2025)
di: Deng, Jiajun, et al.
Pubblicazione: (2025)
IS${}^3$ : Generic Impulsive--Stationary Sound Separation in Acoustic Scenes using Deep Filtering
di: Berger, Clémentine, et al.
Pubblicazione: (2025)
di: Berger, Clémentine, et al.
Pubblicazione: (2025)
BEAST: Online Joint Beat and Downbeat Tracking Based on Streaming Transformer
di: Chang, Chih-Cheng, et al.
Pubblicazione: (2023)
di: Chang, Chih-Cheng, et al.
Pubblicazione: (2023)
Emotion-Aware Speech Self-Supervised Representation Learning with Intensity Knowledge
di: Liu, Rui, et al.
Pubblicazione: (2024)
di: Liu, Rui, et al.
Pubblicazione: (2024)
HingeNet: A Harmonic-Aware Fine-Tuning Approach for Beat Tracking
di: Ru, Ganghui, et al.
Pubblicazione: (2025)
di: Ru, Ganghui, et al.
Pubblicazione: (2025)
MaskBeat: Loopable Drum Beat Generation
di: Lanzendörfer, Luca A., et al.
Pubblicazione: (2025)
di: Lanzendörfer, Luca A., et al.
Pubblicazione: (2025)
The SMC Blind Spot: A Failure Mode Analysis of State-of-the-Art Beat Tracking
di: Ahn, Jaehoon, et al.
Pubblicazione: (2026)
di: Ahn, Jaehoon, et al.
Pubblicazione: (2026)
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval
di: Stewart, Shanti, et al.
Pubblicazione: (2024)
di: Stewart, Shanti, et al.
Pubblicazione: (2024)
Episodic fine-tuning prototypical networks for optimization-based few-shot learning: Application to audio classification
di: Zhuang, Xuanyu, et al.
Pubblicazione: (2024)
di: Zhuang, Xuanyu, et al.
Pubblicazione: (2024)
The Inverse Drum Machine: Source Separation Through Joint Transcription and Analysis-by-Synthesis
di: Torres, Bernardo, et al.
Pubblicazione: (2025)
di: Torres, Bernardo, et al.
Pubblicazione: (2025)
Unsupervised Harmonic Parameter Estimation Using Differentiable DSP and Spectral Optimal Transport
di: Torres, Bernardo, et al.
Pubblicazione: (2023)
di: Torres, Bernardo, et al.
Pubblicazione: (2023)
Leveraging Self-Supervised Learning for Speaker Diarization
di: Han, Jiangyu, et al.
Pubblicazione: (2024)
di: Han, Jiangyu, et al.
Pubblicazione: (2024)
AC-Mix: Self-Supervised Adaptation for Low-Resource Automatic Speech Recognition using Agnostic Contrastive Mixup
di: Carvalho, Carlos, et al.
Pubblicazione: (2024)
di: Carvalho, Carlos, et al.
Pubblicazione: (2024)
Emotion-Coherent Speech Data Augmentation and Self-Supervised Contrastive Style Training for Enhancing Kids's Story Speech Synthesis
di: Chung, Raymond
Pubblicazione: (2026)
di: Chung, Raymond
Pubblicazione: (2026)
Low-Resource Self-Supervised Learning with SSL-Enhanced TTS
di: Hsu, Po-chun, et al.
Pubblicazione: (2023)
di: Hsu, Po-chun, et al.
Pubblicazione: (2023)
Reduction of Nonlinear Distortion in Condenser Microphones Using a Simple Post-Processing Technique
di: Honzík, Petr, et al.
Pubblicazione: (2024)
di: Honzík, Petr, et al.
Pubblicazione: (2024)
The Effect of Batch Size on Contrastive Self-Supervised Speech Representation Learning
di: Vaessen, Nik, et al.
Pubblicazione: (2024)
di: Vaessen, Nik, et al.
Pubblicazione: (2024)
Balancing Information Preservation and Disentanglement in Self-Supervised Music Representation Learning
di: Wilkins, Julia, et al.
Pubblicazione: (2025)
di: Wilkins, Julia, et al.
Pubblicazione: (2025)
Self-Supervised Multi-View Learning for Disentangled Music Audio Representations
di: Wilkins, Julia, et al.
Pubblicazione: (2024)
di: Wilkins, Julia, et al.
Pubblicazione: (2024)
Beat-It: Beat-Synchronized Multi-Condition 3D Dance Generation
di: Huang, Zikai, et al.
Pubblicazione: (2024)
di: Huang, Zikai, et al.
Pubblicazione: (2024)
Stem-JEPA: A Joint-Embedding Predictive Architecture for Musical Stem Compatibility Estimation
di: Riou, Alain, et al.
Pubblicazione: (2024)
di: Riou, Alain, et al.
Pubblicazione: (2024)
Self-Distillation Prototypes Network: Learning Robust Speaker Representations without Supervision
di: Chen, Yafeng, et al.
Pubblicazione: (2024)
di: Chen, Yafeng, et al.
Pubblicazione: (2024)
Self-Distillation Prototypes Network: Learning Robust Speaker Representations without Supervision
di: Chen, Yafeng, et al.
Pubblicazione: (2023)
di: Chen, Yafeng, et al.
Pubblicazione: (2023)
Label-Efficient Self-Supervised Speaker Verification With Information Maximization and Contrastive Learning
di: Lepage, Théo, et al.
Pubblicazione: (2022)
di: Lepage, Théo, et al.
Pubblicazione: (2022)
Additive Margin in Contrastive Self-Supervised Frameworks to Learn Discriminative Speaker Representations
di: Lepage, Theo, et al.
Pubblicazione: (2024)
di: Lepage, Theo, et al.
Pubblicazione: (2024)
Documenti analoghi
-
A Contrastive Self-Supervised Learning scheme for beat tracking amenable to few-shot learning
di: Gagnere, Antonin, et al.
Pubblicazione: (2024) -
Zero-shot Musical Stem Retrieval with Joint-Embedding Predictive Architectures
di: Riou, Alain, et al.
Pubblicazione: (2024) -
Contrastive Knowledge Distillation for Embedding Refinement in Personalized Speech Enhancement
di: Serre, Thomas, et al.
Pubblicazione: (2026) -
A lightweight dual-stage framework for personalized speech enhancement based on DeepFilterNet2
di: Serre, Thomas, et al.
Pubblicazione: (2024) -
Less Forgetting for Better Generalization: Exploring Continual-learning Fine-tuning Methods for Speech Self-supervised Representations
di: Zaiem, Salah, et al.
Pubblicazione: (2024)