Contrastive Learning from Synthetic Audio Doppelgängers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cherep, Manuel, Singh, Nikhil |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Creative Text-to-Audio Generation via Synthesizer Programming
von: Cherep, Manuel, et al.
Veröffentlicht: (2024)
von: Cherep, Manuel, et al.
Veröffentlicht: (2024)
Discovering and Steering Interpretable Concepts in Large Generative Music Models
von: Singh, Nikhil, et al.
Veröffentlicht: (2025)
von: Singh, Nikhil, et al.
Veröffentlicht: (2025)
Towards Robust Few-shot Class Incremental Learning in Audio Classification using Contrastive Representation
von: Singh, Riyansha, et al.
Veröffentlicht: (2024)
von: Singh, Riyansha, et al.
Veröffentlicht: (2024)
COCOLA: Coherence-Oriented Contrastive Learning of Musical Audio Representations
von: Ciranni, Ruben, et al.
Veröffentlicht: (2024)
von: Ciranni, Ruben, et al.
Veröffentlicht: (2024)
On Class Separability Pitfalls In Audio-Text Contrastive Zero-Shot Learning
von: Tavares, Tiago, et al.
Veröffentlicht: (2024)
von: Tavares, Tiago, et al.
Veröffentlicht: (2024)
Patch-Mix Contrastive Learning with Audio Spectrogram Transformer on Respiratory Sound Classification
von: Bae, Sangmin, et al.
Veröffentlicht: (2023)
von: Bae, Sangmin, et al.
Veröffentlicht: (2023)
Can Synthetic Audio From Generative Foundation Models Assist Audio Recognition and Speech Modeling?
von: Feng, Tiantian, et al.
Veröffentlicht: (2024)
von: Feng, Tiantian, et al.
Veröffentlicht: (2024)
CAK: Emergent Audio Effects from Minimal Deep Learning
von: Rockman, Austin
Veröffentlicht: (2025)
von: Rockman, Austin
Veröffentlicht: (2025)
Audio Flamingo: A Novel Audio Language Model with Few-Shot Learning and Dialogue Abilities
von: Kong, Zhifeng, et al.
Veröffentlicht: (2024)
von: Kong, Zhifeng, et al.
Veröffentlicht: (2024)
Learning Spatially-Aware Language and Audio Embeddings
von: Devnani, Bhavika, et al.
Veröffentlicht: (2024)
von: Devnani, Bhavika, et al.
Veröffentlicht: (2024)
Learning to Upsample and Upmix Audio in the Latent Domain
von: Bralios, Dimitrios, et al.
Veröffentlicht: (2025)
von: Bralios, Dimitrios, et al.
Veröffentlicht: (2025)
Learning Music Audio Representations With Limited Data
von: Plachouras, Christos, et al.
Veröffentlicht: (2025)
von: Plachouras, Christos, et al.
Veröffentlicht: (2025)
Learning Disentangled Audio Representations through Controlled Synthesis
von: Brima, Yusuf, et al.
Veröffentlicht: (2024)
von: Brima, Yusuf, et al.
Veröffentlicht: (2024)
A2SB: Audio-to-Audio Schrodinger Bridges
von: Kong, Zhifeng, et al.
Veröffentlicht: (2025)
von: Kong, Zhifeng, et al.
Veröffentlicht: (2025)
Zero Shot Audio to Audio Emotion Transfer With Speaker Disentanglement
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
von: Dutta, Soumya, et al.
Veröffentlicht: (2024)
TACOS: Temporally-aligned Audio CaptiOnS for Language-Audio Pretraining
von: Primus, Paul, et al.
Veröffentlicht: (2025)
von: Primus, Paul, et al.
Veröffentlicht: (2025)
Meta-Learning in Audio and Speech Processing: An End to End Comprehensive Review
von: Raimon, Athul, et al.
Veröffentlicht: (2024)
von: Raimon, Athul, et al.
Veröffentlicht: (2024)
The Rarity of Musical Audio Signals Within the Space of Possible Audio Generation
von: Collins, Nick
Veröffentlicht: (2024)
von: Collins, Nick
Veröffentlicht: (2024)
Estimated Audio-Caption Correspondences Improve Language-Based Audio Retrieval
von: Primus, Paul, et al.
Veröffentlicht: (2024)
von: Primus, Paul, et al.
Veröffentlicht: (2024)
Fusing Audio and Metadata Embeddings Improves Language-based Audio Retrieval
von: Primus, Paul, et al.
Veröffentlicht: (2024)
von: Primus, Paul, et al.
Veröffentlicht: (2024)
Audio Match Cutting: Finding and Creating Matching Audio Transitions in Movies and Videos
von: Fedorishin, Dennis, et al.
Veröffentlicht: (2024)
von: Fedorishin, Dennis, et al.
Veröffentlicht: (2024)
Self-Learning for Personalized Keyword Spotting on Ultra-Low-Power Audio Sensors
von: Rusci, Manuele, et al.
Veröffentlicht: (2024)
von: Rusci, Manuele, et al.
Veröffentlicht: (2024)
CLAP-ART: Automated Audio Captioning with Semantic-rich Audio Representation Tokenizer
von: Takeuchi, Daiki, et al.
Veröffentlicht: (2025)
von: Takeuchi, Daiki, et al.
Veröffentlicht: (2025)
Active Restoration of Lost Audio Signals Using Machine Learning and Latent Information
von: Cheddad, Zohra Adila, et al.
Veröffentlicht: (2021)
von: Cheddad, Zohra Adila, et al.
Veröffentlicht: (2021)
Improving Text-To-Audio Models with Synthetic Captions
von: Kong, Zhifeng, et al.
Veröffentlicht: (2024)
von: Kong, Zhifeng, et al.
Veröffentlicht: (2024)
uaMix-MAE: Efficient Tuning of Pretrained Audio Transformers with Unsupervised Audio Mixtures
von: Tabassum, Afrina, et al.
Veröffentlicht: (2024)
von: Tabassum, Afrina, et al.
Veröffentlicht: (2024)
The Effect of Batch Size on Contrastive Self-Supervised Speech Representation Learning
von: Vaessen, Nik, et al.
Veröffentlicht: (2024)
von: Vaessen, Nik, et al.
Veröffentlicht: (2024)
Domain Adaptation for Contrastive Audio-Language Models
von: Deshmukh, Soham, et al.
Veröffentlicht: (2024)
von: Deshmukh, Soham, et al.
Veröffentlicht: (2024)
CLAD: Robust Audio Deepfake Detection Against Manipulation Attacks with Contrastive Learning
von: Wu, Haolin, et al.
Veröffentlicht: (2024)
von: Wu, Haolin, et al.
Veröffentlicht: (2024)
LC-Protonets: Multi-Label Few-Shot Learning for World Music Audio Tagging
von: Papaioannou, Charilaos, et al.
Veröffentlicht: (2024)
von: Papaioannou, Charilaos, et al.
Veröffentlicht: (2024)
Unsupervised Composable Representations for Audio
von: Bindi, Giovanni, et al.
Veröffentlicht: (2024)
von: Bindi, Giovanni, et al.
Veröffentlicht: (2024)
Diffusion Models for Audio Restoration
von: Lemercier, Jean-Marie, et al.
Veröffentlicht: (2024)
von: Lemercier, Jean-Marie, et al.
Veröffentlicht: (2024)
Multi-bit Audio Watermarking
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2025)
von: Lanzendörfer, Luca A., et al.
Veröffentlicht: (2025)
Instabilities in Convnets for Raw Audio
von: Haider, Daniel, et al.
Veröffentlicht: (2023)
von: Haider, Daniel, et al.
Veröffentlicht: (2023)
Automatic Contextual Audio Denoising
von: Luong, Diep, et al.
Veröffentlicht: (2026)
von: Luong, Diep, et al.
Veröffentlicht: (2026)
Enhancing Audio-Language Models through Self-Supervised Post-Training with Text-Audio Pairs
von: Sinha, Anshuman, et al.
Veröffentlicht: (2024)
von: Sinha, Anshuman, et al.
Veröffentlicht: (2024)
"I am bad": Interpreting Stealthy, Universal and Robust Audio Jailbreaks in Audio-Language Models
von: Gupta, Isha, et al.
Veröffentlicht: (2025)
von: Gupta, Isha, et al.
Veröffentlicht: (2025)
Bringing the Discussion of Minima Sharpness to the Audio Domain: a Filter-Normalised Evaluation for Acoustic Scene Classification
von: Milling, Manuel, et al.
Veröffentlicht: (2023)
von: Milling, Manuel, et al.
Veröffentlicht: (2023)
Additive Margin in Contrastive Self-Supervised Frameworks to Learn Discriminative Speaker Representations
von: Lepage, Theo, et al.
Veröffentlicht: (2024)
von: Lepage, Theo, et al.
Veröffentlicht: (2024)
Label-Efficient Self-Supervised Speaker Verification With Information Maximization and Contrastive Learning
von: Lepage, Théo, et al.
Veröffentlicht: (2022)
von: Lepage, Théo, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
Creative Text-to-Audio Generation via Synthesizer Programming
von: Cherep, Manuel, et al.
Veröffentlicht: (2024) -
Discovering and Steering Interpretable Concepts in Large Generative Music Models
von: Singh, Nikhil, et al.
Veröffentlicht: (2025) -
Towards Robust Few-shot Class Incremental Learning in Audio Classification using Contrastive Representation
von: Singh, Riyansha, et al.
Veröffentlicht: (2024) -
COCOLA: Coherence-Oriented Contrastive Learning of Musical Audio Representations
von: Ciranni, Ruben, et al.
Veröffentlicht: (2024) -
On Class Separability Pitfalls In Audio-Text Contrastive Zero-Shot Learning
von: Tavares, Tiago, et al.
Veröffentlicht: (2024)