Contrastive and Transfer Learning for Effective Audio Fingerprinting through a Real-World Evaluation Protocol
Fuente:
arXiv
Saved in:
| Main Authors: | Nikou, Christos, Giannakopoulos, Theodoros |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Segment Length Matters: A Study of Segment Lengths on Audio Fingerprinting Performance
by: Gong, Ziling, et al.
Published: (2026)
by: Gong, Ziling, et al.
Published: (2026)
Pretrained Conformers for Audio Fingerprinting and Retrieval
by: Altwlkany, Kemal, et al.
Published: (2025)
by: Altwlkany, Kemal, et al.
Published: (2025)
Latent Diffusion Bridges for Unsupervised Musical Audio Timbre Transfer
by: Mancusi, Michele, et al.
Published: (2024)
by: Mancusi, Michele, et al.
Published: (2024)
Evaluation of pretrained language models on music understanding
by: Vasilakis, Yannis, et al.
Published: (2024)
by: Vasilakis, Yannis, et al.
Published: (2024)
DeepSRGM -- Sequence Classification and Ranking in Indian Classical Music with Deep Learning
by: Madhusudhan, Sathwik Tejaswi, et al.
Published: (2024)
by: Madhusudhan, Sathwik Tejaswi, et al.
Published: (2024)
On the Effect of Data-Augmentation on Local Embedding Properties in the Contrastive Learning of Music Audio Representations
by: McCallum, Matthew C., et al.
Published: (2024)
by: McCallum, Matthew C., et al.
Published: (2024)
Music Era Recognition Using Supervised Contrastive Learning and Artist Information
by: He, Qiqi, et al.
Published: (2024)
by: He, Qiqi, et al.
Published: (2024)
Uncertainty Estimation in the Real World: A Study on Music Emotion Recognition
by: Watcharasupat, Karn N., et al.
Published: (2025)
by: Watcharasupat, Karn N., et al.
Published: (2025)
Universal Music Representations? Evaluating Foundation Models on World Music Corpora
by: Papaioannou, Charilaos, et al.
Published: (2025)
by: Papaioannou, Charilaos, et al.
Published: (2025)
A Stem-Agnostic Single-Decoder System for Music Source Separation Beyond Four Stems
by: Watcharasupat, Karn N., et al.
Published: (2024)
by: Watcharasupat, Karn N., et al.
Published: (2024)
Towards Robust Transcription: Exploring Noise Injection Strategies for Training Data Augmentation
by: Kim, Yonghyun, et al.
Published: (2024)
by: Kim, Yonghyun, et al.
Published: (2024)
Transfer Learning with Semi-Supervised Dataset Annotation for Birdcall Classification
by: Miyaguchi, Anthony, et al.
Published: (2023)
by: Miyaguchi, Anthony, et al.
Published: (2023)
Dissecting Temporal Understanding in Text-to-Audio Retrieval
by: Oncescu, Andreea-Maria, et al.
Published: (2024)
by: Oncescu, Andreea-Maria, et al.
Published: (2024)
A Cascaded Architecture for Extractive Summarization of Multimedia Content via Audio-to-Text Alignment
by: Hossain, Tanzir, et al.
Published: (2025)
by: Hossain, Tanzir, et al.
Published: (2025)
EAViT: External Attention Vision Transformer for Audio Classification
by: Iqbal, Aquib, et al.
Published: (2024)
by: Iqbal, Aquib, et al.
Published: (2024)
High-Resolution Sustain Pedal Depth Estimation from Piano Audio Across Room Acoustics
by: Fang, Kun, et al.
Published: (2025)
by: Fang, Kun, et al.
Published: (2025)
A Novel Audio Representation for Music Genre Identification in MIR
by: Kamuni, Navin, et al.
Published: (2024)
by: Kamuni, Navin, et al.
Published: (2024)
Anchor-aware Deep Metric Learning for Audio-visual Retrieval
by: Zeng, Donghuo, et al.
Published: (2024)
by: Zeng, Donghuo, et al.
Published: (2024)
Nested Music Transformer: Sequentially Decoding Compound Tokens in Symbolic Music and Audio Generation
by: Yoo, HaeJun, et al.
Published: (2024)
by: Yoo, HaeJun, et al.
Published: (2024)
Speaker Retrieval in the Wild: Challenges, Effectiveness and Robustness
by: Loweimi, Erfan, et al.
Published: (2025)
by: Loweimi, Erfan, et al.
Published: (2025)
From Real to Cloned Singer Identification
by: Desblancs, Dorian, et al.
Published: (2024)
by: Desblancs, Dorian, et al.
Published: (2024)
Application of Audio Fingerprinting Techniques for Real-Time Scalable Speech Retrieval and Speech Clusterization
by: Altwlkany, Kemal, et al.
Published: (2024)
by: Altwlkany, Kemal, et al.
Published: (2024)
Music Auto-Tagging with Robust Music Representation Learned via Domain Adversarial Training
by: Joung, Haesun, et al.
Published: (2024)
by: Joung, Haesun, et al.
Published: (2024)
wav2graph: A Framework for Supervised Learning Knowledge Graph from Speech
by: Le-Duc, Khai, et al.
Published: (2024)
by: Le-Duc, Khai, et al.
Published: (2024)
ASK: Adaptive Self-improving Knowledge Framework for Audio Text Retrieval
by: Fu, Siyuan, et al.
Published: (2025)
by: Fu, Siyuan, et al.
Published: (2025)
Hybrid Losses for Hierarchical Embedding Learning
by: Tian, Haokun, et al.
Published: (2025)
by: Tian, Haokun, et al.
Published: (2025)
MERGE -- A Bimodal Audio-Lyrics Dataset for Static Music Emotion Recognition
by: Louro, Pedro Lima, et al.
Published: (2024)
by: Louro, Pedro Lima, et al.
Published: (2024)
Deconstructing Jazz Piano Style Using Machine Learning
by: Cheston, Huw, et al.
Published: (2025)
by: Cheston, Huw, et al.
Published: (2025)
Improving Musical Instrument Classification with Advanced Machine Learning Techniques
by: Chulev, Joanikij
Published: (2024)
by: Chulev, Joanikij
Published: (2024)
Advancing the Foundation Model for Music Understanding
by: Jiang, Yi, et al.
Published: (2025)
by: Jiang, Yi, et al.
Published: (2025)
Analyzing Musical Characteristics of National Anthems in Relation to Global Indices
by: Hasan, S M Rakib, et al.
Published: (2024)
by: Hasan, S M Rakib, et al.
Published: (2024)
Distance Sampling-based Paraphraser Leveraging ChatGPT for Text Data Manipulation
by: Oh, Yoori, et al.
Published: (2024)
by: Oh, Yoori, et al.
Published: (2024)
SoundSignature: What Type of Music Do You Like?
by: Carone, Brandon James, et al.
Published: (2024)
by: Carone, Brandon James, et al.
Published: (2024)
Towards Explainable and Interpretable Musical Difficulty Estimation: A Parameter-efficient Approach
by: Ramoneda, Pedro, et al.
Published: (2024)
by: Ramoneda, Pedro, et al.
Published: (2024)
Natural Language Processing Methods for Symbolic Music Generation and Information Retrieval: a Survey
by: Le, Dinh-Viet-Toan, et al.
Published: (2024)
by: Le, Dinh-Viet-Toan, et al.
Published: (2024)
EMelodyGen: Emotion-Conditioned Melody Generation in ABC Notation with the Musical Feature Template
by: Zhou, Monan, et al.
Published: (2023)
by: Zhou, Monan, et al.
Published: (2023)
An approach to hummed-tune and song sequences matching
by: Pham, Loc Bao, et al.
Published: (2024)
by: Pham, Loc Bao, et al.
Published: (2024)
Towards Training Music Taggers on Synthetic Data
by: Kroher, Nadine, et al.
Published: (2024)
by: Kroher, Nadine, et al.
Published: (2024)
KAD: No More FAD! An Effective and Efficient Evaluation Metric for Audio Generation
by: Chung, Yoonjin, et al.
Published: (2025)
by: Chung, Yoonjin, et al.
Published: (2025)
Analyzing and reducing the synthetic-to-real transfer gap in Music Information Retrieval: the task of automatic drum transcription
by: Zehren, Mickaël, et al.
Published: (2024)
by: Zehren, Mickaël, et al.
Published: (2024)
Similar Items
-
Segment Length Matters: A Study of Segment Lengths on Audio Fingerprinting Performance
by: Gong, Ziling, et al.
Published: (2026) -
Pretrained Conformers for Audio Fingerprinting and Retrieval
by: Altwlkany, Kemal, et al.
Published: (2025) -
Latent Diffusion Bridges for Unsupervised Musical Audio Timbre Transfer
by: Mancusi, Michele, et al.
Published: (2024) -
Evaluation of pretrained language models on music understanding
by: Vasilakis, Yannis, et al.
Published: (2024) -
DeepSRGM -- Sequence Classification and Ranking in Indian Classical Music with Deep Learning
by: Madhusudhan, Sathwik Tejaswi, et al.
Published: (2024)