Similar but Faster: Manipulation of Tempo in Music Audio Embeddings for Tempo Prediction and Search
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | McCallum, Matthew C., Henkel, Florian, Kim, Jaehun, Sandberg, Samuel E., Davies, Matthew E. P. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On the Effect of Data-Augmentation on Local Embedding Properties in the Contrastive Learning of Music Audio Representations
von: McCallum, Matthew C., et al.
Veröffentlicht: (2024)
von: McCallum, Matthew C., et al.
Veröffentlicht: (2024)
Tempo estimation as fully self-supervised binary classification
von: Henkel, Florian, et al.
Veröffentlicht: (2024)
von: Henkel, Florian, et al.
Veröffentlicht: (2024)
The GigaMIDI Dataset with Features for Expressive Music Performance Detection
von: Lee, Keon Ju Maverick, et al.
Veröffentlicht: (2025)
von: Lee, Keon Ju Maverick, et al.
Veröffentlicht: (2025)
Multiscale Matching Driven by Cross-Modal Similarity Consistency for Audio-Text Retrieval
von: Wang, Qian, et al.
Veröffentlicht: (2024)
von: Wang, Qian, et al.
Veröffentlicht: (2024)
An Open Research Dataset of the 1932 Cairo Congress of Arab Music
von: Bozkurt, Baris
Veröffentlicht: (2025)
von: Bozkurt, Baris
Veröffentlicht: (2025)
Expressivity-aware Music Performance Retrieval using Mid-level Perceptual Features and Emotion Word Embeddings
von: Chowdhury, Shreyan, et al.
Veröffentlicht: (2024)
von: Chowdhury, Shreyan, et al.
Veröffentlicht: (2024)
FusID: Modality-Fused Semantic IDs for Generative Music Recommendation
von: Kim, Haven, et al.
Veröffentlicht: (2026)
von: Kim, Haven, et al.
Veröffentlicht: (2026)
Nested Music Transformer: Sequentially Decoding Compound Tokens in Symbolic Music and Audio Generation
von: Yoo, HaeJun, et al.
Veröffentlicht: (2024)
von: Yoo, HaeJun, et al.
Veröffentlicht: (2024)
Music Discovery Dialogue Generation Using Human Intent Analysis and Large Language Models
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
Latent Diffusion Bridges for Unsupervised Musical Audio Timbre Transfer
von: Mancusi, Michele, et al.
Veröffentlicht: (2024)
von: Mancusi, Michele, et al.
Veröffentlicht: (2024)
A Novel Audio Representation for Music Genre Identification in MIR
von: Kamuni, Navin, et al.
Veröffentlicht: (2024)
von: Kamuni, Navin, et al.
Veröffentlicht: (2024)
Automatic Estimation of Singing Voice Musical Dynamics
von: Narang, Jyoti, et al.
Veröffentlicht: (2024)
von: Narang, Jyoti, et al.
Veröffentlicht: (2024)
Exploring GPT's Ability as a Judge in Music Understanding
von: Fang, Kun, et al.
Veröffentlicht: (2025)
von: Fang, Kun, et al.
Veröffentlicht: (2025)
Do Captioning Metrics Reflect Music Semantic Alignment?
von: Lee, Jinwoo, et al.
Veröffentlicht: (2024)
von: Lee, Jinwoo, et al.
Veröffentlicht: (2024)
TALKPLAY: Multimodal Music Recommendation with Large Language Models
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
Exploring Diverse Sounds: Identifying Outliers in a Music Corpus
von: Cai, Le, et al.
Veröffentlicht: (2024)
von: Cai, Le, et al.
Veröffentlicht: (2024)
Language-based Audio Retrieval with Co-Attention Networks
von: Sun, Haoran, et al.
Veröffentlicht: (2024)
von: Sun, Haoran, et al.
Veröffentlicht: (2024)
Evaluating Interval-based Tokenization for Pitch Representation in Symbolic Music Analysis
von: Le, Dinh-Viet-Toan, et al.
Veröffentlicht: (2025)
von: Le, Dinh-Viet-Toan, et al.
Veröffentlicht: (2025)
DiffATR: Diffusion-based Generative Modeling for Audio-Text Retrieval
von: Xin, Yifei, et al.
Veröffentlicht: (2024)
von: Xin, Yifei, et al.
Veröffentlicht: (2024)
Personalized Dynamic Music Emotion Recognition with Dual-Scale Attention-Based Meta-Learning
von: Zhang, Dengming, et al.
Veröffentlicht: (2024)
von: Zhang, Dengming, et al.
Veröffentlicht: (2024)
Evaluating High-Resolution Piano Sustain Pedal Depth Estimation with Musically Informed Metrics
von: Zhang, Hanwen, et al.
Veröffentlicht: (2025)
von: Zhang, Hanwen, et al.
Veröffentlicht: (2025)
LARP: Language Audio Relational Pre-training for Cold-Start Playlist Continuation
von: Salganik, Rebecca, et al.
Veröffentlicht: (2024)
von: Salganik, Rebecca, et al.
Veröffentlicht: (2024)
Bridging the Gap Between Semantic and User Preference Spaces for Multi-modal Music Representation Learning
von: Pan, Xiaofeng, et al.
Veröffentlicht: (2025)
von: Pan, Xiaofeng, et al.
Veröffentlicht: (2025)
Enriching Music Descriptions with a Finetuned-LLM and Metadata for Text-to-Music Retrieval
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
GraphMuse: A Library for Symbolic Music Graph Processing
von: Karystinaios, Emmanouil, et al.
Veröffentlicht: (2024)
von: Karystinaios, Emmanouil, et al.
Veröffentlicht: (2024)
Sanidha: A Studio Quality Multi-Modal Dataset for Carnatic Music
von: Krishnan, Venkatakrishnan Vaidyanathapuram, et al.
Veröffentlicht: (2025)
von: Krishnan, Venkatakrishnan Vaidyanathapuram, et al.
Veröffentlicht: (2025)
MERGE -- A Bimodal Audio-Lyrics Dataset for Static Music Emotion Recognition
von: Louro, Pedro Lima, et al.
Veröffentlicht: (2024)
von: Louro, Pedro Lima, et al.
Veröffentlicht: (2024)
TalkPlay-Tools: Conversational Music Recommendation with LLM Tool Calling
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
Track Role Prediction of Single-Instrumental Sequences
von: Han, Changheon, et al.
Veröffentlicht: (2024)
von: Han, Changheon, et al.
Veröffentlicht: (2024)
Music Foundation Model as Generic Booster for Music Downstream Tasks
von: Liao, WeiHsiang, et al.
Veröffentlicht: (2024)
von: Liao, WeiHsiang, et al.
Veröffentlicht: (2024)
Universal Music Representations? Evaluating Foundation Models on World Music Corpora
von: Papaioannou, Charilaos, et al.
Veröffentlicht: (2025)
von: Papaioannou, Charilaos, et al.
Veröffentlicht: (2025)
Diff4Steer: Steerable Diffusion Prior for Generative Music Retrieval with Semantic Guidance
von: Bao, Xuchan, et al.
Veröffentlicht: (2024)
von: Bao, Xuchan, et al.
Veröffentlicht: (2024)
JEPOO: Highly Accurate Joint Estimation of Pitch, Onset and Offset for Music Information Retrieval
von: Wei, Haojie, et al.
Veröffentlicht: (2023)
von: Wei, Haojie, et al.
Veröffentlicht: (2023)
A Manual Bar-by-Bar Tempo Measurement Protocol for Polyphonic Chamber Music Recordings: Design, Validation, and Application to Beethoven's Piano and Cello Sonatas
von: Sole, Ignasi
Veröffentlicht: (2026)
von: Sole, Ignasi
Veröffentlicht: (2026)
Music Auto-Tagging with Robust Music Representation Learned via Domain Adversarial Training
von: Joung, Haesun, et al.
Veröffentlicht: (2024)
von: Joung, Haesun, et al.
Veröffentlicht: (2024)
A SOUND APPROACH: Using Large Language Models to generate audio descriptions for egocentric text-audio retrieval
von: Oncescu, Andreea-Maria, et al.
Veröffentlicht: (2024)
von: Oncescu, Andreea-Maria, et al.
Veröffentlicht: (2024)
A Dataset and Baselines for Measuring and Predicting the Music Piece Memorability
von: Tseng, Li-Yang, et al.
Veröffentlicht: (2024)
von: Tseng, Li-Yang, et al.
Veröffentlicht: (2024)
Analyzing Byte-Pair Encoding on Monophonic and Polyphonic Symbolic Music: A Focus on Musical Phrase Segmentation
von: Le, Dinh-Viet-Toan, et al.
Veröffentlicht: (2024)
von: Le, Dinh-Viet-Toan, et al.
Veröffentlicht: (2024)
Advancing the Foundation Model for Music Understanding
von: Jiang, Yi, et al.
Veröffentlicht: (2025)
von: Jiang, Yi, et al.
Veröffentlicht: (2025)
Pretrained Conformers for Audio Fingerprinting and Retrieval
von: Altwlkany, Kemal, et al.
Veröffentlicht: (2025)
von: Altwlkany, Kemal, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
On the Effect of Data-Augmentation on Local Embedding Properties in the Contrastive Learning of Music Audio Representations
von: McCallum, Matthew C., et al.
Veröffentlicht: (2024) -
Tempo estimation as fully self-supervised binary classification
von: Henkel, Florian, et al.
Veröffentlicht: (2024) -
The GigaMIDI Dataset with Features for Expressive Music Performance Detection
von: Lee, Keon Ju Maverick, et al.
Veröffentlicht: (2025) -
Multiscale Matching Driven by Cross-Modal Similarity Consistency for Audio-Text Retrieval
von: Wang, Qian, et al.
Veröffentlicht: (2024) -
An Open Research Dataset of the 1932 Cairo Congress of Arab Music
von: Bozkurt, Baris
Veröffentlicht: (2025)