Melody-Lyrics Matching with Contrastive Alignment Loss
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Changhong, Olvera, Michel, Richard, Gaël |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EMelodyGen: Emotion-Conditioned Melody Generation in ABC Notation with the Musical Feature Template
von: Zhou, Monan, et al.
Veröffentlicht: (2023)
von: Zhou, Monan, et al.
Veröffentlicht: (2023)
Multiscale Matching Driven by Cross-Modal Similarity Consistency for Audio-Text Retrieval
von: Wang, Qian, et al.
Veröffentlicht: (2024)
von: Wang, Qian, et al.
Veröffentlicht: (2024)
Do Captioning Metrics Reflect Music Semantic Alignment?
von: Lee, Jinwoo, et al.
Veröffentlicht: (2024)
von: Lee, Jinwoo, et al.
Veröffentlicht: (2024)
MERGE -- A Bimodal Audio-Lyrics Dataset for Static Music Emotion Recognition
von: Louro, Pedro Lima, et al.
Veröffentlicht: (2024)
von: Louro, Pedro Lima, et al.
Veröffentlicht: (2024)
Audio Prototypical Network For Controllable Music Recommendation
von: Öncel, Fırat, et al.
Veröffentlicht: (2025)
von: Öncel, Fırat, et al.
Veröffentlicht: (2025)
Multi-Axis Speech Similarity via Factor-Partitioned Embeddings
von: O'Regan, Jim, et al.
Veröffentlicht: (2026)
von: O'Regan, Jim, et al.
Veröffentlicht: (2026)
PolySinger: Singing-Voice to Singing-Voice Translation from English to Japanese
von: Antonisen, Silas, et al.
Veröffentlicht: (2024)
von: Antonisen, Silas, et al.
Veröffentlicht: (2024)
Language-based Audio Retrieval with Co-Attention Networks
von: Sun, Haoran, et al.
Veröffentlicht: (2024)
von: Sun, Haoran, et al.
Veröffentlicht: (2024)
Hybrid Losses for Hierarchical Embedding Learning
von: Tian, Haokun, et al.
Veröffentlicht: (2025)
von: Tian, Haokun, et al.
Veröffentlicht: (2025)
Exploring GPT's Ability as a Judge in Music Understanding
von: Fang, Kun, et al.
Veröffentlicht: (2025)
von: Fang, Kun, et al.
Veröffentlicht: (2025)
Evaluating High-Resolution Piano Sustain Pedal Depth Estimation with Musically Informed Metrics
von: Zhang, Hanwen, et al.
Veröffentlicht: (2025)
von: Zhang, Hanwen, et al.
Veröffentlicht: (2025)
Music Era Recognition Using Supervised Contrastive Learning and Artist Information
von: He, Qiqi, et al.
Veröffentlicht: (2024)
von: He, Qiqi, et al.
Veröffentlicht: (2024)
CLASP: Contrastive Language-Speech Pretraining for Multilingual Multimodal Information Retrieval
von: Abootorabi, Mohammad Mahdi, et al.
Veröffentlicht: (2024)
von: Abootorabi, Mohammad Mahdi, et al.
Veröffentlicht: (2024)
VoxRAG: A Step Toward Transcription-Free RAG Systems in Spoken Question Answering
von: Rackauckas, Zackary, et al.
Veröffentlicht: (2025)
von: Rackauckas, Zackary, et al.
Veröffentlicht: (2025)
Bridging the Gap Between Semantic and User Preference Spaces for Multi-modal Music Representation Learning
von: Pan, Xiaofeng, et al.
Veröffentlicht: (2025)
von: Pan, Xiaofeng, et al.
Veröffentlicht: (2025)
Evaluating Interval-based Tokenization for Pitch Representation in Symbolic Music Analysis
von: Le, Dinh-Viet-Toan, et al.
Veröffentlicht: (2025)
von: Le, Dinh-Viet-Toan, et al.
Veröffentlicht: (2025)
TALKPLAY: Multimodal Music Recommendation with Large Language Models
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
Exploring Diverse Sounds: Identifying Outliers in a Music Corpus
von: Cai, Le, et al.
Veröffentlicht: (2024)
von: Cai, Le, et al.
Veröffentlicht: (2024)
Music Discovery Dialogue Generation Using Human Intent Analysis and Large Language Models
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
Track Role Prediction of Single-Instrumental Sequences
von: Han, Changheon, et al.
Veröffentlicht: (2024)
von: Han, Changheon, et al.
Veröffentlicht: (2024)
Personalized Dynamic Music Emotion Recognition with Dual-Scale Attention-Based Meta-Learning
von: Zhang, Dengming, et al.
Veröffentlicht: (2024)
von: Zhang, Dengming, et al.
Veröffentlicht: (2024)
LARP: Language Audio Relational Pre-training for Cold-Start Playlist Continuation
von: Salganik, Rebecca, et al.
Veröffentlicht: (2024)
von: Salganik, Rebecca, et al.
Veröffentlicht: (2024)
Expressivity-aware Music Performance Retrieval using Mid-level Perceptual Features and Emotion Word Embeddings
von: Chowdhury, Shreyan, et al.
Veröffentlicht: (2024)
von: Chowdhury, Shreyan, et al.
Veröffentlicht: (2024)
DiffATR: Diffusion-based Generative Modeling for Audio-Text Retrieval
von: Xin, Yifei, et al.
Veröffentlicht: (2024)
von: Xin, Yifei, et al.
Veröffentlicht: (2024)
Multi-Sample Dynamic Time Warping for Few-Shot Keyword Spotting
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2024)
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2024)
Automatic Estimation of Singing Voice Musical Dynamics
von: Narang, Jyoti, et al.
Veröffentlicht: (2024)
von: Narang, Jyoti, et al.
Veröffentlicht: (2024)
FusID: Modality-Fused Semantic IDs for Generative Music Recommendation
von: Kim, Haven, et al.
Veröffentlicht: (2026)
von: Kim, Haven, et al.
Veröffentlicht: (2026)
A SOUND APPROACH: Using Large Language Models to generate audio descriptions for egocentric text-audio retrieval
von: Oncescu, Andreea-Maria, et al.
Veröffentlicht: (2024)
von: Oncescu, Andreea-Maria, et al.
Veröffentlicht: (2024)
Engraving Oriented Joint Estimation of Pitch Spelling and Local and Global Keys
von: Bouquillard, Augustin, et al.
Veröffentlicht: (2024)
von: Bouquillard, Augustin, et al.
Veröffentlicht: (2024)
Towards Computational Analysis of Pansori Singing
von: Park, Sangheon, et al.
Veröffentlicht: (2024)
von: Park, Sangheon, et al.
Veröffentlicht: (2024)
REFFLY: Melody-Constrained Lyrics Editing Model
von: Zhao, Songyan, et al.
Veröffentlicht: (2024)
von: Zhao, Songyan, et al.
Veröffentlicht: (2024)
Zema Dataset: A Comprehensive Study of Yaredawi Zema with a Focus on Horologium Chants
von: Muluneh, Mequanent Argaw, et al.
Veröffentlicht: (2024)
von: Muluneh, Mequanent Argaw, et al.
Veröffentlicht: (2024)
A Cascaded Architecture for Extractive Summarization of Multimedia Content via Audio-to-Text Alignment
von: Hossain, Tanzir, et al.
Veröffentlicht: (2025)
von: Hossain, Tanzir, et al.
Veröffentlicht: (2025)
Joint Learning of Wording and Formatting for Singable Melody-to-Lyric Generation
von: Ou, Longshen, et al.
Veröffentlicht: (2023)
von: Ou, Longshen, et al.
Veröffentlicht: (2023)
JEPOO: Highly Accurate Joint Estimation of Pitch, Onset and Offset for Music Information Retrieval
von: Wei, Haojie, et al.
Veröffentlicht: (2023)
von: Wei, Haojie, et al.
Veröffentlicht: (2023)
On the Effect of Data-Augmentation on Local Embedding Properties in the Contrastive Learning of Music Audio Representations
von: McCallum, Matthew C., et al.
Veröffentlicht: (2024)
von: McCallum, Matthew C., et al.
Veröffentlicht: (2024)
SongGLM: Lyric-to-Melody Generation with 2D Alignment Encoding and Multi-Task Pre-Training
von: Yu, Jiaxing, et al.
Veröffentlicht: (2024)
von: Yu, Jiaxing, et al.
Veröffentlicht: (2024)
TalkPlay-Tools: Conversational Music Recommendation with LLM Tool Calling
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
Streaming Piano Transcription Based on Consistent Onset and Offset Decoding with Sustain Pedal Detection
von: Wei, Weixing, et al.
Veröffentlicht: (2025)
von: Wei, Weixing, et al.
Veröffentlicht: (2025)
Diff4Steer: Steerable Diffusion Prior for Generative Music Retrieval with Semantic Guidance
von: Bao, Xuchan, et al.
Veröffentlicht: (2024)
von: Bao, Xuchan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
EMelodyGen: Emotion-Conditioned Melody Generation in ABC Notation with the Musical Feature Template
von: Zhou, Monan, et al.
Veröffentlicht: (2023) -
Multiscale Matching Driven by Cross-Modal Similarity Consistency for Audio-Text Retrieval
von: Wang, Qian, et al.
Veröffentlicht: (2024) -
Do Captioning Metrics Reflect Music Semantic Alignment?
von: Lee, Jinwoo, et al.
Veröffentlicht: (2024) -
MERGE -- A Bimodal Audio-Lyrics Dataset for Static Music Emotion Recognition
von: Louro, Pedro Lima, et al.
Veröffentlicht: (2024) -
Audio Prototypical Network For Controllable Music Recommendation
von: Öncel, Fırat, et al.
Veröffentlicht: (2025)