Enriching Music Descriptions with a Finetuned-LLM and Metadata for Text-to-Music Retrieval
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Doh, SeungHeon, Lee, Minhee, Jeong, Dasaem, Nam, Juhan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Musical Word Embedding for Music Tagging and Retrieval
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
TalkPlay-Tools: Conversational Music Recommendation with LLM Tool Calling
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
Music Discovery Dialogue Generation Using Human Intent Analysis and Large Language Models
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
TalkPlayData 2: An Agentic Synthetic Data Pipeline for Multimodal Conversational Music Recommendation
von: Choi, Keunwoo, et al.
Veröffentlicht: (2025)
von: Choi, Keunwoo, et al.
Veröffentlicht: (2025)
Predicting User Intents and Musical Attributes from Music Discovery Conversations
von: Kwon, Daeyong, et al.
Veröffentlicht: (2024)
von: Kwon, Daeyong, et al.
Veröffentlicht: (2024)
TALKPLAY: Multimodal Music Recommendation with Large Language Models
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
Diff4Steer: Steerable Diffusion Prior for Generative Music Retrieval with Semantic Guidance
von: Bao, Xuchan, et al.
Veröffentlicht: (2024)
von: Bao, Xuchan, et al.
Veröffentlicht: (2024)
JEPOO: Highly Accurate Joint Estimation of Pitch, Onset and Offset for Music Information Retrieval
von: Wei, Haojie, et al.
Veröffentlicht: (2023)
von: Wei, Haojie, et al.
Veröffentlicht: (2023)
Learning Normal Patterns in Musical Loops
von: Dadman, Shayan, et al.
Veröffentlicht: (2025)
von: Dadman, Shayan, et al.
Veröffentlicht: (2025)
Flexible Control in Symbolic Music Generation via Musical Metadata
von: Han, Sangjun, et al.
Veröffentlicht: (2024)
von: Han, Sangjun, et al.
Veröffentlicht: (2024)
A Dataset and Baselines for Measuring and Predicting the Music Piece Memorability
von: Tseng, Li-Yang, et al.
Veröffentlicht: (2024)
von: Tseng, Li-Yang, et al.
Veröffentlicht: (2024)
Music Genre Classification: Ensemble Learning with Subcomponents-level Attention
von: Liu, Yichen, et al.
Veröffentlicht: (2024)
von: Liu, Yichen, et al.
Veröffentlicht: (2024)
Nested Music Transformer: Sequentially Decoding Compound Tokens in Symbolic Music and Audio Generation
von: Yoo, HaeJun, et al.
Veröffentlicht: (2024)
von: Yoo, HaeJun, et al.
Veröffentlicht: (2024)
MERGE -- A Bimodal Audio-Lyrics Dataset for Static Music Emotion Recognition
von: Louro, Pedro Lima, et al.
Veröffentlicht: (2024)
von: Louro, Pedro Lima, et al.
Veröffentlicht: (2024)
On the Effect of Data-Augmentation on Local Embedding Properties in the Contrastive Learning of Music Audio Representations
von: McCallum, Matthew C., et al.
Veröffentlicht: (2024)
von: McCallum, Matthew C., et al.
Veröffentlicht: (2024)
ASK: Adaptive Self-improving Knowledge Framework for Audio Text Retrieval
von: Fu, Siyuan, et al.
Veröffentlicht: (2025)
von: Fu, Siyuan, et al.
Veröffentlicht: (2025)
SteerMusic: Enhanced Musical Consistency for Zero-shot Text-guided and Personalized Music Editing
von: Niu, Xinlei, et al.
Veröffentlicht: (2025)
von: Niu, Xinlei, et al.
Veröffentlicht: (2025)
Streaming Piano Transcription Based on Consistent Onset and Offset Decoding with Sustain Pedal Detection
von: Wei, Weixing, et al.
Veröffentlicht: (2025)
von: Wei, Weixing, et al.
Veröffentlicht: (2025)
Towards Computational Analysis of Pansori Singing
von: Park, Sangheon, et al.
Veröffentlicht: (2024)
von: Park, Sangheon, et al.
Veröffentlicht: (2024)
MusicSem: A Semantically Rich Language--Audio Dataset of Natural Music Descriptions
von: Salganik, Rebecca, et al.
Veröffentlicht: (2026)
von: Salganik, Rebecca, et al.
Veröffentlicht: (2026)
PIAST: A Multimodal Piano Dataset with Audio, Symbolic and Text
von: Bang, Hayeon, et al.
Veröffentlicht: (2024)
von: Bang, Hayeon, et al.
Veröffentlicht: (2024)
Intelligent Text-Conditioned Music Generation
von: Xie, Zhouyao, et al.
Veröffentlicht: (2024)
von: Xie, Zhouyao, et al.
Veröffentlicht: (2024)
Can Large Language Models Predict Audio Effects Parameters from Natural Language?
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
On the de-duplication of the Lakh MIDI dataset
von: Choi, Eunjin, et al.
Veröffentlicht: (2025)
von: Choi, Eunjin, et al.
Veröffentlicht: (2025)
PianoBind: A Multimodal Joint Embedding Model for Pop-piano Music
von: Bang, Hayeon, et al.
Veröffentlicht: (2025)
von: Bang, Hayeon, et al.
Veröffentlicht: (2025)
Anchor-aware Deep Metric Learning for Audio-visual Retrieval
von: Zeng, Donghuo, et al.
Veröffentlicht: (2024)
von: Zeng, Donghuo, et al.
Veröffentlicht: (2024)
Semi-Supervised Contrastive Learning for Controllable Video-to-Music Retrieval
von: Stewart, Shanti, et al.
Veröffentlicht: (2024)
von: Stewart, Shanti, et al.
Veröffentlicht: (2024)
Do Captioning Metrics Reflect Music Semantic Alignment?
von: Lee, Jinwoo, et al.
Veröffentlicht: (2024)
von: Lee, Jinwoo, et al.
Veröffentlicht: (2024)
SynthTab: Leveraging Synthesized Data for Guitar Tablature Transcription
von: Zang, Yongyi, et al.
Veröffentlicht: (2023)
von: Zang, Yongyi, et al.
Veröffentlicht: (2023)
Can Audio Reveal Music Performance Difficulty? Insights from the Piano Syllabus Dataset
von: Ramoneda, Pedro, et al.
Veröffentlicht: (2024)
von: Ramoneda, Pedro, et al.
Veröffentlicht: (2024)
Video-Guided Text-to-Music Generation Using Public Domain Movie Collections
von: Kim, Haven, et al.
Veröffentlicht: (2025)
von: Kim, Haven, et al.
Veröffentlicht: (2025)
Multi-Track MusicLDM: Towards Versatile Music Generation with Latent Diffusion Model
von: Karchkhadze, Tornike, et al.
Veröffentlicht: (2024)
von: Karchkhadze, Tornike, et al.
Veröffentlicht: (2024)
MusicAOG: an Energy-Based Model for Learning and Sampling a Hierarchical Representation of Symbolic Music
von: Qian, Yikai, et al.
Veröffentlicht: (2024)
von: Qian, Yikai, et al.
Veröffentlicht: (2024)
Dance-to-Music Generation with Encoder-based Textual Inversion
von: Li, Sifei, et al.
Veröffentlicht: (2024)
von: Li, Sifei, et al.
Veröffentlicht: (2024)
Expressivity-aware Music Performance Retrieval using Mid-level Perceptual Features and Emotion Word Embeddings
von: Chowdhury, Shreyan, et al.
Veröffentlicht: (2024)
von: Chowdhury, Shreyan, et al.
Veröffentlicht: (2024)
MeloTrans: A Text to Symbolic Music Generation Model Following Human Composition Habit
von: Wang, Yutian, et al.
Veröffentlicht: (2024)
von: Wang, Yutian, et al.
Veröffentlicht: (2024)
Optimizing Feature Extraction for Symbolic Music
von: Simonetta, Federico, et al.
Veröffentlicht: (2023)
von: Simonetta, Federico, et al.
Veröffentlicht: (2023)
LAV: Audio-Driven Dynamic Visual Generation with Neural Compression and StyleGAN2
von: Jung, Jongmin, et al.
Veröffentlicht: (2025)
von: Jung, Jongmin, et al.
Veröffentlicht: (2025)
Automatic Estimation of Singing Voice Musical Dynamics
von: Narang, Jyoti, et al.
Veröffentlicht: (2024)
von: Narang, Jyoti, et al.
Veröffentlicht: (2024)
A Survey on Evaluation Metrics for Music Generation
von: Kader, Faria Binte, et al.
Veröffentlicht: (2025)
von: Kader, Faria Binte, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Musical Word Embedding for Music Tagging and Retrieval
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024) -
TalkPlay-Tools: Conversational Music Recommendation with LLM Tool Calling
von: Doh, Seungheon, et al.
Veröffentlicht: (2025) -
Music Discovery Dialogue Generation Using Human Intent Analysis and Large Language Models
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024) -
TalkPlayData 2: An Agentic Synthetic Data Pipeline for Multimodal Conversational Music Recommendation
von: Choi, Keunwoo, et al.
Veröffentlicht: (2025) -
Predicting User Intents and Musical Attributes from Music Discovery Conversations
von: Kwon, Daeyong, et al.
Veröffentlicht: (2024)