WikiMuTe: A web-sourced dataset of semantic descriptions for music audio
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Weck, Benno, Kirchhoff, Holger, Grosche, Peter, Serra, Xavier |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A SOUND APPROACH: Using Large Language Models to generate audio descriptions for egocentric text-audio retrieval
von: Oncescu, Andreea-Maria, et al.
Veröffentlicht: (2024)
von: Oncescu, Andreea-Maria, et al.
Veröffentlicht: (2024)
CrossMuSim: A Cross-Modal Framework for Music Similarity Retrieval with LLM-Powered Text Description Sourcing and Mining
von: Tsoi, Tristan, et al.
Veröffentlicht: (2025)
von: Tsoi, Tristan, et al.
Veröffentlicht: (2025)
The language of sound search: Examining User Queries in Audio Search Engines
von: Weck, Benno, et al.
Veröffentlicht: (2024)
von: Weck, Benno, et al.
Veröffentlicht: (2024)
Automatic Estimation of Singing Voice Musical Dynamics
von: Narang, Jyoti, et al.
Veröffentlicht: (2024)
von: Narang, Jyoti, et al.
Veröffentlicht: (2024)
Towards Explainable and Interpretable Musical Difficulty Estimation: A Parameter-efficient Approach
von: Ramoneda, Pedro, et al.
Veröffentlicht: (2024)
von: Ramoneda, Pedro, et al.
Veröffentlicht: (2024)
Emergent musical properties of a transformer under contrastive self-supervised learning
von: Kong, Yuexuan, et al.
Veröffentlicht: (2025)
von: Kong, Yuexuan, et al.
Veröffentlicht: (2025)
Equivariance-based self-supervised learning for audio signal recovery from clipped measurements
von: Sechaud, Victor, et al.
Veröffentlicht: (2024)
von: Sechaud, Victor, et al.
Veröffentlicht: (2024)
VoxRAG: A Step Toward Transcription-Free RAG Systems in Spoken Question Answering
von: Rackauckas, Zackary, et al.
Veröffentlicht: (2025)
von: Rackauckas, Zackary, et al.
Veröffentlicht: (2025)
Exploring Diverse Sounds: Identifying Outliers in a Music Corpus
von: Cai, Le, et al.
Veröffentlicht: (2024)
von: Cai, Le, et al.
Veröffentlicht: (2024)
Music Discovery Dialogue Generation Using Human Intent Analysis and Large Language Models
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
Track Role Prediction of Single-Instrumental Sequences
von: Han, Changheon, et al.
Veröffentlicht: (2024)
von: Han, Changheon, et al.
Veröffentlicht: (2024)
Personalized Dynamic Music Emotion Recognition with Dual-Scale Attention-Based Meta-Learning
von: Zhang, Dengming, et al.
Veröffentlicht: (2024)
von: Zhang, Dengming, et al.
Veröffentlicht: (2024)
LARP: Language Audio Relational Pre-training for Cold-Start Playlist Continuation
von: Salganik, Rebecca, et al.
Veröffentlicht: (2024)
von: Salganik, Rebecca, et al.
Veröffentlicht: (2024)
Expressivity-aware Music Performance Retrieval using Mid-level Perceptual Features and Emotion Word Embeddings
von: Chowdhury, Shreyan, et al.
Veröffentlicht: (2024)
von: Chowdhury, Shreyan, et al.
Veröffentlicht: (2024)
DiffATR: Diffusion-based Generative Modeling for Audio-Text Retrieval
von: Xin, Yifei, et al.
Veröffentlicht: (2024)
von: Xin, Yifei, et al.
Veröffentlicht: (2024)
Exploring GPT's Ability as a Judge in Music Understanding
von: Fang, Kun, et al.
Veröffentlicht: (2025)
von: Fang, Kun, et al.
Veröffentlicht: (2025)
Language-based Audio Retrieval with Co-Attention Networks
von: Sun, Haoran, et al.
Veröffentlicht: (2024)
von: Sun, Haoran, et al.
Veröffentlicht: (2024)
Multi-Sample Dynamic Time Warping for Few-Shot Keyword Spotting
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2024)
von: Wilkinghoff, Kevin, et al.
Veröffentlicht: (2024)
Multiscale Matching Driven by Cross-Modal Similarity Consistency for Audio-Text Retrieval
von: Wang, Qian, et al.
Veröffentlicht: (2024)
von: Wang, Qian, et al.
Veröffentlicht: (2024)
Do Captioning Metrics Reflect Music Semantic Alignment?
von: Lee, Jinwoo, et al.
Veröffentlicht: (2024)
von: Lee, Jinwoo, et al.
Veröffentlicht: (2024)
Bridging the Gap Between Semantic and User Preference Spaces for Multi-modal Music Representation Learning
von: Pan, Xiaofeng, et al.
Veröffentlicht: (2025)
von: Pan, Xiaofeng, et al.
Veröffentlicht: (2025)
FusID: Modality-Fused Semantic IDs for Generative Music Recommendation
von: Kim, Haven, et al.
Veröffentlicht: (2026)
von: Kim, Haven, et al.
Veröffentlicht: (2026)
Evaluating Interval-based Tokenization for Pitch Representation in Symbolic Music Analysis
von: Le, Dinh-Viet-Toan, et al.
Veröffentlicht: (2025)
von: Le, Dinh-Viet-Toan, et al.
Veröffentlicht: (2025)
TALKPLAY: Multimodal Music Recommendation with Large Language Models
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
Evaluating High-Resolution Piano Sustain Pedal Depth Estimation with Musically Informed Metrics
von: Zhang, Hanwen, et al.
Veröffentlicht: (2025)
von: Zhang, Hanwen, et al.
Veröffentlicht: (2025)
Engraving Oriented Joint Estimation of Pitch Spelling and Local and Global Keys
von: Bouquillard, Augustin, et al.
Veröffentlicht: (2024)
von: Bouquillard, Augustin, et al.
Veröffentlicht: (2024)
Towards Computational Analysis of Pansori Singing
von: Park, Sangheon, et al.
Veröffentlicht: (2024)
von: Park, Sangheon, et al.
Veröffentlicht: (2024)
Multi-label Cross-lingual automatic music genre classification from lyrics with Sentence BERT
von: Tavares, Tiago Fernandes, et al.
Veröffentlicht: (2025)
von: Tavares, Tiago Fernandes, et al.
Veröffentlicht: (2025)
TalkPlay-Tools: Conversational Music Recommendation with LLM Tool Calling
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
Diff4Steer: Steerable Diffusion Prior for Generative Music Retrieval with Semantic Guidance
von: Bao, Xuchan, et al.
Veröffentlicht: (2024)
von: Bao, Xuchan, et al.
Veröffentlicht: (2024)
Enriching Music Descriptions with a Finetuned-LLM and Metadata for Text-to-Music Retrieval
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
JEPOO: Highly Accurate Joint Estimation of Pitch, Onset and Offset for Music Information Retrieval
von: Wei, Haojie, et al.
Veröffentlicht: (2023)
von: Wei, Haojie, et al.
Veröffentlicht: (2023)
Streaming Piano Transcription Based on Consistent Onset and Offset Decoding with Sustain Pedal Detection
von: Wei, Weixing, et al.
Veröffentlicht: (2025)
von: Wei, Weixing, et al.
Veröffentlicht: (2025)
Evaluation of pretrained language models on music understanding
von: Vasilakis, Yannis, et al.
Veröffentlicht: (2024)
von: Vasilakis, Yannis, et al.
Veröffentlicht: (2024)
Deconstructing Jazz Piano Style Using Machine Learning
von: Cheston, Huw, et al.
Veröffentlicht: (2025)
von: Cheston, Huw, et al.
Veröffentlicht: (2025)
Audio Dialogues: Dialogues dataset for audio and music understanding
von: Goel, Arushi, et al.
Veröffentlicht: (2024)
von: Goel, Arushi, et al.
Veröffentlicht: (2024)
The Role of Large Language Models in Musicology: Are We Ready to Trust the Machines?
von: Ramoneda, Pedro, et al.
Veröffentlicht: (2024)
von: Ramoneda, Pedro, et al.
Veröffentlicht: (2024)
Pretrained Conformers for Audio Fingerprinting and Retrieval
von: Altwlkany, Kemal, et al.
Veröffentlicht: (2025)
von: Altwlkany, Kemal, et al.
Veröffentlicht: (2025)
Advancing the Foundation Model for Music Understanding
von: Jiang, Yi, et al.
Veröffentlicht: (2025)
von: Jiang, Yi, et al.
Veröffentlicht: (2025)
Analyzing and reducing the synthetic-to-real transfer gap in Music Information Retrieval: the task of automatic drum transcription
von: Zehren, Mickaël, et al.
Veröffentlicht: (2024)
von: Zehren, Mickaël, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A SOUND APPROACH: Using Large Language Models to generate audio descriptions for egocentric text-audio retrieval
von: Oncescu, Andreea-Maria, et al.
Veröffentlicht: (2024) -
CrossMuSim: A Cross-Modal Framework for Music Similarity Retrieval with LLM-Powered Text Description Sourcing and Mining
von: Tsoi, Tristan, et al.
Veröffentlicht: (2025) -
The language of sound search: Examining User Queries in Audio Search Engines
von: Weck, Benno, et al.
Veröffentlicht: (2024) -
Automatic Estimation of Singing Voice Musical Dynamics
von: Narang, Jyoti, et al.
Veröffentlicht: (2024) -
Towards Explainable and Interpretable Musical Difficulty Estimation: A Parameter-efficient Approach
von: Ramoneda, Pedro, et al.
Veröffentlicht: (2024)