LabelBuddy: An Open Source Music and Audio Language Annotation Tagging Tool Using AI Assistance
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Prokopiou, Ioannis, Sina, Ioannis, Kounelis, Agisilaos, Vikatos, Pantelis, Stafylakis, Themos |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Latent Space Disentanglement via Activation Steering for Interpretable Attribute Control in Symbolic Music Generation
von: Prokopiou, Ioannis, et al.
Veröffentlicht: (2026)
von: Prokopiou, Ioannis, et al.
Veröffentlicht: (2026)
Linear Complexity Self-Supervised Learning for Music Understanding with Random Quantizer
von: Vavaroutsos, Petros, et al.
Veröffentlicht: (2026)
von: Vavaroutsos, Petros, et al.
Veröffentlicht: (2026)
Synthetic Speech Source Tracing using Metric Learning
von: Koutsianos, Dimitrios, et al.
Veröffentlicht: (2025)
von: Koutsianos, Dimitrios, et al.
Veröffentlicht: (2025)
Scalable Music Cover Retrieval Using Lyrics-Aligned Audio Embeddings
von: Affolter, Joanne, et al.
Veröffentlicht: (2026)
von: Affolter, Joanne, et al.
Veröffentlicht: (2026)
CCMusic: An Open and Diverse Database for Chinese Music Information Retrieval Research
von: Zhou, Monan, et al.
Veröffentlicht: (2025)
von: Zhou, Monan, et al.
Veröffentlicht: (2025)
Music Auto-Tagging with Robust Music Representation Learned via Domain Adversarial Training
von: Joung, Haesun, et al.
Veröffentlicht: (2024)
von: Joung, Haesun, et al.
Veröffentlicht: (2024)
Towards Effective Negation Modeling in Joint Audio-Text Models for Music
von: Vasilakis, Yannis, et al.
Veröffentlicht: (2026)
von: Vasilakis, Yannis, et al.
Veröffentlicht: (2026)
TalkPlay-Tools: Conversational Music Recommendation with LLM Tool Calling
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
DiaPer: End-to-End Neural Diarization with Perceiver-Based Attractors
von: Landini, Federico, et al.
Veröffentlicht: (2023)
von: Landini, Federico, et al.
Veröffentlicht: (2023)
Acoustic Overspecification in Electronic Dance Music Taxonomy
von: Xu, Weilun, et al.
Veröffentlicht: (2025)
von: Xu, Weilun, et al.
Veröffentlicht: (2025)
Nested Music Transformer: Sequentially Decoding Compound Tokens in Symbolic Music and Audio Generation
von: Yoo, HaeJun, et al.
Veröffentlicht: (2024)
von: Yoo, HaeJun, et al.
Veröffentlicht: (2024)
Music Discovery Dialogue Generation Using Human Intent Analysis and Large Language Models
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
Understanding Human Perception of Music Plagiarism Through a Computational Approach
von: Hwang, Daeun, et al.
Veröffentlicht: (2026)
von: Hwang, Daeun, et al.
Veröffentlicht: (2026)
Music4All A+A: A Multimodal Dataset for Music Information Retrieval Tasks
von: Geiger, Jonas, et al.
Veröffentlicht: (2025)
von: Geiger, Jonas, et al.
Veröffentlicht: (2025)
Advancing Multi-Instrument Music Transcription: Results from the 2025 AMT Challenge
von: Chaturvedi, Ojas, et al.
Veröffentlicht: (2026)
von: Chaturvedi, Ojas, et al.
Veröffentlicht: (2026)
Latent Diffusion Bridges for Unsupervised Musical Audio Timbre Transfer
von: Mancusi, Michele, et al.
Veröffentlicht: (2024)
von: Mancusi, Michele, et al.
Veröffentlicht: (2024)
A Novel Audio Representation for Music Genre Identification in MIR
von: Kamuni, Navin, et al.
Veröffentlicht: (2024)
von: Kamuni, Navin, et al.
Veröffentlicht: (2024)
A Music Information Retrieval Approach to Classify Sub-Genres in Role Playing Games
von: Hwang, Daeun, et al.
Veröffentlicht: (2026)
von: Hwang, Daeun, et al.
Veröffentlicht: (2026)
AudioBoost: Increasing Audiobook Retrievability in Spotify Search with Synthetic Query Generation
von: Palumbo, Enrico, et al.
Veröffentlicht: (2025)
von: Palumbo, Enrico, et al.
Veröffentlicht: (2025)
Cosmodoit: A Python Package for Adaptive, Efficient Pipelining of Feature Extraction from Performed Music
von: Guichaoua, Corentin, et al.
Veröffentlicht: (2026)
von: Guichaoua, Corentin, et al.
Veröffentlicht: (2026)
Reddit2Deezer: A Scalable Dataset for Real-World Grounded Conversational Music Recommendation
von: Kim, Haven, et al.
Veröffentlicht: (2026)
von: Kim, Haven, et al.
Veröffentlicht: (2026)
Music Information Retrieval on Representative Mexican Folk Vocal Melodies Through MIDI Feature Extraction
von: Reyes, Mario Alberto Vallejo
Veröffentlicht: (2025)
von: Reyes, Mario Alberto Vallejo
Veröffentlicht: (2025)
TALKPLAY: Multimodal Music Recommendation with Large Language Models
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
von: Doh, Seungheon, et al.
Veröffentlicht: (2025)
Language-based Audio Retrieval with Co-Attention Networks
von: Sun, Haoran, et al.
Veröffentlicht: (2024)
von: Sun, Haoran, et al.
Veröffentlicht: (2024)
PianoBind: A Multimodal Joint Embedding Model for Pop-piano Music
von: Bang, Hayeon, et al.
Veröffentlicht: (2025)
von: Bang, Hayeon, et al.
Veröffentlicht: (2025)
Jamendo-MT-QA: A Benchmark for Multi-Track Comparative Music Question Answering
von: Koh, Junyoung, et al.
Veröffentlicht: (2026)
von: Koh, Junyoung, et al.
Veröffentlicht: (2026)
Audio-to-Image Bird Species Retrieval without Audio-Image Pairs via Text Distillation
von: Moummad, Ilyass, et al.
Veröffentlicht: (2026)
von: Moummad, Ilyass, et al.
Veröffentlicht: (2026)
Separate This, and All of these Things Around It: Music Source Separation via Hyperellipsoidal Queries
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2025)
von: Watcharasupat, Karn N., et al.
Veröffentlicht: (2025)
MERGE -- A Bimodal Audio-Lyrics Dataset for Static Music Emotion Recognition
von: Louro, Pedro Lima, et al.
Veröffentlicht: (2024)
von: Louro, Pedro Lima, et al.
Veröffentlicht: (2024)
LARP: Language Audio Relational Pre-training for Cold-Start Playlist Continuation
von: Salganik, Rebecca, et al.
Veröffentlicht: (2024)
von: Salganik, Rebecca, et al.
Veröffentlicht: (2024)
On the Effect of Data-Augmentation on Local Embedding Properties in the Contrastive Learning of Music Audio Representations
von: McCallum, Matthew C., et al.
Veröffentlicht: (2024)
von: McCallum, Matthew C., et al.
Veröffentlicht: (2024)
Affective Music Recommendation: A Rollout-Based World Model for Offline Preference Optimization
von: Chan, Audrey, et al.
Veröffentlicht: (2026)
von: Chan, Audrey, et al.
Veröffentlicht: (2026)
Automatic Estimation of Singing Voice Musical Dynamics
von: Narang, Jyoti, et al.
Veröffentlicht: (2024)
von: Narang, Jyoti, et al.
Veröffentlicht: (2024)
Similar but Faster: Manipulation of Tempo in Music Audio Embeddings for Tempo Prediction and Search
von: McCallum, Matthew C., et al.
Veröffentlicht: (2024)
von: McCallum, Matthew C., et al.
Veröffentlicht: (2024)
EMO100DB: An Open Dataset of Improvised Songs with Emotion Data
von: Hwang, Daeun, et al.
Veröffentlicht: (2025)
von: Hwang, Daeun, et al.
Veröffentlicht: (2025)
Data-Driven Analysis of Text-Conditioned AI-Generated Music: A Case Study with Suno and Udio
von: Casini, Luca, et al.
Veröffentlicht: (2025)
von: Casini, Luca, et al.
Veröffentlicht: (2025)
Enriching Music Descriptions with a Finetuned-LLM and Metadata for Text-to-Music Retrieval
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
von: Doh, SeungHeon, et al.
Veröffentlicht: (2024)
Exploring GPT's Ability as a Judge in Music Understanding
von: Fang, Kun, et al.
Veröffentlicht: (2025)
von: Fang, Kun, et al.
Veröffentlicht: (2025)
Do Captioning Metrics Reflect Music Semantic Alignment?
von: Lee, Jinwoo, et al.
Veröffentlicht: (2024)
von: Lee, Jinwoo, et al.
Veröffentlicht: (2024)
From Generation to Attribution: Music AI Agent Architectures for the Post-Streaming Era
von: Kim, Wonil, et al.
Veröffentlicht: (2025)
von: Kim, Wonil, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Latent Space Disentanglement via Activation Steering for Interpretable Attribute Control in Symbolic Music Generation
von: Prokopiou, Ioannis, et al.
Veröffentlicht: (2026) -
Linear Complexity Self-Supervised Learning for Music Understanding with Random Quantizer
von: Vavaroutsos, Petros, et al.
Veröffentlicht: (2026) -
Synthetic Speech Source Tracing using Metric Learning
von: Koutsianos, Dimitrios, et al.
Veröffentlicht: (2025) -
Scalable Music Cover Retrieval Using Lyrics-Aligned Audio Embeddings
von: Affolter, Joanne, et al.
Veröffentlicht: (2026) -
CCMusic: An Open and Diverse Database for Chinese Music Information Retrieval Research
von: Zhou, Monan, et al.
Veröffentlicht: (2025)