Are We There Yet? A Brief Survey of Music Emotion Prediction Datasets, Models and Outstanding Challenges
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kang, Jaeyong, Herremans, Dorien |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Unified Music Emotion Recognition across Dimensional and Categorical Models
von: Kang, Jaeyong, et al.
Veröffentlicht: (2025)
von: Kang, Jaeyong, et al.
Veröffentlicht: (2025)
Video2Music: Suitable Music Generation from Videos using an Affective Multimodal Transformer model
von: Kang, Jaeyong, et al.
Veröffentlicht: (2023)
von: Kang, Jaeyong, et al.
Veröffentlicht: (2023)
Aligning Generative Music AI with Human Preferences: Methods and Challenges
von: Herremans, Dorien, et al.
Veröffentlicht: (2025)
von: Herremans, Dorien, et al.
Veröffentlicht: (2025)
BandCondiNet: Parallel Transformers-based Conditional Popular Music Generation with Multi-View Features
von: Luo, Jing, et al.
Veröffentlicht: (2024)
von: Luo, Jing, et al.
Veröffentlicht: (2024)
MelodySim: Measuring Melody-aware Music Similarity for Plagiarism Detection
von: Lu, Tongyu, et al.
Veröffentlicht: (2025)
von: Lu, Tongyu, et al.
Veröffentlicht: (2025)
Leveraging LLM Embeddings for Cross Dataset Label Alignment and Zero Shot Music Emotion Prediction
von: Liu, Renhang, et al.
Veröffentlicht: (2024)
von: Liu, Renhang, et al.
Veröffentlicht: (2024)
Natural Language Processing Methods for Symbolic Music Generation and Information Retrieval: a Survey
von: Le, Dinh-Viet-Toan, et al.
Veröffentlicht: (2024)
von: Le, Dinh-Viet-Toan, et al.
Veröffentlicht: (2024)
A Domain-Knowledge-Inspired Music Embedding Space and a Novel Attention Mechanism for Symbolic Music Modeling
von: Guo, Z., et al.
Veröffentlicht: (2022)
von: Guo, Z., et al.
Veröffentlicht: (2022)
DART: Disentanglement of Accent and Speaker Representation in Multispeaker Text-to-Speech
von: Melechovsky, Jan, et al.
Veröffentlicht: (2024)
von: Melechovsky, Jan, et al.
Veröffentlicht: (2024)
ImprovNet -- Generating Controllable Musical Improvisations with Iterative Corruption Refinement
von: Bhandari, Keshav, et al.
Veröffentlicht: (2025)
von: Bhandari, Keshav, et al.
Veröffentlicht: (2025)
MIRFLEX: Music Information Retrieval Feature Library for Extraction
von: Chopra, Anuradha, et al.
Veröffentlicht: (2024)
von: Chopra, Anuradha, et al.
Veröffentlicht: (2024)
Text2midi: Generating Symbolic Music from Captions
von: Bhandari, Keshav, et al.
Veröffentlicht: (2024)
von: Bhandari, Keshav, et al.
Veröffentlicht: (2024)
Text2midi-InferAlign: Improving Symbolic Music Generation with Inference-Time Alignment
von: Roy, Abhinaba, et al.
Veröffentlicht: (2025)
von: Roy, Abhinaba, et al.
Veröffentlicht: (2025)
DiffRoll: Diffusion-based Generative Music Transcription with Unsupervised Pretraining Capability
von: Cheuk, Kin Wai, et al.
Veröffentlicht: (2022)
von: Cheuk, Kin Wai, et al.
Veröffentlicht: (2022)
The Music Maestro or The Musically Challenged, A Massive Music Evaluation Benchmark for Large Language Models
von: Li, Jiajia, et al.
Veröffentlicht: (2024)
von: Li, Jiajia, et al.
Veröffentlicht: (2024)
Are Inherently Interpretable Models More Robust? A Study In Music Emotion Recognition
von: Hoedt, Katharina, et al.
Veröffentlicht: (2025)
von: Hoedt, Katharina, et al.
Veröffentlicht: (2025)
Joint Learning of Emotions in Music and Generalized Sounds
von: Simonetta, Federico, et al.
Veröffentlicht: (2024)
von: Simonetta, Federico, et al.
Veröffentlicht: (2024)
Wearable Music2Emotion : Assessing Emotions Induced by AI-Generated Music through Portable EEG-fNIRS Fusion
von: Zhao, Sha, et al.
Veröffentlicht: (2025)
von: Zhao, Sha, et al.
Veröffentlicht: (2025)
AImoclips: A Benchmark for Evaluating Emotion Conveyance in Text-to-Music Generation
von: Go, Gyehun, et al.
Veröffentlicht: (2025)
von: Go, Gyehun, et al.
Veröffentlicht: (2025)
Language Model Mapping in Multimodal Music Learning: A Grand Challenge Proposal
von: Chin, Daniel, et al.
Veröffentlicht: (2025)
von: Chin, Daniel, et al.
Veröffentlicht: (2025)
MOSA: Music Motion with Semantic Annotation Dataset for Cross-Modal Music Processing
von: Huang, Yu-Fen, et al.
Veröffentlicht: (2024)
von: Huang, Yu-Fen, et al.
Veröffentlicht: (2024)
Semi-Supervised Self-Learning Enhanced Music Emotion Recognition
von: Sun, Yifu, et al.
Veröffentlicht: (2024)
von: Sun, Yifu, et al.
Veröffentlicht: (2024)
A Survey of Foundation Models for Music Understanding
von: Li, Wenjun, et al.
Veröffentlicht: (2024)
von: Li, Wenjun, et al.
Veröffentlicht: (2024)
Summary of The Inaugural Music Source Restoration Challenge
von: Zang, Yongyi, et al.
Veröffentlicht: (2026)
von: Zang, Yongyi, et al.
Veröffentlicht: (2026)
Emotion-driven Piano Music Generation via Two-stage Disentanglement and Functional Representation
von: Huang, Jingyue, et al.
Veröffentlicht: (2024)
von: Huang, Jingyue, et al.
Veröffentlicht: (2024)
MusER: Musical Element-Based Regularization for Generating Symbolic Music with Emotion
von: Ji, Shulei, et al.
Veröffentlicht: (2023)
von: Ji, Shulei, et al.
Veröffentlicht: (2023)
ABHINAYA -- A System for Speech Emotion Recognition In Naturalistic Conditions Challenge
von: Dutta, Soumya, et al.
Veröffentlicht: (2025)
von: Dutta, Soumya, et al.
Veröffentlicht: (2025)
Disentangling Reasoning in Large Audio-Language Models for Ambiguous Emotion Prediction
von: Yu, Xiaofeng, et al.
Veröffentlicht: (2026)
von: Yu, Xiaofeng, et al.
Veröffentlicht: (2026)
Music Consistency Models
von: Fei, Zhengcong, et al.
Veröffentlicht: (2024)
von: Fei, Zhengcong, et al.
Veröffentlicht: (2024)
SNIPER Training: Single-Shot Sparse Training for Text-to-Speech
von: Lam, Perry, et al.
Veröffentlicht: (2022)
von: Lam, Perry, et al.
Veröffentlicht: (2022)
MixAssist: An Audio-Language Dataset for Co-Creative AI Assistance in Music Mixing
von: Clemens, Michael, et al.
Veröffentlicht: (2025)
von: Clemens, Michael, et al.
Veröffentlicht: (2025)
Exploring Variational Auto-Encoder Architectures, Configurations, and Datasets for Generative Music Explainable AI
von: Bryan-Kinns, Nick, et al.
Veröffentlicht: (2023)
von: Bryan-Kinns, Nick, et al.
Veröffentlicht: (2023)
Zero-shot Musical Stem Retrieval with Joint-Embedding Predictive Architectures
von: Riou, Alain, et al.
Veröffentlicht: (2024)
von: Riou, Alain, et al.
Veröffentlicht: (2024)
Play Me Something Icy: Practical Challenges, Explainability and the Semantic Gap in Generative AI Music
von: Allison, Jesse, et al.
Veröffentlicht: (2024)
von: Allison, Jesse, et al.
Veröffentlicht: (2024)
NotaGen: Advancing Musicality in Symbolic Music Generation with Large Language Model Training Paradigms
von: Wang, Yashan, et al.
Veröffentlicht: (2025)
von: Wang, Yashan, et al.
Veröffentlicht: (2025)
MidiCaps: A large-scale MIDI dataset with text captions
von: Melechovsky, Jan, et al.
Veröffentlicht: (2024)
von: Melechovsky, Jan, et al.
Veröffentlicht: (2024)
WhisQ: Cross-Modal Representation Learning for Text-to-Music MOS Prediction
von: Emon, Jakaria Islam, et al.
Veröffentlicht: (2025)
von: Emon, Jakaria Islam, et al.
Veröffentlicht: (2025)
Unimodal Multi-Task Fusion for Emotional Mimicry Intensity Prediction
von: Hallmen, Tobias, et al.
Veröffentlicht: (2024)
von: Hallmen, Tobias, et al.
Veröffentlicht: (2024)
Exploring Musical Roots: Applying Audio Embeddings to Empower Influence Attribution for a Generative Music Model
von: Barnett, Julia, et al.
Veröffentlicht: (2024)
von: Barnett, Julia, et al.
Veröffentlicht: (2024)
The Interpretation Gap in Text-to-Music Generation Models
von: Zang, Yongyi, et al.
Veröffentlicht: (2024)
von: Zang, Yongyi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Towards Unified Music Emotion Recognition across Dimensional and Categorical Models
von: Kang, Jaeyong, et al.
Veröffentlicht: (2025) -
Video2Music: Suitable Music Generation from Videos using an Affective Multimodal Transformer model
von: Kang, Jaeyong, et al.
Veröffentlicht: (2023) -
Aligning Generative Music AI with Human Preferences: Methods and Challenges
von: Herremans, Dorien, et al.
Veröffentlicht: (2025) -
BandCondiNet: Parallel Transformers-based Conditional Popular Music Generation with Multi-View Features
von: Luo, Jing, et al.
Veröffentlicht: (2024) -
MelodySim: Measuring Melody-aware Music Similarity for Plagiarism Detection
von: Lu, Tongyu, et al.
Veröffentlicht: (2025)