Guardado en:
| Autores principales: | Torrisi, Antonella M. C., Nolasco, Inês, Sgadò, Paola, Versace, Elisabetta, Benetos, Emmanouil |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2601.12203 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Acoustic identification of individual animals with hierarchical contrastive learning
por: Nolasco, Ines, et al.
Publicado: (2024)
por: Nolasco, Ines, et al.
Publicado: (2024)
SAR-LM: Symbolic Audio Reasoning with Large Language Models
por: Taheri, Termeh, et al.
Publicado: (2025)
por: Taheri, Termeh, et al.
Publicado: (2025)
Towards Building an End-to-End Multilingual Automatic Lyrics Transcription Model
por: Huang, Jiawen, et al.
Publicado: (2024)
por: Huang, Jiawen, et al.
Publicado: (2024)
LC-Protonets: Multi-Label Few-Shot Learning for World Music Audio Tagging
por: Papaioannou, Charilaos, et al.
Publicado: (2024)
por: Papaioannou, Charilaos, et al.
Publicado: (2024)
Learning Music Audio Representations With Limited Data
por: Plachouras, Christos, et al.
Publicado: (2025)
por: Plachouras, Christos, et al.
Publicado: (2025)
RUMAA: Repeat-Aware Unified Music Audio Analysis for Score-Performance Alignment, Transcription, and Mistake Detection
por: Chang, Sungkyun, et al.
Publicado: (2025)
por: Chang, Sungkyun, et al.
Publicado: (2025)
Universal Music Representations? Evaluating Foundation Models on World Music Corpora
por: Papaioannou, Charilaos, et al.
Publicado: (2025)
por: Papaioannou, Charilaos, et al.
Publicado: (2025)
NSTR: Neural Spectral Transport Representation for Space-Varying Frequency Fields
por: Versace, Plein
Publicado: (2025)
por: Versace, Plein
Publicado: (2025)
Domain-Invariant Representation Learning of Bird Sounds
por: Moummad, Ilyass, et al.
Publicado: (2024)
por: Moummad, Ilyass, et al.
Publicado: (2024)
GraFPrint: A GNN-Based Approach for Audio Identification
por: Bhattacharjee, Aditya, et al.
Publicado: (2024)
por: Bhattacharjee, Aditya, et al.
Publicado: (2024)
SCRAPL: Scattering Transform with Random Paths for Machine Learning
por: Mitcheltree, Christopher, et al.
Publicado: (2026)
por: Mitcheltree, Christopher, et al.
Publicado: (2026)
YourMT3+: Multi-instrument Music Transcription with Enhanced Transformer Architectures and Cross-dataset Stem Augmentation
por: Chang, Sungkyun, et al.
Publicado: (2024)
por: Chang, Sungkyun, et al.
Publicado: (2024)
LHGNN: Local-Higher Order Graph Neural Networks For Audio Classification and Tagging
por: Singh, Shubhr, et al.
Publicado: (2025)
por: Singh, Shubhr, et al.
Publicado: (2025)
Scalable Evaluation for Audio Identification via Synthetic Latent Fingerprint Generation
por: Bhattacharjee, Aditya, et al.
Publicado: (2025)
por: Bhattacharjee, Aditya, et al.
Publicado: (2025)
Enhancing Lyrics Transcription on Music Mixtures with Consistency Loss
por: Huang, Jiawen, et al.
Publicado: (2025)
por: Huang, Jiawen, et al.
Publicado: (2025)
A Data-Driven Analysis of Robust Automatic Piano Transcription
por: Edwards, Drew, et al.
Publicado: (2024)
por: Edwards, Drew, et al.
Publicado: (2024)
Text2Score: Generating Sheet Music From Textual Prompts
por: Bhandari, Keshav, et al.
Publicado: (2026)
por: Bhandari, Keshav, et al.
Publicado: (2026)
Generalized Multi-Source Inference for Text Conditioned Music Diffusion Models
por: Postolache, Emilian, et al.
Publicado: (2024)
por: Postolache, Emilian, et al.
Publicado: (2024)
Classification of Spontaneous and Scripted Speech for Multilingual Audio
por: Elisha, Shahar, et al.
Publicado: (2024)
por: Elisha, Shahar, et al.
Publicado: (2024)
Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning
por: Tuncay, Ludovic, et al.
Publicado: (2025)
por: Tuncay, Ludovic, et al.
Publicado: (2025)
Rank-based loss for learning hierarchical representations
por: Nolasco, Ines, et al.
Publicado: (2021)
por: Nolasco, Ines, et al.
Publicado: (2021)
CMI-Bench: A Comprehensive Benchmark for Evaluating Music Instruction Following
por: Ma, Yinghao, et al.
Publicado: (2025)
por: Ma, Yinghao, et al.
Publicado: (2025)
Twenty-Five Years of MIR Research: Achievements, Practices, Evaluations, and Future Challenges
por: Peeters, Geoffroy, et al.
Publicado: (2025)
por: Peeters, Geoffroy, et al.
Publicado: (2025)
Refining music sample identification with a self-supervised graph neural network
por: Bhattacharjee, Aditya, et al.
Publicado: (2025)
por: Bhattacharjee, Aditya, et al.
Publicado: (2025)
DB3V: A Dialect Dominated Dataset of Bird Vocalisation for Cross-corpus Bird Species Recognition
por: Jing, Xin, et al.
Publicado: (2024)
por: Jing, Xin, et al.
Publicado: (2024)
ST-ITO: Controlling Audio Effects for Style Transfer with Inference-Time Optimization
por: Steinmetz, Christian J., et al.
Publicado: (2024)
por: Steinmetz, Christian J., et al.
Publicado: (2024)
MuChoMusic: Evaluating Music Understanding in Multimodal Audio-Language Models
por: Weck, Benno, et al.
Publicado: (2024)
por: Weck, Benno, et al.
Publicado: (2024)
A Soft Robotic Interface for Chick-Robot Affective Interactions
por: Chen, Jue, et al.
Publicado: (2026)
por: Chen, Jue, et al.
Publicado: (2026)
Learning to detect an animal sound from five examples
por: Nolasco, Inês, et al.
Publicado: (2023)
por: Nolasco, Inês, et al.
Publicado: (2023)
Quality Audio Prototyping: a prototype system for unified sound retrieval and procedural generation
por: Garcia, Nelly, et al.
Publicado: (2026)
por: Garcia, Nelly, et al.
Publicado: (2026)
Non-Verbal Vocalisations and their Challenges: Emotion, Privacy, Sparseness, and Real Life
por: Batliner, Anton, et al.
Publicado: (2025)
por: Batliner, Anton, et al.
Publicado: (2025)
MusiLingo: Bridging Music and Text with Pre-trained Language Models for Music Captioning and Query Response
por: Deng, Zihao, et al.
Publicado: (2023)
por: Deng, Zihao, et al.
Publicado: (2023)
WeaveMuse: An Open Agentic System for Multimodal Music Understanding and Generation
por: Karystinaios, Emmanouil
Publicado: (2025)
por: Karystinaios, Emmanouil
Publicado: (2025)
Can LLMs "Reason" in Music? An Evaluation of LLMs' Capability of Music Understanding and Generation
por: Zhou, Ziya, et al.
Publicado: (2024)
por: Zhou, Ziya, et al.
Publicado: (2024)
SMUG-Explain: A Framework for Symbolic Music Graph Explanations
por: Karystinaios, Emmanouil, et al.
Publicado: (2024)
por: Karystinaios, Emmanouil, et al.
Publicado: (2024)
Language Models for Music Medicine Generation
por: Nikolakakis, Emmanouil, et al.
Publicado: (2024)
por: Nikolakakis, Emmanouil, et al.
Publicado: (2024)
LyricWhiz: Robust Multilingual Zero-shot Lyrics Transcription by Whispering to ChatGPT
por: Zhuo, Le, et al.
Publicado: (2023)
por: Zhuo, Le, et al.
Publicado: (2023)
MUSE-Explainer: Counterfactual Explanations for Symbolic Music Graph Classification Models
por: Hilaire, Baptiste, et al.
Publicado: (2025)
por: Hilaire, Baptiste, et al.
Publicado: (2025)
Mind the Domain Gap: a Systematic Analysis on Bioacoustic Sound Event Detection
por: Liang, Jinhua, et al.
Publicado: (2024)
por: Liang, Jinhua, et al.
Publicado: (2024)
CMI-RewardBench: Evaluating Music Reward Models with Compositional Multimodal Instruction
por: Ma, Yinghao, et al.
Publicado: (2026)
por: Ma, Yinghao, et al.
Publicado: (2026)
Ejemplares similares
-
Acoustic identification of individual animals with hierarchical contrastive learning
por: Nolasco, Ines, et al.
Publicado: (2024) -
SAR-LM: Symbolic Audio Reasoning with Large Language Models
por: Taheri, Termeh, et al.
Publicado: (2025) -
Towards Building an End-to-End Multilingual Automatic Lyrics Transcription Model
por: Huang, Jiawen, et al.
Publicado: (2024) -
LC-Protonets: Multi-Label Few-Shot Learning for World Music Audio Tagging
por: Papaioannou, Charilaos, et al.
Publicado: (2024) -
Learning Music Audio Representations With Limited Data
por: Plachouras, Christos, et al.
Publicado: (2025)