A New Dataset, Notation Software, and Representation for Computational Schenkerian Analysis
Fuente:
arXiv
Guardado en:
| Autores principales: | Ni-Hahn, Stephen, Xu, Weihan, Yin, Jerry, Zhu, Rico, Mak, Simon, Jiang, Yue, Rudin, Cynthia |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AutoSchA: Automatic Hierarchical Music Representations via Multi-Relational Node Isolation
por: Ni-Hahn, Stephen, et al.
Publicado: (2025)
por: Ni-Hahn, Stephen, et al.
Publicado: (2025)
ProGress: Structured Music Generation via Graph Diffusion and Hierarchical Music Analysis
por: Ni-Hahn, Stephen, et al.
Publicado: (2025)
por: Ni-Hahn, Stephen, et al.
Publicado: (2025)
YNote: A Novel Music Notation for Fine-Tuning LLMs in Music Generation
por: Lu, Shao-Chien, et al.
Publicado: (2025)
por: Lu, Shao-Chien, et al.
Publicado: (2025)
ONOTE: Benchmarking Omnimodal Notation Processing for Expert-level Music Intelligence
por: Ma, Menghe, et al.
Publicado: (2026)
por: Ma, Menghe, et al.
Publicado: (2026)
Via Score to Performance: Efficient Human-Controllable Long Song Generation with Bar-Level Symbolic Notation
por: Wang, Tongxi, et al.
Publicado: (2025)
por: Wang, Tongxi, et al.
Publicado: (2025)
Bridging Biological Hearing and Neuromorphic Computing: End-to-End Time-Domain Audio Signal Processing with Reservoir Computing
por: Sebastian, Rinku, et al.
Publicado: (2026)
por: Sebastian, Rinku, et al.
Publicado: (2026)
EMelodyGen: Emotion-Conditioned Melody Generation in ABC Notation with the Musical Feature Template
por: Zhou, Monan, et al.
Publicado: (2023)
por: Zhou, Monan, et al.
Publicado: (2023)
Aria-MIDI: A Dataset of Piano MIDI Files for Symbolic Music Modeling
por: Bradshaw, Louis, et al.
Publicado: (2025)
por: Bradshaw, Louis, et al.
Publicado: (2025)
Infant Cry Detection Using Causal Temporal Representation
por: Fu, Minghao, et al.
Publicado: (2025)
por: Fu, Minghao, et al.
Publicado: (2025)
What Do Language Models Hear? Probing for Auditory Representations in Language Models
por: Ngo, Jerry, et al.
Publicado: (2024)
por: Ngo, Jerry, et al.
Publicado: (2024)
Evaluation of Deep Audio Representations for Hearables
por: Gröger, Fabian, et al.
Publicado: (2025)
por: Gröger, Fabian, et al.
Publicado: (2025)
NOTA: Multimodal Music Notation Understanding for Visual Large Language Model
por: Tang, Mingni, et al.
Publicado: (2025)
por: Tang, Mingni, et al.
Publicado: (2025)
Quantize More, Lose Less: Autoregressive Generation from Residually Quantized Speech Representations
por: Han, Yichen, et al.
Publicado: (2025)
por: Han, Yichen, et al.
Publicado: (2025)
Tadabur: A Large-Scale Quran Audio Dataset
por: Alherran, Faisal
Publicado: (2026)
por: Alherran, Faisal
Publicado: (2026)
SCDF: A Speaker Characteristics DeepFake Speech Dataset for Bias Analysis
por: Staněk, Vojtěch, et al.
Publicado: (2025)
por: Staněk, Vojtěch, et al.
Publicado: (2025)
Cross-Domain Audio Deepfake Detection: Dataset and Analysis
por: Li, Yuang, et al.
Publicado: (2024)
por: Li, Yuang, et al.
Publicado: (2024)
Deepfake Audio Detection Using Self-supervised Fusion Representations
por: Zaman, Khalid, et al.
Publicado: (2026)
por: Zaman, Khalid, et al.
Publicado: (2026)
Perceptually Aligning Representations of Music via Noise-Augmented Autoencoders
por: Bjare, Mathias Rose, et al.
Publicado: (2025)
por: Bjare, Mathias Rose, et al.
Publicado: (2025)
Enabling Automatic Disordered Speech Recognition: An Impaired Speech Dataset in the Akan Language
por: Wiafe, Isaac, et al.
Publicado: (2026)
por: Wiafe, Isaac, et al.
Publicado: (2026)
HAIM: Human-AI Music Datasets for AI Music Production Tracking Benchmark
por: Go, Seonghyeon, et al.
Publicado: (2026)
por: Go, Seonghyeon, et al.
Publicado: (2026)
Masked Latent Prediction and Classification for Self-Supervised Audio Representation Learning
por: Quelennec, Aurian, et al.
Publicado: (2025)
por: Quelennec, Aurian, et al.
Publicado: (2025)
NSTR: Neural Spectral Transport Representation for Space-Varying Frequency Fields
por: Versace, Plein
Publicado: (2025)
por: Versace, Plein
Publicado: (2025)
Rethinking Leveraging Pre-Trained Multi-Layer Representations for Speaker Verification
por: Kim, Jin Sob, et al.
Publicado: (2025)
por: Kim, Jin Sob, et al.
Publicado: (2025)
Generating Symbolic Music from Natural Language Prompts using an LLM-Enhanced Dataset
por: Xu, Weihan, et al.
Publicado: (2024)
por: Xu, Weihan, et al.
Publicado: (2024)
JamendoMaxCaps: A Large Scale Music-caption Dataset with Imputed Metadata
por: Roy, Abhinaba, et al.
Publicado: (2025)
por: Roy, Abhinaba, et al.
Publicado: (2025)
Structure-Aware Piano Accompaniment via Style Planning and Dataset-Aligned Pattern Retrieval
por: Zang, Wanyu, et al.
Publicado: (2026)
por: Zang, Wanyu, et al.
Publicado: (2026)
Multi-Accent Mandarin Dry-Vocal Singing Dataset: Benchmark for Singing Accent Recognition
por: Wang, Zihao, et al.
Publicado: (2025)
por: Wang, Zihao, et al.
Publicado: (2025)
MATPAC++: Enhanced Masked Latent Prediction for Self-Supervised Audio Representation Learning
por: Quelennec, Aurian, et al.
Publicado: (2025)
por: Quelennec, Aurian, et al.
Publicado: (2025)
Prosodic Boundary-Aware Streaming Generation for LLM-Based TTS with Streaming Text Input
por: Liu, Changsong, et al.
Publicado: (2026)
por: Liu, Changsong, et al.
Publicado: (2026)
UniWhisper: Efficient Continual Multi-task Training for Robust Universal Audio Representation
por: Chen, Yuxuan, et al.
Publicado: (2026)
por: Chen, Yuxuan, et al.
Publicado: (2026)
Cross-Cultural Bias in Mel-Scale Representations: Evidence and Alternatives from Speech and Music
por: Chauhan, Shivam, et al.
Publicado: (2026)
por: Chauhan, Shivam, et al.
Publicado: (2026)
Towards Lightweight and Stable Zero-shot TTS with Self-distilled Representation Disentanglement
por: Chen, Qianniu, et al.
Publicado: (2025)
por: Chen, Qianniu, et al.
Publicado: (2025)
Speech-Forensics: Towards Comprehensive Synthetic Speech Dataset Establishment and Analysis
por: Ji, Zhoulin, et al.
Publicado: (2024)
por: Ji, Zhoulin, et al.
Publicado: (2024)
Improving Anomalous Sound Detection with Attribute-aware Representation from Domain-adaptive Pre-training
por: Fang, Xin, et al.
Publicado: (2025)
por: Fang, Xin, et al.
Publicado: (2025)
A Toolkit for Detecting Spurious Correlations in Speech Datasets
por: Gauder, Lara, et al.
Publicado: (2026)
por: Gauder, Lara, et al.
Publicado: (2026)
AudioMoG: Guiding Audio Generation with Mixture-of-Guidance
por: Wang, Junyou, et al.
Publicado: (2025)
por: Wang, Junyou, et al.
Publicado: (2025)
DAStatFormer: A Hybrid Multibranch Transformer with Statistical Feature Integration for DAS-Based Pattern Recognitions
por: Dione, Michel, et al.
Publicado: (2026)
por: Dione, Michel, et al.
Publicado: (2026)
YMIR: A new Benchmark Dataset and Model for Arabic Yemeni Music Genre Classification Using Convolutional Neural Networks
por: AL-Makhlafi, Moeen, et al.
Publicado: (2026)
por: AL-Makhlafi, Moeen, et al.
Publicado: (2026)
AnalysisGNN: Unified Music Analysis with Graph Neural Networks
por: Karystinaios, Emmanouil, et al.
Publicado: (2025)
por: Karystinaios, Emmanouil, et al.
Publicado: (2025)
Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models
por: Zhou, Yizhi, et al.
Publicado: (2025)
por: Zhou, Yizhi, et al.
Publicado: (2025)
Ejemplares similares
-
AutoSchA: Automatic Hierarchical Music Representations via Multi-Relational Node Isolation
por: Ni-Hahn, Stephen, et al.
Publicado: (2025) -
ProGress: Structured Music Generation via Graph Diffusion and Hierarchical Music Analysis
por: Ni-Hahn, Stephen, et al.
Publicado: (2025) -
YNote: A Novel Music Notation for Fine-Tuning LLMs in Music Generation
por: Lu, Shao-Chien, et al.
Publicado: (2025) -
ONOTE: Benchmarking Omnimodal Notation Processing for Expert-level Music Intelligence
por: Ma, Menghe, et al.
Publicado: (2026) -
Via Score to Performance: Efficient Human-Controllable Long Song Generation with Bar-Level Symbolic Notation
por: Wang, Tongxi, et al.
Publicado: (2025)