Development of Large Annotated Music Datasets using HMM-based Forced Viterbi Alignment
Fuente:
arXiv
Salvato in:
| Autori principali: | Joysingh, S. Johanan, Vijayalakshmi, P., Nagarajan, T. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Feature Engineering Approach for Literary and Colloquial Tamil Speech Classification using 1D-CNN
di: Nanmalar, M., et al.
Pubblicazione: (2024)
di: Nanmalar, M., et al.
Pubblicazione: (2024)
Significance of Chirp MFCC as a Feature in Speech and Audio Applications
di: Joysingh, S. Johanan, et al.
Pubblicazione: (2024)
di: Joysingh, S. Johanan, et al.
Pubblicazione: (2024)
Quartered Spectral Envelope and 1D-CNN-based Classification of Normally Phonated and Whispered Speech
di: Joysingh, S. Johanan, et al.
Pubblicazione: (2024)
di: Joysingh, S. Johanan, et al.
Pubblicazione: (2024)
Quartered Chirp Spectral Envelope for Whispered vs Normal Speech Classification
di: Joysingh, S. Johanan, et al.
Pubblicazione: (2024)
di: Joysingh, S. Johanan, et al.
Pubblicazione: (2024)
Chirp Group Delay based Onset Detection in Instruments with Fast Attack
di: Joysingh, S. Johanan, et al.
Pubblicazione: (2024)
di: Joysingh, S. Johanan, et al.
Pubblicazione: (2024)
MaskCycleGAN-based Whisper to Normal Speech Conversion
di: Gupta, K. Rohith, et al.
Pubblicazione: (2024)
di: Gupta, K. Rohith, et al.
Pubblicazione: (2024)
SynthSOD: Developing an Heterogeneous Dataset for Orchestra Music Source Separation
di: Garcia-Martinez, Jaime, et al.
Pubblicazione: (2024)
di: Garcia-Martinez, Jaime, et al.
Pubblicazione: (2024)
Literary and Colloquial Tamil Dialect Identification
di: Nanmalar, M., et al.
Pubblicazione: (2024)
di: Nanmalar, M., et al.
Pubblicazione: (2024)
Online Symbolic Music Alignment with Offline Reinforcement Learning
di: Peter, Silvan David
Pubblicazione: (2023)
di: Peter, Silvan David
Pubblicazione: (2023)
Leveraging LLM Embeddings for Cross Dataset Label Alignment and Zero Shot Music Emotion Prediction
di: Liu, Renhang, et al.
Pubblicazione: (2024)
di: Liu, Renhang, et al.
Pubblicazione: (2024)
The Spheres Dataset: Multitrack Orchestral Recordings for Music Source Separation and Information Retrieval
di: Garcia-Martinez, Jaime, et al.
Pubblicazione: (2025)
di: Garcia-Martinez, Jaime, et al.
Pubblicazione: (2025)
Analysis-Driven Procedural Generation of an Engine Sound Dataset with Embedded Control Annotations
di: Doerfler, Robin, et al.
Pubblicazione: (2026)
di: Doerfler, Robin, et al.
Pubblicazione: (2026)
Literary and Colloquial Dialect Identification for Tamil using Acoustic Features
di: Nanmalar, M., et al.
Pubblicazione: (2024)
di: Nanmalar, M., et al.
Pubblicazione: (2024)
Discovering and Steering Interpretable Concepts in Large Generative Music Models
di: Singh, Nikhil, et al.
Pubblicazione: (2025)
di: Singh, Nikhil, et al.
Pubblicazione: (2025)
Automatic Identification of Samples in Hip-Hop Music via Multi-Loss Training and an Artificial Dataset
di: Cheston, Huw, et al.
Pubblicazione: (2025)
di: Cheston, Huw, et al.
Pubblicazione: (2025)
JAZZVAR: A Dataset of Variations found within Solo Piano Performances of Jazz Standards for Music Overpainting
di: Row, Eleanor, et al.
Pubblicazione: (2023)
di: Row, Eleanor, et al.
Pubblicazione: (2023)
Audio Processing using Pattern Recognition for Music Genre Classification
di: Chatterjee, Sivangi, et al.
Pubblicazione: (2024)
di: Chatterjee, Sivangi, et al.
Pubblicazione: (2024)
MusicRL: Aligning Music Generation to Human Preferences
di: Cideron, Geoffrey, et al.
Pubblicazione: (2024)
di: Cideron, Geoffrey, et al.
Pubblicazione: (2024)
Subtractive Training for Music Stem Insertion using Latent Diffusion Models
di: Villa-Renteria, Ivan, et al.
Pubblicazione: (2024)
di: Villa-Renteria, Ivan, et al.
Pubblicazione: (2024)
Revisiting Meter Tracking in Carnatic Music using Deep Learning Approaches
di: Prabhu, Satyajeet
Pubblicazione: (2025)
di: Prabhu, Satyajeet
Pubblicazione: (2025)
The Name-Free Gap: Policy-Aware Stylistic Control in Music Generation
di: Nagarajan, Ashwin, et al.
Pubblicazione: (2025)
di: Nagarajan, Ashwin, et al.
Pubblicazione: (2025)
SYMPLEX: Controllable Symbolic Music Generation using Simplex Diffusion with Vocabulary Priors
di: Jonason, Nicolas, et al.
Pubblicazione: (2024)
di: Jonason, Nicolas, et al.
Pubblicazione: (2024)
Multimodal Dataset Normalization and Perceptual Validation for Music-Taste Correspondences
di: Spanio, Matteo, et al.
Pubblicazione: (2026)
di: Spanio, Matteo, et al.
Pubblicazione: (2026)
Anticipatory Music Transformer
di: Thickstun, John, et al.
Pubblicazione: (2023)
di: Thickstun, John, et al.
Pubblicazione: (2023)
Integrating Text-to-Music Models with Language Models: Composing Long Structured Music Pieces
di: Atassi, Lilac
Pubblicazione: (2024)
di: Atassi, Lilac
Pubblicazione: (2024)
ProGress: Structured Music Generation via Graph Diffusion and Hierarchical Music Analysis
di: Ni-Hahn, Stephen, et al.
Pubblicazione: (2025)
di: Ni-Hahn, Stephen, et al.
Pubblicazione: (2025)
Score-informed Music Source Separation: Improving Synthetic-to-real Generalization in Classical Music
di: Tunturi, Eetu, et al.
Pubblicazione: (2025)
di: Tunturi, Eetu, et al.
Pubblicazione: (2025)
Recognizing Ornaments in Vocal Indian Art Music with Active Annotation
di: Kumar, Sumit, et al.
Pubblicazione: (2025)
di: Kumar, Sumit, et al.
Pubblicazione: (2025)
Watermarking Training Data of Music Generation Models
di: Epple, Pascal, et al.
Pubblicazione: (2024)
di: Epple, Pascal, et al.
Pubblicazione: (2024)
Multi-Source Music Generation with Latent Diffusion
di: Xu, Zhongweiyang, et al.
Pubblicazione: (2024)
di: Xu, Zhongweiyang, et al.
Pubblicazione: (2024)
Benchmarking Representations for Speech, Music, and Acoustic Events
di: La Quatra, Moreno, et al.
Pubblicazione: (2024)
di: La Quatra, Moreno, et al.
Pubblicazione: (2024)
Music Genre Classification: Training an AI model
di: Mogonediwa, Keoikantse
Pubblicazione: (2024)
di: Mogonediwa, Keoikantse
Pubblicazione: (2024)
Evaluating Disentangled Representations for Controllable Music Generation
di: Ibáñez-Martínez, Laura, et al.
Pubblicazione: (2026)
di: Ibáñez-Martínez, Laura, et al.
Pubblicazione: (2026)
Tune It Up: Music Genre Transfer and Prediction
di: Samet, Fidan, et al.
Pubblicazione: (2025)
di: Samet, Fidan, et al.
Pubblicazione: (2025)
Learning Music Audio Representations With Limited Data
di: Plachouras, Christos, et al.
Pubblicazione: (2025)
di: Plachouras, Christos, et al.
Pubblicazione: (2025)
Benchmarking Music Autotagging with MGPHot Expert Annotations vs. Generic Tag Datasets
di: Ramoneda, Pedro, et al.
Pubblicazione: (2025)
di: Ramoneda, Pedro, et al.
Pubblicazione: (2025)
CloserMusicDB: A Modern Multipurpose Dataset of High Quality Music
di: Piekarzewicz, Aleksandra, et al.
Pubblicazione: (2024)
di: Piekarzewicz, Aleksandra, et al.
Pubblicazione: (2024)
Generating Music with Structure Using Self-Similarity as Attention
di: Hager, Sophia, et al.
Pubblicazione: (2024)
di: Hager, Sophia, et al.
Pubblicazione: (2024)
Music Emotion Prediction Using Recurrent Neural Networks
di: Chang, Xinyu, et al.
Pubblicazione: (2024)
di: Chang, Xinyu, et al.
Pubblicazione: (2024)
Parameter-Efficient Transfer Learning for Music Foundation Models
di: Ding, Yiwei, et al.
Pubblicazione: (2024)
di: Ding, Yiwei, et al.
Pubblicazione: (2024)
Documenti analoghi
-
A Feature Engineering Approach for Literary and Colloquial Tamil Speech Classification using 1D-CNN
di: Nanmalar, M., et al.
Pubblicazione: (2024) -
Significance of Chirp MFCC as a Feature in Speech and Audio Applications
di: Joysingh, S. Johanan, et al.
Pubblicazione: (2024) -
Quartered Spectral Envelope and 1D-CNN-based Classification of Normally Phonated and Whispered Speech
di: Joysingh, S. Johanan, et al.
Pubblicazione: (2024) -
Quartered Chirp Spectral Envelope for Whispered vs Normal Speech Classification
di: Joysingh, S. Johanan, et al.
Pubblicazione: (2024) -
Chirp Group Delay based Onset Detection in Instruments with Fast Attack
di: Joysingh, S. Johanan, et al.
Pubblicazione: (2024)