Computational music analysis from first principles
Fuente:
arXiv
Guardado en:
| Autores principales: | Tymoczko, Dmitri, Newman, Mark |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
The first Cadenza challenges: using machine learning competitions to improve music for listeners with a hearing loss
por: Dabike, Gerardo Roa, et al.
Publicado: (2024)
por: Dabike, Gerardo Roa, et al.
Publicado: (2024)
Detecting music deepfakes is easy but actually hard
por: Afchar, Darius, et al.
Publicado: (2024)
por: Afchar, Darius, et al.
Publicado: (2024)
Long-form music generation with latent diffusion
por: Evans, Zach, et al.
Publicado: (2024)
por: Evans, Zach, et al.
Publicado: (2024)
StemGen: A music generation model that listens
por: Parker, Julian D., et al.
Publicado: (2023)
por: Parker, Julian D., et al.
Publicado: (2023)
Learning and composing of classical music using restricted Boltzmann machines
por: Kobayashi, Mutsumi, et al.
Publicado: (2025)
por: Kobayashi, Mutsumi, et al.
Publicado: (2025)
Local deployment of large-scale music AI models on commodity hardware
por: Zhou, Xun, et al.
Publicado: (2024)
por: Zhou, Xun, et al.
Publicado: (2024)
SLEEPING-DISCO 9M: A large-scale pre-training dataset for generative music modeling
por: Ahmed, Tawsif, et al.
Publicado: (2025)
por: Ahmed, Tawsif, et al.
Publicado: (2025)
Dynamic Gated Recurrent Neural Network for Compute-efficient Speech Enhancement
por: Cheng, Longbiao, et al.
Publicado: (2024)
por: Cheng, Longbiao, et al.
Publicado: (2024)
SOI: Scaling Down Computational Complexity by Estimating Partial States of the Model
por: Stefański, Grzegorz, et al.
Publicado: (2024)
por: Stefański, Grzegorz, et al.
Publicado: (2024)
A contrastive-learning approach for auditory attention detection
por: Bajestan, Seyed Ali Alavi, et al.
Publicado: (2024)
por: Bajestan, Seyed Ali Alavi, et al.
Publicado: (2024)
Modulating State Space Model with SlowFast Framework for Compute-Efficient Ultra Low-Latency Speech Enhancement
por: Cheng, Longbiao, et al.
Publicado: (2024)
por: Cheng, Longbiao, et al.
Publicado: (2024)
Investigating Confidence Estimation Measures for Speaker Diarization
por: Chowdhury, Anurag, et al.
Publicado: (2024)
por: Chowdhury, Anurag, et al.
Publicado: (2024)
Evaluation of Neural Surrogates for Physical Modelling Synthesis of Nonlinear Elastic Plates
por: Martin, Carlos De La Vega, et al.
Publicado: (2025)
por: Martin, Carlos De La Vega, et al.
Publicado: (2025)
Training Articulatory Inversion Models for Interspeaker Consistency
por: McGhee, Charles, et al.
Publicado: (2025)
por: McGhee, Charles, et al.
Publicado: (2025)
wav2pos: Sound Source Localization using Masked Autoencoders
por: Berg, Axel, et al.
Publicado: (2024)
por: Berg, Axel, et al.
Publicado: (2024)
Sentiment analysis in non-fixed length audios using a Fully Convolutional Neural Network
por: García-Ordás, María Teresa, et al.
Publicado: (2024)
por: García-Ordás, María Teresa, et al.
Publicado: (2024)
Enhancing the analysis of murine neonatal ultrasonic vocalizations: Development, evaluation, and application of different mathematical models
por: Herdt, Rudolf, et al.
Publicado: (2024)
por: Herdt, Rudolf, et al.
Publicado: (2024)
Towards objective and interpretable speech disorder assessment: a comparative analysis of CNN and transformer-based models
por: Maisonneuve, Malo, et al.
Publicado: (2024)
por: Maisonneuve, Malo, et al.
Publicado: (2024)
Sound Tagging in Infant-centric Home Soundscapes
por: Khan, Mohammad Nur Hossain, et al.
Publicado: (2024)
por: Khan, Mohammad Nur Hossain, et al.
Publicado: (2024)
Transcription-Free Fine-Tuning of Speech Separation Models for Noisy and Reverberant Multi-Speaker Automatic Speech Recognition
por: Ravenscroft, William, et al.
Publicado: (2024)
por: Ravenscroft, William, et al.
Publicado: (2024)
Symbotunes: unified hub for symbolic music generative models
por: Skierś, Paweł, et al.
Publicado: (2024)
por: Skierś, Paweł, et al.
Publicado: (2024)
Linear RNNs for autoregressive generation of long music samples
por: Szewczyk, Konrad, et al.
Publicado: (2025)
por: Szewczyk, Konrad, et al.
Publicado: (2025)
Supervised contrastive learning from weakly-labeled audio segments for musical version matching
por: Serrà, Joan, et al.
Publicado: (2025)
por: Serrà, Joan, et al.
Publicado: (2025)
Avoiding an AI-imposed Taylor's Version of all music history
por: Collins, Nick, et al.
Publicado: (2024)
por: Collins, Nick, et al.
Publicado: (2024)
Multi-label Cross-lingual automatic music genre classification from lyrics with Sentence BERT
por: Tavares, Tiago Fernandes, et al.
Publicado: (2025)
por: Tavares, Tiago Fernandes, et al.
Publicado: (2025)
Description on IEEE ICME 2024 Grand Challenge: Semi-supervised Acoustic Scene Classification under Domain Shift
por: Bai, Jisheng, et al.
Publicado: (2024)
por: Bai, Jisheng, et al.
Publicado: (2024)
Emergent musical properties of a transformer under contrastive self-supervised learning
por: Kong, Yuexuan, et al.
Publicado: (2025)
por: Kong, Yuexuan, et al.
Publicado: (2025)
The 2025 PNPL Competition: Speech Detection and Phoneme Classification in the LibriBrain Dataset
por: Landau, Gilad, et al.
Publicado: (2025)
por: Landau, Gilad, et al.
Publicado: (2025)
Contrastive Learning from Synthetic Audio Doppelgängers
por: Cherep, Manuel, et al.
Publicado: (2024)
por: Cherep, Manuel, et al.
Publicado: (2024)
Tessellated Linear Model for Age Prediction from Voice
por: Alharthi, Dareen, et al.
Publicado: (2025)
por: Alharthi, Dareen, et al.
Publicado: (2025)
Switchboard-Affect: Emotion Perception Labels from Conversational Speech
por: Romana, Amrit, et al.
Publicado: (2025)
por: Romana, Amrit, et al.
Publicado: (2025)
CAK: Emergent Audio Effects from Minimal Deep Learning
por: Rockman, Austin
Publicado: (2025)
por: Rockman, Austin
Publicado: (2025)
Foundation Model Hidden Representations for Heart Rate Estimation from Auscultation
por: Nie, Jingping, et al.
Publicado: (2025)
por: Nie, Jingping, et al.
Publicado: (2025)
Voxceleb-ESP: preliminary experiments detecting Spanish celebrities from their voices
por: Labrador, Beltrán, et al.
Publicado: (2023)
por: Labrador, Beltrán, et al.
Publicado: (2023)
Naturalistic Music Decoding from EEG Data via Latent Diffusion Models
por: Postolache, Emilian, et al.
Publicado: (2024)
por: Postolache, Emilian, et al.
Publicado: (2024)
A Recall-First CNN for Sleep Apnea Screening from Snoring Audio
por: Mallick, Anushka, et al.
Publicado: (2025)
por: Mallick, Anushka, et al.
Publicado: (2025)
Context-aware child-directed speech detection from long-form recordings
por: Charlot, Théo, et al.
Publicado: (2026)
por: Charlot, Théo, et al.
Publicado: (2026)
Dementia classification from spontaneous speech using wrapper-based feature selection
por: Niemelä, Marko, et al.
Publicado: (2025)
por: Niemelä, Marko, et al.
Publicado: (2025)
Joint sentiment analysis of lyrics and audio in music
por: Schaab, Lea, et al.
Publicado: (2024)
por: Schaab, Lea, et al.
Publicado: (2024)
Fusion approaches for emotion recognition from speech using acoustic and text-based features
por: Pepino, Leonardo, et al.
Publicado: (2024)
por: Pepino, Leonardo, et al.
Publicado: (2024)
Ejemplares similares
-
The first Cadenza challenges: using machine learning competitions to improve music for listeners with a hearing loss
por: Dabike, Gerardo Roa, et al.
Publicado: (2024) -
Detecting music deepfakes is easy but actually hard
por: Afchar, Darius, et al.
Publicado: (2024) -
Long-form music generation with latent diffusion
por: Evans, Zach, et al.
Publicado: (2024) -
StemGen: A music generation model that listens
por: Parker, Julian D., et al.
Publicado: (2023) -
Learning and composing of classical music using restricted Boltzmann machines
por: Kobayashi, Mutsumi, et al.
Publicado: (2025)