The evolution of inharmonicity and noisiness in contemporary popular music
Fuente:
arXiv
Guardado en:
| Autores principales: | Deruty, Emmanuel, Meredith, David, Lattner, Stefan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Neural Proxies for Sound Synthesizers: Learning Perceptually Informed Preset Representations
por: Combes, Paolo, et al.
Publicado: (2025)
por: Combes, Paolo, et al.
Publicado: (2025)
OBHS: An Optimized Block Huffman Scheme for Real-Time Audio Compression
por: Mahfi, Muntahi Safwan, et al.
Publicado: (2025)
por: Mahfi, Muntahi Safwan, et al.
Publicado: (2025)
Quantum-Enhanced Analysis and Grading of Vocal Performance
por: Agarwal, Rohan
Publicado: (2025)
por: Agarwal, Rohan
Publicado: (2025)
ParaNoise-SV: Integrated Approach for Noise-Robust Speaker Verification with Parallel Joint Learning of Speech Enhancement and Noise Extraction
por: Kim, Minu, et al.
Publicado: (2025)
por: Kim, Minu, et al.
Publicado: (2025)
Joint Estimation of Piano Dynamics and Metrical Structure with a Multi-task Multi-Scale Network
por: He, Zhanhong, et al.
Publicado: (2025)
por: He, Zhanhong, et al.
Publicado: (2025)
Machine Learning Framework for Audio-Based Content Evaluation using MFCC, Chroma, Spectral Contrast, and Temporal Feature Engineering
por: Aristorenas, Aris J.
Publicado: (2024)
por: Aristorenas, Aris J.
Publicado: (2024)
Revisiting SSL for sound event detection: complementary fusion and adaptive post-processing
por: Cui, Hanfang, et al.
Publicado: (2025)
por: Cui, Hanfang, et al.
Publicado: (2025)
Prevailing Research Areas for Music AI in the Era of Foundation Models
por: Wei, Megan, et al.
Publicado: (2024)
por: Wei, Megan, et al.
Publicado: (2024)
Understanding the Algorithm Behind Audio Key Detection
por: Silva, Henrique Perez G.
Publicado: (2025)
por: Silva, Henrique Perez G.
Publicado: (2025)
Improving Cross-Lingual Phonetic Representation of Low-Resource Languages Through Language Similarity Analysis
por: Kim, Minu, et al.
Publicado: (2025)
por: Kim, Minu, et al.
Publicado: (2025)
HELIX: Scaling Raw Audio Understanding with Hybrid Mamba-Attention Beyond the Quadratic Limit
por: Khushiyant, et al.
Publicado: (2026)
por: Khushiyant, et al.
Publicado: (2026)
Matlab-based Epoch Extraction for Speaker Differentiation
por: Li, Kunlun, et al.
Publicado: (2024)
por: Li, Kunlun, et al.
Publicado: (2024)
STRUM: A Spectral Transcription and Rhythm Understanding Model for End-to-End Generation of Playable Rhythm-Game Charts
por: Opria, Joshua
Publicado: (2026)
por: Opria, Joshua
Publicado: (2026)
SeamlessEdit: Background Noise Aware Zero-Shot Speech Editing with in-Context Enhancement
por: Chen, Kuan-Yu, et al.
Publicado: (2025)
por: Chen, Kuan-Yu, et al.
Publicado: (2025)
Dichotic harmony for the musical practice
por: Madgazin, Vadim R.
Publicado: (2010)
por: Madgazin, Vadim R.
Publicado: (2010)
SoundPlot: An Open-Source Framework for Birdsong Acoustic Analysis and Neural Synthesis with Interactive 3D Visualization
por: Mehdi, Naqcho Ali, et al.
Publicado: (2026)
por: Mehdi, Naqcho Ali, et al.
Publicado: (2026)
AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks
por: Maben, Leander Melroy, et al.
Publicado: (2025)
por: Maben, Leander Melroy, et al.
Publicado: (2025)
An introduction to pitch strength in contemporary popular music analysis and production
por: Deruty, Emmanuel
Publicado: (2025)
por: Deruty, Emmanuel
Publicado: (2025)
Reciprocal Latent Fields for Precomputed Sound Propagation
por: Seuté, Hugo, et al.
Publicado: (2026)
por: Seuté, Hugo, et al.
Publicado: (2026)
Window Size Versus Accuracy Experiments in Voice Activity Detectors
por: McKinnon, Max, et al.
Publicado: (2026)
por: McKinnon, Max, et al.
Publicado: (2026)
Generation of Musical Timbres using a Text-Guided Diffusion Model
por: Yuan, Weixuan, et al.
Publicado: (2025)
por: Yuan, Weixuan, et al.
Publicado: (2025)
Self-Improvement for Audio Large Language Model using Unlabeled Speech
por: Wang, Shaowen, et al.
Publicado: (2025)
por: Wang, Shaowen, et al.
Publicado: (2025)
MAIN-VC: Lightweight Speech Representation Disentanglement for One-shot Voice Conversion
por: Li, Pengcheng, et al.
Publicado: (2024)
por: Li, Pengcheng, et al.
Publicado: (2024)
MaskClip: Detachable Clip-on Piezoelectric Sensing of Mask Surface Vibrations for Real-time Noise-Robust Speech Input
por: Hiraki, Hirotaka, et al.
Publicado: (2025)
por: Hiraki, Hirotaka, et al.
Publicado: (2025)
Quantization for OpenAI's Whisper Models: A Comparative Analysis
por: Andreyev, Allison
Publicado: (2025)
por: Andreyev, Allison
Publicado: (2025)
Methods for pitch analysis in contemporary popular music: Vitalic's use of tones that do not operate on the principle of acoustic resonance
por: Deruty, Emmanuel, et al.
Publicado: (2025)
por: Deruty, Emmanuel, et al.
Publicado: (2025)
Real-time Low-latency Music Source Separation using Hybrid Spectrogram-TasNet
por: Venkatesh, Satvik, et al.
Publicado: (2024)
por: Venkatesh, Satvik, et al.
Publicado: (2024)
How much to Dereverberate? Low-Latency Single-Channel Speech Enhancement in Distant Microphone Scenarios
por: Venkatesh, Satvik, et al.
Publicado: (2025)
por: Venkatesh, Satvik, et al.
Publicado: (2025)
Audio Foundation Models Outperform Symbolic Representations for Piano Performance Evaluation
por: Dhiman, Jai
Publicado: (2026)
por: Dhiman, Jai
Publicado: (2026)
Modeling L1 Influence on L2 Pronunciation: An MFCC-Based Framework for Explainable Machine Learning and Pedagogical Feedback
por: Jahanbin, Peyman
Publicado: (2025)
por: Jahanbin, Peyman
Publicado: (2025)
Deep Feed-Forward Neural Network for Bangla Isolated Speech Recognition
por: Bhadra, Dipayan, et al.
Publicado: (2025)
por: Bhadra, Dipayan, et al.
Publicado: (2025)
The Concatenator: A Bayesian Approach To Real Time Concatenative Musaicing
por: Tralie, Christopher, et al.
Publicado: (2024)
por: Tralie, Christopher, et al.
Publicado: (2024)
Delayed Fusion: Integrating Large Language Models into First-Pass Decoding in End-to-end Speech Recognition
por: Hori, Takaaki, et al.
Publicado: (2025)
por: Hori, Takaaki, et al.
Publicado: (2025)
Audio-based Kinship Verification Using Age Domain Conversion
por: Sun, Qiyang, et al.
Publicado: (2024)
por: Sun, Qiyang, et al.
Publicado: (2024)
Make Some Noise: Towards LLM audio reasoning and generation using sound tokens
por: Mehta, Shivam, et al.
Publicado: (2025)
por: Mehta, Shivam, et al.
Publicado: (2025)
SemAlignVC: Enhancing zero-shot timbre conversion using semantic alignment
por: Mehta, Shivam, et al.
Publicado: (2025)
por: Mehta, Shivam, et al.
Publicado: (2025)
Guitar Tone Morphing by Diffusion-based Model
por: Chen, Kuan-Yu, et al.
Publicado: (2025)
por: Chen, Kuan-Yu, et al.
Publicado: (2025)
Bloodroot: When Watermarking Turns Poisonous For Stealthy Backdoor
por: Chen, Kuan-Yu, et al.
Publicado: (2025)
por: Chen, Kuan-Yu, et al.
Publicado: (2025)
Graph Connectionist Temporal Classification for Phoneme Recognition
por: Grafé, Henry, et al.
Publicado: (2025)
por: Grafé, Henry, et al.
Publicado: (2025)
Crossing the Species Divide: Transfer Learning from Speech to Animal Sounds
por: Cauzinille, Jules, et al.
Publicado: (2025)
por: Cauzinille, Jules, et al.
Publicado: (2025)
Ejemplares similares
-
Neural Proxies for Sound Synthesizers: Learning Perceptually Informed Preset Representations
por: Combes, Paolo, et al.
Publicado: (2025) -
OBHS: An Optimized Block Huffman Scheme for Real-Time Audio Compression
por: Mahfi, Muntahi Safwan, et al.
Publicado: (2025) -
Quantum-Enhanced Analysis and Grading of Vocal Performance
por: Agarwal, Rohan
Publicado: (2025) -
ParaNoise-SV: Integrated Approach for Noise-Robust Speaker Verification with Parallel Joint Learning of Speech Enhancement and Noise Extraction
por: Kim, Minu, et al.
Publicado: (2025) -
Joint Estimation of Piano Dynamics and Metrical Structure with a Multi-task Multi-Scale Network
por: He, Zhanhong, et al.
Publicado: (2025)