Salvato in:
| Autore principale: | Zhang, Yawen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2401.12266 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Reinforcement Learning Jazz Improvisation: When Music Meets Game Theory
di: Tapiavala, Vedant, et al.
Pubblicazione: (2024)
di: Tapiavala, Vedant, et al.
Pubblicazione: (2024)
Deconstructing Jazz Piano Style Using Machine Learning
di: Cheston, Huw, et al.
Pubblicazione: (2025)
di: Cheston, Huw, et al.
Pubblicazione: (2025)
Multimodal Assessment of Speech Impairment in ALS Using Audio-Visual and Machine Learning Approaches
di: Pierotti, Francesco, et al.
Pubblicazione: (2025)
di: Pierotti, Francesco, et al.
Pubblicazione: (2025)
Exploring Transformer-Based Music Overpainting for Jazz Piano Variations
di: Row, Eleanor, et al.
Pubblicazione: (2024)
di: Row, Eleanor, et al.
Pubblicazione: (2024)
ImprovNet -- Generating Controllable Musical Improvisations with Iterative Corruption Refinement
di: Bhandari, Keshav, et al.
Pubblicazione: (2025)
di: Bhandari, Keshav, et al.
Pubblicazione: (2025)
Interpreting Graphic Notation with MusicLDM: An AI Improvisation of Cornelius Cardew's Treatise
di: Karchkhadze, Tornike, et al.
Pubblicazione: (2024)
di: Karchkhadze, Tornike, et al.
Pubblicazione: (2024)
Feasibility of Mental Health Triage Call Priority Prediction Using Machine Learning
di: Rana, Rajib, et al.
Pubblicazione: (2024)
di: Rana, Rajib, et al.
Pubblicazione: (2024)
JAZZVAR: A Dataset of Variations found within Solo Piano Performances of Jazz Standards for Music Overpainting
di: Row, Eleanor, et al.
Pubblicazione: (2023)
di: Row, Eleanor, et al.
Pubblicazione: (2023)
Reduction of Nonlinear Distortion in Condenser Microphones Using a Simple Post-Processing Technique
di: Honzík, Petr, et al.
Pubblicazione: (2024)
di: Honzík, Petr, et al.
Pubblicazione: (2024)
Audio-Based Classification of Insect Species Using Machine Learning Models: Cicada, Beetle, Termite, and Cricket
di: Shetty, Manas V, et al.
Pubblicazione: (2025)
di: Shetty, Manas V, et al.
Pubblicazione: (2025)
Water Flow Detection Device Based on Sound Data Analysis and Machine Learning to Detect Water Leakage
di: Pourmehrani, Hossein, et al.
Pubblicazione: (2025)
di: Pourmehrani, Hossein, et al.
Pubblicazione: (2025)
Heart Sound Segmentation Using Deep Learning Techniques
di: Madine, Manas
Pubblicazione: (2024)
di: Madine, Manas
Pubblicazione: (2024)
Fine-Grained Engine Fault Sound Event Detection Using Multimodal Signals
di: Fedorishin, Dennis, et al.
Pubblicazione: (2024)
di: Fedorishin, Dennis, et al.
Pubblicazione: (2024)
Physics-Informed Machine Learning For Sound Field Estimation
di: Koyama, Shoichi, et al.
Pubblicazione: (2024)
di: Koyama, Shoichi, et al.
Pubblicazione: (2024)
SICRN: Advancing Speech Enhancement through State Space Model and Inplace Convolution Techniques
di: Zhao, Changjiang, et al.
Pubblicazione: (2024)
di: Zhao, Changjiang, et al.
Pubblicazione: (2024)
Guitar-TECHS: An Electric Guitar Dataset Covering Techniques, Musical Excerpts, Chords and Scales Using a Diverse Array of Hardware
di: Pedroza, Hegel, et al.
Pubblicazione: (2025)
di: Pedroza, Hegel, et al.
Pubblicazione: (2025)
Learning Physiology-Informed Vocal Spectrotemporal Representations for Speech Emotion Recognition
di: Zhang, Xu, et al.
Pubblicazione: (2026)
di: Zhang, Xu, et al.
Pubblicazione: (2026)
Temporally Heterogeneous Graph Contrastive Learning for Multimodal Acoustic event Classification
di: Chen, Yuanjian, et al.
Pubblicazione: (2025)
di: Chen, Yuanjian, et al.
Pubblicazione: (2025)
Uncertainty Quantification in Machine Learning for Joint Speaker Diarization and Identification
di: McKnight, Simon W., et al.
Pubblicazione: (2023)
di: McKnight, Simon W., et al.
Pubblicazione: (2023)
Audio-Guided Fusion Techniques for Multimodal Emotion Analysis
di: Shi, Pujin, et al.
Pubblicazione: (2024)
di: Shi, Pujin, et al.
Pubblicazione: (2024)
Computer Audition: From Task-Specific Machine Learning to Foundation Models
di: Triantafyllopoulos, Andreas, et al.
Pubblicazione: (2024)
di: Triantafyllopoulos, Andreas, et al.
Pubblicazione: (2024)
Zero-Shot Recognition of Dysarthric Speech Using Commercial Automatic Speech Recognition and Multimodal Large Language Models
di: Alsayegh, Ali, et al.
Pubblicazione: (2025)
di: Alsayegh, Ali, et al.
Pubblicazione: (2025)
Advances in Speech Separation: Techniques, Challenges, and Future Trends
di: Li, Kai, et al.
Pubblicazione: (2025)
di: Li, Kai, et al.
Pubblicazione: (2025)
Teach Me How to ImproVISe: Co-Designing an Augmented Piano Training System for Improvisation
di: Deja, Jordan Aiko, et al.
Pubblicazione: (2024)
di: Deja, Jordan Aiko, et al.
Pubblicazione: (2024)
Stream-based Active Learning for Anomalous Sound Detection in Machine Condition Monitoring
di: Ho, Tuan Vu, et al.
Pubblicazione: (2024)
di: Ho, Tuan Vu, et al.
Pubblicazione: (2024)
Predicting Global HRTFs From Scanned Head Geometry Using Deep Learning and Compact Representations
di: Wang, Yuxiang, et al.
Pubblicazione: (2022)
di: Wang, Yuxiang, et al.
Pubblicazione: (2022)
Accent Normalization Using Self-Supervised Discrete Tokens with Non-Parallel Data
di: Bai, Qibing, et al.
Pubblicazione: (2025)
di: Bai, Qibing, et al.
Pubblicazione: (2025)
Toward Multimodal Industrial Fault Analysis: A Single-Speed Chain Conveyor Dataset with Audio and Vibration Signals
di: Chen, Zhang, et al.
Pubblicazione: (2026)
di: Chen, Zhang, et al.
Pubblicazione: (2026)
The Impact of Frequency Bands on Acoustic Anomaly Detection of Machines using Deep Learning Based Model
di: Nguyen, Tin, et al.
Pubblicazione: (2024)
di: Nguyen, Tin, et al.
Pubblicazione: (2024)
Data-Balanced Curriculum Learning for Audio Question Answering
di: Wijngaard, Gijs, et al.
Pubblicazione: (2025)
di: Wijngaard, Gijs, et al.
Pubblicazione: (2025)
Who Finds This Voice Attractive? A Large-Scale Experiment Using In-the-Wild Data
di: Suda, Hitoshi, et al.
Pubblicazione: (2024)
di: Suda, Hitoshi, et al.
Pubblicazione: (2024)
Deep Audio Watermarks are Shallow: Limitations of Post-Hoc Watermarking Techniques for Speech
di: O'Reilly, Patrick, et al.
Pubblicazione: (2025)
di: O'Reilly, Patrick, et al.
Pubblicazione: (2025)
Machine Learning in Acoustics: A Review and Open-Source Repository
di: McCarthy, Ryan A., et al.
Pubblicazione: (2025)
di: McCarthy, Ryan A., et al.
Pubblicazione: (2025)
Mitigating Category Imbalance: Fosafer System for the Multimodal Emotion and Intent Joint Understanding Challenge
di: Wang, Honghong, et al.
Pubblicazione: (2025)
di: Wang, Honghong, et al.
Pubblicazione: (2025)
Generating Piano Music with Transformers: A Comparative Study of Scale, Data, and Metrics
di: Lehmkuhl, Jonathan, et al.
Pubblicazione: (2025)
di: Lehmkuhl, Jonathan, et al.
Pubblicazione: (2025)
SMSAT: A Multimodal Acoustic Dataset and Deep Contrastive Learning Framework for Affective and Physiological Modeling of Spiritual Meditation
di: Suleman, Ahmad, et al.
Pubblicazione: (2025)
di: Suleman, Ahmad, et al.
Pubblicazione: (2025)
Revisiting and Improving Scoring Fusion for Spoofing-aware Speaker Verification Using Compositional Data Analysis
di: Wang, Xin, et al.
Pubblicazione: (2024)
di: Wang, Xin, et al.
Pubblicazione: (2024)
Exploring the Potential of Data-Driven Spatial Audio Enhancement Using a Single-Channel Model
di: Santos, Arthur N. dos, et al.
Pubblicazione: (2024)
di: Santos, Arthur N. dos, et al.
Pubblicazione: (2024)
$\text{M}^3\text{PDB}$: A Multimodal, Multi-Label, Multilingual Prompt Database for Speech Generation
di: Zhu, Boyu, et al.
Pubblicazione: (2025)
di: Zhu, Boyu, et al.
Pubblicazione: (2025)
A Study of Data Selection Strategies for Pre-training Self-Supervised Speech Models
di: Whetten, Ryan, et al.
Pubblicazione: (2026)
di: Whetten, Ryan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Reinforcement Learning Jazz Improvisation: When Music Meets Game Theory
di: Tapiavala, Vedant, et al.
Pubblicazione: (2024) -
Deconstructing Jazz Piano Style Using Machine Learning
di: Cheston, Huw, et al.
Pubblicazione: (2025) -
Multimodal Assessment of Speech Impairment in ALS Using Audio-Visual and Machine Learning Approaches
di: Pierotti, Francesco, et al.
Pubblicazione: (2025) -
Exploring Transformer-Based Music Overpainting for Jazz Piano Variations
di: Row, Eleanor, et al.
Pubblicazione: (2024) -
ImprovNet -- Generating Controllable Musical Improvisations with Iterative Corruption Refinement
di: Bhandari, Keshav, et al.
Pubblicazione: (2025)