Filling MIDI Velocity using U-Net Image Colorizer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | He, Zhanhong, Cooper, David, Huang, Defeng, Togneri, Roberto |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Score-Informed Transformer for Refining MIDI Velocity in Automatic Music Transcription
von: He, Zhanhong, et al.
Veröffentlicht: (2025)
von: He, Zhanhong, et al.
Veröffentlicht: (2025)
How to Infer Repeat Structures in MIDI Performances
von: Peter, Silvan, et al.
Veröffentlicht: (2025)
von: Peter, Silvan, et al.
Veröffentlicht: (2025)
Pseudo Strong Labels from Frame-Level Predictions for Weakly Supervised Sound Event Detection
von: Zhang, Yuliang, et al.
Veröffentlicht: (2025)
von: Zhang, Yuliang, et al.
Veröffentlicht: (2025)
Impact of Noisy Labels on Sound Event Detection: Deletion Errors Are More Detrimental Than Insertion Errors
von: Zhang, Yuliang, et al.
Veröffentlicht: (2024)
von: Zhang, Yuliang, et al.
Veröffentlicht: (2024)
MIDI-Informed Singing Accompaniment Generation in a Compositional Song Pipeline
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2026)
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2026)
Joint Estimation of Piano Dynamics and Metrical Structure with a Multi-task Multi-Scale Network
von: He, Zhanhong, et al.
Veröffentlicht: (2025)
von: He, Zhanhong, et al.
Veröffentlicht: (2025)
Zero to 16383 Through the Wire: Transmitting High- Resolution MIDI with WebSockets and the Browser
von: McKemie, Daniel
Veröffentlicht: (2025)
von: McKemie, Daniel
Veröffentlicht: (2025)
Dance2MIDI: Dance-driven multi-instruments music generation
von: Han, Bo, et al.
Veröffentlicht: (2023)
von: Han, Bo, et al.
Veröffentlicht: (2023)
Transformer-Based Rhythm Quantization of Performance MIDI Using Beat Annotations
von: Wachter, Maximilian, et al.
Veröffentlicht: (2026)
von: Wachter, Maximilian, et al.
Veröffentlicht: (2026)
Expressive MIDI-format Piano Performance Generation
von: Liu, Jingwei
Veröffentlicht: (2024)
von: Liu, Jingwei
Veröffentlicht: (2024)
End-to-end Piano Performance-MIDI to Score Conversion with Transformers
von: Beyer, Tim, et al.
Veröffentlicht: (2024)
von: Beyer, Tim, et al.
Veröffentlicht: (2024)
Notochord: a Flexible Probabilistic Model for Real-Time MIDI Performance
von: Shepardson, Victor, et al.
Veröffentlicht: (2024)
von: Shepardson, Victor, et al.
Veröffentlicht: (2024)
SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit
von: Huang, Wen-Chin, et al.
Veröffentlicht: (2025)
von: Huang, Wen-Chin, et al.
Veröffentlicht: (2025)
MOS-Bench: Benchmarking Generalization Abilities of Subjective Speech Quality Assessment Models
von: Huang, Wen-Chin, et al.
Veröffentlicht: (2024)
von: Huang, Wen-Chin, et al.
Veröffentlicht: (2024)
CodecMOS-Accent: A MOS Benchmark of Resynthesized and TTS Speech from Neural Codecs Across English Accents
von: Huang, Wen-Chin, et al.
Veröffentlicht: (2026)
von: Huang, Wen-Chin, et al.
Veröffentlicht: (2026)
Moises-Light: Resource-efficient Band-split U-Net For Music Source Separation
von: Yun-Ning, et al.
Veröffentlicht: (2025)
von: Yun-Ning, et al.
Veröffentlicht: (2025)
Beat-Based Rhythm Quantization of MIDI Performances
von: Wachter, Maximilian, et al.
Veröffentlicht: (2025)
von: Wachter, Maximilian, et al.
Veröffentlicht: (2025)
Annotation-Free MIDI-to-Audio Synthesis via Concatenative Synthesis and Generative Refinement
von: Take, Osamu, et al.
Veröffentlicht: (2024)
von: Take, Osamu, et al.
Veröffentlicht: (2024)
BreathNet: Generalizable Audio Deepfake Detection via Breath-Cue-Guided Feature Refinement
von: Ye, Zhe, et al.
Veröffentlicht: (2026)
von: Ye, Zhe, et al.
Veröffentlicht: (2026)
Reproducing the Acoustic Velocity Vectors in a Circular Listening Area
von: Wang, Jiarui, et al.
Veröffentlicht: (2024)
von: Wang, Jiarui, et al.
Veröffentlicht: (2024)
Sub-band and Full-band Interactive U-Net with DPRNN for Demixing Cross-talk Stereo Music
von: Yin, Han, et al.
Veröffentlicht: (2024)
von: Yin, Han, et al.
Veröffentlicht: (2024)
An Adaptive CMSA for Solving the Longest Filled Common Subsequence Problem with an Application in Audio Querying
von: Djukanovic, Marko, et al.
Veröffentlicht: (2025)
von: Djukanovic, Marko, et al.
Veröffentlicht: (2025)
Moonbeam: A MIDI Foundation Model Using Both Absolute and Relative Music Attributes
von: Guo, Zixun, et al.
Veröffentlicht: (2025)
von: Guo, Zixun, et al.
Veröffentlicht: (2025)
COVID-19 Diagnosis from Cough Acoustics using ConvNets and Data Augmentation
von: Mahanta, Saranga Kingkor, et al.
Veröffentlicht: (2021)
von: Mahanta, Saranga Kingkor, et al.
Veröffentlicht: (2021)
MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling
von: Tang, Jingjing, et al.
Veröffentlicht: (2025)
von: Tang, Jingjing, et al.
Veröffentlicht: (2025)
Composer's Assistant 2: Interactive Multi-Track MIDI Infilling with Fine-Grained User Control
von: Malandro, Martin E.
Veröffentlicht: (2024)
von: Malandro, Martin E.
Veröffentlicht: (2024)
Fretting-Transformer: Encoder-Decoder Model for MIDI to Tablature Transcription
von: Hamberger, Anna, et al.
Veröffentlicht: (2025)
von: Hamberger, Anna, et al.
Veröffentlicht: (2025)
SingNet: Towards a Large-Scale, Diverse, and In-the-Wild Singing Voice Dataset
von: Gu, Yicheng, et al.
Veröffentlicht: (2025)
von: Gu, Yicheng, et al.
Veröffentlicht: (2025)
AuralNet: Hierarchical Attention-based 3D Binaural Localization of Overlapping Speakers
von: Fu, Linya, et al.
Veröffentlicht: (2025)
von: Fu, Linya, et al.
Veröffentlicht: (2025)
Fine-Tuning MIDI-to-Audio Alignment using a Neural Network on Piano Roll and CQT Representations
von: Murgul, Sebastian, et al.
Veröffentlicht: (2025)
von: Murgul, Sebastian, et al.
Veröffentlicht: (2025)
RaD-Net: A Repairing and Denoising Network for Speech Signal Improvement
von: Liu, Mingshuai, et al.
Veröffentlicht: (2024)
von: Liu, Mingshuai, et al.
Veröffentlicht: (2024)
TTS-CtrlNet: Time varying emotion aligned text-to-speech generation with ControlNet
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2025)
von: Jeong, Jaeseok, et al.
Veröffentlicht: (2025)
GMM-ResNet2: Ensemble of Group ResNet Networks for Synthetic Speech Detection
von: Lei, Zhenchun, et al.
Veröffentlicht: (2024)
von: Lei, Zhenchun, et al.
Veröffentlicht: (2024)
MidiCaps: A large-scale MIDI dataset with text captions
von: Melechovsky, Jan, et al.
Veröffentlicht: (2024)
von: Melechovsky, Jan, et al.
Veröffentlicht: (2024)
PhiNet: Speaker Verification with Phonetic Interpretability
von: Ma, Yi, et al.
Veröffentlicht: (2026)
von: Ma, Yi, et al.
Veröffentlicht: (2026)
KS-Net: Multi-band joint speech restoration and enhancement network for 2024 ICASSP SSI Challenge
von: Yu, Guochen, et al.
Veröffentlicht: (2024)
von: Yu, Guochen, et al.
Veröffentlicht: (2024)
Beat and Downbeat Tracking in Performance MIDI Using an End-to-End Transformer Architecture
von: Murgul, Sebastian, et al.
Veröffentlicht: (2025)
von: Murgul, Sebastian, et al.
Veröffentlicht: (2025)
Asynchronous Microphone Array Calibration using Hybrid TDOA Information
von: Zhang, Chengjie, et al.
Veröffentlicht: (2024)
von: Zhang, Chengjie, et al.
Veröffentlicht: (2024)
PlumberNet: Fixing interference leakage after GEV beamforming
von: Grondin, François, et al.
Veröffentlicht: (2023)
von: Grondin, François, et al.
Veröffentlicht: (2023)
Inter-channel Conv-TasNet for multichannel speech enhancement
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
von: Lee, Dongheon, et al.
Veröffentlicht: (2021)
Ähnliche Einträge
-
Score-Informed Transformer for Refining MIDI Velocity in Automatic Music Transcription
von: He, Zhanhong, et al.
Veröffentlicht: (2025) -
How to Infer Repeat Structures in MIDI Performances
von: Peter, Silvan, et al.
Veröffentlicht: (2025) -
Pseudo Strong Labels from Frame-Level Predictions for Weakly Supervised Sound Event Detection
von: Zhang, Yuliang, et al.
Veröffentlicht: (2025) -
Impact of Noisy Labels on Sound Event Detection: Deletion Errors Are More Detrimental Than Insertion Errors
von: Zhang, Yuliang, et al.
Veröffentlicht: (2024) -
MIDI-Informed Singing Accompaniment Generation in a Compositional Song Pipeline
von: Tsai, Fang-Duo, et al.
Veröffentlicht: (2026)