A Method for MIDI Velocity Estimation for Piano Performance by a U-Net With Attention and FiLM
Fuente:
Zenodo
Saved in:
| Main Authors: | Hyon Kim, Xavier Serra |
|---|---|
| Format: | Recurso digital |
| Published: |
Zenodo
2024
|
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
STSM-FiLM: A FiLM-Conditioned Neural Architecture for Time-Scale Modification of Speech
by: Wisnu, Dyah A. M. G., et al.
Published: (2025)
by: Wisnu, Dyah A. M. G., et al.
Published: (2025)
FiPA-SR -- FiLM-Conditioned Perceptually Informed Audio Super-Resolution
by: Abreu, Wallace, et al.
Published: (2026)
by: Abreu, Wallace, et al.
Published: (2026)
Head-Pose-Aware Visual Speech Recognition with FiLM Modulation
by: Teng, Matthew Kit Khinn, et al.
Published: (2026)
by: Teng, Matthew Kit Khinn, et al.
Published: (2026)
FiLM-Nav: Efficient and Generalizable Navigation via VLM Fine-tuning
by: Yokoyama, Naoki, et al.
Published: (2025)
by: Yokoyama, Naoki, et al.
Published: (2025)
NudgeVAD: Language-Nudged End-to-End Driving via FiLM Residuals
by: Yang, Chieh-Chi, et al.
Published: (2026)
by: Yang, Chieh-Chi, et al.
Published: (2026)
Expressive MIDI-format Piano Performance Generation
by: Liu, Jingwei
Published: (2024)
by: Liu, Jingwei
Published: (2024)
Filling MIDI Velocity using U-Net Image Colorizer
by: He, Zhanhong, et al.
Published: (2025)
by: He, Zhanhong, et al.
Published: (2025)
PianoCoRe: Combined and Refined Piano MIDI Dataset
by: Borovik, Ilya
Published: (2026)
by: Borovik, Ilya
Published: (2026)
MEDNA-DFM: A Dual-View FiLM-MoE Model for Explainable DNA Methylation Prediction
by: He, Yi, et al.
Published: (2026)
by: He, Yi, et al.
Published: (2026)
End-to-end Piano Performance-MIDI to Score Conversion with Transformers
by: Beyer, Tim, et al.
Published: (2024)
by: Beyer, Tim, et al.
Published: (2024)
Aria-MIDI: A Dataset of Piano MIDI Files for Symbolic Music Modeling
by: Bradshaw, Louis, et al.
Published: (2025)
by: Bradshaw, Louis, et al.
Published: (2025)
Label-Efficient Hyperspectral Image Classification via Spectral FiLM Modulation of Low-Level Pretrained Diffusion Features
by: Hu, Yuzhen, et al.
Published: (2025)
by: Hu, Yuzhen, et al.
Published: (2025)
Prompt-Conditioned FiLM and Multi-Scale Fusion on MedSigLIP for Low-Dose CT Quality Assessment
by: Demiroglu, Tolga, et al.
Published: (2025)
by: Demiroglu, Tolga, et al.
Published: (2025)
Calibration-Conditioned FiLM Decoders for Low-Latency Decoding of Quantum Error Correction Evaluated on IBM Repetition-Code Experiments
by: Stein, Samuel, et al.
Published: (2026)
by: Stein, Samuel, et al.
Published: (2026)
Acoustics-specific Piano Velocity Estimation
by: Simonetta, Federico, et al.
Published: (2022)
by: Simonetta, Federico, et al.
Published: (2022)
MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling
by: Tang, Jingjing, et al.
Published: (2025)
by: Tang, Jingjing, et al.
Published: (2025)
WaveRoll: JavaScript Library for Comparative MIDI Piano-Roll Visualization
by: Park, Hannah, et al.
Published: (2025)
by: Park, Hannah, et al.
Published: (2025)
PianoVAM: A Multimodal Piano Performance Dataset
by: Kim, Yonghyun, et al.
Published: (2025)
by: Kim, Yonghyun, et al.
Published: (2025)
Difficulty-Aware Score Generation for Piano Sight-Reading
by: Ramoneda, Pedro, et al.
Published: (2025)
by: Ramoneda, Pedro, et al.
Published: (2025)
Fine-Tuning MIDI-to-Audio Alignment using a Neural Network on Piano Roll and CQT Representations
by: Murgul, Sebastian, et al.
Published: (2025)
by: Murgul, Sebastian, et al.
Published: (2025)
Can Audio Reveal Music Performance Difficulty? Insights from the Piano Syllabus Dataset
by: Ramoneda, Pedro, et al.
Published: (2024)
by: Ramoneda, Pedro, et al.
Published: (2024)
Score-Informed Transformer for Refining MIDI Velocity in Automatic Music Transcription
by: He, Zhanhong, et al.
Published: (2025)
by: He, Zhanhong, et al.
Published: (2025)
Data, Code, and Visualizations for the Aria-MIDI Digital Humanities Study
by: Fresquet, Xavier
Published: (2026)
by: Fresquet, Xavier
Published: (2026)
VST-Pose: A Velocity-Integrated Spatiotem-poral Attention Network for Human WiFi Pose Estimation
by: Zhang, Xinyu, et al.
Published: (2025)
by: Zhang, Xinyu, et al.
Published: (2025)
Difficulty-Controlled Simplification of Piano Scores with Synthetic Data for Inclusive Music Education
by: Ramoneda, Pedro, et al.
Published: (2025)
by: Ramoneda, Pedro, et al.
Published: (2025)
MIDI-LLM: Adapting Large Language Models for Text-to-MIDI Music Generation
by: Wu, Shih-Lun, et al.
Published: (2025)
by: Wu, Shih-Lun, et al.
Published: (2025)
How to Infer Repeat Structures in MIDI Performances
by: Peter, Silvan, et al.
Published: (2025)
by: Peter, Silvan, et al.
Published: (2025)
Beat-Based Rhythm Quantization of MIDI Performances
by: Wachter, Maximilian, et al.
Published: (2025)
by: Wachter, Maximilian, et al.
Published: (2025)
Learning to Play the Piano in China: Opportunities to Improve the Skills of Piano Playing and Piano Performance Based on the Professional Performance of Modern Pianists
by: Chen Chen, et al.
Published: (2025)
by: Chen Chen, et al.
Published: (2025)
On the de-duplication of the Lakh MIDI dataset
by: Choi, Eunjin, et al.
Published: (2025)
by: Choi, Eunjin, et al.
Published: (2025)
From Audio Encoders to Piano Judges: Benchmarking Performance Understanding for Solo Piano
by: Zhang, Huan, et al.
Published: (2024)
by: Zhang, Huan, et al.
Published: (2024)
MIDI-to-Tab: Guitar Tablature Inference via Masked Language Modeling
by: Edwards, Drew, et al.
Published: (2024)
by: Edwards, Drew, et al.
Published: (2024)
ProFi-Net: Prototype-based Feature Attention with Curriculum Augmentation for WiFi-based Gesture Recognition
by: Cui, Zhe, et al.
Published: (2025)
by: Cui, Zhe, et al.
Published: (2025)
Notochord: a Flexible Probabilistic Model for Real-Time MIDI Performance
by: Shepardson, Victor, et al.
Published: (2024)
by: Shepardson, Victor, et al.
Published: (2024)
Attention-based U-Net Method for Autonomous Lane Detection
by: Tangestanizadeh, Mohammadhamed, et al.
Published: (2024)
by: Tangestanizadeh, Mohammadhamed, et al.
Published: (2024)
The GigaMIDI Dataset with Features for Expressive Music Performance Detection
by: Lee, Keon Ju Maverick, et al.
Published: (2025)
by: Lee, Keon Ju Maverick, et al.
Published: (2025)
PianoMotion10M: Dataset and Benchmark for Hand Motion Generation in Piano Performance
by: Gan, Qijun, et al.
Published: (2024)
by: Gan, Qijun, et al.
Published: (2024)
Two Web Toolkits for Multimodal Piano Performance Dataset Acquisition and Fingering Annotation
by: Park, Junhyung, et al.
Published: (2025)
by: Park, Junhyung, et al.
Published: (2025)
Transformer-Based Rhythm Quantization of Performance MIDI Using Beat Annotations
by: Wachter, Maximilian, et al.
Published: (2026)
by: Wachter, Maximilian, et al.
Published: (2026)
Pay Attention to the Keys: Visual Piano Transcription Using Transformers
by: Zivanovic, Uros, et al.
Published: (2024)
by: Zivanovic, Uros, et al.
Published: (2024)
Similar Items
-
STSM-FiLM: A FiLM-Conditioned Neural Architecture for Time-Scale Modification of Speech
by: Wisnu, Dyah A. M. G., et al.
Published: (2025) -
FiPA-SR -- FiLM-Conditioned Perceptually Informed Audio Super-Resolution
by: Abreu, Wallace, et al.
Published: (2026) -
Head-Pose-Aware Visual Speech Recognition with FiLM Modulation
by: Teng, Matthew Kit Khinn, et al.
Published: (2026) -
FiLM-Nav: Efficient and Generalizable Navigation via VLM Fine-tuning
by: Yokoyama, Naoki, et al.
Published: (2025) -
NudgeVAD: Language-Nudged End-to-End Driving via FiLM Residuals
by: Yang, Chieh-Chi, et al.
Published: (2026)