Guardado en:
| Autores principales: | Simionato, Riccardo, Fasciani, Stefano |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2409.06513 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Modeling Time-Variant Responses of Optical Compressors with Selective State Space Models
por: Simionato, Riccardo, et al.
Publicado: (2024)
por: Simionato, Riccardo, et al.
Publicado: (2024)
MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling
por: Tang, Jingjing, et al.
Publicado: (2025)
por: Tang, Jingjing, et al.
Publicado: (2025)
PianoBART: Symbolic Piano Music Generation and Understanding with Large-Scale Pre-Training
por: Liang, Xiao, et al.
Publicado: (2024)
por: Liang, Xiao, et al.
Publicado: (2024)
Expressive MIDI-format Piano Performance Generation
por: Liu, Jingwei
Publicado: (2024)
por: Liu, Jingwei
Publicado: (2024)
A Holistic Evaluation of Piano Sound Quality
por: Zhou, Monan, et al.
Publicado: (2023)
por: Zhou, Monan, et al.
Publicado: (2023)
Emotion-driven Piano Music Generation via Two-stage Disentanglement and Functional Representation
por: Huang, Jingyue, et al.
Publicado: (2024)
por: Huang, Jingyue, et al.
Publicado: (2024)
End-to-End Real-World Polyphonic Piano Audio-to-Score Transcription with Hierarchical Decoding
por: Zeng, Wei, et al.
Publicado: (2024)
por: Zeng, Wei, et al.
Publicado: (2024)
Dialogue in Resonance: An Interactive Music Piece for Piano and Real-Time Automatic Transcription System
por: Bang, Hayeon, et al.
Publicado: (2025)
por: Bang, Hayeon, et al.
Publicado: (2025)
PIAST: A Multimodal Piano Dataset with Audio, Symbolic and Text
por: Bang, Hayeon, et al.
Publicado: (2024)
por: Bang, Hayeon, et al.
Publicado: (2024)
FürElise: Capturing and Physically Synthesizing Hand Motions of Piano Performance
por: Wang, Ruocheng, et al.
Publicado: (2024)
por: Wang, Ruocheng, et al.
Publicado: (2024)
Disentangling Score Content and Performance Style for Joint Piano Rendering and Transcription
por: Zeng, Wei, et al.
Publicado: (2025)
por: Zeng, Wei, et al.
Publicado: (2025)
Exploring Classical Piano Performance Generation with Expressive Music Variational AutoEncoder
por: Luo, Jing, et al.
Publicado: (2025)
por: Luo, Jing, et al.
Publicado: (2025)
Piano Transcription by Hierarchical Language Modeling with Pretrained Roll-based Encoders
por: Li, Dichucheng, et al.
Publicado: (2025)
por: Li, Dichucheng, et al.
Publicado: (2025)
D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription
por: Kim, Hounsu, et al.
Publicado: (2025)
por: Kim, Hounsu, et al.
Publicado: (2025)
Scaling Self-Supervised Representation Learning for Symbolic Piano Performance
por: Bradshaw, Louis, et al.
Publicado: (2025)
por: Bradshaw, Louis, et al.
Publicado: (2025)
Estimating Musical Surprisal from Audio in Autoregressive Diffusion Model Noise Spaces
por: Bjare, Mathias Rose, et al.
Publicado: (2025)
por: Bjare, Mathias Rose, et al.
Publicado: (2025)
Comparative Study of State-based Neural Networks for Virtual Analog Audio Effects Modeling
por: Simionato, Riccardo, et al.
Publicado: (2024)
por: Simionato, Riccardo, et al.
Publicado: (2024)
BNMusic: Blending Environmental Noises into Personalized Music
por: Zuo, Chi, et al.
Publicado: (2025)
por: Zuo, Chi, et al.
Publicado: (2025)
PianoVAM: A Multimodal Piano Performance Dataset
por: Kim, Yonghyun, et al.
Publicado: (2025)
por: Kim, Yonghyun, et al.
Publicado: (2025)
High-Resolution Sustain Pedal Depth Estimation from Piano Audio Across Room Acoustics
por: Fang, Kun, et al.
Publicado: (2025)
por: Fang, Kun, et al.
Publicado: (2025)
Effects of Dataset Sampling Rate for Noise Cancellation through Deep Learning
por: Colelough, Brandon, et al.
Publicado: (2024)
por: Colelough, Brandon, et al.
Publicado: (2024)
Segment-Factorized Full-Song Generation on Symbolic Piano Music
por: Chen, Ping-Yi, et al.
Publicado: (2025)
por: Chen, Ping-Yi, et al.
Publicado: (2025)
HuBERT-VIC: Improving Noise-Robust Automatic Speech Recognition of Speech Foundation Model via Variance-Invariance-Covariance Regularization
por: Ahn, Hyebin, et al.
Publicado: (2025)
por: Ahn, Hyebin, et al.
Publicado: (2025)
Neuro-MSBG: An End-to-End Neural Model for Hearing Loss Simulation
por: Yuan, Hui-Guan, et al.
Publicado: (2025)
por: Yuan, Hui-Guan, et al.
Publicado: (2025)
Noise-Agnostic Multitask Whisper Training for Reducing False Alarm Errors in Call-for-Help Detection
por: Ryu, Myeonghoon, et al.
Publicado: (2025)
por: Ryu, Myeonghoon, et al.
Publicado: (2025)
Physics-Informed Neural Engine Sound Modeling with Differentiable Pulse-Train Synthesis
por: Doerfler, Robin, et al.
Publicado: (2026)
por: Doerfler, Robin, et al.
Publicado: (2026)
Improving Neural Diarization through Speaker Attribute Attractors and Local Dependency Modeling
por: Palzer, David, et al.
Publicado: (2025)
por: Palzer, David, et al.
Publicado: (2025)
Probing the Information Encoded in Neural-based Acoustic Models of Automatic Speech Recognition Systems
por: Raymondaud, Quentin, et al.
Publicado: (2024)
por: Raymondaud, Quentin, et al.
Publicado: (2024)
HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling
por: Xue, Rongkun, et al.
Publicado: (2025)
por: Xue, Rongkun, et al.
Publicado: (2025)
PianoMotion10M: Dataset and Benchmark for Hand Motion Generation in Piano Performance
por: Gan, Qijun, et al.
Publicado: (2024)
por: Gan, Qijun, et al.
Publicado: (2024)
One-pass Multiple Conformer and Foundation Speech Systems Compression and Quantization Using An All-in-one Neural Model
por: Li, Zhaoqing, et al.
Publicado: (2024)
por: Li, Zhaoqing, et al.
Publicado: (2024)
A Multi-task Learning Balanced Attention Convolutional Neural Network Model for Few-shot Underwater Acoustic Target Recognition
por: Huang, Wei, et al.
Publicado: (2025)
por: Huang, Wei, et al.
Publicado: (2025)
Leveraging Mixture of Experts for Improved Speech Deepfake Detection
por: Negroni, Viola, et al.
Publicado: (2024)
por: Negroni, Viola, et al.
Publicado: (2024)
CONMOD: Controllable Neural Frame-based Modulation Effects
por: Lee, Gyubin, et al.
Publicado: (2024)
por: Lee, Gyubin, et al.
Publicado: (2024)
SPEAR: Receiver-to-Receiver Acoustic Neural Warping Field
por: He, Yuhang, et al.
Publicado: (2024)
por: He, Yuhang, et al.
Publicado: (2024)
Stage-Wise and Prior-Aware Neural Speech Phase Prediction
por: Liu, Fei, et al.
Publicado: (2024)
por: Liu, Fei, et al.
Publicado: (2024)
SpectroStream: A Versatile Neural Codec for General Audio
por: Li, Yunpeng, et al.
Publicado: (2025)
por: Li, Yunpeng, et al.
Publicado: (2025)
How Do Neural Spoofing Countermeasures Detect Partially Spoofed Audio?
por: Liu, Tianchi, et al.
Publicado: (2024)
por: Liu, Tianchi, et al.
Publicado: (2024)
A Real-Time Voice Activity Detection Based On Lightweight Neural
por: Jia, Jidong, et al.
Publicado: (2024)
por: Jia, Jidong, et al.
Publicado: (2024)
Towards Leveraging Contrastively Pretrained Neural Audio Embeddings for Recommender Tasks
por: Grötschla, Florian, et al.
Publicado: (2024)
por: Grötschla, Florian, et al.
Publicado: (2024)
Ejemplares similares
-
Modeling Time-Variant Responses of Optical Compressors with Selective State Space Models
por: Simionato, Riccardo, et al.
Publicado: (2024) -
MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling
por: Tang, Jingjing, et al.
Publicado: (2025) -
PianoBART: Symbolic Piano Music Generation and Understanding with Large-Scale Pre-Training
por: Liang, Xiao, et al.
Publicado: (2024) -
Expressive MIDI-format Piano Performance Generation
por: Liu, Jingwei
Publicado: (2024) -
A Holistic Evaluation of Piano Sound Quality
por: Zhou, Monan, et al.
Publicado: (2023)