Zero to 16383 Through the Wire: Transmitting High- Resolution MIDI with WebSockets and the Browser
Fuente:
arXiv
Guardado en:
| Autor principal: | McKemie, Daniel |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Control Surfaces: Using the Commodore 64 and Analog Synthesizer to Expand Musical Boundaries
por: McKemie, Daniel
Publicado: (2025)
por: McKemie, Daniel
Publicado: (2025)
pyAMPACT: A Score-Audio Alignment Toolkit for Performance Data Estimation and Multi-modal Processing
por: Devaney, Johanna, et al.
Publicado: (2024)
por: Devaney, Johanna, et al.
Publicado: (2024)
How to Infer Repeat Structures in MIDI Performances
por: Peter, Silvan, et al.
Publicado: (2025)
por: Peter, Silvan, et al.
Publicado: (2025)
Filling MIDI Velocity using U-Net Image Colorizer
por: He, Zhanhong, et al.
Publicado: (2025)
por: He, Zhanhong, et al.
Publicado: (2025)
Score-Informed Transformer for Refining MIDI Velocity in Automatic Music Transcription
por: He, Zhanhong, et al.
Publicado: (2025)
por: He, Zhanhong, et al.
Publicado: (2025)
MIDI-Informed Singing Accompaniment Generation in a Compositional Song Pipeline
por: Tsai, Fang-Duo, et al.
Publicado: (2026)
por: Tsai, Fang-Duo, et al.
Publicado: (2026)
Dance2MIDI: Dance-driven multi-instruments music generation
por: Han, Bo, et al.
Publicado: (2023)
por: Han, Bo, et al.
Publicado: (2023)
Transformer-Based Rhythm Quantization of Performance MIDI Using Beat Annotations
por: Wachter, Maximilian, et al.
Publicado: (2026)
por: Wachter, Maximilian, et al.
Publicado: (2026)
MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling
por: Tang, Jingjing, et al.
Publicado: (2025)
por: Tang, Jingjing, et al.
Publicado: (2025)
Expressive MIDI-format Piano Performance Generation
por: Liu, Jingwei
Publicado: (2024)
por: Liu, Jingwei
Publicado: (2024)
RADE: A Neural Codec for Transmitting Speech over HF Radio Channels
por: Rowe, David, et al.
Publicado: (2025)
por: Rowe, David, et al.
Publicado: (2025)
End-to-end Piano Performance-MIDI to Score Conversion with Transformers
por: Beyer, Tim, et al.
Publicado: (2024)
por: Beyer, Tim, et al.
Publicado: (2024)
ZipVoice: Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching
por: Zhu, Han, et al.
Publicado: (2025)
por: Zhu, Han, et al.
Publicado: (2025)
Notochord: a Flexible Probabilistic Model for Real-Time MIDI Performance
por: Shepardson, Victor, et al.
Publicado: (2024)
por: Shepardson, Victor, et al.
Publicado: (2024)
Ambisonics Super-Resolution Using A Waveform-Domain Neural Network
por: Nawfal, Ismael, et al.
Publicado: (2025)
por: Nawfal, Ismael, et al.
Publicado: (2025)
Beat-Based Rhythm Quantization of MIDI Performances
por: Wachter, Maximilian, et al.
Publicado: (2025)
por: Wachter, Maximilian, et al.
Publicado: (2025)
Annotation-Free MIDI-to-Audio Synthesis via Concatenative Synthesis and Generative Refinement
por: Take, Osamu, et al.
Publicado: (2024)
por: Take, Osamu, et al.
Publicado: (2024)
Guitar Pickups I: Analysis of the Effect of Winding and Wire Gauge on Single Coil Electric Guitar Pickups
por: Batchelor, Charles, et al.
Publicado: (2024)
por: Batchelor, Charles, et al.
Publicado: (2024)
Moonbeam: A MIDI Foundation Model Using Both Absolute and Relative Music Attributes
por: Guo, Zixun, et al.
Publicado: (2025)
por: Guo, Zixun, et al.
Publicado: (2025)
STSR: High-Fidelity Speech Super-Resolution via Spectral-Transient Context Modeling
por: Yuan, Jiajun, et al.
Publicado: (2025)
por: Yuan, Jiajun, et al.
Publicado: (2025)
Composer's Assistant 2: Interactive Multi-Track MIDI Infilling with Fine-Grained User Control
por: Malandro, Martin E.
Publicado: (2024)
por: Malandro, Martin E.
Publicado: (2024)
MobileSpeech: A Fast and High-Fidelity Framework for Mobile Zero-Shot Text-to-Speech
por: Ji, Shengpeng, et al.
Publicado: (2024)
por: Ji, Shengpeng, et al.
Publicado: (2024)
Fretting-Transformer: Encoder-Decoder Model for MIDI to Tablature Transcription
por: Hamberger, Anna, et al.
Publicado: (2025)
por: Hamberger, Anna, et al.
Publicado: (2025)
Zero Shot Text to Speech Augmentation for Automatic Speech Recognition on Low-Resource Accented Speech Corpora
por: Nespoli, Francesco, et al.
Publicado: (2024)
por: Nespoli, Francesco, et al.
Publicado: (2024)
GLOBE: A High-quality English Corpus with Global Accents for Zero-shot Speaker Adaptive Text-to-Speech
por: Wang, Wenbin, et al.
Publicado: (2024)
por: Wang, Wenbin, et al.
Publicado: (2024)
HiFi-SR: A Unified Generative Transformer-Convolutional Adversarial Network for High-Fidelity Speech Super-Resolution
por: Zhao, Shengkui, et al.
Publicado: (2025)
por: Zhao, Shengkui, et al.
Publicado: (2025)
MidiCaps: A large-scale MIDI dataset with text captions
por: Melechovsky, Jan, et al.
Publicado: (2024)
por: Melechovsky, Jan, et al.
Publicado: (2024)
StyleTTS-ZS: Efficient High-Quality Zero-Shot Text-to-Speech Synthesis with Distilled Time-Varying Style Diffusion
por: Li, Yinghao Aaron, et al.
Publicado: (2024)
por: Li, Yinghao Aaron, et al.
Publicado: (2024)
A Multi-loudspeaker Binaural Room Impulse Response Dataset with High-Resolution Translational and Rotational Head Coordinates in a Listening Room
por: Qiao, Yue, et al.
Publicado: (2024)
por: Qiao, Yue, et al.
Publicado: (2024)
Beat and Downbeat Tracking in Performance MIDI Using an End-to-End Transformer Architecture
por: Murgul, Sebastian, et al.
Publicado: (2025)
por: Murgul, Sebastian, et al.
Publicado: (2025)
BickGraphing: Web-Based Application for Visual Inspection of Audio Recordings
por: Seow, Kayley, et al.
Publicado: (2026)
por: Seow, Kayley, et al.
Publicado: (2026)
Zero- and Few-shot Sound Event Localization and Detection
por: Shimada, Kazuki, et al.
Publicado: (2023)
por: Shimada, Kazuki, et al.
Publicado: (2023)
Intelli-Z: Toward Intelligible Zero-Shot TTS
por: Jung, Sunghee, et al.
Publicado: (2024)
por: Jung, Sunghee, et al.
Publicado: (2024)
Zero-shot Cross-lingual Voice Transfer for TTS
por: Biadsy, Fadi, et al.
Publicado: (2024)
por: Biadsy, Fadi, et al.
Publicado: (2024)
Joint Fullband-Subband Modeling for High-Resolution SingFake Detection
por: Chen, Xuanjun, et al.
Publicado: (2026)
por: Chen, Xuanjun, et al.
Publicado: (2026)
Improving Rare-Word Recognition of Whisper in Zero-Shot Settings
por: Jogi, Yash, et al.
Publicado: (2025)
por: Jogi, Yash, et al.
Publicado: (2025)
Zero-Shot Audio Captioning Using Soft and Hard Prompts
por: Zhang, Yiming, et al.
Publicado: (2024)
por: Zhang, Yiming, et al.
Publicado: (2024)
MEDIC: Zero-shot Music Editing with Disentangled Inversion Control
por: Liu, Huadai, et al.
Publicado: (2024)
por: Liu, Huadai, et al.
Publicado: (2024)
SF-Speech: Straightened Flow for Zero-Shot Voice Clone
por: Li, Xuyuan, et al.
Publicado: (2024)
por: Li, Xuyuan, et al.
Publicado: (2024)
Zero-Shot Text-to-Speech from Continuous Text Streams
por: Dang, Trung, et al.
Publicado: (2024)
por: Dang, Trung, et al.
Publicado: (2024)
Ejemplares similares
-
Control Surfaces: Using the Commodore 64 and Analog Synthesizer to Expand Musical Boundaries
por: McKemie, Daniel
Publicado: (2025) -
pyAMPACT: A Score-Audio Alignment Toolkit for Performance Data Estimation and Multi-modal Processing
por: Devaney, Johanna, et al.
Publicado: (2024) -
How to Infer Repeat Structures in MIDI Performances
por: Peter, Silvan, et al.
Publicado: (2025) -
Filling MIDI Velocity using U-Net Image Colorizer
por: He, Zhanhong, et al.
Publicado: (2025) -
Score-Informed Transformer for Refining MIDI Velocity in Automatic Music Transcription
por: He, Zhanhong, et al.
Publicado: (2025)