STRUM: A Spectral Transcription and Rhythm Understanding Model for End-to-End Generation of Playable Rhythm-Game Charts
Fuente:
arXiv
Saved in:
| Main Author: | Opria, Joshua |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Objective Gastrointestinal Auscultation: Automated Segmentation and Annotation of Bowel Sound Patterns
by: Mansour, Zahra, et al.
Published: (2026)
by: Mansour, Zahra, et al.
Published: (2026)
Benchmarking machine learning for bowel sound pattern classification from tabular features to pretrained models
by: Mansour, Zahra, et al.
Published: (2025)
by: Mansour, Zahra, et al.
Published: (2025)
Optimizing Hospital Capacity During Pandemics: A Dual-Component Framework for Strategic Patient Relocation
by: Tabatabaee, Sadaf, et al.
Published: (2026)
by: Tabatabaee, Sadaf, et al.
Published: (2026)
Cycling Race Time Prediction: A Personalized Machine Learning Approach Using Route Topology and Training Load
by: Moreno, Francisco Aguilera
Published: (2026)
by: Moreno, Francisco Aguilera
Published: (2026)
OverNaN: NaN-Aware Oversampling for Imbalanced Learning with Meaningful Missingness
by: Barnard, Amanda S
Published: (2026)
by: Barnard, Amanda S
Published: (2026)
Over-Squashing in Graph Neural Networks: A Comprehensive survey
by: Akansha, Singh
Published: (2023)
by: Akansha, Singh
Published: (2023)
CENTS: Generating synthetic electricity consumption time series for rare and unseen scenarios
by: Fuest, Michael, et al.
Published: (2025)
by: Fuest, Michael, et al.
Published: (2025)
Optimal Transport-Guided Safety in Temporal Difference Reinforcement Learning
by: Shahrooei, Zahra, et al.
Published: (2025)
by: Shahrooei, Zahra, et al.
Published: (2025)
FactoryBench: Evaluating Industrial Machine Understanding
by: Merzouki, Yanis, et al.
Published: (2026)
by: Merzouki, Yanis, et al.
Published: (2026)
ParaNoise-SV: Integrated Approach for Noise-Robust Speaker Verification with Parallel Joint Learning of Speech Enhancement and Noise Extraction
by: Kim, Minu, et al.
Published: (2025)
by: Kim, Minu, et al.
Published: (2025)
Fine-Tuning Integrity for Modern Neural Networks: Structured Drift Proofs via Norm, Rank, and Sparsity Certificates
by: Shang, Zhenhang, et al.
Published: (2026)
by: Shang, Zhenhang, et al.
Published: (2026)
Factored Diffusion Policies:Compositionally Generalized Robot Control with a Single Score Network
by: Mitra, Sayan, et al.
Published: (2026)
by: Mitra, Sayan, et al.
Published: (2026)
Semantic Content Determines Algorithmic Performance
by: Ríos-García, Martiño, et al.
Published: (2026)
by: Ríos-García, Martiño, et al.
Published: (2026)
Beyond Tokens in Language Models: Interpreting Activations through Text Genre Chunks
by: Benito-Rodriguez, Éloïse, et al.
Published: (2025)
by: Benito-Rodriguez, Éloïse, et al.
Published: (2025)
Improving Cross-Lingual Phonetic Representation of Low-Resource Languages Through Language Similarity Analysis
by: Kim, Minu, et al.
Published: (2025)
by: Kim, Minu, et al.
Published: (2025)
Accelerating Monte-Carlo Tree Search with Optimized Posterior Policies
by: Frankston, Keith, et al.
Published: (2026)
by: Frankston, Keith, et al.
Published: (2026)
HADA: Human-AI Agent Decision Alignment Architecture
by: Pitkäranta, Tapio, et al.
Published: (2025)
by: Pitkäranta, Tapio, et al.
Published: (2025)
Practical Code RAG at Scale: Task-Aware Retrieval Design Choices under Compute Budgets
by: Galimzyanov, Timur, et al.
Published: (2025)
by: Galimzyanov, Timur, et al.
Published: (2025)
An End-to-End Approach for Korean Wakeword Systems with Speaker Authentication
by: Seo, Geonwoo
Published: (2025)
by: Seo, Geonwoo
Published: (2025)
The Exploration of Neural Collapse under Imbalanced Data
by: Liu, Haixia
Published: (2024)
by: Liu, Haixia
Published: (2024)
Holistic Optimal Label Selection for Robust Prompt Learning under Partial Labels
by: Zhao, Yaqi, et al.
Published: (2026)
by: Zhao, Yaqi, et al.
Published: (2026)
Generation of Musical Timbres using a Text-Guided Diffusion Model
by: Yuan, Weixuan, et al.
Published: (2025)
by: Yuan, Weixuan, et al.
Published: (2025)
Self-Improvement for Audio Large Language Model using Unlabeled Speech
by: Wang, Shaowen, et al.
Published: (2025)
by: Wang, Shaowen, et al.
Published: (2025)
MAIN-VC: Lightweight Speech Representation Disentanglement for One-shot Voice Conversion
by: Li, Pengcheng, et al.
Published: (2024)
by: Li, Pengcheng, et al.
Published: (2024)
Less Stress, More Privacy: Stress Detection on Anonymized Speech of Air Traffic Controllers
by: Viswanathan, Janaki, et al.
Published: (2025)
by: Viswanathan, Janaki, et al.
Published: (2025)
GraphCompNet: A Position-Aware Model for Predicting and Compensating Shape Deviations in 3D Printing
by: Lee, Juheon, et al.
Published: (2025)
by: Lee, Juheon, et al.
Published: (2025)
Breaking the Chains of Probability: Neutrosophic Logic as a New Framework for Epistemic Uncertainty in Large Language Models
by: Leyva-Vázquez, Maikel Yelandi, et al.
Published: (2026)
by: Leyva-Vázquez, Maikel Yelandi, et al.
Published: (2026)
Delayed Fusion: Integrating Large Language Models into First-Pass Decoding in End-to-end Speech Recognition
by: Hori, Takaaki, et al.
Published: (2025)
by: Hori, Takaaki, et al.
Published: (2025)
Quantum-Enhanced Analysis and Grading of Vocal Performance
by: Agarwal, Rohan
Published: (2025)
by: Agarwal, Rohan
Published: (2025)
Thaka at KSAA-2026 Task 2: Regularized Fine-Tuning for Arabic Speech Diacritization
by: Alamr, Meshal, et al.
Published: (2026)
by: Alamr, Meshal, et al.
Published: (2026)
Predicting Upcoming Stuttering Events from Three-Second Audio: Stratified Evaluation Reveals Severity-Selective Precursors, and the Model Deploys Fully On-Device
by: Kozak, Nazar
Published: (2026)
by: Kozak, Nazar
Published: (2026)
SeamlessEdit: Background Noise Aware Zero-Shot Speech Editing with in-Context Enhancement
by: Chen, Kuan-Yu, et al.
Published: (2025)
by: Chen, Kuan-Yu, et al.
Published: (2025)
SFMS-ALR: Script-First Multilingual Speech Synthesis with Adaptive Locale Resolution
by: Donepudi, Dharma Teja
Published: (2025)
by: Donepudi, Dharma Teja
Published: (2025)
Taming Audio VAEs via Target-KL Regularization
by: Seetharaman, Prem, et al.
Published: (2026)
by: Seetharaman, Prem, et al.
Published: (2026)
Machine Learning Framework for Audio-Based Content Evaluation using MFCC, Chroma, Spectral Contrast, and Temporal Feature Engineering
by: Aristorenas, Aris J.
Published: (2024)
by: Aristorenas, Aris J.
Published: (2024)
Emotional Voice Messages (EMOVOME) database: emotion recognition in spontaneous voice messages
by: Zaragozá, Lucía Gómez, et al.
Published: (2024)
by: Zaragozá, Lucía Gómez, et al.
Published: (2024)
Prevailing Research Areas for Music AI in the Era of Foundation Models
by: Wei, Megan, et al.
Published: (2024)
by: Wei, Megan, et al.
Published: (2024)
PerceiverS: A Multi-Scale Perceiver with Effective Segmentation for Long-Term Expressive Symbolic Music Generation
by: Yi, Yungang, et al.
Published: (2024)
by: Yi, Yungang, et al.
Published: (2024)
Revisiting SSL for sound event detection: complementary fusion and adaptive post-processing
by: Cui, Hanfang, et al.
Published: (2025)
by: Cui, Hanfang, et al.
Published: (2025)
Joint Estimation of Piano Dynamics and Metrical Structure with a Multi-task Multi-Scale Network
by: He, Zhanhong, et al.
Published: (2025)
by: He, Zhanhong, et al.
Published: (2025)
Similar Items
-
Towards Objective Gastrointestinal Auscultation: Automated Segmentation and Annotation of Bowel Sound Patterns
by: Mansour, Zahra, et al.
Published: (2026) -
Benchmarking machine learning for bowel sound pattern classification from tabular features to pretrained models
by: Mansour, Zahra, et al.
Published: (2025) -
Optimizing Hospital Capacity During Pandemics: A Dual-Component Framework for Strategic Patient Relocation
by: Tabatabaee, Sadaf, et al.
Published: (2026) -
Cycling Race Time Prediction: A Personalized Machine Learning Approach Using Route Topology and Training Load
by: Moreno, Francisco Aguilera
Published: (2026) -
OverNaN: NaN-Aware Oversampling for Imbalanced Learning with Meaningful Missingness
by: Barnard, Amanda S
Published: (2026)