PANDORA: Diffusion Policy Learning for Dexterous Robotic Piano Playing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Yanjia, Li, Renjie, Tu, Zhengzhong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Review on Sound Source Localization in Robotics: Focusing on Deep Learning Methods
von: Jalayer, Reza, et al.
Veröffentlicht: (2025)
von: Jalayer, Reza, et al.
Veröffentlicht: (2025)
Acoustics-specific Piano Velocity Estimation
von: Simonetta, Federico, et al.
Veröffentlicht: (2022)
von: Simonetta, Federico, et al.
Veröffentlicht: (2022)
PBSCR: The Piano Bootleg Score Composer Recognition Dataset
von: Jain, Arhan, et al.
Veröffentlicht: (2024)
von: Jain, Arhan, et al.
Veröffentlicht: (2024)
Exploring Transformer-Based Music Overpainting for Jazz Piano Variations
von: Row, Eleanor, et al.
Veröffentlicht: (2024)
von: Row, Eleanor, et al.
Veröffentlicht: (2024)
A Data-Driven Analysis of Robust Automatic Piano Transcription
von: Edwards, Drew, et al.
Veröffentlicht: (2024)
von: Edwards, Drew, et al.
Veröffentlicht: (2024)
End-to-end Piano Performance-MIDI to Score Conversion with Transformers
von: Beyer, Tim, et al.
Veröffentlicht: (2024)
von: Beyer, Tim, et al.
Veröffentlicht: (2024)
Exploring System Adaptations For Minimum Latency Real-Time Piano Transcription
von: Hu, Patricia, et al.
Veröffentlicht: (2025)
von: Hu, Patricia, et al.
Veröffentlicht: (2025)
Scoring Time Intervals using Non-Hierarchical Transformer For Automatic Piano Transcription
von: Yan, Yujia, et al.
Veröffentlicht: (2024)
von: Yan, Yujia, et al.
Veröffentlicht: (2024)
Towards Efficient and Real-Time Piano Transcription Using Neural Autoregressive Models
von: Kwon, Taegyun, et al.
Veröffentlicht: (2024)
von: Kwon, Taegyun, et al.
Veröffentlicht: (2024)
A Traditional Approach to Symbolic Piano Continuation
von: Zhou-Zheng, Christian, et al.
Veröffentlicht: (2025)
von: Zhou-Zheng, Christian, et al.
Veröffentlicht: (2025)
From Sound to Setting: AI-Based Equalizer Parameter Prediction for Piano Tone Replication
von: Yu, Song-Ze
Veröffentlicht: (2025)
von: Yu, Song-Ze
Veröffentlicht: (2025)
AMT-APC: Automatic Piano Cover by Fine-Tuning an Automatic Music Transcription Model
von: Komiya, Kazuma, et al.
Veröffentlicht: (2024)
von: Komiya, Kazuma, et al.
Veröffentlicht: (2024)
Etude: Piano Cover Generation with a Three-Stage Approach -- Extract, strucTUralize, and DEcode
von: Chen, Tse-Yang, et al.
Veröffentlicht: (2025)
von: Chen, Tse-Yang, et al.
Veröffentlicht: (2025)
JAZZVAR: A Dataset of Variations found within Solo Piano Performances of Jazz Standards for Music Overpainting
von: Row, Eleanor, et al.
Veröffentlicht: (2023)
von: Row, Eleanor, et al.
Veröffentlicht: (2023)
Siamese Residual Neural Network for Musical Shape Evaluation in Piano Performance Assessment
von: Li, Xiaoquan, et al.
Veröffentlicht: (2024)
von: Li, Xiaoquan, et al.
Veröffentlicht: (2024)
Deconstructing Jazz Piano Style Using Machine Learning
von: Cheston, Huw, et al.
Veröffentlicht: (2025)
von: Cheston, Huw, et al.
Veröffentlicht: (2025)
Scaling Self-Supervised Representation Learning for Symbolic Piano Performance
von: Bradshaw, Louis, et al.
Veröffentlicht: (2025)
von: Bradshaw, Louis, et al.
Veröffentlicht: (2025)
BERT-like Pre-training for Symbolic Piano Music Classification Tasks
von: Chou, Yi-Hui, et al.
Veröffentlicht: (2021)
von: Chou, Yi-Hui, et al.
Veröffentlicht: (2021)
D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription
von: Kim, Hounsu, et al.
Veröffentlicht: (2025)
von: Kim, Hounsu, et al.
Veröffentlicht: (2025)
Long-Term, Store-Front Robotics: Interactive Music for Robotic Arm, Caxixi and Frame Drums
von: Savery, Richard, et al.
Veröffentlicht: (2024)
von: Savery, Richard, et al.
Veröffentlicht: (2024)
Piano Transcription by Hierarchical Language Modeling with Pretrained Roll-based Encoders
von: Li, Dichucheng, et al.
Veröffentlicht: (2025)
von: Li, Dichucheng, et al.
Veröffentlicht: (2025)
Evaluating Speech-in-Speech Perception via a Humanoid Robot
von: Meyer, Luke, et al.
Veröffentlicht: (2023)
von: Meyer, Luke, et al.
Veröffentlicht: (2023)
Theoretical Framework for the Optimization of Microphone Array Configuration for Humanoid Robot Audition
von: Tourbabin, Vladimir, et al.
Veröffentlicht: (2024)
von: Tourbabin, Vladimir, et al.
Veröffentlicht: (2024)
Single-Microphone-Based Sound Source Localization for Mobile Robots in Reverberant Environments
von: Wang, Jiang, et al.
Veröffentlicht: (2025)
von: Wang, Jiang, et al.
Veröffentlicht: (2025)
Direction of Arrival Estimation Using Microphone Array Processing for Moving Humanoid Robots
von: Tourbabin, Vladimir, et al.
Veröffentlicht: (2024)
von: Tourbabin, Vladimir, et al.
Veröffentlicht: (2024)
Symbolic Music Generation with Non-Differentiable Rule Guided Diffusion
von: Huang, Yujia, et al.
Veröffentlicht: (2024)
von: Huang, Yujia, et al.
Veröffentlicht: (2024)
Sim2Real Transfer for Audio-Visual Navigation with Frequency-Adaptive Acoustic Field Prediction
von: Chen, Changan, et al.
Veröffentlicht: (2024)
von: Chen, Changan, et al.
Veröffentlicht: (2024)
Noise-aware Speech Enhancement using Diffusion Probabilistic Model
von: Hu, Yuchen, et al.
Veröffentlicht: (2023)
von: Hu, Yuchen, et al.
Veröffentlicht: (2023)
Extract and Diffuse: Latent Integration for Improved Diffusion-based Speech and Vocal Enhancement
von: Yang, Yudong, et al.
Veröffentlicht: (2024)
von: Yang, Yudong, et al.
Veröffentlicht: (2024)
Diffusion Models for Audio Restoration
von: Lemercier, Jean-Marie, et al.
Veröffentlicht: (2024)
von: Lemercier, Jean-Marie, et al.
Veröffentlicht: (2024)
Subtractive Training for Music Stem Insertion using Latent Diffusion Models
von: Villa-Renteria, Ivan, et al.
Veröffentlicht: (2024)
von: Villa-Renteria, Ivan, et al.
Veröffentlicht: (2024)
TaDiCodec: Text-aware Diffusion Speech Tokenizer for Speech Language Modeling
von: Wang, Yuancheng, et al.
Veröffentlicht: (2025)
von: Wang, Yuancheng, et al.
Veröffentlicht: (2025)
Active Listener: Continuous Generation of Listener's Head Motion Response in Dyadic Interactions
von: Ghosh, Bishal, et al.
Veröffentlicht: (2024)
von: Ghosh, Bishal, et al.
Veröffentlicht: (2024)
Diffusion Buffer for Online Generative Speech Enhancement
von: Lay, Bunlong, et al.
Veröffentlicht: (2025)
von: Lay, Bunlong, et al.
Veröffentlicht: (2025)
An Analysis of the Variance of Diffusion-based Speech Enhancement
von: Lay, Bunlong, et al.
Veröffentlicht: (2024)
von: Lay, Bunlong, et al.
Veröffentlicht: (2024)
Fast Timing-Conditioned Latent Audio Diffusion
von: Evans, Zach, et al.
Veröffentlicht: (2024)
von: Evans, Zach, et al.
Veröffentlicht: (2024)
Zero-shot Voice Conversion with Diffusion Transformers
von: Liu, Songting
Veröffentlicht: (2024)
von: Liu, Songting
Veröffentlicht: (2024)
Multi-Source Music Generation with Latent Diffusion
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2024)
von: Xu, Zhongweiyang, et al.
Veröffentlicht: (2024)
Bass Accompaniment Generation via Latent Diffusion
von: Pasini, Marco, et al.
Veröffentlicht: (2024)
von: Pasini, Marco, et al.
Veröffentlicht: (2024)
Exploring Sentence Type Effects on the Lombard Effect and Intelligibility Enhancement: A Comparative Study of Natural and Grid Sentences
von: Chen, Hongyang, et al.
Veröffentlicht: (2023)
von: Chen, Hongyang, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
A Review on Sound Source Localization in Robotics: Focusing on Deep Learning Methods
von: Jalayer, Reza, et al.
Veröffentlicht: (2025) -
Acoustics-specific Piano Velocity Estimation
von: Simonetta, Federico, et al.
Veröffentlicht: (2022) -
PBSCR: The Piano Bootleg Score Composer Recognition Dataset
von: Jain, Arhan, et al.
Veröffentlicht: (2024) -
Exploring Transformer-Based Music Overpainting for Jazz Piano Variations
von: Row, Eleanor, et al.
Veröffentlicht: (2024) -
A Data-Driven Analysis of Robust Automatic Piano Transcription
von: Edwards, Drew, et al.
Veröffentlicht: (2024)