Peransformer: Improving Low-informed Expressive Performance Rendering with Score-aware Discriminator
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | He, Xian, Zeng, Wei, Wang, Ye |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Disentangling Score Content and Performance Style for Joint Piano Rendering and Transcription
von: Zeng, Wei, et al.
Veröffentlicht: (2025)
von: Zeng, Wei, et al.
Veröffentlicht: (2025)
End-to-End Real-World Polyphonic Piano Audio-to-Score Transcription with Hierarchical Decoding
von: Zeng, Wei, et al.
Veröffentlicht: (2024)
von: Zeng, Wei, et al.
Veröffentlicht: (2024)
Towards An Integrated Approach for Expressive Piano Performance Synthesis from Music Scores
von: Tang, Jingjing, et al.
Veröffentlicht: (2025)
von: Tang, Jingjing, et al.
Veröffentlicht: (2025)
Revisiting and Improving Scoring Fusion for Spoofing-aware Speaker Verification Using Compositional Data Analysis
von: Wang, Xin, et al.
Veröffentlicht: (2024)
von: Wang, Xin, et al.
Veröffentlicht: (2024)
CosyAudio: Improving Audio Generation with Confidence Scores and Synthetic Captions
von: Zhu, Xinfa, et al.
Veröffentlicht: (2025)
von: Zhu, Xinfa, et al.
Veröffentlicht: (2025)
Discriminative-Generative Target Speaker Extraction with Decoder-Only Language Models
von: Zeng, Bang, et al.
Veröffentlicht: (2026)
von: Zeng, Bang, et al.
Veröffentlicht: (2026)
A Study on Synthesizing Expressive Violin Performances: Approaches and Comparisons
von: Hung, Tzu-Yun, et al.
Veröffentlicht: (2024)
von: Hung, Tzu-Yun, et al.
Veröffentlicht: (2024)
Sounding Out Reconstruction Error-Based Evaluation of Generative Models of Expressive Performance
von: Peter, Silvan David, et al.
Veröffentlicht: (2023)
von: Peter, Silvan David, et al.
Veröffentlicht: (2023)
Improving Unsupervised Clean-to-Rendered Guitar Tone Transformation Using GANs and Integrated Unaligned Clean Data
von: Chen, Yu-Hua, et al.
Veröffentlicht: (2024)
von: Chen, Yu-Hua, et al.
Veröffentlicht: (2024)
DiffSound: Differentiable Modal Sound Rendering and Inverse Rendering for Diverse Inference Tasks
von: Jin, Xutong, et al.
Veröffentlicht: (2024)
von: Jin, Xutong, et al.
Veröffentlicht: (2024)
ELGAR: Expressive Cello Performance Motion Generation for Audio Rendition
von: Qiu, Zhiping, et al.
Veröffentlicht: (2025)
von: Qiu, Zhiping, et al.
Veröffentlicht: (2025)
Improving Acoustic Scene Classification in Low-Resource Conditions
von: Chen, Zhi, et al.
Veröffentlicht: (2024)
von: Chen, Zhi, et al.
Veröffentlicht: (2024)
CONTUNER: Singing Voice Beautifying with Pitch and Expressiveness Condition
von: Wang, Jianzong, et al.
Veröffentlicht: (2024)
von: Wang, Jianzong, et al.
Veröffentlicht: (2024)
Hierarchical Control of Emotion Rendering in Speech Synthesis
von: Inoue, Sho, et al.
Veröffentlicht: (2024)
von: Inoue, Sho, et al.
Veröffentlicht: (2024)
Expressivity-aware Music Performance Retrieval using Mid-level Perceptual Features and Emotion Word Embeddings
von: Chowdhury, Shreyan, et al.
Veröffentlicht: (2024)
von: Chowdhury, Shreyan, et al.
Veröffentlicht: (2024)
Comparative Analysis Of Discriminative Deep Learning-Based Noise Reduction Methods In Low SNR Scenarios
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2024)
von: Shetu, Shrishti Saha, et al.
Veröffentlicht: (2024)
MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling
von: Tang, Jingjing, et al.
Veröffentlicht: (2025)
von: Tang, Jingjing, et al.
Veröffentlicht: (2025)
Towards Expressive Zero-Shot Speech Synthesis with Hierarchical Prosody Modeling
von: Jiang, Yuepeng, et al.
Veröffentlicht: (2024)
von: Jiang, Yuepeng, et al.
Veröffentlicht: (2024)
Melodic and Metrical Elements of Expressiveness in Hindustani Vocal Music
von: Bhake, Yash, et al.
Veröffentlicht: (2025)
von: Bhake, Yash, et al.
Veröffentlicht: (2025)
Learning Expressive Disentangled Speech Representations with Soft Speech Units and Adversarial Style Augmentation
von: Deng, Yimin, et al.
Veröffentlicht: (2024)
von: Deng, Yimin, et al.
Veröffentlicht: (2024)
MSceneSpeech: A Multi-Scene Speech Dataset For Expressive Speech Synthesis
von: Yang, Qian, et al.
Veröffentlicht: (2024)
von: Yang, Qian, et al.
Veröffentlicht: (2024)
Score-Informed Transformer for Refining MIDI Velocity in Automatic Music Transcription
von: He, Zhanhong, et al.
Veröffentlicht: (2025)
von: He, Zhanhong, et al.
Veröffentlicht: (2025)
MelodyT5: A Unified Score-to-Score Transformer for Symbolic Music Processing
von: Wu, Shangda, et al.
Veröffentlicht: (2024)
von: Wu, Shangda, et al.
Veröffentlicht: (2024)
Acoustic Volume Rendering for Neural Impulse Response Fields
von: Lan, Zitong, et al.
Veröffentlicht: (2024)
von: Lan, Zitong, et al.
Veröffentlicht: (2024)
Expressive paragraph text-to-speech synthesis with multi-step variational autoencoder
von: Li, Xuyuan, et al.
Veröffentlicht: (2023)
von: Li, Xuyuan, et al.
Veröffentlicht: (2023)
MM-TTS: Multi-modal Prompt based Style Transfer for Expressive Text-to-Speech Synthesis
von: Guan, Wenhao, et al.
Veröffentlicht: (2023)
von: Guan, Wenhao, et al.
Veröffentlicht: (2023)
Pureformer-VC: Non-parallel Voice Conversion with Pure Stylized Transformer Blocks and Triplet Discriminative Training
von: Yao, Wenhan, et al.
Veröffentlicht: (2025)
von: Yao, Wenhan, et al.
Veröffentlicht: (2025)
Ambisonics Binaural Rendering via Masked Magnitude Least Squares
von: Berebi, Or, et al.
Veröffentlicht: (2025)
von: Berebi, Or, et al.
Veröffentlicht: (2025)
Expressive MIDI-format Piano Performance Generation
von: Liu, Jingwei
Veröffentlicht: (2024)
von: Liu, Jingwei
Veröffentlicht: (2024)
DrawSpeech: Expressive Speech Synthesis Using Prosodic Sketches as Control Conditions
von: Chen, Weidong, et al.
Veröffentlicht: (2025)
von: Chen, Weidong, et al.
Veröffentlicht: (2025)
Enhancing Expressive Voice Conversion with Discrete Pitch-Conditioned Flow Matching Model
von: Zuo, Jialong, et al.
Veröffentlicht: (2025)
von: Zuo, Jialong, et al.
Veröffentlicht: (2025)
Boosting Multi-Speaker Expressive Speech Synthesis with Semi-supervised Contrastive Learning
von: Zhu, Xinfa, et al.
Veröffentlicht: (2023)
von: Zhu, Xinfa, et al.
Veröffentlicht: (2023)
Score-informed Music Source Separation: Improving Synthetic-to-real Generalization in Classical Music
von: Tunturi, Eetu, et al.
Veröffentlicht: (2025)
von: Tunturi, Eetu, et al.
Veröffentlicht: (2025)
Chain-Talker: Chain Understanding and Rendering for Empathetic Conversational Speech Synthesis
von: Hu, Yifan, et al.
Veröffentlicht: (2025)
von: Hu, Yifan, et al.
Veröffentlicht: (2025)
SHroom: A Python Framework for Ambisonics Room Acoustics Simulation and Binaural Rendering
von: Gayer, Yhonatan
Veröffentlicht: (2026)
von: Gayer, Yhonatan
Veröffentlicht: (2026)
SaD: A Scenario-Aware Discriminator for Speech Enhancement
von: Yuan, Xihao, et al.
Veröffentlicht: (2025)
von: Yuan, Xihao, et al.
Veröffentlicht: (2025)
A Hybrid Discriminative and Generative System for Universal Speech Enhancement
von: Liu, Yinghao, et al.
Veröffentlicht: (2026)
von: Liu, Yinghao, et al.
Veröffentlicht: (2026)
Streaming Keyword Spotting Boosted by Cross-layer Discrimination Consistency
von: Xi, Yu, et al.
Veröffentlicht: (2024)
von: Xi, Yu, et al.
Veröffentlicht: (2024)
Converting Anyone's Voice: End-to-End Expressive Voice Conversion with a Conditional Diffusion Model
von: Du, Zongyang, et al.
Veröffentlicht: (2024)
von: Du, Zongyang, et al.
Veröffentlicht: (2024)
VoxInstruct: Expressive Human Instruction-to-Speech Generation with Unified Multilingual Codec Language Modelling
von: Zhou, Yixuan, et al.
Veröffentlicht: (2024)
von: Zhou, Yixuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Disentangling Score Content and Performance Style for Joint Piano Rendering and Transcription
von: Zeng, Wei, et al.
Veröffentlicht: (2025) -
End-to-End Real-World Polyphonic Piano Audio-to-Score Transcription with Hierarchical Decoding
von: Zeng, Wei, et al.
Veröffentlicht: (2024) -
Towards An Integrated Approach for Expressive Piano Performance Synthesis from Music Scores
von: Tang, Jingjing, et al.
Veröffentlicht: (2025) -
Revisiting and Improving Scoring Fusion for Spoofing-aware Speaker Verification Using Compositional Data Analysis
von: Wang, Xin, et al.
Veröffentlicht: (2024) -
CosyAudio: Improving Audio Generation with Confidence Scores and Synthetic Captions
von: Zhu, Xinfa, et al.
Veröffentlicht: (2025)