Harmonizing Pixels and Melodies: Maestro-Guided Film Score Generation and Composition Style Transfer
Fuente:
arXiv
Guardado en:
| Autores principales: | Qi, F., Ni, L., Xu, C. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
StyleSpeaker: Audio-Enhanced Fine-Grained Style Modeling for Speech-Driven 3D Facial Animation
por: Yang, An, et al.
Publicado: (2025)
por: Yang, An, et al.
Publicado: (2025)
Training-and-Prompt-Free General Painterly Harmonization via Zero-Shot Disentenglement on Style and Content References
por: Hsiao, Teng-Fang, et al.
Publicado: (2024)
por: Hsiao, Teng-Fang, et al.
Publicado: (2024)
PixelThink: Towards Efficient Chain-of-Pixel Reasoning
por: Wang, Song, et al.
Publicado: (2025)
por: Wang, Song, et al.
Publicado: (2025)
DiffuseST: Unleashing the Capability of the Diffusion Model for Style Transfer
por: Hu, Ying, et al.
Publicado: (2024)
por: Hu, Ying, et al.
Publicado: (2024)
Neural Style Transfer for Audio Spectograms
por: Verma, Prateek, et al.
Publicado: (2018)
por: Verma, Prateek, et al.
Publicado: (2018)
PixelatedScatter: Arbitrary-level Visual Abstraction for Large-scale Multiclass Scatterplots
por: Guo, Ziheng, et al.
Publicado: (2025)
por: Guo, Ziheng, et al.
Publicado: (2025)
MusicWeaver: Composer-Style Structural Editing and Minute-Scale Coherent Music Generation
por: Wang, Xuanchen, et al.
Publicado: (2025)
por: Wang, Xuanchen, et al.
Publicado: (2025)
Inter-Frame Coding for Dynamic Meshes via Coarse-to-Fine Anchor Mesh Generation
por: Huang, He, et al.
Publicado: (2024)
por: Huang, He, et al.
Publicado: (2024)
Harnessing the Latent Diffusion Model for Training-Free Image Style Transfer
por: Masui, Kento, et al.
Publicado: (2024)
por: Masui, Kento, et al.
Publicado: (2024)
Controllable Dance Generation with Style-Guided Motion Diffusion
por: Wang, Hongsong, et al.
Publicado: (2024)
por: Wang, Hongsong, et al.
Publicado: (2024)
Controllable Text-to-Speech Synthesis with Masked-Autoencoded Style-Rich Representation
por: Wang, Yongqi, et al.
Publicado: (2025)
por: Wang, Yongqi, et al.
Publicado: (2025)
StePO-Rec: Towards Personalized Outfit Styling Assistant via Knowledge-Guided Multi-Step Reasoning
por: Bi, Yuxi, et al.
Publicado: (2025)
por: Bi, Yuxi, et al.
Publicado: (2025)
Detecting Notational Errors in Digital Music Scores
por: Léo, Géré, et al.
Publicado: (2025)
por: Léo, Géré, et al.
Publicado: (2025)
Disentangling Score Content and Performance Style for Joint Piano Rendering and Transcription
por: Zeng, Wei, et al.
Publicado: (2025)
por: Zeng, Wei, et al.
Publicado: (2025)
Latent Feature-Guided Conditional Diffusion for Generative Image Semantic Communication
por: Chen, Zehao, et al.
Publicado: (2025)
por: Chen, Zehao, et al.
Publicado: (2025)
Harmonizing Attention: Training-free Texture-aware Geometry Transfer
por: Ikuta, Eito, et al.
Publicado: (2024)
por: Ikuta, Eito, et al.
Publicado: (2024)
Contribution-Guided Asymmetric Learning for Robust Multimodal Fusion under Imbalance and Noise
por: Xu, Zijing, et al.
Publicado: (2025)
por: Xu, Zijing, et al.
Publicado: (2025)
Enhancing Expressiveness in Dance Generation via Integrating Frequency and Music Style Information
por: Huang, Qiaochu, et al.
Publicado: (2024)
por: Huang, Qiaochu, et al.
Publicado: (2024)
AeroLite: Tag-Guided Lightweight Generation of Aerial Image Captions
por: Zi, Xing, et al.
Publicado: (2025)
por: Zi, Xing, et al.
Publicado: (2025)
Multi-scale Attention Guided Pose Transfer
por: Roy, Prasun, et al.
Publicado: (2022)
por: Roy, Prasun, et al.
Publicado: (2022)
Enhancing Film Grain Coding in VVC: Improving Encoding Quality and Efficiency
por: Menon, Vignesh V, et al.
Publicado: (2024)
por: Menon, Vignesh V, et al.
Publicado: (2024)
Protégé: Learn and Generate Basic Makeup Styles with Generative Adversarial Networks (GANs)
por: Sii, Jia Wei, et al.
Publicado: (2024)
por: Sii, Jia Wei, et al.
Publicado: (2024)
MusicScore: A Dataset for Music Score Modeling and Generation
por: Lin, Yuheng, et al.
Publicado: (2024)
por: Lin, Yuheng, et al.
Publicado: (2024)
Gain of Grain: A Film Grain Handling Toolchain for VVC-based Open Implementations
por: Menon, Vignesh V, et al.
Publicado: (2024)
por: Menon, Vignesh V, et al.
Publicado: (2024)
LLM2Manim: Pedagogy-Aware AI Generation of STEM Animations
por: Joshi, Aastha, et al.
Publicado: (2026)
por: Joshi, Aastha, et al.
Publicado: (2026)
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation
por: Wu, Yi, et al.
Publicado: (2025)
por: Wu, Yi, et al.
Publicado: (2025)
Virbo: Multimodal Multilingual Avatar Video Generation in Digital Marketing
por: Zhang, Juan, et al.
Publicado: (2024)
por: Zhang, Juan, et al.
Publicado: (2024)
ChoreoMuse: Robust Music-to-Dance Video Generation with Style Transfer and Beat-Adherent Motion
por: Wang, Xuanchen, et al.
Publicado: (2025)
por: Wang, Xuanchen, et al.
Publicado: (2025)
Generative AI-enabled Mobile Tactical Multimedia Networks: Distribution, Generation, and Perception
por: Xu, Minrui, et al.
Publicado: (2024)
por: Xu, Minrui, et al.
Publicado: (2024)
MORE-R1: Guiding LVLM for Multimodal Object-Entity Relation Extraction via Stepwise Reasoning with Reinforcement Learning
por: Yuan, Xiang, et al.
Publicado: (2026)
por: Yuan, Xiang, et al.
Publicado: (2026)
HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer
por: Cai, Qi, et al.
Publicado: (2026)
por: Cai, Qi, et al.
Publicado: (2026)
SIG-Chat: Spatial Intent-Guided Conversational Gesture Generation Involving How, When and Where
por: Huang, Yiheng, et al.
Publicado: (2025)
por: Huang, Yiheng, et al.
Publicado: (2025)
An Inverse Partial Optimal Transport Framework for Music-guided Movie Trailer Generation
por: Wang, Yutong, et al.
Publicado: (2024)
por: Wang, Yutong, et al.
Publicado: (2024)
Design of a 5G Multimedia Broadcast Application Function Supporting Adaptive Error Recovery
por: Lentisco, C. M., et al.
Publicado: (2024)
por: Lentisco, C. M., et al.
Publicado: (2024)
Reducing Latency for Multimedia Broadcast Services Over Mobile Networks
por: Lentisco, C. M., et al.
Publicado: (2024)
por: Lentisco, C. M., et al.
Publicado: (2024)
Exploring Transferability of Multimodal Adversarial Samples for Vision-Language Pre-training Models with Contrastive Learning
por: Wang, Youze, et al.
Publicado: (2023)
por: Wang, Youze, et al.
Publicado: (2023)
ReSyncer: Rewiring Style-based Generator for Unified Audio-Visually Synced Facial Performer
por: Guan, Jiazhi, et al.
Publicado: (2024)
por: Guan, Jiazhi, et al.
Publicado: (2024)
CAMMSR: Category-Guided Attentive Mixture of Experts for Multimodal Sequential Recommendation
por: Xu, Jinfeng, et al.
Publicado: (2026)
por: Xu, Jinfeng, et al.
Publicado: (2026)
Diffusion Model-Based Size Variable Virtual Try-On Technology and Evaluation Method
por: Zhang, Shufang, et al.
Publicado: (2025)
por: Zhang, Shufang, et al.
Publicado: (2025)
HarmonyIQA: Pioneering Benchmark and Model for Image Harmonization Quality Assessment
por: Xu, Zitong, et al.
Publicado: (2025)
por: Xu, Zitong, et al.
Publicado: (2025)
Ejemplares similares
-
StyleSpeaker: Audio-Enhanced Fine-Grained Style Modeling for Speech-Driven 3D Facial Animation
por: Yang, An, et al.
Publicado: (2025) -
Training-and-Prompt-Free General Painterly Harmonization via Zero-Shot Disentenglement on Style and Content References
por: Hsiao, Teng-Fang, et al.
Publicado: (2024) -
PixelThink: Towards Efficient Chain-of-Pixel Reasoning
por: Wang, Song, et al.
Publicado: (2025) -
DiffuseST: Unleashing the Capability of the Diffusion Model for Style Transfer
por: Hu, Ying, et al.
Publicado: (2024) -
Neural Style Transfer for Audio Spectograms
por: Verma, Prateek, et al.
Publicado: (2018)