HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Hu, Teng, Yu, Zhentao, Zhou, Zhengguang, Liang, Sen, Zhou, Yuan, Lin, Qin, Lu, Qinglin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters
por: Chen, Yi, et al.
Publicado: (2025)
por: Chen, Yi, et al.
Publicado: (2025)
OmniV2V: Versatile Video Generation and Editing via Dynamic Content Manipulation
por: Liang, Sen, et al.
Publicado: (2025)
por: Liang, Sen, et al.
Publicado: (2025)
HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation
por: Huang, Ziyao, et al.
Publicado: (2025)
por: Huang, Ziyao, et al.
Publicado: (2025)
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement
por: Hu, Teng, et al.
Publicado: (2025)
por: Hu, Teng, et al.
Publicado: (2025)
Harmony: Harmonizing Audio and Video Generation through Cross-Task Synergy
por: Hu, Teng, et al.
Publicado: (2025)
por: Hu, Teng, et al.
Publicado: (2025)
HunyuanPortrait: Implicit Condition Control for Enhanced Portrait Animation
por: Xu, Zunnan, et al.
Publicado: (2025)
por: Xu, Zunnan, et al.
Publicado: (2025)
Hunyuan-GameCraft: High-dynamic Interactive Game Video Generation with Hybrid History Condition
por: Li, Jiaqi, et al.
Publicado: (2025)
por: Li, Jiaqi, et al.
Publicado: (2025)
HunyuanVideo: A Systematic Framework For Large Video Generative Models
por: Kong, Weijie, et al.
Publicado: (2024)
por: Kong, Weijie, et al.
Publicado: (2024)
Making Avatars Interact: Towards Text-Driven Human-Object Interaction for Controllable Talking Avatars
por: Zhang, Youliang, et al.
Publicado: (2026)
por: Zhang, Youliang, et al.
Publicado: (2026)
CustomCrafter: Customized Video Generation with Preserving Motion and Concept Composition Abilities
por: Wu, Tao, et al.
Publicado: (2024)
por: Wu, Tao, et al.
Publicado: (2024)
CustomVideo: Customizing Text-to-Video Generation with Multiple Subjects
por: Wang, Zhao, et al.
Publicado: (2024)
por: Wang, Zhao, et al.
Publicado: (2024)
Omni-Customizer: End-to-End MultiModal Customization for Joint Audio-Video Generation
por: Chen, Yuheng, et al.
Publicado: (2026)
por: Chen, Yuheng, et al.
Publicado: (2026)
Still-Moving: Customized Video Generation without Customized Video Data
por: Chefer, Hila, et al.
Publicado: (2024)
por: Chefer, Hila, et al.
Publicado: (2024)
HunyuanVideo 1.5 Technical Report
por: Wu, Bing, et al.
Publicado: (2025)
por: Wu, Bing, et al.
Publicado: (2025)
Customize-A-Video: One-Shot Motion Customization of Text-to-Video Diffusion Models
por: Ren, Yixuan, et al.
Publicado: (2024)
por: Ren, Yixuan, et al.
Publicado: (2024)
EffectMaker: Unifying Reasoning and Generation for Customized Visual Effect Creation
por: Yang, Shiyuan, et al.
Publicado: (2026)
por: Yang, Shiyuan, et al.
Publicado: (2026)
CustomTTT: Motion and Appearance Customized Video Generation via Test-Time Training
por: Bi, Xiuli, et al.
Publicado: (2024)
por: Bi, Xiuli, et al.
Publicado: (2024)
OmniVCus: Feedforward Subject-driven Video Customization with Multimodal Control Conditions
por: Cai, Yuanhao, et al.
Publicado: (2025)
por: Cai, Yuanhao, et al.
Publicado: (2025)
SUGAR: Subject-Driven Video Customization in a Zero-Shot Manner
por: Zhou, Yufan, et al.
Publicado: (2024)
por: Zhou, Yufan, et al.
Publicado: (2024)
CustomVideoX: 3D Reference Attention Driven Dynamic Adaptation for Zero-Shot Customized Video Diffusion Transformers
por: She, D., et al.
Publicado: (2025)
por: She, D., et al.
Publicado: (2025)
UniAVGen: Unified Audio and Video Generation with Asymmetric Cross-Modal Interactions
por: Zhang, Guozhen, et al.
Publicado: (2025)
por: Zhang, Guozhen, et al.
Publicado: (2025)
Hunyuan-Game: Industrial-grade Intelligent Game Creation Model
por: Li, Ruihuang, et al.
Publicado: (2025)
por: Li, Ruihuang, et al.
Publicado: (2025)
MotionBooth: Motion-Aware Customized Text-to-Video Generation
por: Wu, Jianzong, et al.
Publicado: (2024)
por: Wu, Jianzong, et al.
Publicado: (2024)
Motion Inversion for Video Customization
por: Wang, Luozhou, et al.
Publicado: (2024)
por: Wang, Luozhou, et al.
Publicado: (2024)
Customization Assistant for Text-to-image Generation
por: Zhou, Yufan, et al.
Publicado: (2023)
por: Zhou, Yufan, et al.
Publicado: (2023)
ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts
por: Ge, Yuying, et al.
Publicado: (2025)
por: Ge, Yuying, et al.
Publicado: (2025)
Non-confusing Generation of Customized Concepts in Diffusion Models
por: Lin, Wang, et al.
Publicado: (2024)
por: Lin, Wang, et al.
Publicado: (2024)
Multimodal LLMs as Customized Reward Models for Text-to-Image Generation
por: Zhou, Shijie, et al.
Publicado: (2025)
por: Zhou, Shijie, et al.
Publicado: (2025)
FlexIP: Dynamic Control of Preservation and Personality for Customized Image Generation
por: Huang, Linyan, et al.
Publicado: (2025)
por: Huang, Linyan, et al.
Publicado: (2025)
Hunyuan-GameCraft-2: Instruction-following Interactive Game World Model
por: Tang, Junshu, et al.
Publicado: (2025)
por: Tang, Junshu, et al.
Publicado: (2025)
PhyCustom: Towards Realistic Physical Customization in Text-to-Image Generation
por: Wu, Fan, et al.
Publicado: (2025)
por: Wu, Fan, et al.
Publicado: (2025)
BridgeIV: Bridging Customized Image and Video Generation through Test-Time Autoregressive Identity Propagation
por: Hu, Panwen, et al.
Publicado: (2025)
por: Hu, Panwen, et al.
Publicado: (2025)
Arbitrary Generative Video Interpolation
por: Zhang, Guozhen, et al.
Publicado: (2025)
por: Zhang, Guozhen, et al.
Publicado: (2025)
DreamVideo-2: Zero-Shot Subject-Driven Video Customization with Precise Motion Control
por: Wei, Yujie, et al.
Publicado: (2024)
por: Wei, Yujie, et al.
Publicado: (2024)
CoMo: Compositional Motion Customization for Text-to-Video Generation
por: Xu, Youcan, et al.
Publicado: (2025)
por: Xu, Youcan, et al.
Publicado: (2025)
USV: Unified Sparsification for Accelerating Video Diffusion Models
por: Wu, Xinjian, et al.
Publicado: (2025)
por: Wu, Xinjian, et al.
Publicado: (2025)
HunyuanVideo-Foley: Multimodal Diffusion with Representation Alignment for High-Fidelity Foley Audio Generation
por: Shan, Sizhe, et al.
Publicado: (2025)
por: Shan, Sizhe, et al.
Publicado: (2025)
DisenStudio: Customized Multi-subject Text-to-Video Generation with Disentangled Spatial Control
por: Chen, Hong, et al.
Publicado: (2024)
por: Chen, Hong, et al.
Publicado: (2024)
Tora2: Motion and Appearance Customized Diffusion Transformer for Multi-Entity Video Generation
por: Zhang, Zhenghao, et al.
Publicado: (2025)
por: Zhang, Zhenghao, et al.
Publicado: (2025)
DreamRelation: Relation-Centric Video Customization
por: Wei, Yujie, et al.
Publicado: (2025)
por: Wei, Yujie, et al.
Publicado: (2025)
Ejemplares similares
-
HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters
por: Chen, Yi, et al.
Publicado: (2025) -
OmniV2V: Versatile Video Generation and Editing via Dynamic Content Manipulation
por: Liang, Sen, et al.
Publicado: (2025) -
HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation
por: Huang, Ziyao, et al.
Publicado: (2025) -
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement
por: Hu, Teng, et al.
Publicado: (2025) -
Harmony: Harmonizing Audio and Video Generation through Cross-Task Synergy
por: Hu, Teng, et al.
Publicado: (2025)