From Parts to Whole: A Unified Reference Framework for Controllable Human Image Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Huang, Zehuan, Fan, Hongxing, Wang, Lipeng, Sheng, Lu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
InterMoE: Individual-Specific 3D Human Interaction Generation via Dynamic Temporal-Selective MoE
di: Wang, Lipeng, et al.
Pubblicazione: (2025)
di: Wang, Lipeng, et al.
Pubblicazione: (2025)
Multi-Agent Amodal Completion: Direct Synthesis with Fine-Grained Semantic Guidance
di: Fan, Hongxing, et al.
Pubblicazione: (2025)
di: Fan, Hongxing, et al.
Pubblicazione: (2025)
Ouroboros3D: Image-to-3D Generation via 3D-aware Recursive Diffusion
di: Wen, Hao, et al.
Pubblicazione: (2024)
di: Wen, Hao, et al.
Pubblicazione: (2024)
SegviGen: Repurposing 3D Generative Model for Part Segmentation
di: Li, Lin, et al.
Pubblicazione: (2026)
di: Li, Lin, et al.
Pubblicazione: (2026)
MV-Adapter: Multi-view Consistent Image Generation Made Easy
di: Huang, Zehuan, et al.
Pubblicazione: (2024)
di: Huang, Zehuan, et al.
Pubblicazione: (2024)
UniMLVG: Unified Framework for Multi-view Long Video Generation with Comprehensive Control Capabilities for Autonomous Driving
di: Chen, Rui, et al.
Pubblicazione: (2024)
di: Chen, Rui, et al.
Pubblicazione: (2024)
Reasoning-Driven Amodal Completion: Collaborative Agents and Perceptual Evaluation
di: Fan, Hongxing, et al.
Pubblicazione: (2025)
di: Fan, Hongxing, et al.
Pubblicazione: (2025)
RealisHuman: A Two-Stage Approach for Refining Malformed Human Parts in Generated Images
di: Wang, Benzhi, et al.
Pubblicazione: (2024)
di: Wang, Benzhi, et al.
Pubblicazione: (2024)
Personalize Anything for Free with Diffusion Transformer
di: Feng, Haoran, et al.
Pubblicazione: (2025)
di: Feng, Haoran, et al.
Pubblicazione: (2025)
Repurposing 3D Generative Model for Autoregressive Layout Generation
di: Feng, Haoran, et al.
Pubblicazione: (2026)
di: Feng, Haoran, et al.
Pubblicazione: (2026)
UniCustom: Unified Visual Conditioning for Multi-Reference Image Generation
di: Xu, Yiyan, et al.
Pubblicazione: (2026)
di: Xu, Yiyan, et al.
Pubblicazione: (2026)
From Wardrobe to Canvas: Wardrobe Polyptych LoRA for Part-level Controllable Human Image Generation
di: Kim, Jeongho, et al.
Pubblicazione: (2025)
di: Kim, Jeongho, et al.
Pubblicazione: (2025)
A Unified and Controllable Framework for Layered Image Generation with Visual Effects
di: Yang, Jinrui, et al.
Pubblicazione: (2026)
di: Yang, Jinrui, et al.
Pubblicazione: (2026)
TELA: Text to Layer-wise 3D Clothed Human Generation
di: Dong, Junting, et al.
Pubblicazione: (2024)
di: Dong, Junting, et al.
Pubblicazione: (2024)
A Unified Transformer-Based Framework with Pretraining For Whole Body Grasping Motion Generation
di: Effendy, Edward, et al.
Pubblicazione: (2025)
di: Effendy, Edward, et al.
Pubblicazione: (2025)
MIDI: Multi-Instance Diffusion for Single Image to 3D Scene Generation
di: Huang, Zehuan, et al.
Pubblicazione: (2024)
di: Huang, Zehuan, et al.
Pubblicazione: (2024)
AnimaX: Animating the Inanimate in 3D with Joint Video-Pose Diffusion Models
di: Huang, Zehuan, et al.
Pubblicazione: (2025)
di: Huang, Zehuan, et al.
Pubblicazione: (2025)
Evaluating and Predicting Distorted Human Body Parts for Generated Images
di: Ma, Lu, et al.
Pubblicazione: (2025)
di: Ma, Lu, et al.
Pubblicazione: (2025)
RefHCM: A Unified Model for Referring Perceptions in Human-Centric Scenarios
di: Huang, Jie, et al.
Pubblicazione: (2024)
di: Huang, Jie, et al.
Pubblicazione: (2024)
From Part to Whole: 3D Generative World Model with an Adaptive Structural Hierarchy
di: Du, Bi'an, et al.
Pubblicazione: (2026)
di: Du, Bi'an, et al.
Pubblicazione: (2026)
OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation
di: Wang, Junke, et al.
Pubblicazione: (2024)
di: Wang, Junke, et al.
Pubblicazione: (2024)
Assessment of Multimodal Large Language Models in Alignment with Human Values
di: Shi, Zhelun, et al.
Pubblicazione: (2024)
di: Shi, Zhelun, et al.
Pubblicazione: (2024)
UniView: Enhancing Novel View Synthesis From A Single Image By Unifying Reference Features
di: Cui, Haowang, et al.
Pubblicazione: (2025)
di: Cui, Haowang, et al.
Pubblicazione: (2025)
MoGen: A Unified Collaborative Framework for Controllable Multi-Object Image Generation
di: Li, Yanfeng, et al.
Pubblicazione: (2026)
di: Li, Yanfeng, et al.
Pubblicazione: (2026)
Lumina-Image 2.0: A Unified and Efficient Image Generative Framework
di: Qin, Qi, et al.
Pubblicazione: (2025)
di: Qin, Qi, et al.
Pubblicazione: (2025)
RefDrop: Controllable Consistency in Image or Video Generation via Reference Feature Guidance
di: Fan, Jiaojiao, et al.
Pubblicazione: (2024)
di: Fan, Jiaojiao, et al.
Pubblicazione: (2024)
PixelRefer: A Unified Framework for Spatio-Temporal Object Referring with Arbitrary Granularity
di: Yuan, Yuqian, et al.
Pubblicazione: (2025)
di: Yuan, Yuqian, et al.
Pubblicazione: (2025)
DreamID-Omni: Unified Framework for Controllable Human-Centric Audio-Video Generation
di: Guo, Xu, et al.
Pubblicazione: (2026)
di: Guo, Xu, et al.
Pubblicazione: (2026)
TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation
di: Qu, Liao, et al.
Pubblicazione: (2024)
di: Qu, Liao, et al.
Pubblicazione: (2024)
UniGenDet: A Unified Generative-Discriminative Framework for Co-Evolutionary Image Generation and Generated Image Detection
di: Zhang, Yanran, et al.
Pubblicazione: (2026)
di: Zhang, Yanran, et al.
Pubblicazione: (2026)
MRStyle: A Unified Framework for Color Style Transfer with Multi-Modality Reference
di: Huang, Jiancheng, et al.
Pubblicazione: (2024)
di: Huang, Jiancheng, et al.
Pubblicazione: (2024)
Liquid: Language Models are Scalable and Unified Multi-modal Generators
di: Wu, Junfeng, et al.
Pubblicazione: (2024)
di: Wu, Junfeng, et al.
Pubblicazione: (2024)
A Unified Agentic Framework for Evaluating Conditional Image Generation
di: Wang, Jifang, et al.
Pubblicazione: (2025)
di: Wang, Jifang, et al.
Pubblicazione: (2025)
InfinityStar: Unified Spacetime AutoRegressive Modeling for Visual Generation
di: Liu, Jinlai, et al.
Pubblicazione: (2025)
di: Liu, Jinlai, et al.
Pubblicazione: (2025)
MultiRef: Controllable Image Generation with Multiple Visual References
di: Chen, Ruoxi, et al.
Pubblicazione: (2025)
di: Chen, Ruoxi, et al.
Pubblicazione: (2025)
BrainSegNet: A Novel Framework for Whole-Brain MRI Parcellation Enhanced by Large Models
di: Li, Yucheng, et al.
Pubblicazione: (2026)
di: Li, Yucheng, et al.
Pubblicazione: (2026)
EchoMotion: Unified Human Video and Motion Generation via Dual-Modality Diffusion Transformer
di: Yang, Yuxiao, et al.
Pubblicazione: (2025)
di: Yang, Yuxiao, et al.
Pubblicazione: (2025)
FuseAnyPart: Diffusion-Driven Facial Parts Swapping via Multiple Reference Images
di: Yu, Zheng, et al.
Pubblicazione: (2024)
di: Yu, Zheng, et al.
Pubblicazione: (2024)
OmniCamera: A Unified Framework for Multi-task Video Generation with Arbitrary Camera Control
di: Wang, Yukun, et al.
Pubblicazione: (2026)
di: Wang, Yukun, et al.
Pubblicazione: (2026)
ChatSearch: a Dataset and a Generative Retrieval Model for General Conversational Image Retrieval
di: Zhao, Zijia, et al.
Pubblicazione: (2024)
di: Zhao, Zijia, et al.
Pubblicazione: (2024)
Documenti analoghi
-
InterMoE: Individual-Specific 3D Human Interaction Generation via Dynamic Temporal-Selective MoE
di: Wang, Lipeng, et al.
Pubblicazione: (2025) -
Multi-Agent Amodal Completion: Direct Synthesis with Fine-Grained Semantic Guidance
di: Fan, Hongxing, et al.
Pubblicazione: (2025) -
Ouroboros3D: Image-to-3D Generation via 3D-aware Recursive Diffusion
di: Wen, Hao, et al.
Pubblicazione: (2024) -
SegviGen: Repurposing 3D Generative Model for Part Segmentation
di: Li, Lin, et al.
Pubblicazione: (2026) -
MV-Adapter: Multi-view Consistent Image Generation Made Easy
di: Huang, Zehuan, et al.
Pubblicazione: (2024)