PSDiffusion: Harmonized Multi-Layer Image Generation via Layout and Appearance Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Dingbang, Li, Wenbo, Zhao, Yifei, Pan, Xinyu, Wang, Chun, Zeng, Yanhong, Dai, Bo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LayerFusion: Harmonized Multi-Layer Text-to-Image Generation with Generative Priors
von: Dalva, Yusuf, et al.
Veröffentlicht: (2024)
von: Dalva, Yusuf, et al.
Veröffentlicht: (2024)
Layout2Scene: 3D Semantic Layout Guided Scene Generation via Geometry and Appearance Diffusion Priors
von: Chen, Minglin, et al.
Veröffentlicht: (2025)
von: Chen, Minglin, et al.
Veröffentlicht: (2025)
CHROMA: Consistent Harmonization of Multi-View Appearance via Bilateral Grid Prediction
von: Shin, Jisu, et al.
Veröffentlicht: (2025)
von: Shin, Jisu, et al.
Veröffentlicht: (2025)
LocRef-Diffusion:Tuning-Free Layout and Appearance-Guided Generation
von: Deng, Fan, et al.
Veröffentlicht: (2024)
von: Deng, Fan, et al.
Veröffentlicht: (2024)
Multi-identity Human Image Animation with Structural Video Diffusion
von: Wang, Zhenzhi, et al.
Veröffentlicht: (2025)
von: Wang, Zhenzhi, et al.
Veröffentlicht: (2025)
Synergizing Motion and Appearance: Multi-Scale Compensatory Codebooks for Talking Head Video Generation
von: Zhao, Shuling, et al.
Veröffentlicht: (2024)
von: Zhao, Shuling, et al.
Veröffentlicht: (2024)
ToLo: A Two-Stage, Training-Free Layout-To-Image Generation Framework For High-Overlap Layouts
von: Huang, Linhao, et al.
Veröffentlicht: (2025)
von: Huang, Linhao, et al.
Veröffentlicht: (2025)
AeSlides: Incentivizing Aesthetic Layout in LLM-Based Slide Generation via Verifiable Rewards
von: Pan, Yiming, et al.
Veröffentlicht: (2026)
von: Pan, Yiming, et al.
Veröffentlicht: (2026)
UltrasoundAgents: Hierarchical Multi-Agent Evidence-Chain Reasoning for Breast Ultrasound Diagnosis
von: Zhu, Yali, et al.
Veröffentlicht: (2026)
von: Zhu, Yali, et al.
Veröffentlicht: (2026)
Learning to Align Generative Appearance Priors for Fine-grained Image Retrieval
von: Wang, Shijie, et al.
Veröffentlicht: (2026)
von: Wang, Shijie, et al.
Veröffentlicht: (2026)
DreamLayer: Simultaneous Multi-Layer Generation via Diffusion Mode
von: Huang, Junjia, et al.
Veröffentlicht: (2025)
von: Huang, Junjia, et al.
Veröffentlicht: (2025)
Manga Generation via Layout-controllable Diffusion
von: Chen, Siyu, et al.
Veröffentlicht: (2024)
von: Chen, Siyu, et al.
Veröffentlicht: (2024)
A Task is Worth One Word: Learning with Task Prompts for High-Quality Versatile Image Inpainting
von: Zhuang, Junhao, et al.
Veröffentlicht: (2023)
von: Zhuang, Junhao, et al.
Veröffentlicht: (2023)
HYDRA: Unifying Multi-modal Generation and Understanding via Representation-Harmonized Tokenization
von: Qiu, Xuerui, et al.
Veröffentlicht: (2026)
von: Qiu, Xuerui, et al.
Veröffentlicht: (2026)
Improving Layout Representation Learning Across Inconsistently Annotated Datasets via Agentic Harmonization
von: Li, Renyu, et al.
Veröffentlicht: (2026)
von: Li, Renyu, et al.
Veröffentlicht: (2026)
Hierarchical Concept-to-Appearance Guidance for Multi-Subject Image Generation
von: Xu, Yijia, et al.
Veröffentlicht: (2026)
von: Xu, Yijia, et al.
Veröffentlicht: (2026)
SpotActor: Training-Free Layout-Controlled Consistent Image Generation
von: Wang, Jiahao, et al.
Veröffentlicht: (2024)
von: Wang, Jiahao, et al.
Veröffentlicht: (2024)
CreatiLayout: Siamese Multimodal Diffusion Transformer for Creative Layout-to-Image Generation
von: Zhang, Hui, et al.
Veröffentlicht: (2024)
von: Zhang, Hui, et al.
Veröffentlicht: (2024)
STAY Diffusion: Styled Layout Diffusion Model for Diverse Layout-to-Image Generation
von: Wang, Ruyu, et al.
Veröffentlicht: (2025)
von: Wang, Ruyu, et al.
Veröffentlicht: (2025)
Direct Numerical Layout Generation for 3D Indoor Scene Synthesis via Spatial Reasoning
von: Ran, Xingjian, et al.
Veröffentlicht: (2025)
von: Ran, Xingjian, et al.
Veröffentlicht: (2025)
Rethinking The Training And Evaluation of Rich-Context Layout-to-Image Generation
von: Cheng, Jiaxin, et al.
Veröffentlicht: (2024)
von: Cheng, Jiaxin, et al.
Veröffentlicht: (2024)
StructLayoutFormer:Conditional Structured Layout Generation via Structure Serialization and Disentanglement
von: Hu, Xin, et al.
Veröffentlicht: (2025)
von: Hu, Xin, et al.
Veröffentlicht: (2025)
DiffX: Guide Your Layout to Cross-Modal Generative Modeling
von: Wang, Zeyu, et al.
Veröffentlicht: (2024)
von: Wang, Zeyu, et al.
Veröffentlicht: (2024)
IPDreamer: Appearance-Controllable 3D Object Generation with Complex Image Prompts
von: Zeng, Bohan, et al.
Veröffentlicht: (2023)
von: Zeng, Bohan, et al.
Veröffentlicht: (2023)
Layout-Guided Controllable Pathology Image Generation with In-Context Diffusion Transformers
von: Shou, Yuntao, et al.
Veröffentlicht: (2026)
von: Shou, Yuntao, et al.
Veröffentlicht: (2026)
OmniDocLayout: Towards Diverse Document Layout Generation via Coarse-to-Fine LLM Learning
von: Kang, Hengrui, et al.
Veröffentlicht: (2025)
von: Kang, Hengrui, et al.
Veröffentlicht: (2025)
Alleviating Distortion in Image Generation via Multi-Resolution Diffusion Models and Time-Dependent Layer Normalization
von: Liu, Qihao, et al.
Veröffentlicht: (2024)
von: Liu, Qihao, et al.
Veröffentlicht: (2024)
MS-Diffusion: Multi-subject Zero-shot Image Personalization with Layout Guidance
von: Wang, Xierui, et al.
Veröffentlicht: (2024)
von: Wang, Xierui, et al.
Veröffentlicht: (2024)
Once Is Enough: Lightweight DiT-Based Video Virtual Try-On via One-Time Garment Appearance Injection
von: Pan, Yanjie, et al.
Veröffentlicht: (2025)
von: Pan, Yanjie, et al.
Veröffentlicht: (2025)
InstanceAssemble: Layout-Aware Image Generation via Instance Assembling Attention
von: Xiang, Qiang, et al.
Veröffentlicht: (2025)
von: Xiang, Qiang, et al.
Veröffentlicht: (2025)
CustomTTT: Motion and Appearance Customized Video Generation via Test-Time Training
von: Bi, Xiuli, et al.
Veröffentlicht: (2024)
von: Bi, Xiuli, et al.
Veröffentlicht: (2024)
Multi-Scale Diffusion: Enhancing Spatial Layout in High-Resolution Panoramic Image Generation
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2024)
LAMIC: Layout-Aware Multi-Image Composition via Scalability of Multimodal Diffusion Transformer
von: Chen, Yuzhuo, et al.
Veröffentlicht: (2025)
von: Chen, Yuzhuo, et al.
Veröffentlicht: (2025)
Layout-Conditioned Autoregressive Text-to-Image Generation via Structured Masking
von: Zheng, Zirui, et al.
Veröffentlicht: (2025)
von: Zheng, Zirui, et al.
Veröffentlicht: (2025)
ReLayout: Integrating Relation Reasoning for Content-aware Layout Generation with Multi-modal Large Language Models
von: Tian, Jiaxu, et al.
Veröffentlicht: (2025)
von: Tian, Jiaxu, et al.
Veröffentlicht: (2025)
GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion
von: Zhu, Hanxin, et al.
Veröffentlicht: (2026)
von: Zhu, Hanxin, et al.
Veröffentlicht: (2026)
GA-Drive: Geometry-Appearance Decoupled Modeling for Free-viewpoint Driving Scene Generation
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
LayoutFlow: Flow Matching for Layout Generation
von: Guerreiro, Julian Jorge Andrade, et al.
Veröffentlicht: (2024)
von: Guerreiro, Julian Jorge Andrade, et al.
Veröffentlicht: (2024)
Griffin: Generative Reference and Layout Guided Image Composition
von: Mikaeili, Aryan, et al.
Veröffentlicht: (2025)
von: Mikaeili, Aryan, et al.
Veröffentlicht: (2025)
EchoGen: Cycle-Consistent Learning for Unified Layout-Image Generation and Understanding
von: Zou, Kai, et al.
Veröffentlicht: (2026)
von: Zou, Kai, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
LayerFusion: Harmonized Multi-Layer Text-to-Image Generation with Generative Priors
von: Dalva, Yusuf, et al.
Veröffentlicht: (2024) -
Layout2Scene: 3D Semantic Layout Guided Scene Generation via Geometry and Appearance Diffusion Priors
von: Chen, Minglin, et al.
Veröffentlicht: (2025) -
CHROMA: Consistent Harmonization of Multi-View Appearance via Bilateral Grid Prediction
von: Shin, Jisu, et al.
Veröffentlicht: (2025) -
LocRef-Diffusion:Tuning-Free Layout and Appearance-Guided Generation
von: Deng, Fan, et al.
Veröffentlicht: (2024) -
Multi-identity Human Image Animation with Structural Video Diffusion
von: Wang, Zhenzhi, et al.
Veröffentlicht: (2025)