Enhancing Object Coherence in Layout-to-Image Synthesis
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Yibin, Zhou, Changhai, Xu, Honghui |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DreamText: High Fidelity Scene Text Synthesis
por: Wang, Yibin, et al.
Publicado: (2024)
por: Wang, Yibin, et al.
Publicado: (2024)
Training-free Composite Scene Generation for Layout-to-Image Synthesis
por: Liu, Jiaqi, et al.
Publicado: (2024)
por: Liu, Jiaqi, et al.
Publicado: (2024)
MagicFace: Training-free Universal-Style Human Image Customized Synthesis
por: Wang, Yibin, et al.
Publicado: (2024)
por: Wang, Yibin, et al.
Publicado: (2024)
High-fidelity Person-centric Subject-to-Image Synthesis
por: Wang, Yibin, et al.
Publicado: (2023)
por: Wang, Yibin, et al.
Publicado: (2023)
LTOS: Layout-controllable Text-Object Synthesis via Adaptive Cross-attention Fusions
por: Zhao, Xiaoran, et al.
Publicado: (2024)
por: Zhao, Xiaoran, et al.
Publicado: (2024)
A Framework For Image Synthesis Using Supervised Contrastive Learning
por: Liu, Yibin, et al.
Publicado: (2024)
por: Liu, Yibin, et al.
Publicado: (2024)
Layout-Conditioned Autoregressive Text-to-Image Generation via Structured Masking
por: Zheng, Zirui, et al.
Publicado: (2025)
por: Zheng, Zirui, et al.
Publicado: (2025)
RSGen: Enhancing Layout-Driven Remote Sensing Image Generation with Diverse Edge Guidance
por: Hou, Xianbao, et al.
Publicado: (2026)
por: Hou, Xianbao, et al.
Publicado: (2026)
TextDiffuser-RL: Efficient and Robust Text Layout Optimization for High-Fidelity Text-to-Image Synthesis
por: Rahman, Kazi Mahathir, et al.
Publicado: (2025)
por: Rahman, Kazi Mahathir, et al.
Publicado: (2025)
POCI-Diff: Position Objects Consistently and Interactively with 3D-Layout Guided Diffusion
por: Rigo, Andrea, et al.
Publicado: (2026)
por: Rigo, Andrea, et al.
Publicado: (2026)
C3S3: Complementary Competition and Contrastive Selection for Semi-Supervised Medical Image Segmentation
por: He, Jiaying, et al.
Publicado: (2025)
por: He, Jiaying, et al.
Publicado: (2025)
Direct Numerical Layout Generation for 3D Indoor Scene Synthesis via Spatial Reasoning
por: Ran, Xingjian, et al.
Publicado: (2025)
por: Ran, Xingjian, et al.
Publicado: (2025)
Enhance Multimodal Consistency and Coherence for Text-Image Plan Generation
por: Lu, Xiaoxin, et al.
Publicado: (2025)
por: Lu, Xiaoxin, et al.
Publicado: (2025)
PrimeComposer: Faster Progressively Combined Diffusion for Image Composition with Attention Steering
por: Wang, Yibin, et al.
Publicado: (2024)
por: Wang, Yibin, et al.
Publicado: (2024)
Segmentor-Guided Counterfactual Fine-Tuning for Locally Coherent and Targeted Image Synthesis
por: Xia, Tian, et al.
Publicado: (2025)
por: Xia, Tian, et al.
Publicado: (2025)
Sketch-to-Layout: Sketch-Guided Multimodal Layout Generation
por: Brioschi, Riccardo, et al.
Publicado: (2025)
por: Brioschi, Riccardo, et al.
Publicado: (2025)
ICDAR 2025 Competition on End-to-End Document Image Machine Translation Towards Complex Layouts
por: Zhang, Yaping, et al.
Publicado: (2026)
por: Zhang, Yaping, et al.
Publicado: (2026)
SSMG: Spatial-Semantic Map Guided Diffusion Model for Free-form Layout-to-Image Generation
por: Jia, Chengyou, et al.
Publicado: (2023)
por: Jia, Chengyou, et al.
Publicado: (2023)
A Semantically Enhanced Generative Foundation Model Improves Pathological Image Synthesis
por: Guan, Xianchao, et al.
Publicado: (2025)
por: Guan, Xianchao, et al.
Publicado: (2025)
EMIFF: Enhanced Multi-scale Image Feature Fusion for Vehicle-Infrastructure Cooperative 3D Object Detection
por: Wang, Zhe, et al.
Publicado: (2024)
por: Wang, Zhe, et al.
Publicado: (2024)
IV-Mixed Sampler: Leveraging Image Diffusion Models for Enhanced Video Synthesis
por: Shao, Shitong, et al.
Publicado: (2024)
por: Shao, Shitong, et al.
Publicado: (2024)
ArtAug: Enhancing Text-to-Image Generation through Synthesis-Understanding Interaction
por: Duan, Zhongjie, et al.
Publicado: (2024)
por: Duan, Zhongjie, et al.
Publicado: (2024)
Interact-Custom: Customized Human Object Interaction Image Generation
por: Xu, Zhu, et al.
Publicado: (2025)
por: Xu, Zhu, et al.
Publicado: (2025)
LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
por: Sun, Fan-Yun, et al.
Publicado: (2024)
por: Sun, Fan-Yun, et al.
Publicado: (2024)
SpikeNVS: Enhancing Novel View Synthesis from Blurry Images via Spike Camera
por: Dai, Gaole, et al.
Publicado: (2024)
por: Dai, Gaole, et al.
Publicado: (2024)
Layout-and-Retouch: A Dual-stage Framework for Improving Diversity in Personalized Image Generation
por: Kim, Kangyeol, et al.
Publicado: (2024)
por: Kim, Kangyeol, et al.
Publicado: (2024)
Point2RBox-v2: Rethinking Point-supervised Oriented Object Detection with Spatial Layout Among Instances
por: Yu, Yi, et al.
Publicado: (2025)
por: Yu, Yi, et al.
Publicado: (2025)
Order Is Not Layout: Order-to-Space Bias in Image Generation
por: Zhang, Yongkang, et al.
Publicado: (2026)
por: Zhang, Yongkang, et al.
Publicado: (2026)
SOWing Information: Cultivating Contextual Coherence with MLLMs in Image Generation
por: Pei, Yuhan, et al.
Publicado: (2024)
por: Pei, Yuhan, et al.
Publicado: (2024)
Evaluating Text-to-Image Generative Models: An Empirical Study on Human Image Synthesis
por: Chen, Muxi, et al.
Publicado: (2024)
por: Chen, Muxi, et al.
Publicado: (2024)
OAD-Promoter: Enhancing Zero-shot VQA using Large Language Models with Object Attribute Description
por: Xu, Quanxing, et al.
Publicado: (2025)
por: Xu, Quanxing, et al.
Publicado: (2025)
Dark-ISP: Enhancing RAW Image Processing for Low-Light Object Detection
por: Guo, Jiasheng, et al.
Publicado: (2025)
por: Guo, Jiasheng, et al.
Publicado: (2025)
PP-DocLayout: A Unified Document Layout Detection Model to Accelerate Large-Scale Data Construction
por: Sun, Ting, et al.
Publicado: (2025)
por: Sun, Ting, et al.
Publicado: (2025)
HouseLayout3D: A Benchmark and Training-Free Baseline for 3D Layout Estimation in the Wild
por: Bieri, Valentin, et al.
Publicado: (2025)
por: Bieri, Valentin, et al.
Publicado: (2025)
LAYOUTDREAMER: Physics-guided Layout for Text-to-3D Compositional Scene Generation
por: Zhou, Yang, et al.
Publicado: (2025)
por: Zhou, Yang, et al.
Publicado: (2025)
LocRef-Diffusion:Tuning-Free Layout and Appearance-Guided Generation
por: Deng, Fan, et al.
Publicado: (2024)
por: Deng, Fan, et al.
Publicado: (2024)
Batch-Instructed Gradient for Prompt Evolution:Systematic Prompt Optimization for Enhanced Text-to-Image Synthesis
por: Yang, Xinrui, et al.
Publicado: (2024)
por: Yang, Xinrui, et al.
Publicado: (2024)
GenHOI: Generalizing Text-driven 4D Human-Object Interaction Synthesis for Unseen Objects
por: Li, Shujia, et al.
Publicado: (2025)
por: Li, Shujia, et al.
Publicado: (2025)
Multimodal-Enhanced Objectness Learner for Corner Case Detection in Autonomous Driving
por: Xiao, Lixing, et al.
Publicado: (2024)
por: Xiao, Lixing, et al.
Publicado: (2024)
Mitigating Hallucinations on Object Attributes using Multiview Images and Negative Instructions
por: Tan, Zhijie, et al.
Publicado: (2025)
por: Tan, Zhijie, et al.
Publicado: (2025)
Ejemplares similares
-
DreamText: High Fidelity Scene Text Synthesis
por: Wang, Yibin, et al.
Publicado: (2024) -
Training-free Composite Scene Generation for Layout-to-Image Synthesis
por: Liu, Jiaqi, et al.
Publicado: (2024) -
MagicFace: Training-free Universal-Style Human Image Customized Synthesis
por: Wang, Yibin, et al.
Publicado: (2024) -
High-fidelity Person-centric Subject-to-Image Synthesis
por: Wang, Yibin, et al.
Publicado: (2023) -
LTOS: Layout-controllable Text-Object Synthesis via Adaptive Cross-attention Fusions
por: Zhao, Xiaoran, et al.
Publicado: (2024)