Training-Free Layout-to-Image Generation with Marginal Attention Constraints
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Huancheng, Li, Jingtao, Zhuang, Weiming, Vikalo, Haris, Lyu, Lingjuan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Replay-Free Continual Low-Rank Adaptation with Dynamic Memory
by: Chen, Huancheng, et al.
Published: (2024)
by: Chen, Huancheng, et al.
Published: (2024)
FEDEXCHANGE: Bridging the Domain Gap in Federated Object Detection for Free
by: Yuan, Haolin, et al.
Published: (2025)
by: Yuan, Haolin, et al.
Published: (2025)
On the Limits of Token Reduction for Efficient Unified Vision Language Training
by: Chen, Siyi, et al.
Published: (2026)
by: Chen, Siyi, et al.
Published: (2026)
UniCompress: Token Compression for Unified Vision-Language Understanding and Generation
by: Wang, Ziyao, et al.
Published: (2026)
by: Wang, Ziyao, et al.
Published: (2026)
COALA: A Practical and Vision-Centric Federated Learning Platform
by: Zhuang, Weiming, et al.
Published: (2024)
by: Zhuang, Weiming, et al.
Published: (2024)
VibeToken: Scaling 1D Image Tokenizers and Autoregressive Models for Dynamic Resolution Generations
by: Patel, Maitreya, et al.
Published: (2026)
by: Patel, Maitreya, et al.
Published: (2026)
Closer to Reality: Practical Semi-Supervised Federated Learning for Foundation Model Adaptation
by: Sun, Guangyu, et al.
Published: (2025)
by: Sun, Guangyu, et al.
Published: (2025)
UNIFORM: Unifying Knowledge from Large-scale and Diverse Pre-trained Models
by: Wang, Yimu, et al.
Published: (2025)
by: Wang, Yimu, et al.
Published: (2025)
FedWon: Triumphing Multi-domain Federated Learning Without Normalization
by: Zhuang, Weiming, et al.
Published: (2023)
by: Zhuang, Weiming, et al.
Published: (2023)
Towards Fundamentally Scalable Model Selection: Asymptotically Fast Update and Selection
by: Wang, Wenxiao, et al.
Published: (2024)
by: Wang, Wenxiao, et al.
Published: (2024)
A Simple Background Augmentation Method for Object Detection with Diffusion Model
by: Li, Yuhang, et al.
Published: (2024)
by: Li, Yuhang, et al.
Published: (2024)
Is Synthetic Image Useful for Transfer Learning? An Investigation into Data Generation, Volume, and Utilization
by: Li, Yuhang, et al.
Published: (2024)
by: Li, Yuhang, et al.
Published: (2024)
Empirical Recipes for Efficient and Compact Vision-Language Models
by: Huang, Jiabo, et al.
Published: (2026)
by: Huang, Jiabo, et al.
Published: (2026)
Stretching Each Dollar: Diffusion Training from Scratch on a Micro-Budget
by: Sehwag, Vikash, et al.
Published: (2024)
by: Sehwag, Vikash, et al.
Published: (2024)
SpotActor: Training-Free Layout-Controlled Consistent Image Generation
by: Wang, Jiahao, et al.
Published: (2024)
by: Wang, Jiahao, et al.
Published: (2024)
ToLo: A Two-Stage, Training-Free Layout-To-Image Generation Framework For High-Overlap Layouts
by: Huang, Linhao, et al.
Published: (2025)
by: Huang, Linhao, et al.
Published: (2025)
ExpVG: Investigating the Design Space of Visual Grounding in Multimodal Large Language Model
by: Kang, Weitai, et al.
Published: (2025)
by: Kang, Weitai, et al.
Published: (2025)
LoCo: Locally Constrained Training-Free Layout-to-Image Synthesis
by: Zhao, Peiang, et al.
Published: (2023)
by: Zhao, Peiang, et al.
Published: (2023)
Check, Locate, Rectify: A Training-Free Layout Calibration System for Text-to-Image Generation
by: Gong, Biao, et al.
Published: (2023)
by: Gong, Biao, et al.
Published: (2023)
InstanceAssemble: Layout-Aware Image Generation via Instance Assembling Attention
by: Xiang, Qiang, et al.
Published: (2025)
by: Xiang, Qiang, et al.
Published: (2025)
Zero-Painter: Training-Free Layout Control for Text-to-Image Synthesis
by: Ohanyan, Marianna, et al.
Published: (2024)
by: Ohanyan, Marianna, et al.
Published: (2024)
Seeing Further on the Shoulders of Giants: Knowledge Inheritance for Vision Foundation Models
by: Huang, Jiabo, et al.
Published: (2025)
by: Huang, Jiabo, et al.
Published: (2025)
Rethinking The Training And Evaluation of Rich-Context Layout-to-Image Generation
by: Cheng, Jiaxin, et al.
Published: (2024)
by: Cheng, Jiaxin, et al.
Published: (2024)
Evaluating and Mitigating IP Infringement in Visual Generative AI
by: Wang, Zhenting, et al.
Published: (2024)
by: Wang, Zhenting, et al.
Published: (2024)
One-to-More: High-Fidelity Training-Free Anomaly Generation with Attention Control
by: Rao, Haoxiang, et al.
Published: (2026)
by: Rao, Haoxiang, et al.
Published: (2026)
PerSense: Training-Free Personalized Instance Segmentation in Dense Images
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2024)
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2024)
FreeText: Training-Free Text Rendering in Diffusion Transformers via Attention Localization and Spectral Glyph Injection
by: Zhang, Ruiqiang, et al.
Published: (2026)
by: Zhang, Ruiqiang, et al.
Published: (2026)
AnyMS: Bottom-up Attention Decoupling for Layout-guided and Training-free Multi-subject Customization
by: Yu, Binhe, et al.
Published: (2025)
by: Yu, Binhe, et al.
Published: (2025)
LAMIC: Layout-Aware Multi-Image Composition via Scalability of Multimodal Diffusion Transformer
by: Chen, Yuzhuo, et al.
Published: (2025)
by: Chen, Yuzhuo, et al.
Published: (2025)
Detecting, Explaining, and Mitigating Memorization in Diffusion Models
by: Wen, Yuxin, et al.
Published: (2024)
by: Wen, Yuxin, et al.
Published: (2024)
Layout-to-Image Generation with Localized Descriptions using ControlNet with Cross-Attention Control
by: Lukovnikov, Denis, et al.
Published: (2024)
by: Lukovnikov, Denis, et al.
Published: (2024)
E-SAM: Training-Free Segment Every Entity Model
by: Zhang, Weiming, et al.
Published: (2025)
by: Zhang, Weiming, et al.
Published: (2025)
CO-SPY: Combining Semantic and Pixel Features to Detect Synthetic Images by AI
by: Cheng, Siyuan, et al.
Published: (2025)
by: Cheng, Siyuan, et al.
Published: (2025)
Control and Realism: Best of Both Worlds in Layout-to-Image without Training
by: Li, Bonan, et al.
Published: (2025)
by: Li, Bonan, et al.
Published: (2025)
Towards PerSense++: Advancing Training-Free Personalized Instance Segmentation in Dense Images
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
by: Siddiqui, Muhammad Ibraheem, et al.
Published: (2025)
IP-Prompter: Training-Free Theme-Specific Image Generation via Dynamic Visual Prompting
by: Zhang, Yuxin, et al.
Published: (2025)
by: Zhang, Yuxin, et al.
Published: (2025)
Training-free Composite Scene Generation for Layout-to-Image Synthesis
by: Liu, Jiaqi, et al.
Published: (2024)
by: Liu, Jiaqi, et al.
Published: (2024)
CreatiLayout: Siamese Multimodal Diffusion Transformer for Creative Layout-to-Image Generation
by: Zhang, Hui, et al.
Published: (2024)
by: Zhang, Hui, et al.
Published: (2024)
STAY Diffusion: Styled Layout Diffusion Model for Diverse Layout-to-Image Generation
by: Wang, Ruyu, et al.
Published: (2025)
by: Wang, Ruyu, et al.
Published: (2025)
Paying More Attention to Image: A Training-Free Method for Alleviating Hallucination in LVLMs
by: Liu, Shi, et al.
Published: (2024)
by: Liu, Shi, et al.
Published: (2024)
Similar Items
-
Replay-Free Continual Low-Rank Adaptation with Dynamic Memory
by: Chen, Huancheng, et al.
Published: (2024) -
FEDEXCHANGE: Bridging the Domain Gap in Federated Object Detection for Free
by: Yuan, Haolin, et al.
Published: (2025) -
On the Limits of Token Reduction for Efficient Unified Vision Language Training
by: Chen, Siyi, et al.
Published: (2026) -
UniCompress: Token Compression for Unified Vision-Language Understanding and Generation
by: Wang, Ziyao, et al.
Published: (2026) -
COALA: A Practical and Vision-Centric Federated Learning Platform
by: Zhuang, Weiming, et al.
Published: (2024)