EliGen: Entity-Level Controlled Image Generation with Regional Attention
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Hong, Duan, Zhongjie, Wang, Xingjun, Chen, Yingda, Zhang, Yu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Nexus-Gen: Unified Image Understanding, Generation, and Editing via Prefilled Autoregression in Shared Embedding Space
by: Zhang, Hong, et al.
Published: (2025)
by: Zhang, Hong, et al.
Published: (2025)
Diffusion Templates: A Unified Plugin Framework for Controllable Diffusion
by: Duan, Zhongjie, et al.
Published: (2026)
by: Duan, Zhongjie, et al.
Published: (2026)
AutoLoRA: Automatic LoRA Retrieval and Fine-Grained Gated Fusion for Text-to-Image Generation
by: Li, Zhiwen, et al.
Published: (2025)
by: Li, Zhiwen, et al.
Published: (2025)
AttriCtrl: Fine-Grained Control of Aesthetic Attribute Intensity in Diffusion Models
by: Chen, Die, et al.
Published: (2025)
by: Chen, Die, et al.
Published: (2025)
ArtAug: Enhancing Text-to-Image Generation through Synthesis-Understanding Interaction
by: Duan, Zhongjie, et al.
Published: (2024)
by: Duan, Zhongjie, et al.
Published: (2024)
Spectral Evolution Search: Efficient Inference-Time Scaling for Reward-Aligned Image Generation
by: Ye, Jinyan, et al.
Published: (2026)
by: Ye, Jinyan, et al.
Published: (2026)
VIRAL: Visual In-Context Reasoning via Analogy in Diffusion Transformers
by: Li, Zhiwen, et al.
Published: (2026)
by: Li, Zhiwen, et al.
Published: (2026)
UltraGen: Efficient Ultra-High-Resolution Image Generation with Hierarchical Local Attention
by: Zhang, Yuyao, et al.
Published: (2025)
by: Zhang, Yuyao, et al.
Published: (2025)
AlignedGen: Aligning Style Across Generated Images
by: Zhang, Jiexuan, et al.
Published: (2025)
by: Zhang, Jiexuan, et al.
Published: (2025)
MaskAttn-SDXL: Controllable Region-Level Text-To-Image Generation
by: Chang, Yu, et al.
Published: (2025)
by: Chang, Yu, et al.
Published: (2025)
CaliTex: Geometry-Calibrated Attention for View-Coherent 3D Texture Generation
by: Liu, Chenyu, et al.
Published: (2025)
by: Liu, Chenyu, et al.
Published: (2025)
UltraGen: High-Resolution Video Generation with Hierarchical Attention
by: Hu, Teng, et al.
Published: (2025)
by: Hu, Teng, et al.
Published: (2025)
Breaking the Modality Barrier: Universal Embedding Learning with Multimodal LLMs
by: Gu, Tiancheng, et al.
Published: (2025)
by: Gu, Tiancheng, et al.
Published: (2025)
UniGenDet: A Unified Generative-Discriminative Framework for Co-Evolutionary Image Generation and Generated Image Detection
by: Zhang, Yanran, et al.
Published: (2026)
by: Zhang, Yanran, et al.
Published: (2026)
SerialGen: Personalized Image Generation by First Standardization Then Personalization
by: Xie, Cong, et al.
Published: (2024)
by: Xie, Cong, et al.
Published: (2024)
From Text to Mask: Localizing Entities Using the Attention of Text-to-Image Diffusion Models
by: Xiao, Changming, et al.
Published: (2023)
by: Xiao, Changming, et al.
Published: (2023)
Gen-Searcher: Reinforcing Agentic Search for Image Generation
by: Feng, Kaituo, et al.
Published: (2026)
by: Feng, Kaituo, et al.
Published: (2026)
GDGS: 3D Gaussian Splatting Via Geometry-Guided Initialization And Dynamic Density Control
by: Wang, Xingjun, et al.
Published: (2025)
by: Wang, Xingjun, et al.
Published: (2025)
RecipeGen: A Benchmark for Real-World Recipe Image Generation
by: Zhang, Ruoxuan, et al.
Published: (2025)
by: Zhang, Ruoxuan, et al.
Published: (2025)
RA-Det: Towards Universal Detection of AI-Generated Images via Robustness Asymmetry
by: Wang, Xinchang, et al.
Published: (2026)
by: Wang, Xinchang, et al.
Published: (2026)
GenSpace: Benchmarking Spatially-Aware Image Generation
by: Wang, Zehan, et al.
Published: (2025)
by: Wang, Zehan, et al.
Published: (2025)
TextCenGen: Attention-Guided Text-Centric Background Adaptation for Text-to-Image Generation
by: Liang, Tianyi, et al.
Published: (2024)
by: Liang, Tianyi, et al.
Published: (2024)
LSReGen: Large-Scale Regional Generator via Backward Guidance Framework
by: Zhang, Bowen, et al.
Published: (2024)
by: Zhang, Bowen, et al.
Published: (2024)
BadPatch: Diffusion-Based Generation of Physical Adversarial Patches
by: Wang, Zhixiang, et al.
Published: (2024)
by: Wang, Zhixiang, et al.
Published: (2024)
GenAgent: Scaling Text-to-Image Generation via Agentic Multimodal Reasoning
by: Jiang, Kaixun, et al.
Published: (2026)
by: Jiang, Kaixun, et al.
Published: (2026)
ArtGen: Conditional Generative Modeling of Articulated Objects in Arbitrary Part-Level States
by: Wang, Haowen, et al.
Published: (2025)
by: Wang, Haowen, et al.
Published: (2025)
RAGSR: Regional Attention Guided Diffusion for Image Super-Resolution
by: He, Haodong, et al.
Published: (2025)
by: He, Haodong, et al.
Published: (2025)
RevSAM2: Prompt SAM2 for Medical Image Segmentation via Reverse-Propagation without Fine-tuning
by: Bai, Yunhao, et al.
Published: (2024)
by: Bai, Yunhao, et al.
Published: (2024)
DragEntity: Trajectory Guided Video Generation using Entity and Positional Relationships
by: Wan, Zhang, et al.
Published: (2024)
by: Wan, Zhang, et al.
Published: (2024)
AIM: Additional Image Guided Generation of Transferable Adversarial Attacks
by: Li, Teng, et al.
Published: (2025)
by: Li, Teng, et al.
Published: (2025)
FOCUS: Optimal Control for Multi-Entity World Modeling in Text-to-Image Generation
by: Bill, Eric Tillmann, et al.
Published: (2025)
by: Bill, Eric Tillmann, et al.
Published: (2025)
SafeCtrl: Region-Aware Safety Control for Text-to-Image Diffusion via Detect-Then-Suppress
by: Zhang, Lingyun, et al.
Published: (2026)
by: Zhang, Lingyun, et al.
Published: (2026)
GenMix: Combining Generative and Mixture Data Augmentation for Medical Image Classification
by: Lee, Hansang, et al.
Published: (2024)
by: Lee, Hansang, et al.
Published: (2024)
SafeCtrl: Region-Based Safety Control for Text-to-Image Diffusion via Detect-Then-Suppress
by: Zhang, Lingyun, et al.
Published: (2025)
by: Zhang, Lingyun, et al.
Published: (2025)
MathGen: Revealing the Illusion of Mathematical Competence through Text-to-Image Generation
by: Liu, Ruiyao, et al.
Published: (2026)
by: Liu, Ruiyao, et al.
Published: (2026)
Diffutoon: High-Resolution Editable Toon Shading via Diffusion Models
by: Duan, Zhongjie, et al.
Published: (2024)
by: Duan, Zhongjie, et al.
Published: (2024)
GUSLO: General and Unified Structured Light Optimization
by: Wan, Tinglei, et al.
Published: (2025)
by: Wan, Tinglei, et al.
Published: (2025)
WeditGAN: Few-Shot Image Generation via Latent Space Relocation
by: Duan, Yuxuan, et al.
Published: (2023)
by: Duan, Yuxuan, et al.
Published: (2023)
GenRC: Generative 3D Room Completion from Sparse Image Collections
by: Li, Ming-Feng, et al.
Published: (2024)
by: Li, Ming-Feng, et al.
Published: (2024)
TabletopGen: Instance-Level Interactive 3D Tabletop Scene Generation from Text or Single Image
by: Wang, Ziqian, et al.
Published: (2025)
by: Wang, Ziqian, et al.
Published: (2025)
Similar Items
-
Nexus-Gen: Unified Image Understanding, Generation, and Editing via Prefilled Autoregression in Shared Embedding Space
by: Zhang, Hong, et al.
Published: (2025) -
Diffusion Templates: A Unified Plugin Framework for Controllable Diffusion
by: Duan, Zhongjie, et al.
Published: (2026) -
AutoLoRA: Automatic LoRA Retrieval and Fine-Grained Gated Fusion for Text-to-Image Generation
by: Li, Zhiwen, et al.
Published: (2025) -
AttriCtrl: Fine-Grained Control of Aesthetic Attribute Intensity in Diffusion Models
by: Chen, Die, et al.
Published: (2025) -
ArtAug: Enhancing Text-to-Image Generation through Synthesis-Understanding Interaction
by: Duan, Zhongjie, et al.
Published: (2024)