EliGen: Entity-Level Controlled Image Generation with Regional Attention
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Hong, Duan, Zhongjie, Wang, Xingjun, Chen, Yingda, Zhang, Yu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Nexus-Gen: Unified Image Understanding, Generation, and Editing via Prefilled Autoregression in Shared Embedding Space
von: Zhang, Hong, et al.
Veröffentlicht: (2025)
von: Zhang, Hong, et al.
Veröffentlicht: (2025)
Diffusion Templates: A Unified Plugin Framework for Controllable Diffusion
von: Duan, Zhongjie, et al.
Veröffentlicht: (2026)
von: Duan, Zhongjie, et al.
Veröffentlicht: (2026)
AutoLoRA: Automatic LoRA Retrieval and Fine-Grained Gated Fusion for Text-to-Image Generation
von: Li, Zhiwen, et al.
Veröffentlicht: (2025)
von: Li, Zhiwen, et al.
Veröffentlicht: (2025)
AttriCtrl: Fine-Grained Control of Aesthetic Attribute Intensity in Diffusion Models
von: Chen, Die, et al.
Veröffentlicht: (2025)
von: Chen, Die, et al.
Veröffentlicht: (2025)
ArtAug: Enhancing Text-to-Image Generation through Synthesis-Understanding Interaction
von: Duan, Zhongjie, et al.
Veröffentlicht: (2024)
von: Duan, Zhongjie, et al.
Veröffentlicht: (2024)
Spectral Evolution Search: Efficient Inference-Time Scaling for Reward-Aligned Image Generation
von: Ye, Jinyan, et al.
Veröffentlicht: (2026)
von: Ye, Jinyan, et al.
Veröffentlicht: (2026)
VIRAL: Visual In-Context Reasoning via Analogy in Diffusion Transformers
von: Li, Zhiwen, et al.
Veröffentlicht: (2026)
von: Li, Zhiwen, et al.
Veröffentlicht: (2026)
UltraGen: Efficient Ultra-High-Resolution Image Generation with Hierarchical Local Attention
von: Zhang, Yuyao, et al.
Veröffentlicht: (2025)
von: Zhang, Yuyao, et al.
Veröffentlicht: (2025)
AlignedGen: Aligning Style Across Generated Images
von: Zhang, Jiexuan, et al.
Veröffentlicht: (2025)
von: Zhang, Jiexuan, et al.
Veröffentlicht: (2025)
MaskAttn-SDXL: Controllable Region-Level Text-To-Image Generation
von: Chang, Yu, et al.
Veröffentlicht: (2025)
von: Chang, Yu, et al.
Veröffentlicht: (2025)
CaliTex: Geometry-Calibrated Attention for View-Coherent 3D Texture Generation
von: Liu, Chenyu, et al.
Veröffentlicht: (2025)
von: Liu, Chenyu, et al.
Veröffentlicht: (2025)
UltraGen: High-Resolution Video Generation with Hierarchical Attention
von: Hu, Teng, et al.
Veröffentlicht: (2025)
von: Hu, Teng, et al.
Veröffentlicht: (2025)
Breaking the Modality Barrier: Universal Embedding Learning with Multimodal LLMs
von: Gu, Tiancheng, et al.
Veröffentlicht: (2025)
von: Gu, Tiancheng, et al.
Veröffentlicht: (2025)
UniGenDet: A Unified Generative-Discriminative Framework for Co-Evolutionary Image Generation and Generated Image Detection
von: Zhang, Yanran, et al.
Veröffentlicht: (2026)
von: Zhang, Yanran, et al.
Veröffentlicht: (2026)
SerialGen: Personalized Image Generation by First Standardization Then Personalization
von: Xie, Cong, et al.
Veröffentlicht: (2024)
von: Xie, Cong, et al.
Veröffentlicht: (2024)
From Text to Mask: Localizing Entities Using the Attention of Text-to-Image Diffusion Models
von: Xiao, Changming, et al.
Veröffentlicht: (2023)
von: Xiao, Changming, et al.
Veröffentlicht: (2023)
Gen-Searcher: Reinforcing Agentic Search for Image Generation
von: Feng, Kaituo, et al.
Veröffentlicht: (2026)
von: Feng, Kaituo, et al.
Veröffentlicht: (2026)
GDGS: 3D Gaussian Splatting Via Geometry-Guided Initialization And Dynamic Density Control
von: Wang, Xingjun, et al.
Veröffentlicht: (2025)
von: Wang, Xingjun, et al.
Veröffentlicht: (2025)
RecipeGen: A Benchmark for Real-World Recipe Image Generation
von: Zhang, Ruoxuan, et al.
Veröffentlicht: (2025)
von: Zhang, Ruoxuan, et al.
Veröffentlicht: (2025)
RA-Det: Towards Universal Detection of AI-Generated Images via Robustness Asymmetry
von: Wang, Xinchang, et al.
Veröffentlicht: (2026)
von: Wang, Xinchang, et al.
Veröffentlicht: (2026)
GenSpace: Benchmarking Spatially-Aware Image Generation
von: Wang, Zehan, et al.
Veröffentlicht: (2025)
von: Wang, Zehan, et al.
Veröffentlicht: (2025)
TextCenGen: Attention-Guided Text-Centric Background Adaptation for Text-to-Image Generation
von: Liang, Tianyi, et al.
Veröffentlicht: (2024)
von: Liang, Tianyi, et al.
Veröffentlicht: (2024)
LSReGen: Large-Scale Regional Generator via Backward Guidance Framework
von: Zhang, Bowen, et al.
Veröffentlicht: (2024)
von: Zhang, Bowen, et al.
Veröffentlicht: (2024)
BadPatch: Diffusion-Based Generation of Physical Adversarial Patches
von: Wang, Zhixiang, et al.
Veröffentlicht: (2024)
von: Wang, Zhixiang, et al.
Veröffentlicht: (2024)
GenAgent: Scaling Text-to-Image Generation via Agentic Multimodal Reasoning
von: Jiang, Kaixun, et al.
Veröffentlicht: (2026)
von: Jiang, Kaixun, et al.
Veröffentlicht: (2026)
ArtGen: Conditional Generative Modeling of Articulated Objects in Arbitrary Part-Level States
von: Wang, Haowen, et al.
Veröffentlicht: (2025)
von: Wang, Haowen, et al.
Veröffentlicht: (2025)
RAGSR: Regional Attention Guided Diffusion for Image Super-Resolution
von: He, Haodong, et al.
Veröffentlicht: (2025)
von: He, Haodong, et al.
Veröffentlicht: (2025)
RevSAM2: Prompt SAM2 for Medical Image Segmentation via Reverse-Propagation without Fine-tuning
von: Bai, Yunhao, et al.
Veröffentlicht: (2024)
von: Bai, Yunhao, et al.
Veröffentlicht: (2024)
DragEntity: Trajectory Guided Video Generation using Entity and Positional Relationships
von: Wan, Zhang, et al.
Veröffentlicht: (2024)
von: Wan, Zhang, et al.
Veröffentlicht: (2024)
AIM: Additional Image Guided Generation of Transferable Adversarial Attacks
von: Li, Teng, et al.
Veröffentlicht: (2025)
von: Li, Teng, et al.
Veröffentlicht: (2025)
FOCUS: Optimal Control for Multi-Entity World Modeling in Text-to-Image Generation
von: Bill, Eric Tillmann, et al.
Veröffentlicht: (2025)
von: Bill, Eric Tillmann, et al.
Veröffentlicht: (2025)
SafeCtrl: Region-Aware Safety Control for Text-to-Image Diffusion via Detect-Then-Suppress
von: Zhang, Lingyun, et al.
Veröffentlicht: (2026)
von: Zhang, Lingyun, et al.
Veröffentlicht: (2026)
GenMix: Combining Generative and Mixture Data Augmentation for Medical Image Classification
von: Lee, Hansang, et al.
Veröffentlicht: (2024)
von: Lee, Hansang, et al.
Veröffentlicht: (2024)
SafeCtrl: Region-Based Safety Control for Text-to-Image Diffusion via Detect-Then-Suppress
von: Zhang, Lingyun, et al.
Veröffentlicht: (2025)
von: Zhang, Lingyun, et al.
Veröffentlicht: (2025)
MathGen: Revealing the Illusion of Mathematical Competence through Text-to-Image Generation
von: Liu, Ruiyao, et al.
Veröffentlicht: (2026)
von: Liu, Ruiyao, et al.
Veröffentlicht: (2026)
Diffutoon: High-Resolution Editable Toon Shading via Diffusion Models
von: Duan, Zhongjie, et al.
Veröffentlicht: (2024)
von: Duan, Zhongjie, et al.
Veröffentlicht: (2024)
GUSLO: General and Unified Structured Light Optimization
von: Wan, Tinglei, et al.
Veröffentlicht: (2025)
von: Wan, Tinglei, et al.
Veröffentlicht: (2025)
WeditGAN: Few-Shot Image Generation via Latent Space Relocation
von: Duan, Yuxuan, et al.
Veröffentlicht: (2023)
von: Duan, Yuxuan, et al.
Veröffentlicht: (2023)
GenRC: Generative 3D Room Completion from Sparse Image Collections
von: Li, Ming-Feng, et al.
Veröffentlicht: (2024)
von: Li, Ming-Feng, et al.
Veröffentlicht: (2024)
TabletopGen: Instance-Level Interactive 3D Tabletop Scene Generation from Text or Single Image
von: Wang, Ziqian, et al.
Veröffentlicht: (2025)
von: Wang, Ziqian, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Nexus-Gen: Unified Image Understanding, Generation, and Editing via Prefilled Autoregression in Shared Embedding Space
von: Zhang, Hong, et al.
Veröffentlicht: (2025) -
Diffusion Templates: A Unified Plugin Framework for Controllable Diffusion
von: Duan, Zhongjie, et al.
Veröffentlicht: (2026) -
AutoLoRA: Automatic LoRA Retrieval and Fine-Grained Gated Fusion for Text-to-Image Generation
von: Li, Zhiwen, et al.
Veröffentlicht: (2025) -
AttriCtrl: Fine-Grained Control of Aesthetic Attribute Intensity in Diffusion Models
von: Chen, Die, et al.
Veröffentlicht: (2025) -
ArtAug: Enhancing Text-to-Image Generation through Synthesis-Understanding Interaction
von: Duan, Zhongjie, et al.
Veröffentlicht: (2024)