Group Relative Attention Guidance for Image Editing
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Xuanpu, Niu, Xuesong, Chen, Ruidong, Song, Dan, Zeng, Jianhao, Du, Penghui, Cao, Haoxiang, Wu, Kai, Liu, An-an |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Layer-wise Instance Binding for Regional and Occlusion Control in Text-to-Image Diffusion Transformers
by: Chen, Ruidong, et al.
Published: (2026)
by: Chen, Ruidong, et al.
Published: (2026)
Better Fit: Accommodate Variations in Clothing Types for Virtual Try-on
by: Song, Dan, et al.
Published: (2024)
by: Song, Dan, et al.
Published: (2024)
Eevee: Towards Close-up High-resolution Video-based Virtual Try-on
by: Zeng, Jianhao, et al.
Published: (2025)
by: Zeng, Jianhao, et al.
Published: (2025)
BooW-VTON: Boosting In-the-Wild Virtual Try-On via Mask-Free Pseudo Data Training
by: Zhang, Xuanpu, et al.
Published: (2024)
by: Zhang, Xuanpu, et al.
Published: (2024)
Normalized Attention Guidance: Universal Negative Guidance for Diffusion Models
by: Chen, Dar-Yen, et al.
Published: (2025)
by: Chen, Dar-Yen, et al.
Published: (2025)
Image-Based Virtual Try-On: A Survey
by: Song, Dan, et al.
Published: (2023)
by: Song, Dan, et al.
Published: (2023)
Restorer: Removing Multi-Degradation with All-Axis Attention and Prompt Guidance
by: Mao, Jiawei, et al.
Published: (2024)
by: Mao, Jiawei, et al.
Published: (2024)
E4C: Enhance Editability for Text-Based Image Editing by Harnessing Efficient CLIP Guidance
by: Huang, Tianrui, et al.
Published: (2024)
by: Huang, Tianrui, et al.
Published: (2024)
Semantics-Aware Attention Guidance for Diagnosing Whole Slide Images
by: Liu, Kechun, et al.
Published: (2024)
by: Liu, Kechun, et al.
Published: (2024)
Boosting Latent Diffusion Models via Disentangled Representation Alignment
by: Page, John, et al.
Published: (2026)
by: Page, John, et al.
Published: (2026)
From Competition to Coopetition: Coopetitive Training-Free Image Editing Based on Text Guidance
by: Shen, Jinhao, et al.
Published: (2026)
by: Shen, Jinhao, et al.
Published: (2026)
VAGS: Velocity Adaptive Guidance Scale for Image Editing and Generation
by: Luo, Yan, et al.
Published: (2026)
by: Luo, Yan, et al.
Published: (2026)
Guidance Free Image Editing via Explicit Conditioning
by: Noroozi, Mehdi, et al.
Published: (2025)
by: Noroozi, Mehdi, et al.
Published: (2025)
Dual-Channel Attention Guidance for Training-Free Image Editing Control in Diffusion Transformers
by: Li, Guandong
Published: (2026)
by: Li, Guandong
Published: (2026)
Region-Constrained Group Relative Policy Optimization for Flow-Based Image Editing
by: Ouyang, Zhuohan, et al.
Published: (2026)
by: Ouyang, Zhuohan, et al.
Published: (2026)
Towards Understanding Cross and Self-Attention in Stable Diffusion for Text-Guided Image Editing
by: Liu, Bingyan, et al.
Published: (2024)
by: Liu, Bingyan, et al.
Published: (2024)
Unified Diffusion-Based Rigid and Non-Rigid Editing with Text and Image Guidance
by: Wang, Jiacheng, et al.
Published: (2024)
by: Wang, Jiacheng, et al.
Published: (2024)
CAT-DM: Controllable Accelerated Virtual Try-on with Diffusion Model
by: Zeng, Jianhao, et al.
Published: (2023)
by: Zeng, Jianhao, et al.
Published: (2023)
UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models
by: Jiang, Hong, et al.
Published: (2026)
by: Jiang, Hong, et al.
Published: (2026)
Group Editing: Edit Multiple Images in One Go
by: Ma, Yue, et al.
Published: (2026)
by: Ma, Yue, et al.
Published: (2026)
Tuning-Free Image Customization with Image and Text Guidance
by: Li, Pengzhi, et al.
Published: (2024)
by: Li, Pengzhi, et al.
Published: (2024)
Visual-CoG: Stage-Aware Reinforcement Learning with Chain of Guidance for Text-to-Image Generation
by: Li, Yaqi, et al.
Published: (2025)
by: Li, Yaqi, et al.
Published: (2025)
Diffusion-Based Conditional Image Editing through Optimized Inference with Guidance
by: Lee, Hyunsoo, et al.
Published: (2024)
by: Lee, Hyunsoo, et al.
Published: (2024)
Noise Map Guidance: Inversion with Spatial Context for Real Image Editing
by: Cho, Hansam, et al.
Published: (2024)
by: Cho, Hansam, et al.
Published: (2024)
Motion Guidance: Diffusion-Based Image Editing with Differentiable Motion Estimators
by: Geng, Daniel, et al.
Published: (2024)
by: Geng, Daniel, et al.
Published: (2024)
GroupDiff: Diffusion-based Group Portrait Editing
by: Jiang, Yuming, et al.
Published: (2024)
by: Jiang, Yuming, et al.
Published: (2024)
SmokeNet: Efficient Smoke Segmentation Leveraging Multiscale Convolutions and Multiview Attention Mechanisms
by: Liu, Xuesong, et al.
Published: (2025)
by: Liu, Xuesong, et al.
Published: (2025)
Point-to-Point: Sparse Motion Guidance for Controllable Video Editing
by: Song, Yeji, et al.
Published: (2025)
by: Song, Yeji, et al.
Published: (2025)
Patch-Based Stochastic Attention for Image Editing
by: Cherel, Nicolas, et al.
Published: (2022)
by: Cherel, Nicolas, et al.
Published: (2022)
MedDiff-FT: Data-Efficient Diffusion Model Fine-tuning with Structural Guidance for Controllable Medical Image Synthesis
by: Xie, Jianhao, et al.
Published: (2025)
by: Xie, Jianhao, et al.
Published: (2025)
A Style is Worth One Code: Unlocking Code-to-Style Image Generation with Discrete Style Space
by: Liu, Huijie, et al.
Published: (2025)
by: Liu, Huijie, et al.
Published: (2025)
GaussianPainter: Painting Point Cloud into 3D Gaussians with Normal Guidance
by: Zhou, Jingqiu, et al.
Published: (2024)
by: Zhou, Jingqiu, et al.
Published: (2024)
TextCtrl: Diffusion-based Scene Text Editing with Prior Guidance Control
by: Zeng, Weichao, et al.
Published: (2024)
by: Zeng, Weichao, et al.
Published: (2024)
EditEmoTalk: Controllable Speech-Driven 3D Facial Animation with Continuous Expression Editing
by: Jiang, Diqiong, et al.
Published: (2026)
by: Jiang, Diqiong, et al.
Published: (2026)
Tuning-free Instruction-based Video Editing Via Structural Noise Initialization and Guidance
by: Wu, Song, et al.
Published: (2026)
by: Wu, Song, et al.
Published: (2026)
FocusDiT: Masking Queries in Diffusion Transformers for Fine-grained Image Generation
by: Fang, Xueji, et al.
Published: (2026)
by: Fang, Xueji, et al.
Published: (2026)
Training-free Subject-Enhanced Attention Guidance for Compositional Text-to-image Generation
by: Liu, Shengyuan, et al.
Published: (2024)
by: Liu, Shengyuan, et al.
Published: (2024)
StableDrag: Stable Dragging for Point-based Image Editing
by: Cui, Yutao, et al.
Published: (2024)
by: Cui, Yutao, et al.
Published: (2024)
Kiwi-Edit: Versatile Video Editing via Instruction and Reference Guidance
by: Lin, Yiqi, et al.
Published: (2026)
by: Lin, Yiqi, et al.
Published: (2026)
Re-Align: Structured Reasoning-guided Alignment for In-Context Image Generation and Editing
by: He, Runze, et al.
Published: (2026)
by: He, Runze, et al.
Published: (2026)
Similar Items
-
Layer-wise Instance Binding for Regional and Occlusion Control in Text-to-Image Diffusion Transformers
by: Chen, Ruidong, et al.
Published: (2026) -
Better Fit: Accommodate Variations in Clothing Types for Virtual Try-on
by: Song, Dan, et al.
Published: (2024) -
Eevee: Towards Close-up High-resolution Video-based Virtual Try-on
by: Zeng, Jianhao, et al.
Published: (2025) -
BooW-VTON: Boosting In-the-Wild Virtual Try-On via Mask-Free Pseudo Data Training
by: Zhang, Xuanpu, et al.
Published: (2024) -
Normalized Attention Guidance: Universal Negative Guidance for Diffusion Models
by: Chen, Dar-Yen, et al.
Published: (2025)