Group Relative Attention Guidance for Image Editing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Xuanpu, Niu, Xuesong, Chen, Ruidong, Song, Dan, Zeng, Jianhao, Du, Penghui, Cao, Haoxiang, Wu, Kai, Liu, An-an |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Layer-wise Instance Binding for Regional and Occlusion Control in Text-to-Image Diffusion Transformers
von: Chen, Ruidong, et al.
Veröffentlicht: (2026)
von: Chen, Ruidong, et al.
Veröffentlicht: (2026)
Better Fit: Accommodate Variations in Clothing Types for Virtual Try-on
von: Song, Dan, et al.
Veröffentlicht: (2024)
von: Song, Dan, et al.
Veröffentlicht: (2024)
Eevee: Towards Close-up High-resolution Video-based Virtual Try-on
von: Zeng, Jianhao, et al.
Veröffentlicht: (2025)
von: Zeng, Jianhao, et al.
Veröffentlicht: (2025)
BooW-VTON: Boosting In-the-Wild Virtual Try-On via Mask-Free Pseudo Data Training
von: Zhang, Xuanpu, et al.
Veröffentlicht: (2024)
von: Zhang, Xuanpu, et al.
Veröffentlicht: (2024)
Normalized Attention Guidance: Universal Negative Guidance for Diffusion Models
von: Chen, Dar-Yen, et al.
Veröffentlicht: (2025)
von: Chen, Dar-Yen, et al.
Veröffentlicht: (2025)
Image-Based Virtual Try-On: A Survey
von: Song, Dan, et al.
Veröffentlicht: (2023)
von: Song, Dan, et al.
Veröffentlicht: (2023)
Restorer: Removing Multi-Degradation with All-Axis Attention and Prompt Guidance
von: Mao, Jiawei, et al.
Veröffentlicht: (2024)
von: Mao, Jiawei, et al.
Veröffentlicht: (2024)
E4C: Enhance Editability for Text-Based Image Editing by Harnessing Efficient CLIP Guidance
von: Huang, Tianrui, et al.
Veröffentlicht: (2024)
von: Huang, Tianrui, et al.
Veröffentlicht: (2024)
Semantics-Aware Attention Guidance for Diagnosing Whole Slide Images
von: Liu, Kechun, et al.
Veröffentlicht: (2024)
von: Liu, Kechun, et al.
Veröffentlicht: (2024)
Boosting Latent Diffusion Models via Disentangled Representation Alignment
von: Page, John, et al.
Veröffentlicht: (2026)
von: Page, John, et al.
Veröffentlicht: (2026)
From Competition to Coopetition: Coopetitive Training-Free Image Editing Based on Text Guidance
von: Shen, Jinhao, et al.
Veröffentlicht: (2026)
von: Shen, Jinhao, et al.
Veröffentlicht: (2026)
VAGS: Velocity Adaptive Guidance Scale for Image Editing and Generation
von: Luo, Yan, et al.
Veröffentlicht: (2026)
von: Luo, Yan, et al.
Veröffentlicht: (2026)
Guidance Free Image Editing via Explicit Conditioning
von: Noroozi, Mehdi, et al.
Veröffentlicht: (2025)
von: Noroozi, Mehdi, et al.
Veröffentlicht: (2025)
Dual-Channel Attention Guidance for Training-Free Image Editing Control in Diffusion Transformers
von: Li, Guandong
Veröffentlicht: (2026)
von: Li, Guandong
Veröffentlicht: (2026)
Region-Constrained Group Relative Policy Optimization for Flow-Based Image Editing
von: Ouyang, Zhuohan, et al.
Veröffentlicht: (2026)
von: Ouyang, Zhuohan, et al.
Veröffentlicht: (2026)
Towards Understanding Cross and Self-Attention in Stable Diffusion for Text-Guided Image Editing
von: Liu, Bingyan, et al.
Veröffentlicht: (2024)
von: Liu, Bingyan, et al.
Veröffentlicht: (2024)
Unified Diffusion-Based Rigid and Non-Rigid Editing with Text and Image Guidance
von: Wang, Jiacheng, et al.
Veröffentlicht: (2024)
von: Wang, Jiacheng, et al.
Veröffentlicht: (2024)
CAT-DM: Controllable Accelerated Virtual Try-on with Diffusion Model
von: Zeng, Jianhao, et al.
Veröffentlicht: (2023)
von: Zeng, Jianhao, et al.
Veröffentlicht: (2023)
UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models
von: Jiang, Hong, et al.
Veröffentlicht: (2026)
von: Jiang, Hong, et al.
Veröffentlicht: (2026)
Group Editing: Edit Multiple Images in One Go
von: Ma, Yue, et al.
Veröffentlicht: (2026)
von: Ma, Yue, et al.
Veröffentlicht: (2026)
Tuning-Free Image Customization with Image and Text Guidance
von: Li, Pengzhi, et al.
Veröffentlicht: (2024)
von: Li, Pengzhi, et al.
Veröffentlicht: (2024)
Visual-CoG: Stage-Aware Reinforcement Learning with Chain of Guidance for Text-to-Image Generation
von: Li, Yaqi, et al.
Veröffentlicht: (2025)
von: Li, Yaqi, et al.
Veröffentlicht: (2025)
Diffusion-Based Conditional Image Editing through Optimized Inference with Guidance
von: Lee, Hyunsoo, et al.
Veröffentlicht: (2024)
von: Lee, Hyunsoo, et al.
Veröffentlicht: (2024)
Noise Map Guidance: Inversion with Spatial Context for Real Image Editing
von: Cho, Hansam, et al.
Veröffentlicht: (2024)
von: Cho, Hansam, et al.
Veröffentlicht: (2024)
Motion Guidance: Diffusion-Based Image Editing with Differentiable Motion Estimators
von: Geng, Daniel, et al.
Veröffentlicht: (2024)
von: Geng, Daniel, et al.
Veröffentlicht: (2024)
GroupDiff: Diffusion-based Group Portrait Editing
von: Jiang, Yuming, et al.
Veröffentlicht: (2024)
von: Jiang, Yuming, et al.
Veröffentlicht: (2024)
SmokeNet: Efficient Smoke Segmentation Leveraging Multiscale Convolutions and Multiview Attention Mechanisms
von: Liu, Xuesong, et al.
Veröffentlicht: (2025)
von: Liu, Xuesong, et al.
Veröffentlicht: (2025)
Point-to-Point: Sparse Motion Guidance for Controllable Video Editing
von: Song, Yeji, et al.
Veröffentlicht: (2025)
von: Song, Yeji, et al.
Veröffentlicht: (2025)
Patch-Based Stochastic Attention for Image Editing
von: Cherel, Nicolas, et al.
Veröffentlicht: (2022)
von: Cherel, Nicolas, et al.
Veröffentlicht: (2022)
MedDiff-FT: Data-Efficient Diffusion Model Fine-tuning with Structural Guidance for Controllable Medical Image Synthesis
von: Xie, Jianhao, et al.
Veröffentlicht: (2025)
von: Xie, Jianhao, et al.
Veröffentlicht: (2025)
A Style is Worth One Code: Unlocking Code-to-Style Image Generation with Discrete Style Space
von: Liu, Huijie, et al.
Veröffentlicht: (2025)
von: Liu, Huijie, et al.
Veröffentlicht: (2025)
GaussianPainter: Painting Point Cloud into 3D Gaussians with Normal Guidance
von: Zhou, Jingqiu, et al.
Veröffentlicht: (2024)
von: Zhou, Jingqiu, et al.
Veröffentlicht: (2024)
TextCtrl: Diffusion-based Scene Text Editing with Prior Guidance Control
von: Zeng, Weichao, et al.
Veröffentlicht: (2024)
von: Zeng, Weichao, et al.
Veröffentlicht: (2024)
EditEmoTalk: Controllable Speech-Driven 3D Facial Animation with Continuous Expression Editing
von: Jiang, Diqiong, et al.
Veröffentlicht: (2026)
von: Jiang, Diqiong, et al.
Veröffentlicht: (2026)
Tuning-free Instruction-based Video Editing Via Structural Noise Initialization and Guidance
von: Wu, Song, et al.
Veröffentlicht: (2026)
von: Wu, Song, et al.
Veröffentlicht: (2026)
FocusDiT: Masking Queries in Diffusion Transformers for Fine-grained Image Generation
von: Fang, Xueji, et al.
Veröffentlicht: (2026)
von: Fang, Xueji, et al.
Veröffentlicht: (2026)
Training-free Subject-Enhanced Attention Guidance for Compositional Text-to-image Generation
von: Liu, Shengyuan, et al.
Veröffentlicht: (2024)
von: Liu, Shengyuan, et al.
Veröffentlicht: (2024)
StableDrag: Stable Dragging for Point-based Image Editing
von: Cui, Yutao, et al.
Veröffentlicht: (2024)
von: Cui, Yutao, et al.
Veröffentlicht: (2024)
Kiwi-Edit: Versatile Video Editing via Instruction and Reference Guidance
von: Lin, Yiqi, et al.
Veröffentlicht: (2026)
von: Lin, Yiqi, et al.
Veröffentlicht: (2026)
Re-Align: Structured Reasoning-guided Alignment for In-Context Image Generation and Editing
von: He, Runze, et al.
Veröffentlicht: (2026)
von: He, Runze, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Layer-wise Instance Binding for Regional and Occlusion Control in Text-to-Image Diffusion Transformers
von: Chen, Ruidong, et al.
Veröffentlicht: (2026) -
Better Fit: Accommodate Variations in Clothing Types for Virtual Try-on
von: Song, Dan, et al.
Veröffentlicht: (2024) -
Eevee: Towards Close-up High-resolution Video-based Virtual Try-on
von: Zeng, Jianhao, et al.
Veröffentlicht: (2025) -
BooW-VTON: Boosting In-the-Wild Virtual Try-On via Mask-Free Pseudo Data Training
von: Zhang, Xuanpu, et al.
Veröffentlicht: (2024) -
Normalized Attention Guidance: Universal Negative Guidance for Diffusion Models
von: Chen, Dar-Yen, et al.
Veröffentlicht: (2025)