The Devil is in Attention Sharing: Improving Complex Non-rigid Image Editing Faithfulness via Attention Synergy
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Zhuo, Wei, Fanyue, Xu, Runze, Li, Jingjing, Duan, Lixin, Yao, Angela, Li, Wen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Devil is in the Spurious Correlations: Boosting Moment Retrieval with Dynamic Learning
by: Zhou, Xinyang, et al.
Published: (2025)
by: Zhou, Xinyang, et al.
Published: (2025)
Powerful and Flexible: Personalized Text-to-Image Generation via Reinforcement Learning
by: Wei, Fanyue, et al.
Published: (2024)
by: Wei, Fanyue, et al.
Published: (2024)
ResCLIP: Residual Attention for Training-free Dense Vision-language Inference
by: Yang, Yuhang, et al.
Published: (2024)
by: Yang, Yuhang, et al.
Published: (2024)
SynergyWarpNet: Attention-Guided Cooperative Warping for Neural Portrait Animation
by: Li, Shihang, et al.
Published: (2025)
by: Li, Shihang, et al.
Published: (2025)
Improving Diffusion-Based Image Editing Faithfulness via Guidance and Scheduling
by: Cho, Hansam, et al.
Published: (2025)
by: Cho, Hansam, et al.
Published: (2025)
Faithful-MR1: Faithful Multimodal Reasoning via Anchoring and Reinforcing Visual Attention
by: Tian, Changyuan, et al.
Published: (2026)
by: Tian, Changyuan, et al.
Published: (2026)
GateAttentionPose: Enhancing Pose Estimation with Agent Attention and Improved Gated Convolutions
by: Feng, Liang, et al.
Published: (2024)
by: Feng, Liang, et al.
Published: (2024)
ImVideoEdit: Image-learning Video Editing via 2D Spatial Difference Attention Blocks
by: Xu, Jiayang, et al.
Published: (2026)
by: Xu, Jiayang, et al.
Published: (2026)
Devil is in the Detail: Towards Injecting Fine Details of Image Prompt in Image Generation via Conflict-free Guidance and Stratified Attention
by: Jo, Kyungmin, et al.
Published: (2025)
by: Jo, Kyungmin, et al.
Published: (2025)
Style Aligned Image Generation via Shared Attention
by: Hertz, Amir, et al.
Published: (2023)
by: Hertz, Amir, et al.
Published: (2023)
Immunizing Images from Text to Image Editing via Adversarial Cross-Attention
by: Trippodo, Matteo, et al.
Published: (2025)
by: Trippodo, Matteo, et al.
Published: (2025)
Group Relative Attention Guidance for Image Editing
by: Zhang, Xuanpu, et al.
Published: (2025)
by: Zhang, Xuanpu, et al.
Published: (2025)
Patch-Based Stochastic Attention for Image Editing
by: Cherel, Nicolas, et al.
Published: (2022)
by: Cherel, Nicolas, et al.
Published: (2022)
Scale Contrastive Learning with Selective Attentions for Blind Image Quality Assessment
by: Hu, Runze, et al.
Published: (2024)
by: Hu, Runze, et al.
Published: (2024)
Faithful Attention Explainer: Verbalizing Decisions Based on Discriminative Features
by: Rong, Yao, et al.
Published: (2024)
by: Rong, Yao, et al.
Published: (2024)
LIME: Localized Image Editing via Attention Regularization in Diffusion Models
by: Simsar, Enis, et al.
Published: (2023)
by: Simsar, Enis, et al.
Published: (2023)
Uniform Attention Maps: Boosting Image Fidelity in Reconstruction and Editing
by: Mo, Wenyi, et al.
Published: (2024)
by: Mo, Wenyi, et al.
Published: (2024)
InstructBrush: Learning Attention-based Instruction Optimization for Image Editing
by: Zhao, Ruoyu, et al.
Published: (2024)
by: Zhao, Ruoyu, et al.
Published: (2024)
Devils in Middle Layers of Large Vision-Language Models: Interpreting, Detecting and Mitigating Object Hallucinations via Attention Lens
by: Jiang, Zhangqi, et al.
Published: (2024)
by: Jiang, Zhangqi, et al.
Published: (2024)
LIPE: Learning Personalized Identity Prior for Non-rigid Image Editing
by: Liu, Aoyang, et al.
Published: (2024)
by: Liu, Aoyang, et al.
Published: (2024)
AID: Attention Interpolation of Text-to-Image Diffusion
by: He, Qiyuan, et al.
Published: (2024)
by: He, Qiyuan, et al.
Published: (2024)
Qffusion: Controllable Portrait Video Editing via Quadrant-Grid Attention Learning
by: Li, Maomao, et al.
Published: (2025)
by: Li, Maomao, et al.
Published: (2025)
AttnRouter: Per-Category Attention Routing for Training-Free Image Editing on MMDiT
by: Li, Guandong, et al.
Published: (2026)
by: Li, Guandong, et al.
Published: (2026)
Mixture-of-Top-k Attention: Efficient Attention via Scalable Fast Weights
by: Wen, Qishuai, et al.
Published: (2026)
by: Wen, Qishuai, et al.
Published: (2026)
MedFlowSeg: Flow Matching for Medical Image Segmentation with Frequency-Aware Attention
by: Chen, Zhi, et al.
Published: (2026)
by: Chen, Zhi, et al.
Published: (2026)
Dual-Channel Attention Guidance for Training-Free Image Editing Control in Diffusion Transformers
by: Li, Guandong
Published: (2026)
by: Li, Guandong
Published: (2026)
InstanceAssemble: Layout-Aware Image Generation via Instance Assembling Attention
by: Xiang, Qiang, et al.
Published: (2025)
by: Xiang, Qiang, et al.
Published: (2025)
SpikeVideoFormer: An Efficient Spike-Driven Video Transformer with Hamming Attention and $\mathcal{O}(T)$ Complexity
by: Zou, Shihao, et al.
Published: (2025)
by: Zou, Shihao, et al.
Published: (2025)
Improving Brain-to-Image Reconstruction via Fine-Grained Text Bridging
by: Xia, Runze, et al.
Published: (2025)
by: Xia, Runze, et al.
Published: (2025)
FIA-Edit: Frequency-Interactive Attention for Efficient and High-Fidelity Inversion-Free Text-Guided Image Editing
by: Yang, Kaixiang, et al.
Published: (2025)
by: Yang, Kaixiang, et al.
Published: (2025)
Consolidating Attention Features for Multi-view Image Editing
by: Patashnik, Or, et al.
Published: (2024)
by: Patashnik, Or, et al.
Published: (2024)
Nodule-DETR: A Novel DETR Architecture with Frequency-Channel Attention for Ultrasound Thyroid Nodule Detection
by: Wang, Jingjing, et al.
Published: (2026)
by: Wang, Jingjing, et al.
Published: (2026)
Re-Attentional Controllable Video Diffusion Editing
by: Wang, Yuanzhi, et al.
Published: (2024)
by: Wang, Yuanzhi, et al.
Published: (2024)
Learning Semantic Latent Directions for Accurate and Controllable Human Motion Prediction
by: Xu, Guowei, et al.
Published: (2024)
by: Xu, Guowei, et al.
Published: (2024)
FMDConv: Fast Multi-Attention Dynamic Convolution via Speed-Accuracy Trade-off
by: Zhang, Tianyu, et al.
Published: (2025)
by: Zhang, Tianyu, et al.
Published: (2025)
Native 3D Editing with Full Attention
by: Cai, Weiwei, et al.
Published: (2025)
by: Cai, Weiwei, et al.
Published: (2025)
MATCNN: Infrared and Visible Image Fusion Method Based on Multi-scale CNN with Attention Transformer
by: Liu, Jingjing, et al.
Published: (2025)
by: Liu, Jingjing, et al.
Published: (2025)
NOAH: Learning Pairwise Object Category Attentions for Image Classification
by: Li, Chao, et al.
Published: (2024)
by: Li, Chao, et al.
Published: (2024)
RealViformer: Investigating Attention for Real-World Video Super-Resolution
by: Zhang, Yuehan, et al.
Published: (2024)
by: Zhang, Yuehan, et al.
Published: (2024)
LIVEditor-14B: Lightning Unified Video Editing via In-Context Sparse Attention
by: Shao, Shitong, et al.
Published: (2026)
by: Shao, Shitong, et al.
Published: (2026)
Similar Items
-
The Devil is in the Spurious Correlations: Boosting Moment Retrieval with Dynamic Learning
by: Zhou, Xinyang, et al.
Published: (2025) -
Powerful and Flexible: Personalized Text-to-Image Generation via Reinforcement Learning
by: Wei, Fanyue, et al.
Published: (2024) -
ResCLIP: Residual Attention for Training-free Dense Vision-language Inference
by: Yang, Yuhang, et al.
Published: (2024) -
SynergyWarpNet: Attention-Guided Cooperative Warping for Neural Portrait Animation
by: Li, Shihang, et al.
Published: (2025) -
Improving Diffusion-Based Image Editing Faithfulness via Guidance and Scheduling
by: Cho, Hansam, et al.
Published: (2025)