Uniform Attention Maps: Boosting Image Fidelity in Reconstruction and Editing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mo, Wenyi, Zhang, Tianyu, Bai, Yalong, Su, Bing, Wen, Ji-Rong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Dynamic Prompt Optimizing for Text-to-Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
von: Mo, Wenyi, et al.
Veröffentlicht: (2024)
PrefGen: Multimodal Preference Learning for Preference-Conditioned Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
Enhancing Reward Models for High-quality Image Generation: Beyond Text-Image Alignment
von: Ba, Ying, et al.
Veröffentlicht: (2025)
von: Ba, Ying, et al.
Veröffentlicht: (2025)
Pareto-Guided Optimal Transport for Multi-Reward Alignment
von: Ba, Ying, et al.
Veröffentlicht: (2026)
von: Ba, Ying, et al.
Veröffentlicht: (2026)
V2Flow: Unifying Visual Tokenization and Large Language Model Vocabularies for Autoregressive Image Generation
von: Zhang, Guiwei, et al.
Veröffentlicht: (2025)
von: Zhang, Guiwei, et al.
Veröffentlicht: (2025)
Learning User Preferences for Image Generation Model
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
von: Mo, Wenyi, et al.
Veröffentlicht: (2025)
Hyperspherical Autoencoder for High-Fidelity Image Reconstruction and Generation
von: Chang, Hun, et al.
Veröffentlicht: (2026)
von: Chang, Hun, et al.
Veröffentlicht: (2026)
Scratching Visual Transformer's Back with Uniform Attention
von: Hyeon-Woo, Nam, et al.
Veröffentlicht: (2022)
von: Hyeon-Woo, Nam, et al.
Veröffentlicht: (2022)
Not All Attention Heads Are What You Need: Refining CLIP's Image Representation with Attention Ablation
von: Lin, Feng, et al.
Veröffentlicht: (2025)
von: Lin, Feng, et al.
Veröffentlicht: (2025)
Do Sparse Subnetworks Exhibit Cognitively Aligned Attention? Effects of Pruning on Saliency Map Fidelity, Sparsity, and Concept Coherence
von: Suwal, Sanish, et al.
Veröffentlicht: (2025)
von: Suwal, Sanish, et al.
Veröffentlicht: (2025)
Text-Driven Image Editing via Learnable Regions
von: Lin, Yuanze, et al.
Veröffentlicht: (2023)
von: Lin, Yuanze, et al.
Veröffentlicht: (2023)
ConsistDreamer: 3D-Consistent 2D Diffusion for High-Fidelity Scene Editing
von: Chen, Jun-Kun, et al.
Veröffentlicht: (2024)
von: Chen, Jun-Kun, et al.
Veröffentlicht: (2024)
Attention in Diffusion Model: A Survey
von: Hua, Litao, et al.
Veröffentlicht: (2025)
von: Hua, Litao, et al.
Veröffentlicht: (2025)
SuperEdit: Rectifying and Facilitating Supervision for Instruction-Based Image Editing
von: Li, Ming, et al.
Veröffentlicht: (2025)
von: Li, Ming, et al.
Veröffentlicht: (2025)
MagicBrush: A Manually Annotated Dataset for Instruction-Guided Image Editing
von: Zhang, Kai, et al.
Veröffentlicht: (2023)
von: Zhang, Kai, et al.
Veröffentlicht: (2023)
Towards Efficient Diffusion-Based Image Editing with Instant Attention Masks
von: Zou, Siyu, et al.
Veröffentlicht: (2024)
von: Zou, Siyu, et al.
Veröffentlicht: (2024)
Pathways on the Image Manifold: Image Editing via Video Generation
von: Rotstein, Noam, et al.
Veröffentlicht: (2024)
von: Rotstein, Noam, et al.
Veröffentlicht: (2024)
Class-Discriminative Attention Maps for Vision Transformers
von: Brocki, Lennart, et al.
Veröffentlicht: (2023)
von: Brocki, Lennart, et al.
Veröffentlicht: (2023)
Attention-based Shape-Deformation Networks for Artifact-Free Geometry Reconstruction of Lumbar Spine from MR Images
von: Qian, Linchen, et al.
Veröffentlicht: (2024)
von: Qian, Linchen, et al.
Veröffentlicht: (2024)
EditWorld: Simulating World Dynamics for Instruction-Following Image Editing
von: Yang, Ling, et al.
Veröffentlicht: (2024)
von: Yang, Ling, et al.
Veröffentlicht: (2024)
ImageEdit-R1: Boosting Multi-Agent Image Editing via Reinforcement Learning
von: Zhao, Yiran, et al.
Veröffentlicht: (2026)
von: Zhao, Yiran, et al.
Veröffentlicht: (2026)
Editing Massive Concepts in Text-to-Image Diffusion Models
von: Xiong, Tianwei, et al.
Veröffentlicht: (2024)
von: Xiong, Tianwei, et al.
Veröffentlicht: (2024)
Camouflaged Image Synthesis Is All You Need to Boost Camouflaged Detection
von: Zhang, Haichao, et al.
Veröffentlicht: (2023)
von: Zhang, Haichao, et al.
Veröffentlicht: (2023)
Generalized Neighborhood Attention: Multi-dimensional Sparse Attention at the Speed of Light
von: Hassani, Ali, et al.
Veröffentlicht: (2025)
von: Hassani, Ali, et al.
Veröffentlicht: (2025)
Optimizing Negative Prompts for Enhanced Aesthetics and Fidelity in Text-To-Image Generation
von: Ogezi, Michael, et al.
Veröffentlicht: (2024)
von: Ogezi, Michael, et al.
Veröffentlicht: (2024)
Preserving Product Fidelity in Large Scale Image Recontextualization with Diffusion Models
von: Malhi, Ishaan, et al.
Veröffentlicht: (2025)
von: Malhi, Ishaan, et al.
Veröffentlicht: (2025)
MASC: Boosting Autoregressive Image Generation with a Manifold-Aligned Semantic Clustering
von: He, Lixuan, et al.
Veröffentlicht: (2025)
von: He, Lixuan, et al.
Veröffentlicht: (2025)
Edit Away and My Face Will not Stay: Personal Biometric Defense against Malicious Generative Editing
von: Wang, Hanhui, et al.
Veröffentlicht: (2024)
von: Wang, Hanhui, et al.
Veröffentlicht: (2024)
Scaling Diffusion Mamba with Bidirectional SSMs for Efficient Image and Video Generation
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
Specify and Edit: Overcoming Ambiguity in Text-Based Image Editing
von: Iakovleva, Ekaterina, et al.
Veröffentlicht: (2024)
von: Iakovleva, Ekaterina, et al.
Veröffentlicht: (2024)
CoDi: Conditional Diffusion Distillation for Higher-Fidelity and Faster Image Generation
von: Mei, Kangfu, et al.
Veröffentlicht: (2023)
von: Mei, Kangfu, et al.
Veröffentlicht: (2023)
Diffusion Model-Based Video Editing: A Survey
von: Sun, Wenhao, et al.
Veröffentlicht: (2024)
von: Sun, Wenhao, et al.
Veröffentlicht: (2024)
Meta-CoT: Enhancing Granularity and Generalization in Image Editing
von: Zhang, Shiyi, et al.
Veröffentlicht: (2026)
von: Zhang, Shiyi, et al.
Veröffentlicht: (2026)
Enhancing Multi-task Learning Capability of Medical Generalist Foundation Model via Image-centric Multi-annotation Data
von: Zhu, Xun, et al.
Veröffentlicht: (2025)
von: Zhu, Xun, et al.
Veröffentlicht: (2025)
Geometrically Constrained Stenosis Editing in Coronary Angiography via Entropic Optimal Transport
von: Li, Jialin, et al.
Veröffentlicht: (2026)
von: Li, Jialin, et al.
Veröffentlicht: (2026)
MagDiff: Multi-Alignment Diffusion for High-Fidelity Video Generation and Editing
von: Zhao, Haoyu, et al.
Veröffentlicht: (2023)
von: Zhao, Haoyu, et al.
Veröffentlicht: (2023)
ICE-G: Image Conditional Editing of 3D Gaussian Splats
von: Jaganathan, Vishnu, et al.
Veröffentlicht: (2024)
von: Jaganathan, Vishnu, et al.
Veröffentlicht: (2024)
Improving Diffusion-Based Image Editing Faithfulness via Guidance and Scheduling
von: Cho, Hansam, et al.
Veröffentlicht: (2025)
von: Cho, Hansam, et al.
Veröffentlicht: (2025)
Contrastive Denoising Score for Text-guided Latent Diffusion Image Editing
von: Nam, Hyelin, et al.
Veröffentlicht: (2023)
von: Nam, Hyelin, et al.
Veröffentlicht: (2023)
MapDiffusion: Generative Diffusion for Vectorized Online HD Map Construction and Uncertainty Estimation in Autonomous Driving
von: Monninger, Thomas, et al.
Veröffentlicht: (2025)
von: Monninger, Thomas, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Dynamic Prompt Optimizing for Text-to-Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2024) -
PrefGen: Multimodal Preference Learning for Preference-Conditioned Image Generation
von: Mo, Wenyi, et al.
Veröffentlicht: (2025) -
Enhancing Reward Models for High-quality Image Generation: Beyond Text-Image Alignment
von: Ba, Ying, et al.
Veröffentlicht: (2025) -
Pareto-Guided Optimal Transport for Multi-Reward Alignment
von: Ba, Ying, et al.
Veröffentlicht: (2026) -
V2Flow: Unifying Visual Tokenization and Large Language Model Vocabularies for Autoregressive Image Generation
von: Zhang, Guiwei, et al.
Veröffentlicht: (2025)