Towards Robust Sequential Decomposition for Complex Image Editing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zeng, Zilai, Cao, Mingdeng, Li, Zijie, Lian, Xiaochen, Shi, Yichun, Zhu, Peihao, Sun, Chen, Wang, Peng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings
von: Li, Zijie, et al.
Veröffentlicht: (2026)
von: Li, Zijie, et al.
Veröffentlicht: (2026)
SeedEdit 3.0: Fast and High-Quality Generative Image Editing
von: Wang, Peng, et al.
Veröffentlicht: (2025)
von: Wang, Peng, et al.
Veröffentlicht: (2025)
ByteMorph: Benchmarking Instruction-Guided Image Editing with Non-Rigid Motions
von: Chang, Di, et al.
Veröffentlicht: (2025)
von: Chang, Di, et al.
Veröffentlicht: (2025)
SeedEdit: Align Image Re-Generation to Image Editing
von: Shi, Yichun, et al.
Veröffentlicht: (2024)
von: Shi, Yichun, et al.
Veröffentlicht: (2024)
Text-Aware Diffusion for Policy Learning
von: Luo, Calvin, et al.
Veröffentlicht: (2024)
von: Luo, Calvin, et al.
Veröffentlicht: (2024)
HQ-Edit: A High-Quality Dataset for Instruction-based Image Editing
von: Hui, Mude, et al.
Veröffentlicht: (2024)
von: Hui, Mude, et al.
Veröffentlicht: (2024)
VideoPainter: Any-length Video Inpainting and Editing with Plug-and-Play Context Control
von: Bian, Yuxuan, et al.
Veröffentlicht: (2025)
von: Bian, Yuxuan, et al.
Veröffentlicht: (2025)
Dual Diffusion for Unified Image Generation and Understanding
von: Li, Zijie, et al.
Veröffentlicht: (2024)
von: Li, Zijie, et al.
Veröffentlicht: (2024)
VEU-Bench: Towards Comprehensive Understanding of Video Editing
von: Li, Bozheng, et al.
Veröffentlicht: (2025)
von: Li, Bozheng, et al.
Veröffentlicht: (2025)
TAGE: Trustworthy Attribute Group Editing for Stable Few-shot Image Generation
von: Zhang, Ruicheng, et al.
Veröffentlicht: (2024)
von: Zhang, Ruicheng, et al.
Veröffentlicht: (2024)
SCOPE: Structured Decomposition and Conditional Skill Orchestration for Complex Image Generation
von: Ren, Tianfei, et al.
Veröffentlicht: (2026)
von: Ren, Tianfei, et al.
Veröffentlicht: (2026)
DLEBench: Evaluating Small-scale Object Editing Ability for Instruction-based Image Editing Model
von: Hong, Shibo, et al.
Veröffentlicht: (2026)
von: Hong, Shibo, et al.
Veröffentlicht: (2026)
RegionE: Adaptive Region-Aware Generation for Efficient Image Editing
von: Chen, Pengtao, et al.
Veröffentlicht: (2025)
von: Chen, Pengtao, et al.
Veröffentlicht: (2025)
Instruction-based Image Manipulation by Watching How Things Move
von: Cao, Mingdeng, et al.
Veröffentlicht: (2024)
von: Cao, Mingdeng, et al.
Veröffentlicht: (2024)
Towards Efficient Diffusion-Based Image Editing with Instant Attention Masks
von: Zou, Siyu, et al.
Veröffentlicht: (2024)
von: Zou, Siyu, et al.
Veröffentlicht: (2024)
$\texttt{Complex-Edit}$: CoT-Like Instruction Generation for Complexity-Controllable Image Editing Benchmark
von: Yang, Siwei, et al.
Veröffentlicht: (2025)
von: Yang, Siwei, et al.
Veröffentlicht: (2025)
Taming Lookup Tables for Efficient Image Retouching
von: Yang, Sidi, et al.
Veröffentlicht: (2024)
von: Yang, Sidi, et al.
Veröffentlicht: (2024)
MCIE: Multimodal LLM-Driven Complex Instruction Image Editing with Spatial Guidance
von: Bai, Xuehai, et al.
Veröffentlicht: (2026)
von: Bai, Xuehai, et al.
Veröffentlicht: (2026)
PICABench: How Far Are We from Physically Realistic Image Editing?
von: Pu, Yuandong, et al.
Veröffentlicht: (2025)
von: Pu, Yuandong, et al.
Veröffentlicht: (2025)
OpenGPT-4o-Image: A Comprehensive Dataset for Advanced Image Generation and Editing
von: Chen, Zhihong, et al.
Veröffentlicht: (2025)
von: Chen, Zhihong, et al.
Veröffentlicht: (2025)
WorldEdit: Towards Open-World Image Editing with a Knowledge-Informed Benchmark
von: Lin, Wang, et al.
Veröffentlicht: (2026)
von: Lin, Wang, et al.
Veröffentlicht: (2026)
QDFormer: Towards Robust Audiovisual Segmentation in Complex Environments with Quantization-based Semantic Decomposition
von: Li, Xiang, et al.
Veröffentlicht: (2023)
von: Li, Xiang, et al.
Veröffentlicht: (2023)
Flash-DMD: Towards High-Fidelity Few-Step Image Generation with Efficient Distillation and Joint Reinforcement Learning
von: Chen, Guanjie, et al.
Veröffentlicht: (2025)
von: Chen, Guanjie, et al.
Veröffentlicht: (2025)
AdvI2I: Adversarial Image Attack on Image-to-Image Diffusion models
von: Zeng, Yaopei, et al.
Veröffentlicht: (2024)
von: Zeng, Yaopei, et al.
Veröffentlicht: (2024)
VINCIE: Unlocking In-context Image Editing from Video
von: Qu, Leigang, et al.
Veröffentlicht: (2025)
von: Qu, Leigang, et al.
Veröffentlicht: (2025)
Learning Feature-Preserving Portrait Editing from Generated Pairs
von: Chen, Bowei, et al.
Veröffentlicht: (2024)
von: Chen, Bowei, et al.
Veröffentlicht: (2024)
UM-Text: A Unified Multimodal Model for Image Understanding and Visual Text Editing
von: Ma, Lichen, et al.
Veröffentlicht: (2026)
von: Ma, Lichen, et al.
Veröffentlicht: (2026)
Robust Watermarking Using Generative Priors Against Image Editing: From Benchmarking to Advances
von: Lu, Shilin, et al.
Veröffentlicht: (2024)
von: Lu, Shilin, et al.
Veröffentlicht: (2024)
Interpreting Global Perturbation Robustness of Image Models using Axiomatic Spectral Importance Decomposition
von: Luo, Róisín, et al.
Veröffentlicht: (2024)
von: Luo, Róisín, et al.
Veröffentlicht: (2024)
DPDEdit: Detail-Preserved Diffusion Models for Multimodal Fashion Image Editing
von: Wang, Xiaolong, et al.
Veröffentlicht: (2024)
von: Wang, Xiaolong, et al.
Veröffentlicht: (2024)
Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling
von: Wang, Yuran, et al.
Veröffentlicht: (2025)
von: Wang, Yuran, et al.
Veröffentlicht: (2025)
FAME: Fairness-aware Attention-modulated Video Editing
von: Wu, Zhangkai, et al.
Veröffentlicht: (2025)
von: Wu, Zhangkai, et al.
Veröffentlicht: (2025)
Draw-In-Mind: Rebalancing Designer-Painter Roles in Unified Multimodal Models Benefits Image Editing
von: Zeng, Ziyun, et al.
Veröffentlicht: (2025)
von: Zeng, Ziyun, et al.
Veröffentlicht: (2025)
Rolling Shutter Correction with Intermediate Distortion Flow Estimation
von: Cao, Mingdeng, et al.
Veröffentlicht: (2024)
von: Cao, Mingdeng, et al.
Veröffentlicht: (2024)
Navigating Beyond Dropout: An Intriguing Solution Towards Generalizable Image Super Resolution
von: Wang, Hongjun, et al.
Veröffentlicht: (2024)
von: Wang, Hongjun, et al.
Veröffentlicht: (2024)
DeltaSpace: A Semantic-aligned Feature Space for Flexible Text-guided Image Editing
von: Lyu, Yueming, et al.
Veröffentlicht: (2023)
von: Lyu, Yueming, et al.
Veröffentlicht: (2023)
XHand: Real-time Expressive Hand Avatar
von: Gan, Qijun, et al.
Veröffentlicht: (2024)
von: Gan, Qijun, et al.
Veröffentlicht: (2024)
Inline Critic Steers Image Editing
von: Kang, Weitai, et al.
Veröffentlicht: (2026)
von: Kang, Weitai, et al.
Veröffentlicht: (2026)
Looking Back and Forth: Cross-Image Attention Calibration and Attentive Preference Learning for Multi-Image Hallucination Mitigation
von: Yang, Xiaochen, et al.
Veröffentlicht: (2026)
von: Yang, Xiaochen, et al.
Veröffentlicht: (2026)
UniDemoiré: Towards Universal Image Demoiréing with Data Generation and Synthesis
von: Yang, Zemin, et al.
Veröffentlicht: (2025)
von: Yang, Zemin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings
von: Li, Zijie, et al.
Veröffentlicht: (2026) -
SeedEdit 3.0: Fast and High-Quality Generative Image Editing
von: Wang, Peng, et al.
Veröffentlicht: (2025) -
ByteMorph: Benchmarking Instruction-Guided Image Editing with Non-Rigid Motions
von: Chang, Di, et al.
Veröffentlicht: (2025) -
SeedEdit: Align Image Re-Generation to Image Editing
von: Shi, Yichun, et al.
Veröffentlicht: (2024) -
Text-Aware Diffusion for Policy Learning
von: Luo, Calvin, et al.
Veröffentlicht: (2024)