WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Hui, Liu, Juntao, Liu, Zongkai, Niu, Liqiang, Meng, Fandong, Wu, Zuxuan, Jiang, Yu-Gang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GlyphMastero: A Glyph Encoder for High-Fidelity Scene Text Editing
von: Wang, Tong, et al.
Veröffentlicht: (2025)
von: Wang, Tong, et al.
Veröffentlicht: (2025)
LaCo: Efficient Layer-wise Compression of Visual Tokens for Multimodal Large Language Models
von: Liu, Juntao, et al.
Veröffentlicht: (2025)
von: Liu, Juntao, et al.
Veröffentlicht: (2025)
FLEKE: Federated Locate-then-Edit Knowledge Editing
von: Zhao, Zongkai, et al.
Veröffentlicht: (2025)
von: Zhao, Zongkai, et al.
Veröffentlicht: (2025)
S$^2$Edit: Text-Guided Image Editing with Precise Semantic and Spatial Control
von: Liu, Xudong, et al.
Veröffentlicht: (2025)
von: Liu, Xudong, et al.
Veröffentlicht: (2025)
ImgEdit: A Unified Image Editing Dataset and Benchmark
von: Ye, Yang, et al.
Veröffentlicht: (2025)
von: Ye, Yang, et al.
Veröffentlicht: (2025)
EditInspector: A Benchmark for Evaluation of Text-Guided Image Edits
von: Yosef, Ron, et al.
Veröffentlicht: (2025)
von: Yosef, Ron, et al.
Veröffentlicht: (2025)
MaskMamba: A Hybrid Mamba-Transformer Model for Masked Image Generation
von: Chen, Wenchao, et al.
Veröffentlicht: (2024)
von: Chen, Wenchao, et al.
Veröffentlicht: (2024)
EditReward: A Human-Aligned Reward Model for Instruction-Guided Image Editing
von: Wu, Keming, et al.
Veröffentlicht: (2025)
von: Wu, Keming, et al.
Veröffentlicht: (2025)
GlyphWeaver: Unlocking Glyph Design Creativity with Uniform Glyph DSL and AI
von: Liu, Can, et al.
Veröffentlicht: (2025)
von: Liu, Can, et al.
Veröffentlicht: (2025)
LOGO: Video Text Spotting with Language Collaboration and Glyph Perception Model
von: Liu, Hongen, et al.
Veröffentlicht: (2024)
von: Liu, Hongen, et al.
Veröffentlicht: (2024)
FlexEdit: Flexible and Controllable Diffusion-based Object-centric Image Editing
von: Nguyen, Trong-Tung, et al.
Veröffentlicht: (2024)
von: Nguyen, Trong-Tung, et al.
Veröffentlicht: (2024)
ArrowGEV: Grounding Events in Video via Learning the Arrow of Time
von: Yu, Fangxu, et al.
Veröffentlicht: (2026)
von: Yu, Fangxu, et al.
Veröffentlicht: (2026)
GaussEdit: Adaptive 3D Scene Editing with Text and Image Prompts
von: Shu, Zhenyu, et al.
Veröffentlicht: (2025)
von: Shu, Zhenyu, et al.
Veröffentlicht: (2025)
TextMaster: A Unified Framework for Realistic Text Editing via Glyph-Style Dual-Control
von: Yan, Zhenyu, et al.
Veröffentlicht: (2024)
von: Yan, Zhenyu, et al.
Veröffentlicht: (2024)
MotionEdit: Benchmarking and Learning Motion-Centric Image Editing
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
von: Wan, Yixin, et al.
Veröffentlicht: (2025)
A Survey of Multimodal-Guided Image Editing with Text-to-Image Diffusion Models
von: Shuai, Xincheng, et al.
Veröffentlicht: (2024)
von: Shuai, Xincheng, et al.
Veröffentlicht: (2024)
Are We Evaluating the Edit Locality of LLM Model Editing Properly?
von: Liu, Wei, et al.
Veröffentlicht: (2026)
von: Liu, Wei, et al.
Veröffentlicht: (2026)
Translatotron-V(ison): An End-to-End Model for In-Image Machine Translation
von: Lan, Zhibin, et al.
Veröffentlicht: (2024)
von: Lan, Zhibin, et al.
Veröffentlicht: (2024)
D2C: Unlocking the Potential of Continuous Autoregressive Image Generation with Discrete Tokens
von: Wang, Panpan, et al.
Veröffentlicht: (2025)
von: Wang, Panpan, et al.
Veröffentlicht: (2025)
LMM4Edit: Benchmarking and Evaluating Multimodal Image Editing with LMMs
von: Xu, Zitong, et al.
Veröffentlicht: (2025)
von: Xu, Zitong, et al.
Veröffentlicht: (2025)
Step1X-Edit: A Practical Framework for General Image Editing
von: Liu, Shiyu, et al.
Veröffentlicht: (2025)
von: Liu, Shiyu, et al.
Veröffentlicht: (2025)
Edit-Compass & EditReward-Compass: A Unified Benchmark for Image Editing and Reward Modeling
von: Bai, Xuehai, et al.
Veröffentlicht: (2026)
von: Bai, Xuehai, et al.
Veröffentlicht: (2026)
InstantEdit: Text-Guided Few-Step Image Editing with Piecewise Rectified Flow
von: Gong, Yiming, et al.
Veröffentlicht: (2025)
von: Gong, Yiming, et al.
Veröffentlicht: (2025)
SpatialEdit: Benchmarking Fine-Grained Image Spatial Editing
von: Xiao, Yicheng, et al.
Veröffentlicht: (2026)
von: Xiao, Yicheng, et al.
Veröffentlicht: (2026)
WiseEdit: Benchmarking Cognition- and Creativity-Informed Image Editing
von: Pan, Kaihang, et al.
Veröffentlicht: (2025)
von: Pan, Kaihang, et al.
Veröffentlicht: (2025)
AdapEdit: Spatio-Temporal Guided Adaptive Editing Algorithm for Text-Based Continuity-Sensitive Image Editing
von: Ma, Zhiyuan, et al.
Veröffentlicht: (2023)
von: Ma, Zhiyuan, et al.
Veröffentlicht: (2023)
UME-R1: Exploring Reasoning-Driven Generative Multimodal Embeddings
von: Lan, Zhibin, et al.
Veröffentlicht: (2025)
von: Lan, Zhibin, et al.
Veröffentlicht: (2025)
LLaVE: Large Language and Vision Embedding Models with Hardness-Weighted Contrastive Learning
von: Lan, Zhibin, et al.
Veröffentlicht: (2025)
von: Lan, Zhibin, et al.
Veröffentlicht: (2025)
ReasonEdit: Towards Reasoning-Enhanced Image Editing Models
von: Yin, Fukun, et al.
Veröffentlicht: (2025)
von: Yin, Fukun, et al.
Veröffentlicht: (2025)
InterEdit: Navigating Text-Guided Multi-Human 3D Motion Editing
von: Yang, Yebin, et al.
Veröffentlicht: (2026)
von: Yang, Yebin, et al.
Veröffentlicht: (2026)
Exploring Text-Guided Single Image Editing for Remote Sensing Images
von: Han, Fangzhou, et al.
Veröffentlicht: (2024)
von: Han, Fangzhou, et al.
Veröffentlicht: (2024)
HQ-Edit: A High-Quality Dataset for Instruction-based Image Editing
von: Hui, Mude, et al.
Veröffentlicht: (2024)
von: Hui, Mude, et al.
Veröffentlicht: (2024)
A$^2$-Edit: Precise Reference-Guided Image Editing of Arbitrary Objects and Ambiguous Masks
von: Zheng, Huayu, et al.
Veröffentlicht: (2026)
von: Zheng, Huayu, et al.
Veröffentlicht: (2026)
SwiftEdit: Lightning Fast Text-Guided Image Editing via One-Step Diffusion
von: Nguyen, Trong-Tung, et al.
Veröffentlicht: (2024)
von: Nguyen, Trong-Tung, et al.
Veröffentlicht: (2024)
LocateEdit-Bench: A Benchmark for Instruction-Based Editing Localization
von: Wu, Shiyu, et al.
Veröffentlicht: (2026)
von: Wu, Shiyu, et al.
Veröffentlicht: (2026)
CaTok: Taming Mean Flows for One-Dimensional Causal Image Tokenization
von: Chen, Yitong, et al.
Veröffentlicht: (2026)
von: Chen, Yitong, et al.
Veröffentlicht: (2026)
ScribbleEdit: Synthetic Data for Image Editing with Scribbles and Text
von: Ji, Anya, et al.
Veröffentlicht: (2026)
von: Ji, Anya, et al.
Veröffentlicht: (2026)
Specify and Edit: Overcoming Ambiguity in Text-Based Image Editing
von: Iakovleva, Ekaterina, et al.
Veröffentlicht: (2024)
von: Iakovleva, Ekaterina, et al.
Veröffentlicht: (2024)
MotionFollower: Editing Video Motion via Lightweight Score-Guided Diffusion
von: Tu, Shuyuan, et al.
Veröffentlicht: (2024)
von: Tu, Shuyuan, et al.
Veröffentlicht: (2024)
UniEditBench: A Unified and Cost-Effective Benchmark for Image and Video Editing via Distilled MLLMs
von: Jiang, Lifan, et al.
Veröffentlicht: (2026)
von: Jiang, Lifan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
GlyphMastero: A Glyph Encoder for High-Fidelity Scene Text Editing
von: Wang, Tong, et al.
Veröffentlicht: (2025) -
LaCo: Efficient Layer-wise Compression of Visual Tokens for Multimodal Large Language Models
von: Liu, Juntao, et al.
Veröffentlicht: (2025) -
FLEKE: Federated Locate-then-Edit Knowledge Editing
von: Zhao, Zongkai, et al.
Veröffentlicht: (2025) -
S$^2$Edit: Text-Guided Image Editing with Precise Semantic and Spatial Control
von: Liu, Xudong, et al.
Veröffentlicht: (2025) -
ImgEdit: A Unified Image Editing Dataset and Benchmark
von: Ye, Yang, et al.
Veröffentlicht: (2025)