Early Semantic Grounding in Image Editing Models for Zero-Shot Referring Image Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | He, Jingxuan, Wang, Xiyu, Wang, Yunke, Zheng, Mengyu, Xu, Chang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rethinking Where to Edit: Task-Aware Localization for Instruction-Based Image Editing
by: He, Jingxuan, et al.
Published: (2026)
by: He, Jingxuan, et al.
Published: (2026)
LGD: Leveraging Generative Descriptions for Zero-Shot Referring Image Segmentation
by: Li, Jiachen, et al.
Published: (2025)
by: Li, Jiachen, et al.
Published: (2025)
Zero-shot Image Editing with Reference Imitation
by: Chen, Xi, et al.
Published: (2024)
by: Chen, Xi, et al.
Published: (2024)
Mask Grounding for Referring Image Segmentation
by: Chng, Yong Xien, et al.
Published: (2023)
by: Chng, Yong Xien, et al.
Published: (2023)
Early Timestep Zero-Shot Candidate Selection for Instruction-Guided Image Editing
by: Kim, Joowon, et al.
Published: (2025)
by: Kim, Joowon, et al.
Published: (2025)
Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models
by: Wang, Wen, et al.
Published: (2023)
by: Wang, Wen, et al.
Published: (2023)
Boundary-Aware Test-Time Adaptation for Zero-Shot Medical Image Segmentation
by: Xu, Chenlin, et al.
Published: (2025)
by: Xu, Chenlin, et al.
Published: (2025)
Marine Saliency Segmenter: Object-Focused Conditional Diffusion with Region-Level Semantic Knowledge Distillation
by: Chang, Laibin, et al.
Published: (2025)
by: Chang, Laibin, et al.
Published: (2025)
Open-Source Image Editing Models Are Zero-Shot Vision Learners
by: Liu, Wei, et al.
Published: (2026)
by: Liu, Wei, et al.
Published: (2026)
Generative Editing in the Joint Vision-Language Space for Zero-Shot Composed Image Retrieval
by: Wang, Xin, et al.
Published: (2025)
by: Wang, Xin, et al.
Published: (2025)
Unified and Semantically Grounded Domain Adaptation for Medical Image Segmentation
by: Wang, Xin, et al.
Published: (2025)
by: Wang, Xin, et al.
Published: (2025)
Beyond Visual Cues: Leveraging General Semantics as Support for Few-Shot Segmentation
by: Wang, Jin, et al.
Published: (2025)
by: Wang, Jin, et al.
Published: (2025)
Zero-to-Hero: Zero-Shot Initialization Empowering Reference-Based Video Appearance Editing
by: Su, Tongtong, et al.
Published: (2025)
by: Su, Tongtong, et al.
Published: (2025)
MeshSegmenter: Zero-Shot Mesh Semantic Segmentation via Texture Synthesis
by: Zhong, Ziming, et al.
Published: (2024)
by: Zhong, Ziming, et al.
Published: (2024)
Zero-Shot Video Semantic Segmentation based on Pre-Trained Diffusion Models
by: Wang, Qian, et al.
Published: (2024)
by: Wang, Qian, et al.
Published: (2024)
Spectral Prompt Tuning:Unveiling Unseen Classes for Zero-Shot Semantic Segmentation
by: Xu, Wenhao, et al.
Published: (2023)
by: Xu, Wenhao, et al.
Published: (2023)
Image-Conditioned Instance Prompt Network for Referring Remote Sensing Image Segmentation
by: Ren, Biaoyu, et al.
Published: (2026)
by: Ren, Biaoyu, et al.
Published: (2026)
Semantic Localization Guiding Segment Anything Model For Reference Remote Sensing Image Segmentation
by: Li, Shuyang, et al.
Published: (2025)
by: Li, Shuyang, et al.
Published: (2025)
Model Synthesis for Zero-Shot Model Attribution
by: Yang, Tianyun, et al.
Published: (2023)
by: Yang, Tianyun, et al.
Published: (2023)
InteractEdit: Zero-Shot Editing of Human-Object Interactions in Images
by: Hoe, Jiun Tian, et al.
Published: (2025)
by: Hoe, Jiun Tian, et al.
Published: (2025)
CPAM: Context-Preserving Adaptive Manipulation for Zero-Shot Real Image Editing
by: Vo, Dinh-Khoi, et al.
Published: (2025)
by: Vo, Dinh-Khoi, et al.
Published: (2025)
Negative Entity Suppression for Zero-Shot Captioning with Synthetic Images
by: Lu, Zimao, et al.
Published: (2025)
by: Lu, Zimao, et al.
Published: (2025)
Omni-Referring Image Segmentation
by: Zheng, Qiancheng, et al.
Published: (2025)
by: Zheng, Qiancheng, et al.
Published: (2025)
Zero-Shot Interpretable Image Steganalysis for Invertible Image Hiding
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
Hybrid Global-Local Representation with Augmented Spatial Guidance for Zero-Shot Referring Image Segmentation
by: Liu, Ting, et al.
Published: (2025)
by: Liu, Ting, et al.
Published: (2025)
Slicedit: Zero-Shot Video Editing With Text-to-Image Diffusion Models Using Spatio-Temporal Slices
by: Cohen, Nathaniel, et al.
Published: (2024)
by: Cohen, Nathaniel, et al.
Published: (2024)
Structure-Preserving Zero-Shot Image Editing via Stage-Wise Latent Injection in Diffusion Models
by: Jeong, Dasol, et al.
Published: (2025)
by: Jeong, Dasol, et al.
Published: (2025)
Deep Semantic-Visual Alignment for Zero-Shot Remote Sensing Image Scene Classification
by: Xu, Wenjia, et al.
Published: (2024)
by: Xu, Wenjia, et al.
Published: (2024)
FireFlow: Fast Inversion of Rectified Flow for Image Semantic Editing
by: Deng, Yingying, et al.
Published: (2024)
by: Deng, Yingying, et al.
Published: (2024)
Spatial Semantic Recurrent Mining for Referring Image Segmentation
by: Yang, Jiaxing, et al.
Published: (2024)
by: Yang, Jiaxing, et al.
Published: (2024)
Are Image-to-Video Models Good Zero-Shot Image Editors?
by: Zhang, Zechuan, et al.
Published: (2025)
by: Zhang, Zechuan, et al.
Published: (2025)
Studying Image Diffusion Features for Zero-Shot Video Object Segmentation
by: Delatolas, Thanos, et al.
Published: (2025)
by: Delatolas, Thanos, et al.
Published: (2025)
EAVL: Explicitly Align Vision and Language for Referring Image Segmentation
by: Yan, Yichen, et al.
Published: (2023)
by: Yan, Yichen, et al.
Published: (2023)
ImaGGen: Zero-Shot Generation of Co-Speech Semantic Gestures Grounded in Language and Image Input
by: Voss, Hendric, et al.
Published: (2025)
by: Voss, Hendric, et al.
Published: (2025)
MAEDiff: Masked Autoencoder-enhanced Diffusion Models for Unsupervised Anomaly Detection in Brain Images
by: Xu, Rui, et al.
Published: (2024)
by: Xu, Rui, et al.
Published: (2024)
Progressive Prompt-Guided Cross-Modal Reasoning for Referring Image Segmentation
by: Li, Jiachen, et al.
Published: (2026)
by: Li, Jiachen, et al.
Published: (2026)
Latent Expression Generation for Referring Image Segmentation and Grounding
by: Yu, Seonghoon, et al.
Published: (2025)
by: Yu, Seonghoon, et al.
Published: (2025)
Simultaneous Image-to-Zero and Zero-to-Noise: Diffusion Models with Analytical Image Attenuation
by: Huang, Yuhang, et al.
Published: (2023)
by: Huang, Yuhang, et al.
Published: (2023)
Generative Visual Chain-of-Thought for Image Editing
by: Yin, Zijin, et al.
Published: (2026)
by: Yin, Zijin, et al.
Published: (2026)
Self-supervised Dynamic Heterogeneous Degradation Modeling for Unified Zero-Shot Image Restoration
by: Hu, XiaoWan, et al.
Published: (2026)
by: Hu, XiaoWan, et al.
Published: (2026)
Similar Items
-
Rethinking Where to Edit: Task-Aware Localization for Instruction-Based Image Editing
by: He, Jingxuan, et al.
Published: (2026) -
LGD: Leveraging Generative Descriptions for Zero-Shot Referring Image Segmentation
by: Li, Jiachen, et al.
Published: (2025) -
Zero-shot Image Editing with Reference Imitation
by: Chen, Xi, et al.
Published: (2024) -
Mask Grounding for Referring Image Segmentation
by: Chng, Yong Xien, et al.
Published: (2023) -
Early Timestep Zero-Shot Candidate Selection for Instruction-Guided Image Editing
by: Kim, Joowon, et al.
Published: (2025)