VIPA: Visual Informative Part Attention for Referring Image Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Cho, Yubin, Yu, Hyunwoo, Kong, Kyeongbo, Sohn, Kyomin, Hyun, Bongjoon, Kang, Suk-Ju |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cross-aware Early Fusion with Stage-divided Vision and Language Transformer Encoders for Referring Image Segmentation
by: Cho, Yubin, et al.
Published: (2024)
by: Cho, Yubin, et al.
Published: (2024)
Embedding-Free Transformer with Inference Spatial Reduction for Efficient Semantic Segmentation
by: Yu, Hyunwoo, et al.
Published: (2024)
by: Yu, Hyunwoo, et al.
Published: (2024)
AttentionHand: Text-driven Controllable Hand Image Generation for 3D Hand Reconstruction in the Wild
by: Park, Junho, et al.
Published: (2024)
by: Park, Junho, et al.
Published: (2024)
MetaSeg: MetaFormer-based Global Contexts-aware Network for Efficient Semantic Segmentation
by: Kang, Beoungwoo, et al.
Published: (2024)
by: Kang, Beoungwoo, et al.
Published: (2024)
Programmable-Room: Interactive Textured 3D Room Meshes Generation Empowered by Large Language Models
by: Kim, Jihyun, et al.
Published: (2025)
by: Kim, Jihyun, et al.
Published: (2025)
HOIGS: Human-Object Interaction Gaussian Splatting
by: Kim, Taewoo, et al.
Published: (2026)
by: Kim, Taewoo, et al.
Published: (2026)
AgilePruner: An Empirical Study of Attention and Diversity for Adaptive Visual Token Pruning in Large Vision-Language Models
by: Baek, Changwoo, et al.
Published: (2026)
by: Baek, Changwoo, et al.
Published: (2026)
Focus Matters: Phase-Aware Suppression for Hallucination in Vision-Language Models
by: Kim, Sohyeon, et al.
Published: (2026)
by: Kim, Sohyeon, et al.
Published: (2026)
ReflectCAP: Detailed Image Captioning with Reflective Memory
by: Min, Kyungmin, et al.
Published: (2026)
by: Min, Kyungmin, et al.
Published: (2026)
Training-Free Global Geometric Association for 4D LiDAR Panoptic Segmentation
by: Oh, Gyeongrok, et al.
Published: (2025)
by: Oh, Gyeongrok, et al.
Published: (2025)
Multi-scale Information Sharing and Selection Network with Boundary Attention for Polyp Segmentation
by: Kang, Xiaolu, et al.
Published: (2024)
by: Kang, Xiaolu, et al.
Published: (2024)
Latent Expression Generation for Referring Image Segmentation and Grounding
by: Yu, Seonghoon, et al.
Published: (2025)
by: Yu, Seonghoon, et al.
Published: (2025)
Re-purposing SAM into Efficient Visual Projectors for MLLM-Based Referring Image Segmentation
by: Yang, Xiaobo, et al.
Published: (2025)
by: Yang, Xiaobo, et al.
Published: (2025)
Editable Noise Map Inversion: Encoding Target-image into Noise For High-Fidelity Image Manipulation
by: Kang, Mingyu, et al.
Published: (2025)
by: Kang, Mingyu, et al.
Published: (2025)
Thinking Diffusion: Penalize and Guide Visual-Grounded Reasoning in Diffusion Multimodal Language Models
by: Kim, Keuntae, et al.
Published: (2026)
by: Kim, Keuntae, et al.
Published: (2026)
Pseudo-RIS: Distinctive Pseudo-supervision Generation for Referring Image Segmentation
by: Yu, Seonghoon, et al.
Published: (2024)
by: Yu, Seonghoon, et al.
Published: (2024)
Ref-AVS: Refer and Segment Objects in Audio-Visual Scenes
by: Wang, Yaoting, et al.
Published: (2024)
by: Wang, Yaoting, et al.
Published: (2024)
Refer to Any Segmentation Mask Group With Vision-Language Prompts
by: Cao, Shengcao, et al.
Published: (2025)
by: Cao, Shengcao, et al.
Published: (2025)
ReMamber: Referring Image Segmentation with Mamba Twister
by: Yang, Yuhuan, et al.
Published: (2024)
by: Yang, Yuhuan, et al.
Published: (2024)
Attention at Rest Stays at Rest: Breaking Visual Inertia for Cognitive Hallucination Mitigation
by: Gong, Boyang, et al.
Published: (2026)
by: Gong, Boyang, et al.
Published: (2026)
EdgeSRIE: A hybrid deep learning framework for real-time speckle reduction and image enhancement on portable ultrasound systems
by: Cho, Hyunwoo, et al.
Published: (2025)
by: Cho, Hyunwoo, et al.
Published: (2025)
LivingWorld: Interactive 4D World Generation with Environmental Dynamics
by: Mun, Hyeongju, et al.
Published: (2026)
by: Mun, Hyeongju, et al.
Published: (2026)
Semantic Localization Guiding Segment Anything Model For Reference Remote Sensing Image Segmentation
by: Li, Shuyang, et al.
Published: (2025)
by: Li, Shuyang, et al.
Published: (2025)
AMLRIS: Alignment-aware Masked Learning for Referring Image Segmentation
by: Chen, Tongfei, et al.
Published: (2026)
by: Chen, Tongfei, et al.
Published: (2026)
HARIS: Human-Like Attention for Reference Image Segmentation
by: Zhang, Mengxi, et al.
Published: (2024)
by: Zhang, Mengxi, et al.
Published: (2024)
UniPixel: Unified Object Referring and Segmentation for Pixel-Level Visual Reasoning
by: Liu, Ye, et al.
Published: (2025)
by: Liu, Ye, et al.
Published: (2025)
Unlocking the Potential of Unlabeled Data in Semi-Supervised Domain Generalization
by: Lee, Dongkwan, et al.
Published: (2025)
by: Lee, Dongkwan, et al.
Published: (2025)
In Defense of Lazy Visual Grounding for Open-Vocabulary Semantic Segmentation
by: Kang, Dahyun, et al.
Published: (2024)
by: Kang, Dahyun, et al.
Published: (2024)
Retinal Layer Segmentation in OCT Images With 2.5D Cross-slice Feature Fusion Module for Glaucoma Assessment
by: Kim, Hyunwoo, et al.
Published: (2026)
by: Kim, Hyunwoo, et al.
Published: (2026)
DarkQA: Benchmarking Vision-Language Models on Visual-Primitive Question Answering in Low-Light Indoor Scenes
by: Park, Yohan, et al.
Published: (2025)
by: Park, Yohan, et al.
Published: (2025)
MatchSeg: Towards Better Segmentation via Reference Image Matching
by: Huo, Jiayu, et al.
Published: (2024)
by: Huo, Jiayu, et al.
Published: (2024)
Few Shot Part Segmentation Reveals Compositional Logic for Industrial Anomaly Detection
by: Kim, Soopil, et al.
Published: (2023)
by: Kim, Soopil, et al.
Published: (2023)
EMRA-proxy: Enhancing Multi-Class Region Semantic Segmentation in Remote Sensing Images with Attention Proxy
by: Yu, Yichun, et al.
Published: (2025)
by: Yu, Yichun, et al.
Published: (2025)
A Text-Image Fusion Method with Data Augmentation Capabilities for Referring Medical Image Segmentation
by: Chai, Shurong, et al.
Published: (2025)
by: Chai, Shurong, et al.
Published: (2025)
TFANet: Three-Stage Image-Text Feature Alignment Network for Robust Referring Image Segmentation
by: Lu, Qianqi, et al.
Published: (2025)
by: Lu, Qianqi, et al.
Published: (2025)
Error as Signal: Stiffness-Aware Diffusion Sampling via Embedded Runge-Kutta Guidance
by: Kong, Inho, et al.
Published: (2026)
by: Kong, Inho, et al.
Published: (2026)
Cross-Modal Bidirectional Interaction Model for Referring Remote Sensing Image Segmentation
by: Dong, Zhe, et al.
Published: (2024)
by: Dong, Zhe, et al.
Published: (2024)
Referring Remote Sensing Image Segmentation with Cross-view Semantics Interaction Network
by: Yang, Jiaxing, et al.
Published: (2025)
by: Yang, Jiaxing, et al.
Published: (2025)
CroBIM-U: Uncertainty-Driven Referring Remote Sensing Image Segmentation
by: Sun, Yuzhe, et al.
Published: (2026)
by: Sun, Yuzhe, et al.
Published: (2026)
SaFiRe: Saccade-Fixation Reiteration with Mamba for Referring Image Segmentation
by: Mao, Zhenjie, et al.
Published: (2025)
by: Mao, Zhenjie, et al.
Published: (2025)
Similar Items
-
Cross-aware Early Fusion with Stage-divided Vision and Language Transformer Encoders for Referring Image Segmentation
by: Cho, Yubin, et al.
Published: (2024) -
Embedding-Free Transformer with Inference Spatial Reduction for Efficient Semantic Segmentation
by: Yu, Hyunwoo, et al.
Published: (2024) -
AttentionHand: Text-driven Controllable Hand Image Generation for 3D Hand Reconstruction in the Wild
by: Park, Junho, et al.
Published: (2024) -
MetaSeg: MetaFormer-based Global Contexts-aware Network for Efficient Semantic Segmentation
by: Kang, Beoungwoo, et al.
Published: (2024) -
Programmable-Room: Interactive Textured 3D Room Meshes Generation Empowered by Large Language Models
by: Kim, Jihyun, et al.
Published: (2025)