Looking Beyond the Window: Global-Local Aligned CLIP for Training-free Open-Vocabulary Semantic Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, ByeongCheol, Seong, Hyun Seok, Hyun, Sangeek, Park, Gilhan, Moon, WonJun, Heo, Jae-Pil |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Progressive Proxy Anchor Propagation for Unsupervised Semantic Segmentation
by: Seong, Hyun Seok, et al.
Published: (2024)
by: Seong, Hyun Seok, et al.
Published: (2024)
From Vicious to Virtuous Cycles: Synergistic Representation Learning for Unsupervised Video Object-Centric Learning
by: Seong, Hyun Seok, et al.
Published: (2026)
by: Seong, Hyun Seok, et al.
Published: (2026)
Selective Contrastive Learning for Weakly Supervised Affordance Grounding
by: Moon, WonJun, et al.
Published: (2025)
by: Moon, WonJun, et al.
Published: (2025)
Reconstruction-Guided Slot Curriculum: Addressing Object Over-Fragmentation in Video Object-Centric Learning
by: Moon, WonJun, et al.
Published: (2026)
by: Moon, WonJun, et al.
Published: (2026)
Correlation-Guided Query-Dependency Calibration for Video Temporal Grounding
by: Moon, WonJun, et al.
Published: (2023)
by: Moon, WonJun, et al.
Published: (2023)
Mitigating Background Shift in Class-Incremental Semantic Segmentation
by: Park, Gilhan, et al.
Published: (2024)
by: Park, Gilhan, et al.
Published: (2024)
Temporal Alignment-Free Video Matching for Few-shot Action Recognition
by: Lee, SuBeen, et al.
Published: (2025)
by: Lee, SuBeen, et al.
Published: (2025)
GSGAN: Adversarial Learning for Hierarchical Generation of 3D Gaussian Splats
by: Hyun, Sangeek, et al.
Published: (2024)
by: Hyun, Sangeek, et al.
Published: (2024)
Cross-scale Aligned Supervision for Training GANs
by: Hyun, Sangeek, et al.
Published: (2026)
by: Hyun, Sangeek, et al.
Published: (2026)
Mitigating Semantic Collapse in Partially Relevant Video Retrieval
by: Moon, WonJun, et al.
Published: (2025)
by: Moon, WonJun, et al.
Published: (2025)
Style Injection in Diffusion: A Training-free Approach for Adapting Large-scale Diffusion Models for Style Transfer
by: Chung, Jiwoo, et al.
Published: (2023)
by: Chung, Jiwoo, et al.
Published: (2023)
Auxiliary Descriptive Knowledge for Few-Shot Adaptation of Vision-Language Model
by: Lee, SuBeen, et al.
Published: (2025)
by: Lee, SuBeen, et al.
Published: (2025)
Scalable GANs with Transformers
by: Hyun, Sangeek, et al.
Published: (2025)
by: Hyun, Sangeek, et al.
Published: (2025)
Disambiguating 2D-3D Correspondences in Gaussian Splatting-based Feature Fields for Visual Localization
by: Lee, Miso, et al.
Published: (2026)
by: Lee, Miso, et al.
Published: (2026)
VLCounter: Text-aware Visual Representation for Zero-Shot Object Counting
by: Kang, Seunggu, et al.
Published: (2023)
by: Kang, Seunggu, et al.
Published: (2023)
Masking Matters: Unlocking the Spatial Reasoning Capabilities of LLMs for 3D Scene-Language Understanding
by: Jeon, Yerim, et al.
Published: (2025)
by: Jeon, Yerim, et al.
Published: (2025)
Diversity-aware Channel Pruning for StyleGAN Compression
by: Chung, Jiwoo, et al.
Published: (2024)
by: Chung, Jiwoo, et al.
Published: (2024)
Auto-Encoded Supervision for Perceptual Image Super-Resolution
by: Lee, MinKyu, et al.
Published: (2024)
by: Lee, MinKyu, et al.
Published: (2024)
PDF-GS: Progressive Distractor Filtering for Robust 3D Gaussian Splatting
by: Seo, Kangmin, et al.
Published: (2026)
by: Seo, Kangmin, et al.
Published: (2026)
Foreground-Covering Prototype Generation and Matching for SAM-Aided Few-Shot Segmentation
by: Park, Suho, et al.
Published: (2025)
by: Park, Suho, et al.
Published: (2025)
Fine-Tuning Visual Autoregressive Models for Subject-Driven Generation
by: Chung, Jiwoo, et al.
Published: (2025)
by: Chung, Jiwoo, et al.
Published: (2025)
Analyzing the Training Dynamics of Image Restoration Transformers: A Revisit to Layer Normalization
by: Lee, MinKyu, et al.
Published: (2025)
by: Lee, MinKyu, et al.
Published: (2025)
Prototypes are Balanced Units for Efficient and Effective Partially Relevant Video Retrieval
by: Moon, WonJun, et al.
Published: (2025)
by: Moon, WonJun, et al.
Published: (2025)
Rethinking the Global Knowledge of CLIP in Training-Free Open-Vocabulary Semantic Segmentation
by: Wang, Jingyun, et al.
Published: (2025)
by: Wang, Jingyun, et al.
Published: (2025)
SeaCache: Spectral-Evolution-Aware Cache for Accelerating Diffusion Models
by: Chung, Jiwoo, et al.
Published: (2026)
by: Chung, Jiwoo, et al.
Published: (2026)
CorrCLIP: Reconstructing Patch Correlations in CLIP for Open-Vocabulary Semantic Segmentation
by: Zhang, Dengke, et al.
Published: (2024)
by: Zhang, Dengke, et al.
Published: (2024)
Noise-free Optimization in Early Training Steps for Image Super-Resolution
by: Lee, MinKyu, et al.
Published: (2023)
by: Lee, MinKyu, et al.
Published: (2023)
Distilling Spectral Graph for Object-Context Aware Open-Vocabulary Semantic Segmentation
by: Kim, Chanyoung, et al.
Published: (2024)
by: Kim, Chanyoung, et al.
Published: (2024)
TagCLIP: Improving Discrimination Ability of Open-Vocabulary Semantic Segmentation
by: Li, Jingyao, et al.
Published: (2023)
by: Li, Jingyao, et al.
Published: (2023)
Explore the Potential of CLIP for Training-Free Open Vocabulary Semantic Segmentation
by: Shao, Tong, et al.
Published: (2024)
by: Shao, Tong, et al.
Published: (2024)
White Aggregation and Restoration for Few-shot 3D Point Cloud Semantic Segmentation
by: Im, Jiyun, et al.
Published: (2025)
by: Im, Jiyun, et al.
Published: (2025)
OV-Stitcher: A Global Context-Aware Framework for Training-Free Open-Vocabulary Semantic Segmentation
by: Moon, Seungjae, et al.
Published: (2026)
by: Moon, Seungjae, et al.
Published: (2026)
18‐1: Composite UTG Cover Window Selectively Reinforced with Glass‐Cloth for Improvement of Both Pen‐Drop Resistance and Foldability
by: Hyun Seok Kang, et al.
Published: (2024)
by: Hyun Seok Kang, et al.
Published: (2024)
TAG: Guidance-free Open-Vocabulary Semantic Segmentation
by: Kawano, Yasufumi, et al.
Published: (2024)
by: Kawano, Yasufumi, et al.
Published: (2024)
Improving Visual Discriminability of CLIP for Training-Free Open-Vocabulary Semantic Segmentation
by: Zhou, Jinxin, et al.
Published: (2025)
by: Zhou, Jinxin, et al.
Published: (2025)
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation
by: Sun, Lin, et al.
Published: (2024)
by: Sun, Lin, et al.
Published: (2024)
PEARL: Geometry Aligns Semantics for Training-Free Open-Vocabulary Semantic Segmentation
by: Pei, Gensheng, et al.
Published: (2026)
by: Pei, Gensheng, et al.
Published: (2026)
ProxyCLIP: Proxy Attention Improves CLIP for Open-Vocabulary Segmentation
by: Lan, Mengcheng, et al.
Published: (2024)
by: Lan, Mengcheng, et al.
Published: (2024)
CLIP-VIS: Adapting CLIP for Open-Vocabulary Video Instance Segmentation
by: Zhu, Wenqi, et al.
Published: (2024)
by: Zhu, Wenqi, et al.
Published: (2024)
Plug-in Feedback Self-adaptive Attention in CLIP for Training-free Open-Vocabulary Segmentation
by: Chi, Zhixiang, et al.
Published: (2025)
by: Chi, Zhixiang, et al.
Published: (2025)
Similar Items
-
Progressive Proxy Anchor Propagation for Unsupervised Semantic Segmentation
by: Seong, Hyun Seok, et al.
Published: (2024) -
From Vicious to Virtuous Cycles: Synergistic Representation Learning for Unsupervised Video Object-Centric Learning
by: Seong, Hyun Seok, et al.
Published: (2026) -
Selective Contrastive Learning for Weakly Supervised Affordance Grounding
by: Moon, WonJun, et al.
Published: (2025) -
Reconstruction-Guided Slot Curriculum: Addressing Object Over-Fragmentation in Video Object-Centric Learning
by: Moon, WonJun, et al.
Published: (2026) -
Correlation-Guided Query-Dependency Calibration for Video Temporal Grounding
by: Moon, WonJun, et al.
Published: (2023)