Emergent Open-Vocabulary Semantic Segmentation from Off-the-shelf Vision-Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Luo, Jiayun, Khandelwal, Siddhesh, Sigal, Leonid, Li, Boyang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
To Sink or Not to Sink: Visual Information Pathways in Large Vision-Language Models
por: Luo, Jiayun, et al.
Publicado: (2025)
por: Luo, Jiayun, et al.
Publicado: (2025)
Decomposed Vision-Language Alignment for Fine-Grained Open-Vocabulary Segmentation
por: Wang, Chenhao, et al.
Publicado: (2026)
por: Wang, Chenhao, et al.
Publicado: (2026)
MROVSeg: Breaking the Resolution Curse of Vision-Language Models in Open-Vocabulary Image Segmentation
por: Zhu, Yuanbing, et al.
Publicado: (2024)
por: Zhu, Yuanbing, et al.
Publicado: (2024)
Open-Vocabulary Remote Sensing Image Semantic Segmentation
por: Cao, Qinglong, et al.
Publicado: (2024)
por: Cao, Qinglong, et al.
Publicado: (2024)
MARIS: Marine Open-Vocabulary Instance Segmentation with Geometric Enhancement and Semantic Alignment
por: Li, Bingyu, et al.
Publicado: (2025)
por: Li, Bingyu, et al.
Publicado: (2025)
From Open Vocabulary to Open World: Teaching Vision Language Models to Detect Novel Objects
por: Li, Zizhao, et al.
Publicado: (2024)
por: Li, Zizhao, et al.
Publicado: (2024)
dinov3.seg: Open-Vocabulary Semantic Segmentation with DINOv3
por: Dutta, Saikat, et al.
Publicado: (2026)
por: Dutta, Saikat, et al.
Publicado: (2026)
GeoGuide: Hierarchical Geometric Guidance for Open-Vocabulary 3D Semantic Segmentation
por: Tao, Xujing, et al.
Publicado: (2026)
por: Tao, Xujing, et al.
Publicado: (2026)
Seeking Consensus: Geometric-Semantic On-the-Fly Recalibration for Open-Vocabulary Remote Sensing Semantic Segmentation
por: Wang, Guanchun, et al.
Publicado: (2026)
por: Wang, Guanchun, et al.
Publicado: (2026)
AttrSeg: Open-Vocabulary Semantic Segmentation via Attribute Decomposition-Aggregation
por: Ma, Chaofan, et al.
Publicado: (2023)
por: Ma, Chaofan, et al.
Publicado: (2023)
Black Swan: Abductive and Defeasible Video Reasoning in Unpredictable Events
por: Chinchure, Aditya, et al.
Publicado: (2024)
por: Chinchure, Aditya, et al.
Publicado: (2024)
Tinted Frames: Question Framing Blinds Vision-Language Models
por: Fan, Wan-Cyuan, et al.
Publicado: (2026)
por: Fan, Wan-Cyuan, et al.
Publicado: (2026)
Talking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabulary Segmentation
por: Barsellotti, Luca, et al.
Publicado: (2024)
por: Barsellotti, Luca, et al.
Publicado: (2024)
Open-Vocabulary Panoptic Segmentation Using BERT Pre-Training of Vision-Language Multiway Transformer Model
por: Chen, Yi-Chia, et al.
Publicado: (2024)
por: Chen, Yi-Chia, et al.
Publicado: (2024)
CoCo-SAM3: Harnessing Concept Conflict in Open-Vocabulary Semantic Segmentation
por: Chen, Yanhui, et al.
Publicado: (2026)
por: Chen, Yanhui, et al.
Publicado: (2026)
Exploring Efficient Open-Vocabulary Segmentation in the Remote Sensing
por: Li, Bingyu, et al.
Publicado: (2025)
por: Li, Bingyu, et al.
Publicado: (2025)
MMFactory: A Universal Solution Search Engine for Vision-Language Tasks
por: Fan, Wan-Cyuan, et al.
Publicado: (2024)
por: Fan, Wan-Cyuan, et al.
Publicado: (2024)
Vocabulary-Free 3D Instance Segmentation with Vision and Language Assistant
por: Mei, Guofeng, et al.
Publicado: (2024)
por: Mei, Guofeng, et al.
Publicado: (2024)
Looking Beyond the Window: Global-Local Aligned CLIP for Training-free Open-Vocabulary Semantic Segmentation
por: Lee, ByeongCheol, et al.
Publicado: (2026)
por: Lee, ByeongCheol, et al.
Publicado: (2026)
Structure-Aware Feature Rectification with Region Adjacency Graphs for Training-Free Open-Vocabulary Semantic Segmentation
por: Huang, Qiming, et al.
Publicado: (2025)
por: Huang, Qiming, et al.
Publicado: (2025)
XMask3D: Cross-modal Mask Reasoning for Open Vocabulary 3D Semantic Segmentation
por: Wang, Ziyi, et al.
Publicado: (2024)
por: Wang, Ziyi, et al.
Publicado: (2024)
SHOE: Semantic HOI Open-Vocabulary Evaluation Metric
por: Noack, Maja, et al.
Publicado: (2026)
por: Noack, Maja, et al.
Publicado: (2026)
Open-Vocabulary Segmentation with Unpaired Mask-Text Supervision
por: Wang, Zhaoqing, et al.
Publicado: (2024)
por: Wang, Zhaoqing, et al.
Publicado: (2024)
OVS-DINO: Open-Vocabulary Segmentation via Structure-Aligned SAM-DINO with Language Guidance
por: Zeng, Haoxi, et al.
Publicado: (2026)
por: Zeng, Haoxi, et al.
Publicado: (2026)
World-to-Words: Grounded Open Vocabulary Acquisition through Fast Mapping in Vision-Language Models
por: Ma, Ziqiao, et al.
Publicado: (2023)
por: Ma, Ziqiao, et al.
Publicado: (2023)
Urban Socio-Semantic Segmentation with Vision-Language Reasoning
por: Wang, Yu, et al.
Publicado: (2026)
por: Wang, Yu, et al.
Publicado: (2026)
WAVER: Writing-style Agnostic Text-Video Retrieval via Distilling Vision-Language Models Through Open-Vocabulary Knowledge
por: Le, Huy, et al.
Publicado: (2023)
por: Le, Huy, et al.
Publicado: (2023)
Vision-Language Model Purified Semi-Supervised Semantic Segmentation for Remote Sensing Images
por: Wang, Shanwen, et al.
Publicado: (2026)
por: Wang, Shanwen, et al.
Publicado: (2026)
A Study on Unsupervised Domain Adaptation for Semantic Segmentation in the Era of Vision-Language Models
por: Schwonberg, Manuel, et al.
Publicado: (2024)
por: Schwonberg, Manuel, et al.
Publicado: (2024)
Leveraging Out-of-Distribution Unlabeled Images: Semi-Supervised Semantic Segmentation with an Open-Vocabulary Model
por: Shin, Wooseok, et al.
Publicado: (2025)
por: Shin, Wooseok, et al.
Publicado: (2025)
On the Fairness, Diversity and Reliability of Text-to-Image Generative Models
por: Vice, Jordan, et al.
Publicado: (2024)
por: Vice, Jordan, et al.
Publicado: (2024)
ExtrinSplat: Decoupling Geometry and Semantics for Open-Vocabulary Understanding in 3D Gaussian Splatting
por: Ding, Jiayu, et al.
Publicado: (2025)
por: Ding, Jiayu, et al.
Publicado: (2025)
AerOSeg: Harnessing SAM for Open-Vocabulary Segmentation in Remote Sensing Images
por: Dutta, Saikat, et al.
Publicado: (2025)
por: Dutta, Saikat, et al.
Publicado: (2025)
On Pre-training of Multimodal Language Models Customized for Chart Understanding
por: Fan, Wan-Cyuan, et al.
Publicado: (2024)
por: Fan, Wan-Cyuan, et al.
Publicado: (2024)
VLMs meet UDA: Boosting Transferability of Open Vocabulary Segmentation with Unsupervised Domain Adaptation
por: Alcover-Couso, Roberto, et al.
Publicado: (2024)
por: Alcover-Couso, Roberto, et al.
Publicado: (2024)
OV-Stitcher: A Global Context-Aware Framework for Training-Free Open-Vocabulary Semantic Segmentation
por: Moon, Seungjae, et al.
Publicado: (2026)
por: Moon, Seungjae, et al.
Publicado: (2026)
Test-Time Adaptation of Vision-Language Models for Open-Vocabulary Semantic Segmentation
por: Noori, Mehrdad, et al.
Publicado: (2025)
por: Noori, Mehrdad, et al.
Publicado: (2025)
Adapting Vision-Language Model with Fine-grained Semantics for Open-Vocabulary Segmentation
por: Chng, Yong Xien, et al.
Publicado: (2024)
por: Chng, Yong Xien, et al.
Publicado: (2024)
Open-Set Domain Adaptation for Semantic Segmentation
por: Choe, Seun-An, et al.
Publicado: (2024)
por: Choe, Seun-An, et al.
Publicado: (2024)
Towards Open Vocabulary Learning: A Survey
por: Wu, Jianzong, et al.
Publicado: (2023)
por: Wu, Jianzong, et al.
Publicado: (2023)
Ejemplares similares
-
To Sink or Not to Sink: Visual Information Pathways in Large Vision-Language Models
por: Luo, Jiayun, et al.
Publicado: (2025) -
Decomposed Vision-Language Alignment for Fine-Grained Open-Vocabulary Segmentation
por: Wang, Chenhao, et al.
Publicado: (2026) -
MROVSeg: Breaking the Resolution Curse of Vision-Language Models in Open-Vocabulary Image Segmentation
por: Zhu, Yuanbing, et al.
Publicado: (2024) -
Open-Vocabulary Remote Sensing Image Semantic Segmentation
por: Cao, Qinglong, et al.
Publicado: (2024) -
MARIS: Marine Open-Vocabulary Instance Segmentation with Geometric Enhancement and Semantic Alignment
por: Li, Bingyu, et al.
Publicado: (2025)