Boosting Segment Anything Model Towards Open-Vocabulary Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Han, Xumeng, Wei, Longhui, Yu, Xuehui, Dou, Zhiyang, He, Xin, Wang, Kuiran, Sun, Yingfei, Han, Zhenjun, Tian, Qi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ViMoE: An Empirical Study of Designing Vision Mixture-of-Experts
von: Han, Xumeng, et al.
Veröffentlicht: (2024)
von: Han, Xumeng, et al.
Veröffentlicht: (2024)
P2Object: Single Point Supervised Object Detection and Instance Segmentation
von: Chen, Pengfei, et al.
Veröffentlicht: (2025)
von: Chen, Pengfei, et al.
Veröffentlicht: (2025)
CPR++: Object Localization via Single Coarse Point Supervision
von: Yu, Xuehui, et al.
Veröffentlicht: (2024)
von: Yu, Xuehui, et al.
Veröffentlicht: (2024)
GaGA: Towards Interactive Global Geolocation Assistant
von: Dou, Zhiyang, et al.
Veröffentlicht: (2024)
von: Dou, Zhiyang, et al.
Veröffentlicht: (2024)
Mixpert: Mitigating Multimodal Learning Conflicts with Efficient Mixture-of-Vision-Experts
von: He, Xin, et al.
Veröffentlicht: (2025)
von: He, Xin, et al.
Veröffentlicht: (2025)
SAPNet++: Evolving Point-Prompted Instance Segmentation with Semantic and Spatial Awareness
von: Wei, Zhaoyang, et al.
Veröffentlicht: (2026)
von: Wei, Zhaoyang, et al.
Veröffentlicht: (2026)
SAM-CP: Marrying SAM with Composable Prompts for Versatile Segmentation
von: Chen, Pengfei, et al.
Veröffentlicht: (2024)
von: Chen, Pengfei, et al.
Veröffentlicht: (2024)
P2Seg: Pointly-supervised Segmentation via Mutual Distillation
von: Wang, Zipeng, et al.
Veröffentlicht: (2024)
von: Wang, Zipeng, et al.
Veröffentlicht: (2024)
AD^2-Bench: A Hierarchical CoT Benchmark for MLLM in Autonomous Driving under Adverse Conditions
von: Wei, Zhaoyang, et al.
Veröffentlicht: (2025)
von: Wei, Zhaoyang, et al.
Veröffentlicht: (2025)
ClickTrack: Towards Real-time Interactive Single Object Tracking
von: Wang, Kuiran, et al.
Veröffentlicht: (2024)
von: Wang, Kuiran, et al.
Veröffentlicht: (2024)
HeroGS: Hierarchical Guidance for Robust 3D Gaussian Splatting under Sparse Views
von: Li, Jiashu, et al.
Veröffentlicht: (2026)
von: Li, Jiashu, et al.
Veröffentlicht: (2026)
OVMR: Open-Vocabulary Recognition with Multi-Modal References
von: Ma, Zehong, et al.
Veröffentlicht: (2024)
von: Ma, Zehong, et al.
Veröffentlicht: (2024)
P2RBox: Point Prompt Oriented Object Detection with SAM
von: Cao, Guangming, et al.
Veröffentlicht: (2023)
von: Cao, Guangming, et al.
Veröffentlicht: (2023)
Rethinking Sampling Strategies for Unsupervised Person Re-identification
von: Han, Xumeng, et al.
Veröffentlicht: (2021)
von: Han, Xumeng, et al.
Veröffentlicht: (2021)
Semantic-aware SAM for Point-Prompted Instance Segmentation
von: Wei, Zhaoyang, et al.
Veröffentlicht: (2023)
von: Wei, Zhaoyang, et al.
Veröffentlicht: (2023)
USE: Universal Segment Embeddings for Open-Vocabulary Image Segmentation
von: Wang, Xiaoqi, et al.
Veröffentlicht: (2024)
von: Wang, Xiaoqi, et al.
Veröffentlicht: (2024)
Generalization Boosted Adapter for Open-Vocabulary Segmentation
von: Xu, Wenhao, et al.
Veröffentlicht: (2024)
von: Xu, Wenhao, et al.
Veröffentlicht: (2024)
Causal Prompt Calibration Guided Segment Anything Model for Open-Vocabulary Multi-Entity Segmentation
von: Wang, Jingyao, et al.
Veröffentlicht: (2025)
von: Wang, Jingyao, et al.
Veröffentlicht: (2025)
Incorporating Visual Experts to Resolve the Information Loss in Multimodal Large Language Models
von: He, Xin, et al.
Veröffentlicht: (2024)
von: He, Xin, et al.
Veröffentlicht: (2024)
VER-Bench: Evaluating MLLMs on Reasoning with Fine-Grained Visual Evidence
von: Qiang, Chenhui, et al.
Veröffentlicht: (2025)
von: Qiang, Chenhui, et al.
Veröffentlicht: (2025)
Boosting Few-Shot Semantic Segmentation Via Segment Anything Model
von: Feng, Chen-Bin, et al.
Veröffentlicht: (2024)
von: Feng, Chen-Bin, et al.
Veröffentlicht: (2024)
OVS Meets Continual Learning: Towards Sustainable Open-Vocabulary Segmentation
von: Hwang, Dongjun, et al.
Veröffentlicht: (2024)
von: Hwang, Dongjun, et al.
Veröffentlicht: (2024)
RepViT-SAM: Towards Real-Time Segmenting Anything
von: Wang, Ao, et al.
Veröffentlicht: (2023)
von: Wang, Ao, et al.
Veröffentlicht: (2023)
Learning to Prompt Segment Anything Models
von: Huang, Jiaxing, et al.
Veröffentlicht: (2024)
von: Huang, Jiaxing, et al.
Veröffentlicht: (2024)
Segment Anything, Even Occluded
von: Tai, Wei-En, et al.
Veröffentlicht: (2025)
von: Tai, Wei-En, et al.
Veröffentlicht: (2025)
Self-Calibrated CLIP for Training-Free Open-Vocabulary Segmentation
von: Bai, Sule, et al.
Veröffentlicht: (2024)
von: Bai, Sule, et al.
Veröffentlicht: (2024)
Prompt-Guided Mask Proposal for Two-Stage Open-Vocabulary Segmentation
von: Li, Yu-Jhe, et al.
Veröffentlicht: (2024)
von: Li, Yu-Jhe, et al.
Veröffentlicht: (2024)
Segment Anything in Medical Images
von: Ma, Jun, et al.
Veröffentlicht: (2023)
von: Ma, Jun, et al.
Veröffentlicht: (2023)
Segment and Caption Anything
von: Huang, Xiaoke, et al.
Veröffentlicht: (2023)
von: Huang, Xiaoke, et al.
Veröffentlicht: (2023)
Towards Real-Time Open-Vocabulary Video Instance Segmentation
von: Yan, Bin, et al.
Veröffentlicht: (2024)
von: Yan, Bin, et al.
Veröffentlicht: (2024)
Towards Open-Vocabulary Semantic Segmentation Without Semantic Labels
von: Shin, Heeseong, et al.
Veröffentlicht: (2024)
von: Shin, Heeseong, et al.
Veröffentlicht: (2024)
Efficient Multi-modal Long Context Learning for Training-free Adaptation
von: Ma, Zehong, et al.
Veröffentlicht: (2025)
von: Ma, Zehong, et al.
Veröffentlicht: (2025)
Segment Anything in 3D with Radiance Fields
von: Cen, Jiazhong, et al.
Veröffentlicht: (2023)
von: Cen, Jiazhong, et al.
Veröffentlicht: (2023)
ODOV: Benchmark the Open-Domain Open-Vocabulary Object Detection
von: Zhang, Yupeng, et al.
Veröffentlicht: (2025)
von: Zhang, Yupeng, et al.
Veröffentlicht: (2025)
OpenTrack3D: Towards Accurate and Generalizable Open-Vocabulary 3D Instance Segmentation
von: Zhou, Zhishan, et al.
Veröffentlicht: (2025)
von: Zhou, Zhishan, et al.
Veröffentlicht: (2025)
From Open-Vocabulary to Vocabulary-Free Semantic Segmentation
von: Reichard, Klara, et al.
Veröffentlicht: (2025)
von: Reichard, Klara, et al.
Veröffentlicht: (2025)
ASAM: Boosting Segment Anything Model with Adversarial Tuning
von: Li, Bo, et al.
Veröffentlicht: (2024)
von: Li, Bo, et al.
Veröffentlicht: (2024)
Open-Vocabulary Domain Generalization in Urban-Scene Segmentation
von: Zhao, Dong, et al.
Veröffentlicht: (2026)
von: Zhao, Dong, et al.
Veröffentlicht: (2026)
OpenSeg-R: Improving Open-Vocabulary Segmentation via Step-by-Step Visual Reasoning
von: Han, Zongyan, et al.
Veröffentlicht: (2025)
von: Han, Zongyan, et al.
Veröffentlicht: (2025)
Matching Anything by Segmenting Anything
von: Li, Siyuan, et al.
Veröffentlicht: (2024)
von: Li, Siyuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
ViMoE: An Empirical Study of Designing Vision Mixture-of-Experts
von: Han, Xumeng, et al.
Veröffentlicht: (2024) -
P2Object: Single Point Supervised Object Detection and Instance Segmentation
von: Chen, Pengfei, et al.
Veröffentlicht: (2025) -
CPR++: Object Localization via Single Coarse Point Supervision
von: Yu, Xuehui, et al.
Veröffentlicht: (2024) -
GaGA: Towards Interactive Global Geolocation Assistant
von: Dou, Zhiyang, et al.
Veröffentlicht: (2024) -
Mixpert: Mitigating Multimodal Learning Conflicts with Efficient Mixture-of-Vision-Experts
von: He, Xin, et al.
Veröffentlicht: (2025)