Taming SAM3 in the Wild: A Concept Bank for Open-Vocabulary Segmentation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pei, Gensheng, Jiang, Xiruo, Yao, Yazhou, Shu, Xiangbo, Shen, Fumin, Jeon, Byeungwoo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PEARL: Geometry Aligns Semantics for Training-Free Open-Vocabulary Semantic Segmentation
von: Pei, Gensheng, et al.
Veröffentlicht: (2026)
von: Pei, Gensheng, et al.
Veröffentlicht: (2026)
PCA-Seg: Revisiting Cost Aggregation for Open-Vocabulary Semantic and Part Segmentation
von: Yin, Jianjian, et al.
Veröffentlicht: (2026)
von: Yin, Jianjian, et al.
Veröffentlicht: (2026)
VideoMAC: Video Masked Autoencoders Meet ConvNets
von: Pei, Gensheng, et al.
Veröffentlicht: (2024)
von: Pei, Gensheng, et al.
Veröffentlicht: (2024)
A Light-weight Transformer-based Self-supervised Matching Network for Heterogeneous Images
von: Zhang, Wang, et al.
Veröffentlicht: (2024)
von: Zhang, Wang, et al.
Veröffentlicht: (2024)
Towards Remote Sensing Change Detection with Neural Memory
von: Yang, Zhenyu, et al.
Veröffentlicht: (2026)
von: Yang, Zhenyu, et al.
Veröffentlicht: (2026)
Iris: Bringing Real-World Priors into Diffusion Model for Monocular Depth Estimation
von: Cai, Xinhao, et al.
Veröffentlicht: (2026)
von: Cai, Xinhao, et al.
Veröffentlicht: (2026)
Anti-Collapse Loss for Deep Metric Learning Based on Coding Rate Metric
von: Jiang, Xiruo, et al.
Veröffentlicht: (2024)
von: Jiang, Xiruo, et al.
Veröffentlicht: (2024)
Beyond Quadratic: Linear-Time Change Detection with RWKV
von: Yang, Zhenyu, et al.
Veröffentlicht: (2026)
von: Yang, Zhenyu, et al.
Veröffentlicht: (2026)
Seeing What Matters: Empowering CLIP with Patch Generation-to-Selection
von: Pei, Gensheng, et al.
Veröffentlicht: (2025)
von: Pei, Gensheng, et al.
Veröffentlicht: (2025)
Efficiency Follows Global-Local Decoupling
von: Yang, Zhenyu, et al.
Veröffentlicht: (2026)
von: Yang, Zhenyu, et al.
Veröffentlicht: (2026)
Knowledge Transfer with Simulated Inter-Image Erasing for Weakly Supervised Semantic Segmentation
von: Chen, Tao, et al.
Veröffentlicht: (2024)
von: Chen, Tao, et al.
Veröffentlicht: (2024)
OmniGaze: Reward-inspired Generalizable Gaze Estimation In The Wild
von: Qu, Hongyu, et al.
Veröffentlicht: (2025)
von: Qu, Hongyu, et al.
Veröffentlicht: (2025)
Dynamic in Static: Hybrid Visual Correspondence for Self-Supervised Video Object Segmentation
von: Pei, Gensheng, et al.
Veröffentlicht: (2024)
von: Pei, Gensheng, et al.
Veröffentlicht: (2024)
CoCo-SAM3: Harnessing Concept Conflict in Open-Vocabulary Semantic Segmentation
von: Chen, Yanhui, et al.
Veröffentlicht: (2026)
von: Chen, Yanhui, et al.
Veröffentlicht: (2026)
Semi-supervised Semantic Segmentation with Multi-Constraint Consistency Learning
von: Yin, Jianjian, et al.
Veröffentlicht: (2025)
von: Yin, Jianjian, et al.
Veröffentlicht: (2025)
Effective SAM Combination for Open-Vocabulary Semantic Segmentation
von: Lee, Minhyeok, et al.
Veröffentlicht: (2024)
von: Lee, Minhyeok, et al.
Veröffentlicht: (2024)
3DTeethSAM: Taming SAM2 for 3D Teeth Segmentation
von: Lu, Zhiguo, et al.
Veröffentlicht: (2025)
von: Lu, Zhiguo, et al.
Veröffentlicht: (2025)
SAM-MI: A Mask-Injected Framework for Enhancing Open-Vocabulary Semantic Segmentation with SAM
von: Chen, Lin, et al.
Veröffentlicht: (2025)
von: Chen, Lin, et al.
Veröffentlicht: (2025)
Open-Vocabulary SAM: Segment and Recognize Twenty-thousand Classes Interactively
von: Yuan, Haobo, et al.
Veröffentlicht: (2024)
von: Yuan, Haobo, et al.
Veröffentlicht: (2024)
PKINet-v2: Towards Powerful and Efficient Poly-Kernel Remote Sensing Object Detection
von: Cai, Xinhao, et al.
Veröffentlicht: (2026)
von: Cai, Xinhao, et al.
Veröffentlicht: (2026)
LoGoSeg: Integrating Local and Global Features for Open-Vocabulary Semantic Segmentation
von: Chen, Junyang, et al.
Veröffentlicht: (2026)
von: Chen, Junyang, et al.
Veröffentlicht: (2026)
FTMoMamba: Motion Generation with Frequency and Text State Space Models
von: Li, Chengjian, et al.
Veröffentlicht: (2024)
von: Li, Chengjian, et al.
Veröffentlicht: (2024)
Open-Vocabulary SAM3D: Towards Training-free Open-Vocabulary 3D Scene Understanding
von: Tai, Hanchen, et al.
Veröffentlicht: (2024)
von: Tai, Hanchen, et al.
Veröffentlicht: (2024)
Learning 3D Representations for Spatial Intelligence from Unposed Multi-View Images
von: Zhou, Bo, et al.
Veröffentlicht: (2026)
von: Zhou, Bo, et al.
Veröffentlicht: (2026)
Taming Self-Training for Open-Vocabulary Object Detection
von: Zhao, Shiyu, et al.
Veröffentlicht: (2023)
von: Zhao, Shiyu, et al.
Veröffentlicht: (2023)
SegEarth-OV3: Exploring SAM 3 for Open-Vocabulary Semantic Segmentation in Remote Sensing Images
von: Li, Kaiyu, et al.
Veröffentlicht: (2025)
von: Li, Kaiyu, et al.
Veröffentlicht: (2025)
Personalized OVSS: Understanding Personal Concept in Open-Vocabulary Semantic Segmentation
von: Park, Sunghyun, et al.
Veröffentlicht: (2025)
von: Park, Sunghyun, et al.
Veröffentlicht: (2025)
WildOS: Open-Vocabulary Object Search in the Wild
von: Shah, Hardik, et al.
Veröffentlicht: (2026)
von: Shah, Hardik, et al.
Veröffentlicht: (2026)
Unbiased Object Detection Beyond Frequency with Visually Prompted Image Synthesis
von: Cai, Xinhao, et al.
Veröffentlicht: (2025)
von: Cai, Xinhao, et al.
Veröffentlicht: (2025)
OVS-DINO: Open-Vocabulary Segmentation via Structure-Aligned SAM-DINO with Language Guidance
von: Zeng, Haoxi, et al.
Veröffentlicht: (2026)
von: Zeng, Haoxi, et al.
Veröffentlicht: (2026)
SAM 3: Segment Anything with Concepts
von: Carion, Nicolas, et al.
Veröffentlicht: (2025)
von: Carion, Nicolas, et al.
Veröffentlicht: (2025)
AerOSeg: Harnessing SAM for Open-Vocabulary Segmentation in Remote Sensing Images
von: Dutta, Saikat, et al.
Veröffentlicht: (2025)
von: Dutta, Saikat, et al.
Veröffentlicht: (2025)
Combating Noisy Labels through Fostering Self- and Neighbor-Consistency
von: Sun, Zeren, et al.
Veröffentlicht: (2026)
von: Sun, Zeren, et al.
Veröffentlicht: (2026)
OVLW-DETR: Open-Vocabulary Light-Weighted Detection Transformer
von: Wang, Yu, et al.
Veröffentlicht: (2024)
von: Wang, Yu, et al.
Veröffentlicht: (2024)
Reconstructing In-the-Wild Open-Vocabulary Human-Object Interactions
von: Wen, Boran, et al.
Veröffentlicht: (2025)
von: Wen, Boran, et al.
Veröffentlicht: (2025)
FOLK: Fast Open-Vocabulary 3D Instance Segmentation via Label-guided Knowledge Distillation
von: Wu, Hongrui, et al.
Veröffentlicht: (2025)
von: Wu, Hongrui, et al.
Veröffentlicht: (2025)
EfficientSAM3: Progressive Hierarchical Distillation for Video Concept Segmentation from SAM1, 2, and 3
von: Zeng, Chengxi, et al.
Veröffentlicht: (2025)
von: Zeng, Chengxi, et al.
Veröffentlicht: (2025)
Relating CNN-Transformer Fusion Network for Change Detection
von: Gao, Yuhao, et al.
Veröffentlicht: (2024)
von: Gao, Yuhao, et al.
Veröffentlicht: (2024)
MedSAM3: Delving into Segment Anything with Medical Concepts
von: Liu, Anglin, et al.
Veröffentlicht: (2025)
von: Liu, Anglin, et al.
Veröffentlicht: (2025)
AdaFPP: Adapt-Focused Bi-Propagating Prototype Learning for Panoramic Activity Recognition
von: Cao, Meiqi, et al.
Veröffentlicht: (2024)
von: Cao, Meiqi, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
PEARL: Geometry Aligns Semantics for Training-Free Open-Vocabulary Semantic Segmentation
von: Pei, Gensheng, et al.
Veröffentlicht: (2026) -
PCA-Seg: Revisiting Cost Aggregation for Open-Vocabulary Semantic and Part Segmentation
von: Yin, Jianjian, et al.
Veröffentlicht: (2026) -
VideoMAC: Video Masked Autoencoders Meet ConvNets
von: Pei, Gensheng, et al.
Veröffentlicht: (2024) -
A Light-weight Transformer-based Self-supervised Matching Network for Heterogeneous Images
von: Zhang, Wang, et al.
Veröffentlicht: (2024) -
Towards Remote Sensing Change Detection with Neural Memory
von: Yang, Zhenyu, et al.
Veröffentlicht: (2026)