Segment Any Vehicle: Semantic and Visual Context Driven SAM and A Benchmark
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Xiao, Wang, Ziwen, Wu, Wentao, Wang, Anjie, Wu, Jiashu, Pan, Yantao, Li, Chenglong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Vehicle-centric Perception via Multimodal Structured Pre-training
by: Wu, Wentao, et al.
Published: (2025)
by: Wu, Wentao, et al.
Published: (2025)
T2I-VeRW: Part-level Fine-grained Perception for Text-to-Image Vehicle Retrieval
by: Wang, Xiao, et al.
Published: (2026)
by: Wang, Xiao, et al.
Published: (2026)
X-SAM: From Segment Anything to Any Segmentation
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
SAM4UDASS: When SAM Meets Unsupervised Domain Adaptive Semantic Segmentation in Intelligent Vehicles
by: Yan, Weihao, et al.
Published: (2023)
by: Yan, Weihao, et al.
Published: (2023)
Compress Any Segment Anything Model (SAM)
by: Fan, Juntong, et al.
Published: (2025)
by: Fan, Juntong, et al.
Published: (2025)
X2SAM: Any Segmentation in Images and Videos
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
StitchFusion: Weaving Any Visual Modalities to Enhance Multimodal Semantic Segmentation
by: Li, Bingyu, et al.
Published: (2024)
by: Li, Bingyu, et al.
Published: (2024)
ConformalSAM: Unlocking the Potential of Foundational Segmentation Models in Semi-Supervised Semantic Segmentation with Conformal Prediction
by: Chen, Danhui, et al.
Published: (2025)
by: Chen, Danhui, et al.
Published: (2025)
VFM-Det: Towards High-Performance Vehicle Detection via Large Foundation Models
by: Wu, Wentao, et al.
Published: (2024)
by: Wu, Wentao, et al.
Published: (2024)
Large Language Model Guided Progressive Feature Alignment for Multimodal UAV Object Detection
by: Wu, Wentao, et al.
Published: (2025)
by: Wu, Wentao, et al.
Published: (2025)
SAM 2: Segment Anything in Images and Videos
by: Ravi, Nikhila, et al.
Published: (2024)
by: Ravi, Nikhila, et al.
Published: (2024)
CRISP-SAM2: SAM2 with Cross-Modal Interaction and Semantic Prompting for Multi-Organ Segmentation
by: Yu, Xinlei, et al.
Published: (2025)
by: Yu, Xinlei, et al.
Published: (2025)
LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts
by: Xiao, Yijia, et al.
Published: (2024)
by: Xiao, Yijia, et al.
Published: (2024)
Multimodal SAM-adapter for Semantic Segmentation
by: Curti, Iacopo, et al.
Published: (2025)
by: Curti, Iacopo, et al.
Published: (2025)
IPixMatch: Boost Semi-supervised Semantic Segmentation with Inter-Pixel Relation
by: Wu, Kebin, et al.
Published: (2024)
by: Wu, Kebin, et al.
Published: (2024)
SAM Guided Semantic and Motion Changed Region Mining for Remote Sensing Change Captioning
by: Wang, Futian, et al.
Published: (2025)
by: Wang, Futian, et al.
Published: (2025)
Adversarial Semantic and Label Perturbation Attack for Pedestrian Attribute Recognition
by: Kong, Weizhe, et al.
Published: (2025)
by: Kong, Weizhe, et al.
Published: (2025)
UnSAMv2: Self-Supervised Learning Enables Segment Anything at Any Granularity
by: Yu, Junwei, et al.
Published: (2025)
by: Yu, Junwei, et al.
Published: (2025)
CoCo-SAM3: Harnessing Concept Conflict in Open-Vocabulary Semantic Segmentation
by: Chen, Yanhui, et al.
Published: (2026)
by: Chen, Yanhui, et al.
Published: (2026)
Generate Any Scene: Scene Graph Driven Data Synthesis for Visual Generation Training
by: Gao, Ziqi, et al.
Published: (2024)
by: Gao, Ziqi, et al.
Published: (2024)
SeqSAM: Autoregressive Multiple Hypothesis Prediction for Medical Image Segmentation using SAM
by: Towle, Benjamin, et al.
Published: (2025)
by: Towle, Benjamin, et al.
Published: (2025)
FloorSAM: SAM-Guided Floorplan Reconstruction with Semantic-Geometric Fusion
by: Ye, Han, et al.
Published: (2025)
by: Ye, Han, et al.
Published: (2025)
DeiSAM: Segment Anything with Deictic Prompting
by: Shindo, Hikaru, et al.
Published: (2024)
by: Shindo, Hikaru, et al.
Published: (2024)
TopoLoRA-SAM: Topology-Aware Parameter-Efficient Adaptation of Foundation Segmenters for Thin-Structure and Cross-Domain Binary Semantic Segmentation
by: Khazem, Salim
Published: (2026)
by: Khazem, Salim
Published: (2026)
CM3AE: A Unified RGB Frame and Event-Voxel/-Frame Pre-training Framework
by: Wu, Wentao, et al.
Published: (2025)
by: Wu, Wentao, et al.
Published: (2025)
Semore: VLM-guided Enhanced Semantic Motion Representations for Visual Reinforcement Learning
by: Wang, Wentao, et al.
Published: (2025)
by: Wang, Wentao, et al.
Published: (2025)
ProMISe: Promptable Medical Image Segmentation using SAM
by: Wang, Jinfeng, et al.
Published: (2024)
by: Wang, Jinfeng, et al.
Published: (2024)
Restoration Adaptation for Semantic Segmentation on Low Quality Images
by: Guan, Kai, et al.
Published: (2026)
by: Guan, Kai, et al.
Published: (2026)
SGAT4PASS: Spherical Geometry-Aware Transformer for PAnoramic Semantic Segmentation
by: Li, Xuewei, et al.
Published: (2023)
by: Li, Xuewei, et al.
Published: (2023)
The Perceptual Bandwidth Bottleneck in Vision-Language Models: Active Visual Reasoning via Sequential Experimental Design
by: Liu, Anjie, et al.
Published: (2026)
by: Liu, Anjie, et al.
Published: (2026)
Visual Prompt Selection for In-Context Learning Segmentation
by: Suo, Wei, et al.
Published: (2024)
by: Suo, Wei, et al.
Published: (2024)
LIX: Implicitly Infusing Spatial Geometric Prior Knowledge into Visual Semantic Segmentation for Autonomous Driving
by: Guo, Sicen, et al.
Published: (2024)
by: Guo, Sicen, et al.
Published: (2024)
Language-Driven Visual Consensus for Zero-Shot Semantic Segmentation
by: Zhang, Zicheng, et al.
Published: (2024)
by: Zhang, Zicheng, et al.
Published: (2024)
Learning Local and Global Temporal Contexts for Video Semantic Segmentation
by: Sun, Guolei, et al.
Published: (2022)
by: Sun, Guolei, et al.
Published: (2022)
How Do Optical Flow and Textual Prompts Collaborate to Assist in Audio-Visual Semantic Segmentation?
by: Lee, Yujian, et al.
Published: (2026)
by: Lee, Yujian, et al.
Published: (2026)
ScSAM: Debiasing Morphology and Distributional Variability in Subcellular Semantic Segmentation
by: Fang, Bo, et al.
Published: (2025)
by: Fang, Bo, et al.
Published: (2025)
MIAS-SAM: Medical Image Anomaly Segmentation without thresholding
by: Colussi, Marco, et al.
Published: (2025)
by: Colussi, Marco, et al.
Published: (2025)
Robust Box Prompt based SAM for Medical Image Segmentation
by: Huang, Yuhao, et al.
Published: (2024)
by: Huang, Yuhao, et al.
Published: (2024)
PointSSC: A Cooperative Vehicle-Infrastructure Point Cloud Benchmark for Semantic Scene Completion
by: Yan, Yuxiang, et al.
Published: (2023)
by: Yan, Yuxiang, et al.
Published: (2023)
Medical SAM3: A Foundation Model for Universal Prompt-Driven Medical Image Segmentation
by: Jiang, Chongcong, et al.
Published: (2026)
by: Jiang, Chongcong, et al.
Published: (2026)
Similar Items
-
Vehicle-centric Perception via Multimodal Structured Pre-training
by: Wu, Wentao, et al.
Published: (2025) -
T2I-VeRW: Part-level Fine-grained Perception for Text-to-Image Vehicle Retrieval
by: Wang, Xiao, et al.
Published: (2026) -
X-SAM: From Segment Anything to Any Segmentation
by: Wang, Hao, et al.
Published: (2025) -
SAM4UDASS: When SAM Meets Unsupervised Domain Adaptive Semantic Segmentation in Intelligent Vehicles
by: Yan, Weihao, et al.
Published: (2023) -
Compress Any Segment Anything Model (SAM)
by: Fan, Juntong, et al.
Published: (2025)