SAM 2 in Robotic Surgery: An Empirical Evaluation for Robustness and Generalization in Surgical Video Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Jieming, Wang, An, Dong, Wenzhen, Xu, Mengya, Islam, Mobarakol, Wang, Jie, Bai, Long, Ren, Hongliang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adapting SAM for Surgical Instrument Tracking and Segmentation in Endoscopic Submucosal Dissection Videos
by: Yu, Jieming, et al.
Published: (2024)
by: Yu, Jieming, et al.
Published: (2024)
Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery
by: Wang, Guankun, et al.
Published: (2024)
by: Wang, Guankun, et al.
Published: (2024)
Web-based Augmented Reality with Auto-Scaling and Real-Time Head Tracking towards Markerless Neurointerventional Preoperative Planning and Training of Head-mounted Robotic Needle Insertion
by: Ho, Hon Lung, et al.
Published: (2024)
by: Ho, Hon Lung, et al.
Published: (2024)
EndoDAC: Efficient Adapting Foundation Model for Self-Supervised Depth Estimation from Any Endoscopic Camera
by: Cui, Beilei, et al.
Published: (2024)
by: Cui, Beilei, et al.
Published: (2024)
Surgical-VQLA++: Adversarial Contrastive Learning for Calibrated Robust Visual Question-Localized Answering in Robotic Surgery
by: Bai, Long, et al.
Published: (2024)
by: Bai, Long, et al.
Published: (2024)
Surgical SAM 2: Real-time Segment Anything in Surgical Video by Efficient Frame Pruning
by: Liu, Haofeng, et al.
Published: (2024)
by: Liu, Haofeng, et al.
Published: (2024)
More than Segmentation: Benchmarking SAM 3 for Segmentation, 3D Perception, and Reconstruction in Robotic Surgery
by: Dong, Wenzhen, et al.
Published: (2025)
by: Dong, Wenzhen, et al.
Published: (2025)
Transferring Knowledge from High-Quality to Low-Quality MRI for Adult Glioma Diagnosis
by: Zhao, Yanguang, et al.
Published: (2024)
by: Zhao, Yanguang, et al.
Published: (2024)
Privacy-Preserving Synthetic Continual Semantic Segmentation for Robotic Surgery
by: Xu, Mengya, et al.
Published: (2024)
by: Xu, Mengya, et al.
Published: (2024)
VISTA: A Benchmark for Real-Time Video Streaming under Network Impairments in Surgical Teleoperation
by: Deng, Zexin, et al.
Published: (2026)
by: Deng, Zexin, et al.
Published: (2026)
OSSAR: Towards Open-Set Surgical Activity Recognition in Robot-assisted Surgery
by: Bai, Long, et al.
Published: (2024)
by: Bai, Long, et al.
Published: (2024)
Illumination Histogram Consistency Metric for Quantitative Assessment of Video Sequences
by: Chen, Long, et al.
Published: (2024)
by: Chen, Long, et al.
Published: (2024)
Benchmarking Robustness of Endoscopic Depth Estimation with Synthetically Corrupted Data
by: Wang, An, et al.
Published: (2024)
by: Wang, An, et al.
Published: (2024)
Robotic CBCT Meets Robotic Ultrasound
by: Li, Feng, et al.
Published: (2025)
by: Li, Feng, et al.
Published: (2025)
QuaDreamer: Controllable Panoramic Video Generation for Quadruped Robots
by: Wu, Sheng, et al.
Published: (2025)
by: Wu, Sheng, et al.
Published: (2025)
D-Compress: Detail-Preserving LiDAR Range Image Compression for Real-Time Streaming on Resource-Constrained Robots
by: Wang, Shengqian, et al.
Published: (2026)
by: Wang, Shengqian, et al.
Published: (2026)
A Hybrid-Layered System for Image-Guided Navigation and Robot Assisted Spine Surgeries
by: T, Suhail Ansari, et al.
Published: (2024)
by: T, Suhail Ansari, et al.
Published: (2024)
Surface-Enhanced Raman Spectroscopy and Transfer Learning Toward Accurate Reconstruction of the Surgical Zone
by: Raman, Ashutosh, et al.
Published: (2024)
by: Raman, Ashutosh, et al.
Published: (2024)
Performance and Non-adversarial Robustness of the Segment Anything Model 2 in Surgical Video Segmentation
by: Shen, Yiqing, et al.
Published: (2024)
by: Shen, Yiqing, et al.
Published: (2024)
SELECT: A Submodular Approach for Active LiDAR Semantic Segmentation
by: Mao, Ruiyu, et al.
Published: (2025)
by: Mao, Ruiyu, et al.
Published: (2025)
MonoSIM: An open source SIL framework for Ackermann Vehicular Systems with Monocular Vision
by: Rahman, Shantanu, et al.
Published: (2026)
by: Rahman, Shantanu, et al.
Published: (2026)
Noise Analysis and Modeling of the PMD Flexx2 Depth Camera for Robotic Applications
by: Cai, Yuke, et al.
Published: (2024)
by: Cai, Yuke, et al.
Published: (2024)
Training-Free Robot Pose Estimation using Off-the-Shelf Foundational Models
by: Liang, Laurence
Published: (2025)
by: Liang, Laurence
Published: (2025)
Invascal: Inverse-Vacuity Self-Calibration for Uncertainty-Aware LiDAR Range-View Semantic Segmentation
by: Turacan, Kerim, et al.
Published: (2026)
by: Turacan, Kerim, et al.
Published: (2026)
SceneVGGT: VGGT-based online 3D semantic SLAM for indoor scene understanding and navigation
by: Gelencsér-Horváth, Anna, et al.
Published: (2026)
by: Gelencsér-Horváth, Anna, et al.
Published: (2026)
Force-EvT: A Closer Look at Robotic Gripper Force Measurement with Event-based Vision Transformer
by: Guo, Qianyu, et al.
Published: (2024)
by: Guo, Qianyu, et al.
Published: (2024)
Object-Centric Representations Improve Policy Generalization in Robot Manipulation
by: Chapin, Alexandre, et al.
Published: (2025)
by: Chapin, Alexandre, et al.
Published: (2025)
Segment-to-Act: Label-Noise-Robust Action-Prompted Video Segmentation Towards Embodied Intelligence
by: Li, Wenxin, et al.
Published: (2025)
by: Li, Wenxin, et al.
Published: (2025)
MediViSTA: Medical Video Segmentation via Temporal Fusion SAM Adaptation for Echocardiography
by: Kim, Sekeun, et al.
Published: (2023)
by: Kim, Sekeun, et al.
Published: (2023)
Sensorless Remote Center of Motion Misalignment Estimation
by: Yang, Hao, et al.
Published: (2025)
by: Yang, Hao, et al.
Published: (2025)
SAM2S: Segment Anything in Surgical Videos via Semantic Long-term Tracking
by: Liu, Haofeng, et al.
Published: (2025)
by: Liu, Haofeng, et al.
Published: (2025)
Towards Robotic Knee Arthroscopy: Multi-Scale Network for Tissue-Tool Segmentation
by: Ali, Shahnewaz, et al.
Published: (2021)
by: Ali, Shahnewaz, et al.
Published: (2021)
CaveSeg: Deep Semantic Segmentation and Scene Parsing for Autonomous Underwater Cave Exploration
by: Abdullah, A., et al.
Published: (2023)
by: Abdullah, A., et al.
Published: (2023)
A Review of 3D Reconstruction Techniques for Deformable Tissues in Robotic Surgery
by: Xu, Mengya, et al.
Published: (2024)
by: Xu, Mengya, et al.
Published: (2024)
VRUD: A Drone Dataset for Complex Vehicle-VRU Interactions within Mixed Traffic
by: Wang, Ziyu, et al.
Published: (2026)
by: Wang, Ziyu, et al.
Published: (2026)
MBA-Net: SAM-driven Bidirectional Aggregation Network for Ovarian Tumor Segmentation
by: Gao, Yifan, et al.
Published: (2024)
by: Gao, Yifan, et al.
Published: (2024)
OneOcc: Semantic Occupancy Prediction for Legged Robots with a Single Panoramic Camera
by: Shi, Hao, et al.
Published: (2025)
by: Shi, Hao, et al.
Published: (2025)
Optimizing Prompt Strategies for SAM: Advancing lesion Segmentation Across Diverse Medical Imaging Modalities
by: Wang, Yuli, et al.
Published: (2024)
by: Wang, Yuli, et al.
Published: (2024)
Visible Light Communication using Led-Based AR Markers for Robot Localization
by: Uemura, Wataru, et al.
Published: (2026)
by: Uemura, Wataru, et al.
Published: (2026)
Benchmarking the Robustness of Optical Flow Estimation to Corruptions
by: Yi, Zhonghua, et al.
Published: (2024)
by: Yi, Zhonghua, et al.
Published: (2024)
Similar Items
-
Adapting SAM for Surgical Instrument Tracking and Segmentation in Endoscopic Submucosal Dissection Videos
by: Yu, Jieming, et al.
Published: (2024) -
Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery
by: Wang, Guankun, et al.
Published: (2024) -
Web-based Augmented Reality with Auto-Scaling and Real-Time Head Tracking towards Markerless Neurointerventional Preoperative Planning and Training of Head-mounted Robotic Needle Insertion
by: Ho, Hon Lung, et al.
Published: (2024) -
EndoDAC: Efficient Adapting Foundation Model for Self-Supervised Depth Estimation from Any Endoscopic Camera
by: Cui, Beilei, et al.
Published: (2024) -
Surgical-VQLA++: Adversarial Contrastive Learning for Calibrated Robust Visual Question-Localized Answering in Robotic Surgery
by: Bai, Long, et al.
Published: (2024)