Text Promptable Surgical Instrument Segmentation with Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Zijian, Alabi, Oluwatosin, Wei, Meng, Vercauteren, Tom, Shi, Miaojing |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multitask Learning in Minimally Invasive Surgical Vision: A Review
by: Alabi, Oluwatosin, et al.
Published: (2024)
by: Alabi, Oluwatosin, et al.
Published: (2024)
Grounding Surgical Action Triplets with Instrument Instance Segmentation: A Dataset and Target-Aware Fusion Approach
by: Alabi, Oluwatosin, et al.
Published: (2025)
by: Alabi, Oluwatosin, et al.
Published: (2025)
SurgPIS: Surgical-instrument-level Instances and Part-level Semantics for Weakly-supervised Part-aware Instance Segmentation
by: Wei, Meng, et al.
Published: (2025)
by: Wei, Meng, et al.
Published: (2025)
Large Model driven Radiology Report Generation with Clinical Quality Reinforcement Learning
by: Zhou, Zijian, et al.
Published: (2024)
by: Zhou, Zijian, et al.
Published: (2024)
CholecInstanceSeg: A Tool Instance Segmentation Dataset for Laparoscopic Surgery
by: Alabi, Oluwatosin, et al.
Published: (2024)
by: Alabi, Oluwatosin, et al.
Published: (2024)
Rethinking Text-Promptable Surgical Instrument Segmentation with Robust Framework
by: Choi, Tae-Min, et al.
Published: (2024)
by: Choi, Tae-Min, et al.
Published: (2024)
VLPrompt: Vision-Language Prompting for Panoptic Scene Graph Generation
by: Zhou, Zijian, et al.
Published: (2023)
by: Zhou, Zijian, et al.
Published: (2023)
Where It Moves, It Matters: Referring Surgical Instrument Segmentation via Motion
by: Wei, Meng, et al.
Published: (2026)
by: Wei, Meng, et al.
Published: (2026)
Transferring Relative Monocular Depth to Surgical Vision with Temporal Consistency
by: Budd, Charlie, et al.
Published: (2024)
by: Budd, Charlie, et al.
Published: (2024)
ROBUST-MIPS: A Combined Skeletal Pose and Instance Segmentation Dataset for Laparoscopic Surgical Instruments
by: Han, Zhe, et al.
Published: (2025)
by: Han, Zhe, et al.
Published: (2025)
Enhancing Generalized Few-Shot Semantic Segmentation via Effective Knowledge Transfer
by: Chen, Xinyue, et al.
Published: (2024)
by: Chen, Xinyue, et al.
Published: (2024)
OpenPSG: Open-set Panoptic Scene Graph Generation via Large Multimodal Models
by: Zhou, Zijian, et al.
Published: (2024)
by: Zhou, Zijian, et al.
Published: (2024)
SegMatch: A semi-supervised learning method for surgical instrument segmentation
by: Wei, Meng, et al.
Published: (2023)
by: Wei, Meng, et al.
Published: (2023)
Weakly-Supervised Referring Video Object Segmentation through Text Supervision
by: Shi, Miaojing, et al.
Published: (2026)
by: Shi, Miaojing, et al.
Published: (2026)
Text-Promptable Propagation for Referring Medical Image Sequence Segmentation
by: Yuan, Runtian, et al.
Published: (2025)
by: Yuan, Runtian, et al.
Published: (2025)
SEG-SAM: Semantic-Guided SAM for Unified Medical Image Segmentation
by: Huang, Shuangping, et al.
Published: (2024)
by: Huang, Shuangping, et al.
Published: (2024)
LoSh: Long-Short Text Joint Prediction Network for Referring Video Object Segmentation
by: Yuan, Linfeng, et al.
Published: (2023)
by: Yuan, Linfeng, et al.
Published: (2023)
Augmenting Efficient Real-time Surgical Instrument Segmentation in Video with Point Tracking and Segment Anything
by: Wu, Zijian, et al.
Published: (2024)
by: Wu, Zijian, et al.
Published: (2024)
Memory-guided Network with Uncertainty-based Feature Augmentation for Few-shot Semantic Segmentation
by: Chen, Xinyue, et al.
Published: (2024)
by: Chen, Xinyue, et al.
Published: (2024)
ToolTipNet: A Segmentation-Driven Deep Learning Baseline for Surgical Instrument Tip Detection
by: Wu, Zijian, et al.
Published: (2025)
by: Wu, Zijian, et al.
Published: (2025)
Robust Promptable Video Object Segmentation
by: Lee, Sohyun, et al.
Published: (2026)
by: Lee, Sohyun, et al.
Published: (2026)
SegSLR: Promptable Video Segmentation for Isolated Sign Language Recognition
by: Schreiber, Sven, et al.
Published: (2025)
by: Schreiber, Sven, et al.
Published: (2025)
Enhancing Space-time Video Super-resolution via Spatial-temporal Feature Interaction
by: Yue, Zijie, et al.
Published: (2022)
by: Yue, Zijie, et al.
Published: (2022)
Unifying 3D Vision-Language Understanding via Promptable Queries
by: Zhu, Ziyu, et al.
Published: (2024)
by: Zhu, Ziyu, et al.
Published: (2024)
LEMON: A Large Endoscopic MONocular Dataset and Foundation Model for Perception in Surgical Settings
by: Che, Chengan, et al.
Published: (2025)
by: Che, Chengan, et al.
Published: (2025)
PRISM: A Promptable and Robust Interactive Segmentation Model with Visual Prompts
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
MATIS: Masked-Attention Transformers for Surgical Instrument Segmentation
by: Ayobi, Nicolás, et al.
Published: (2023)
by: Ayobi, Nicolás, et al.
Published: (2023)
nnInteractive: Redefining 3D Promptable Segmentation
by: Isensee, Fabian, et al.
Published: (2025)
by: Isensee, Fabian, et al.
Published: (2025)
Point-SAM: Promptable 3D Segmentation Model for Point Clouds
by: Zhou, Yuchen, et al.
Published: (2024)
by: Zhou, Yuchen, et al.
Published: (2024)
Bootstrapping Vision-language Models for Self-supervised Remote Physiological Measurement
by: Yue, Zijie, et al.
Published: (2024)
by: Yue, Zijie, et al.
Published: (2024)
Towards Interactive Lesion Segmentation in Whole-Body PET/CT with Promptable Models
by: Rokuss, Maximilian, et al.
Published: (2025)
by: Rokuss, Maximilian, et al.
Published: (2025)
Instrument-Splatting: Controllable Photorealistic Reconstruction of Surgical Instruments Using Gaussian Splatting
by: Yang, Shuojue, et al.
Published: (2025)
by: Yang, Shuojue, et al.
Published: (2025)
Event-Level Detection of Surgical Instrument Handovers in Videos with Interpretable Vision Models
by: Katsarou, Katerina, et al.
Published: (2026)
by: Katsarou, Katerina, et al.
Published: (2026)
Promptable Anomaly Segmentation with SAM Through Self-Perception Tuning
by: Yang, Hui-Yue, et al.
Published: (2024)
by: Yang, Hui-Yue, et al.
Published: (2024)
Leveraging Hallucinations to Reduce Manual Prompt Dependency in Promptable Segmentation
by: Hu, Jian, et al.
Published: (2024)
by: Hu, Jian, et al.
Published: (2024)
The World is Your Canvas: Painting Promptable Events with Reference Images, Trajectories, and Text
by: Wang, Hanlin, et al.
Published: (2025)
by: Wang, Hanlin, et al.
Published: (2025)
Vision-Language Models Provide Promptable Representations for Reinforcement Learning
by: Chen, William, et al.
Published: (2024)
by: Chen, William, et al.
Published: (2024)
MPDrive: Improving Spatial Understanding with Marker-Based Prompt Learning for Autonomous Driving
by: Zhang, Zhiyuan, et al.
Published: (2025)
by: Zhang, Zhiyuan, et al.
Published: (2025)
PartSAM: A Scalable Promptable Part Segmentation Model Trained on Native 3D Data
by: Zhu, Zhe, et al.
Published: (2025)
by: Zhu, Zhe, et al.
Published: (2025)
VoxTell: Free-Text Promptable Universal 3D Medical Image Segmentation
by: Rokuss, Maximilian, et al.
Published: (2025)
by: Rokuss, Maximilian, et al.
Published: (2025)
Similar Items
-
Multitask Learning in Minimally Invasive Surgical Vision: A Review
by: Alabi, Oluwatosin, et al.
Published: (2024) -
Grounding Surgical Action Triplets with Instrument Instance Segmentation: A Dataset and Target-Aware Fusion Approach
by: Alabi, Oluwatosin, et al.
Published: (2025) -
SurgPIS: Surgical-instrument-level Instances and Part-level Semantics for Weakly-supervised Part-aware Instance Segmentation
by: Wei, Meng, et al.
Published: (2025) -
Large Model driven Radiology Report Generation with Clinical Quality Reinforcement Learning
by: Zhou, Zijian, et al.
Published: (2024) -
CholecInstanceSeg: A Tool Instance Segmentation Dataset for Laparoscopic Surgery
by: Alabi, Oluwatosin, et al.
Published: (2024)