GroundedSurg: A Multi-Procedure Benchmark for Language-Conditioned Surgical Tool Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Ashraf, Tajamul, Riyaz, Abrar Ul, Tak, Wasif, Tariq, Tavaheed, Yadav, Sonia, Abdar, Moloud, Bashir, Janibul |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
QTrack: Query-Driven Reasoning for Multi-modal MOT
by: Ashraf, Tajamul, et al.
Published: (2026)
by: Ashraf, Tajamul, et al.
Published: (2026)
Context Aware Grounded Teacher for Source Free Object Detection
by: Ashraf, Tajamul, et al.
Published: (2025)
by: Ashraf, Tajamul, et al.
Published: (2025)
TITAN: Query-Token based Domain Adaptive Adversarial Learning
by: Ashraf, Tajamul, et al.
Published: (2025)
by: Ashraf, Tajamul, et al.
Published: (2025)
FATE: Focal-modulated Attention Encoder for Multivariate Time-series Forecasting
by: Ashraf, Tajamul, et al.
Published: (2024)
by: Ashraf, Tajamul, et al.
Published: (2024)
ATR-Bench: A Federated Learning Benchmark for Adaptation, Trust, and Reasoning
by: Ashraf, Tajamul, et al.
Published: (2025)
by: Ashraf, Tajamul, et al.
Published: (2025)
Bolbosh: Script-Aware Flow Matching for Kashmiri Text-to-Speech
by: Ashraf, Tajamul, et al.
Published: (2026)
by: Ashraf, Tajamul, et al.
Published: (2026)
HF-Fed: Hierarchical based customized Federated Learning Framework for X-Ray Imaging
by: Ashraf, Tajamul, et al.
Published: (2024)
by: Ashraf, Tajamul, et al.
Published: (2024)
BayTTA: Uncertainty-aware medical image classification with optimized test-time augmentation using Bayesian model averaging
by: Sherkatghanad, Zeinab, et al.
Published: (2024)
by: Sherkatghanad, Zeinab, et al.
Published: (2024)
FOCUS: Bridging Fine-Grained Recognition and Open-World Discovery across Domains
by: Rathore, Vaibhav, et al.
Published: (2026)
by: Rathore, Vaibhav, et al.
Published: (2026)
MedSPOT: A Workflow-Aware Sequential Grounding Benchmark for Clinical GUI
by: Shakeel, Rozain, et al.
Published: (2026)
by: Shakeel, Rozain, et al.
Published: (2026)
Generalizable Federated Learning using Client Adaptive Focal Modulation
by: Ashraf, Tajamul, et al.
Published: (2025)
by: Ashraf, Tajamul, et al.
Published: (2025)
Diffusion Models for Influence Maximization on Temporal Networks: A Guide to Make the Best Choice
by: Zahoor, Aaqib, et al.
Published: (2025)
by: Zahoor, Aaqib, et al.
Published: (2025)
Debiasing CLIP: Interpreting and Correcting Bias in Attention Heads
by: Yeo, Wei Jie, et al.
Published: (2025)
by: Yeo, Wei Jie, et al.
Published: (2025)
Intuitive Surgical SurgToolLoc and SurgVU Challenges Results: 2022-2025
by: Zia, Aneeq, et al.
Published: (2023)
by: Zia, Aneeq, et al.
Published: (2023)
Unknown Prompt, the only Lacuna: Unveiling CLIP's Potential for Open Domain Generalization
by: Singha, Mainak, et al.
Published: (2024)
by: Singha, Mainak, et al.
Published: (2024)
RoboSurg-VQA: A Multimodal Benchmark for Surgical Segmentation-Aware Visual Question Answering
by: Zhang, Chengyi, et al.
Published: (2026)
by: Zhang, Chengyi, et al.
Published: (2026)
Phase-Informed Tool Segmentation for Manual Small-Incision Cataract Surgery
by: Sachdeva, Bhuvan, et al.
Published: (2024)
by: Sachdeva, Bhuvan, et al.
Published: (2024)
SurgWound-Bench: A Benchmark for Surgical Wound Diagnosis
by: Xu, Jiahao, et al.
Published: (2025)
by: Xu, Jiahao, et al.
Published: (2025)
FedMVP: Federated Multimodal Visual Prompt Tuning for Vision-Language Models
by: Singha, Mainak, et al.
Published: (2025)
by: Singha, Mainak, et al.
Published: (2025)
SurgRIPE challenge: Benchmark of Surgical Robot Instrument Pose Estimation
by: Xu, Haozheng, et al.
Published: (2025)
by: Xu, Haozheng, et al.
Published: (2025)
SurgViVQA: Temporally-Grounded Video Question Answering for Surgical Scene Understanding
by: Drago, Mauro Orazio, et al.
Published: (2025)
by: Drago, Mauro Orazio, et al.
Published: (2025)
A Novel Reduced Switch Five‐Level Common Ground Transformerless Inverter Feeding an Isolated Load
by: Tajamul Hayat Parray, et al.
Published: (2026)
by: Tajamul Hayat Parray, et al.
Published: (2026)
SurgPose: a Dataset for Articulated Robotic Surgical Tool Pose Estimation and Tracking
by: Wu, Zijian, et al.
Published: (2025)
by: Wu, Zijian, et al.
Published: (2025)
MATRIX: Multimodal Agent Tuning for Robust Tool-Use Reasoning
by: Ashraf, Tajamul, et al.
Published: (2025)
by: Ashraf, Tajamul, et al.
Published: (2025)
SurgBench: A Unified Large-Scale Benchmark for Surgical Video Analysis
by: Wei, Jianhui, et al.
Published: (2025)
by: Wei, Jianhui, et al.
Published: (2025)
Enhance Hyperbolic Representation Learning via Second-order Pooling
by: Song, Kun, et al.
Published: (2024)
by: Song, Kun, et al.
Published: (2024)
UniSurgSAM: A Unified Promptable Model for Reliable Surgical Video Segmentation
by: Liu, Haofeng, et al.
Published: (2026)
by: Liu, Haofeng, et al.
Published: (2026)
D-MASTER: Mask Annealed Transformer for Unsupervised Domain Adaptation in Breast Cancer Detection from Mammograms
by: Ashraf, Tajamul, et al.
Published: (2024)
by: Ashraf, Tajamul, et al.
Published: (2024)
Surgical Visual Understanding (SurgVU) Dataset
by: Zia, Aneeq, et al.
Published: (2025)
by: Zia, Aneeq, et al.
Published: (2025)
Surg-SegFormer: A Dual Transformer-Based Model for Holistic Surgical Scene Segmentation
by: Ahmed, Fatimaelzahraa, et al.
Published: (2025)
by: Ahmed, Fatimaelzahraa, et al.
Published: (2025)
SurgCoT: Advancing Spatiotemporal Reasoning in Surgical Videos through a Chain-of-Thought Benchmark
by: Wang, Gui, et al.
Published: (2026)
by: Wang, Gui, et al.
Published: (2026)
SurgMLLMBench: A Multimodal Large Language Model Benchmark Dataset for Surgical Scene Understanding
by: Choi, Tae-Min, et al.
Published: (2025)
by: Choi, Tae-Min, et al.
Published: (2025)
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence
by: Zeng, Zhitao, et al.
Published: (2025)
by: Zeng, Zhitao, et al.
Published: (2025)
MIRA: A Novel Framework for Fusing Modalities in Medical RAG
by: Wang, Jinhong, et al.
Published: (2025)
by: Wang, Jinhong, et al.
Published: (2025)
SurgFed: Language-guided Multi-Task Federated Learning for Surgical Video Understanding
by: Fang, Zheng, et al.
Published: (2026)
by: Fang, Zheng, et al.
Published: (2026)
SpinalNet: Deep Neural Network with Gradual Input
by: Kabir, H M Dipu, et al.
Published: (2020)
by: Kabir, H M Dipu, et al.
Published: (2020)
Analog-digital Scheduling for Federated Learning: A Communication-Efficient Approach
by: Abrar, Muhammad Faraz Ul, et al.
Published: (2024)
by: Abrar, Muhammad Faraz Ul, et al.
Published: (2024)
Biased Over-the-Air Federated Learning under Wireless Heterogeneity
by: Abrar, Muhammad Faraz Ul, et al.
Published: (2024)
by: Abrar, Muhammad Faraz Ul, et al.
Published: (2024)
Biased Federated Learning under Wireless Heterogeneity
by: Abrar, Muhammad Faraz Ul, et al.
Published: (2025)
by: Abrar, Muhammad Faraz Ul, et al.
Published: (2025)
Non-Convex Over-the-Air Heterogeneous Federated Learning: A Bias-Variance Trade-off
by: Abrar, Muhammad Faraz Ul, et al.
Published: (2025)
by: Abrar, Muhammad Faraz Ul, et al.
Published: (2025)
Similar Items
-
QTrack: Query-Driven Reasoning for Multi-modal MOT
by: Ashraf, Tajamul, et al.
Published: (2026) -
Context Aware Grounded Teacher for Source Free Object Detection
by: Ashraf, Tajamul, et al.
Published: (2025) -
TITAN: Query-Token based Domain Adaptive Adversarial Learning
by: Ashraf, Tajamul, et al.
Published: (2025) -
FATE: Focal-modulated Attention Encoder for Multivariate Time-series Forecasting
by: Ashraf, Tajamul, et al.
Published: (2024) -
ATR-Bench: A Federated Learning Benchmark for Adaptation, Trust, and Reasoning
by: Ashraf, Tajamul, et al.
Published: (2025)