SurgAtt-Tracker: Online Surgical Attention Tracking via Temporal Proposal Reranking and Motion-Aware Refinement
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Rulin, Wang, Guankun, Wang, An, Ma, Yujie, Ouyang, Lixin, Cui, Bolin, Li, Junyan, Zhu, Chaowei, Li, Mingyang, Chen, Ming, Zhong, Xiaopin, Lu, Peng, Wang, Jiankun, Liu, Xianming, Ren, Hongliang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Bridging Vision and Language for Robust Context-Aware Surgical Point Tracking: The VL-SurgPT Dataset and Benchmark
von: Zhou, Rulin, et al.
Veröffentlicht: (2025)
von: Zhou, Rulin, et al.
Veröffentlicht: (2025)
TSUBF-Net: Trans-Spatial UNet-like Network with Bi-direction Fusion for Segmentation of Adenoid Hypertrophy in CT
von: Zhou, Rulin, et al.
Veröffentlicht: (2024)
von: Zhou, Rulin, et al.
Veröffentlicht: (2024)
Adapting SAM for Surgical Instrument Tracking and Segmentation in Endoscopic Submucosal Dissection Videos
von: Yu, Jieming, et al.
Veröffentlicht: (2024)
von: Yu, Jieming, et al.
Veröffentlicht: (2024)
SurgTPGS: Semantic 3D Surgical Scene Understanding with Text Promptable Gaussian Splatting
von: Huang, Yiming, et al.
Veröffentlicht: (2025)
von: Huang, Yiming, et al.
Veröffentlicht: (2025)
Mask Focal Loss: A unifying framework for dense crowd counting with canonical object detection networks
von: Zhong, Xiaopin, et al.
Veröffentlicht: (2022)
von: Zhong, Xiaopin, et al.
Veröffentlicht: (2022)
Endo-TTAP: Robust Endoscopic Tissue Tracking via Multi-Facet Guided Attention and Hybrid Flow-point Supervision
von: Zhou, Rulin, et al.
Veröffentlicht: (2025)
von: Zhou, Rulin, et al.
Veröffentlicht: (2025)
SurgVidLM: Towards Multi-grained Surgical Video Understanding with Large Language Model
von: Wang, Guankun, et al.
Veröffentlicht: (2025)
von: Wang, Guankun, et al.
Veröffentlicht: (2025)
Surgical-VQLA++: Adversarial Contrastive Learning for Calibrated Robust Visual Question-Localized Answering in Robotic Surgery
von: Bai, Long, et al.
Veröffentlicht: (2024)
von: Bai, Long, et al.
Veröffentlicht: (2024)
How can reasoning capability empower the AI copilot robot in endoscopic surgery
von: Wang, Guankun, et al.
Veröffentlicht: (2026)
von: Wang, Guankun, et al.
Veröffentlicht: (2026)
CoPESD: A Multi-Level Surgical Motion Dataset for Training Large Vision-Language Models to Co-Pilot Endoscopic Submucosal Dissection
von: Wang, Guankun, et al.
Veröffentlicht: (2024)
von: Wang, Guankun, et al.
Veröffentlicht: (2024)
Geo-RepNet: Geometry-Aware Representation Learning for Surgical Phase Recognition in Endoscopic Submucosal Dissection
von: Tang, Rui, et al.
Veröffentlicht: (2025)
von: Tang, Rui, et al.
Veröffentlicht: (2025)
EndoControlMag: Robust Endoscopic Vascular Motion Magnification with Periodic Reference Resetting and Hierarchical Tissue-aware Dual-Mask Control
von: Wang, An, et al.
Veröffentlicht: (2025)
von: Wang, An, et al.
Veröffentlicht: (2025)
SurgSora: Object-Aware Diffusion Model for Controllable Surgical Video Generation
von: Chen, Tong, et al.
Veröffentlicht: (2024)
von: Chen, Tong, et al.
Veröffentlicht: (2024)
OSSAR: Towards Open-Set Surgical Activity Recognition in Robot-assisted Surgery
von: Bai, Long, et al.
Veröffentlicht: (2024)
von: Bai, Long, et al.
Veröffentlicht: (2024)
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos
von: Wu, Jinlin, et al.
Veröffentlicht: (2026)
von: Wu, Jinlin, et al.
Veröffentlicht: (2026)
BleedOrigin: Dynamic Bleeding Source Localization in Endoscopic Submucosal Dissection via Dual-Stage Detection and Tracking
von: Xu, Mengya, et al.
Veröffentlicht: (2025)
von: Xu, Mengya, et al.
Veröffentlicht: (2025)
LightFC-X: Lightweight Convolutional Tracker for RGB-X Tracking
von: Li, Yunfeng, et al.
Veröffentlicht: (2025)
von: Li, Yunfeng, et al.
Veröffentlicht: (2025)
DeTracker: Motion-decoupled Vehicle Detection and Tracking in Unstabilized Satellite Videos
von: Chen, Jiajun, et al.
Veröffentlicht: (2026)
von: Chen, Jiajun, et al.
Veröffentlicht: (2026)
Coarse-to-Fine Proposal Refinement Framework for Audio Temporal Forgery Detection and Localization
von: Wu, Junyan, et al.
Veröffentlicht: (2024)
von: Wu, Junyan, et al.
Veröffentlicht: (2024)
UniTracker: Learning Universal Whole-Body Motion Tracker for Humanoid Robots
von: Yin, Kangning, et al.
Veröffentlicht: (2025)
von: Yin, Kangning, et al.
Veröffentlicht: (2025)
RGB-Sonar Tracking Benchmark and Spatial Cross-Attention Transformer Tracker
von: Li, Yunfeng, et al.
Veröffentlicht: (2024)
von: Li, Yunfeng, et al.
Veröffentlicht: (2024)
SurgTrack: CAD-Free 3D Tracking of Real-world Surgical Instruments
von: Guo, Wenwu, et al.
Veröffentlicht: (2024)
von: Guo, Wenwu, et al.
Veröffentlicht: (2024)
EndoOOD: Uncertainty-aware Out-of-distribution Detection in Capsule Endoscopy Diagnosis
von: Tan, Qiaozhi, et al.
Veröffentlicht: (2024)
von: Tan, Qiaozhi, et al.
Veröffentlicht: (2024)
TMR-VLA:Vision-Language-Action Model for Magnetic Motion Control of Tri-leg Silicone-based Soft Robot
von: Tang, Ruijie, et al.
Veröffentlicht: (2026)
von: Tang, Ruijie, et al.
Veröffentlicht: (2026)
Dual-Rerank: Fusing Causality and Utility for Industrial Generative Reranking
von: Zhang, Chao, et al.
Veröffentlicht: (2026)
von: Zhang, Chao, et al.
Veröffentlicht: (2026)
EndoVLA: Dual-Phase Vision-Language-Action Model for Autonomous Tracking in Endoscopy
von: Ng, Chi Kit, et al.
Veröffentlicht: (2025)
von: Ng, Chi Kit, et al.
Veröffentlicht: (2025)
SurgPLAN++: Universal Surgical Phase Localization Network for Online and Offline Inference
von: Chen, Zhen, et al.
Veröffentlicht: (2024)
von: Chen, Zhen, et al.
Veröffentlicht: (2024)
OmniTracker: Unifying Object Tracking by Tracking-with-Detection
von: Wang, Junke, et al.
Veröffentlicht: (2023)
von: Wang, Junke, et al.
Veröffentlicht: (2023)
Surgical Visual Understanding (SurgVU) Dataset
von: Zia, Aneeq, et al.
Veröffentlicht: (2025)
von: Zia, Aneeq, et al.
Veröffentlicht: (2025)
ReSurgSAM2: Referring Segment Anything in Surgical Video via Credible Long-term Tracking
von: Liu, Haofeng, et al.
Veröffentlicht: (2025)
von: Liu, Haofeng, et al.
Veröffentlicht: (2025)
Intuitive Surgical SurgToolLoc and SurgVU Challenges Results: 2022-2025
von: Zia, Aneeq, et al.
Veröffentlicht: (2023)
von: Zia, Aneeq, et al.
Veröffentlicht: (2023)
Multimodal Graph Representation Learning for Robust Surgical Workflow Recognition with Adversarial Feature Disentanglement
von: Bai, Long, et al.
Veröffentlicht: (2025)
von: Bai, Long, et al.
Veröffentlicht: (2025)
Surgical-LVLM: Learning to Adapt Large Vision-Language Model for Grounded Visual Question Answering in Robotic Surgery
von: Wang, Guankun, et al.
Veröffentlicht: (2024)
von: Wang, Guankun, et al.
Veröffentlicht: (2024)
PDZSeg: Adapting the Foundation Model for Dissection Zone Segmentation with Visual Prompts in Robot-assisted Endoscopic Submucosal Dissection
von: Xu, Mengya, et al.
Veröffentlicht: (2024)
von: Xu, Mengya, et al.
Veröffentlicht: (2024)
TimeTracker: Event-based Continuous Point Tracking for Video Frame Interpolation with Non-linear Motion
von: Liu, Haoyue, et al.
Veröffentlicht: (2025)
von: Liu, Haoyue, et al.
Veröffentlicht: (2025)
SurgPose: a Dataset for Articulated Robotic Surgical Tool Pose Estimation and Tracking
von: Wu, Zijian, et al.
Veröffentlicht: (2025)
von: Wu, Zijian, et al.
Veröffentlicht: (2025)
Semantic Ensemble Loss and Latent Refinement for High-Fidelity Neural Image Compression
von: Li, Daxin, et al.
Veröffentlicht: (2024)
von: Li, Daxin, et al.
Veröffentlicht: (2024)
Collision Risk Quantification and Conflict Resolution in Trajectory Tracking for Acceleration-Actuated Multi-Robot Systems
von: Li, Xiaoxiao, et al.
Veröffentlicht: (2025)
von: Li, Xiaoxiao, et al.
Veröffentlicht: (2025)
EndoARSS: Adapting Spatially-Aware Foundation Model for Efficient Activity Recognition and Semantic Segmentation in Endoscopic Surgery
von: Wang, Guankun, et al.
Veröffentlicht: (2025)
von: Wang, Guankun, et al.
Veröffentlicht: (2025)
EndoARSS: Adapting Spatially Aware Foundation Model for Efficient Activity Recognition and Semantic Segmentation in Endoscopic Surgery
von: Guankun Wang, et al.
Veröffentlicht: (2025)
von: Guankun Wang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Bridging Vision and Language for Robust Context-Aware Surgical Point Tracking: The VL-SurgPT Dataset and Benchmark
von: Zhou, Rulin, et al.
Veröffentlicht: (2025) -
TSUBF-Net: Trans-Spatial UNet-like Network with Bi-direction Fusion for Segmentation of Adenoid Hypertrophy in CT
von: Zhou, Rulin, et al.
Veröffentlicht: (2024) -
Adapting SAM for Surgical Instrument Tracking and Segmentation in Endoscopic Submucosal Dissection Videos
von: Yu, Jieming, et al.
Veröffentlicht: (2024) -
SurgTPGS: Semantic 3D Surgical Scene Understanding with Text Promptable Gaussian Splatting
von: Huang, Yiming, et al.
Veröffentlicht: (2025) -
Mask Focal Loss: A unifying framework for dense crowd counting with canonical object detection networks
von: Zhong, Xiaopin, et al.
Veröffentlicht: (2022)