Training-free Temporal Object Tracking in Surgical Videos
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Koley, Subhadeep, Kadkhodamohammadi, Abdolrahim, Barbarisi, Santiago, Stoyanov, Danail, Luengo, Imanol |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Learning from Single Timestamps: Complexity Estimation in Laparoscopic Cholecystectomy
von: Anastasiou, Dimitrios, et al.
Veröffentlicht: (2025)
von: Anastasiou, Dimitrios, et al.
Veröffentlicht: (2025)
Graph Neural Networks for Surgical Scene Segmentation
von: Li, Yihan, et al.
Veröffentlicht: (2025)
von: Li, Yihan, et al.
Veröffentlicht: (2025)
Temporal Cluster Assignment for Efficient Real-Time Video Segmentation
von: Yung, Ka-Wai, et al.
Veröffentlicht: (2025)
von: Yung, Ka-Wai, et al.
Veröffentlicht: (2025)
Depth Augmented and FE Free 3D/2D Liver Registration for Laparoscopic Liver AR
von: Zhang, Hanyuan, et al.
Veröffentlicht: (2026)
von: Zhang, Hanyuan, et al.
Veröffentlicht: (2026)
Warm-Started Reinforcement Learning for Iterative 3D/2D Liver Registration
von: Zhang, Hanyuan, et al.
Veröffentlicht: (2026)
von: Zhang, Hanyuan, et al.
Veröffentlicht: (2026)
TemporalDoRA: Temporal PEFT for Robust Surgical Video Question Answering
von: Carlini, Luca, et al.
Veröffentlicht: (2026)
von: Carlini, Luca, et al.
Veröffentlicht: (2026)
Zero-shot Monocular Metric Depth for Endoscopic Images
von: Toussaint, Nicolas, et al.
Veröffentlicht: (2025)
von: Toussaint, Nicolas, et al.
Veröffentlicht: (2025)
SurgViVQA: Temporally-Grounded Video Question Answering for Surgical Scene Understanding
von: Drago, Mauro Orazio, et al.
Veröffentlicht: (2025)
von: Drago, Mauro Orazio, et al.
Veröffentlicht: (2025)
Think Step by Step: Chain-of-Gesture Prompting for Error Detection in Robotic Surgical Videos
von: Shao, Zhimin, et al.
Veröffentlicht: (2024)
von: Shao, Zhimin, et al.
Veröffentlicht: (2024)
Diff2DGS: Reliable Reconstruction of Occluded Surgical Scenes via 2D Gaussian Splatting
von: Song, Tianyi, et al.
Veröffentlicht: (2026)
von: Song, Tianyi, et al.
Veröffentlicht: (2026)
SHADeS: Self-supervised Monocular Depth Estimation Through Non-Lambertian Image Decomposition
von: Daher, Rema, et al.
Veröffentlicht: (2025)
von: Daher, Rema, et al.
Veröffentlicht: (2025)
SurgicalGS: Dynamic 3D Gaussian Splatting for Accurate Robotic-Assisted Surgical Scene Reconstruction
von: Chen, Jialei, et al.
Veröffentlicht: (2024)
von: Chen, Jialei, et al.
Veröffentlicht: (2024)
Confidence-aware Monocular Depth Estimation for Minimally Invasive Surgery
von: Asad, Muhammad, et al.
Veröffentlicht: (2026)
von: Asad, Muhammad, et al.
Veröffentlicht: (2026)
DreamColour: Controllable Video Colour Editing without Training
von: Utintu, Chaitat, et al.
Veröffentlicht: (2024)
von: Utintu, Chaitat, et al.
Veröffentlicht: (2024)
Tracking Everything in Robotic-Assisted Surgery
von: Zhan, Bohan, et al.
Veröffentlicht: (2024)
von: Zhan, Bohan, et al.
Veröffentlicht: (2024)
Temporally Guided Articulated Hand Pose Tracking in Surgical Videos
von: Louis, Nathan, et al.
Veröffentlicht: (2021)
von: Louis, Nathan, et al.
Veröffentlicht: (2021)
SurgAnt-ViVQA: Learning to Anticipate Surgical Events through GRU-Driven Temporal Cross-Attention
von: Dhake, Shreyas C., et al.
Veröffentlicht: (2025)
von: Dhake, Shreyas C., et al.
Veröffentlicht: (2025)
Surgical AI Copilot: Energy-Based Fourier Gradient Low-Rank Adaptation for Surgical LLM Agent Reasoning and Planning
von: Huang, Jiayuan, et al.
Veröffentlicht: (2025)
von: Huang, Jiayuan, et al.
Veröffentlicht: (2025)
Multi-Modal Monocular Endoscopic Depth and Pose Estimation with Edge-Guided Self-Supervision
von: Ju, Xinwei, et al.
Veröffentlicht: (2026)
von: Ju, Xinwei, et al.
Veröffentlicht: (2026)
CoRe-DA: Contrastive Regression for Unsupervised Domain Adaptation in Surgical Skill Assessment
von: Anastasiou, Dimitrios, et al.
Veröffentlicht: (2026)
von: Anastasiou, Dimitrios, et al.
Veröffentlicht: (2026)
SEDMamba: Enhancing Selective State Space Modelling with Bottleneck Mechanism and Fine-to-Coarse Temporal Fusion for Efficient Error Detection in Robot-Assisted Surgery
von: Xu, Jialang, et al.
Veröffentlicht: (2024)
von: Xu, Jialang, et al.
Veröffentlicht: (2024)
EndoLRMGS: Complete Endoscopic Scene Reconstruction combining Large Reconstruction Modelling and Gaussian Splatting
von: Wang, Xu, et al.
Veröffentlicht: (2025)
von: Wang, Xu, et al.
Veröffentlicht: (2025)
TSMS-SAM2: Multi-scale Temporal Sampling Augmentation and Memory-Splitting Pruning for Promptable Video Object Segmentation and Tracking in Surgical Scenarios
von: Xu, Guoping, et al.
Veröffentlicht: (2025)
von: Xu, Guoping, et al.
Veröffentlicht: (2025)
SegCol Challenge: Semantic Segmentation for Tools and Fold Edges in Colonoscopy data
von: Ju, Xinwei, et al.
Veröffentlicht: (2024)
von: Ju, Xinwei, et al.
Veröffentlicht: (2024)
High-fidelity Endoscopic Image Synthesis by Utilizing Depth-guided Neural Surfaces
von: Huang, Baoru, et al.
Veröffentlicht: (2024)
von: Huang, Baoru, et al.
Veröffentlicht: (2024)
Gaussian Pancakes: Geometrically-Regularized 3D Gaussian Splatting for Realistic Endoscopic Reconstruction
von: Bonilla, Sierra, et al.
Veröffentlicht: (2024)
von: Bonilla, Sierra, et al.
Veröffentlicht: (2024)
Multimodal Optimal Transport for Training-free Temporal Segmentation in Surgical Robotics
von: Mohamed, Omar, et al.
Veröffentlicht: (2026)
von: Mohamed, Omar, et al.
Veröffentlicht: (2026)
MOVi: Training-free Text-conditioned Multi-Object Video Generation
von: Rahman, Aimon, et al.
Veröffentlicht: (2025)
von: Rahman, Aimon, et al.
Veröffentlicht: (2025)
RGB to Hyperspectral: Spectral Reconstruction for Enhanced Surgical Imaging
von: Czempiel, Tobias, et al.
Veröffentlicht: (2024)
von: Czempiel, Tobias, et al.
Veröffentlicht: (2024)
SURGIVID: Annotation-Efficient Surgical Video Object Discovery
von: Köksal, Çağhan, et al.
Veröffentlicht: (2024)
von: Köksal, Çağhan, et al.
Veröffentlicht: (2024)
When to Trust the Answer: Question-Aligned Semantic Nearest Neighbor Entropy for Safer Surgical VQA
von: Carlini, Luca, et al.
Veröffentlicht: (2025)
von: Carlini, Luca, et al.
Veröffentlicht: (2025)
Freeview Sketching: View-Aware Fine-Grained Sketch-Based Image Retrieval
von: Sain, Aneeshan, et al.
Veröffentlicht: (2024)
von: Sain, Aneeshan, et al.
Veröffentlicht: (2024)
SurgiTrack: Fine-Grained Multi-Class Multi-Tool Tracking in Surgical Videos
von: Nwoye, Chinedu Innocent, et al.
Veröffentlicht: (2024)
von: Nwoye, Chinedu Innocent, et al.
Veröffentlicht: (2024)
Training-free Video Temporal Grounding using Large-scale Pre-trained Models
von: Zheng, Minghang, et al.
Veröffentlicht: (2024)
von: Zheng, Minghang, et al.
Veröffentlicht: (2024)
SMTrack: End-to-End Trained Spiking Neural Networks for Multi-Object Tracking in RGB Videos
von: Zhong, Pengzhi, et al.
Veröffentlicht: (2025)
von: Zhong, Pengzhi, et al.
Veröffentlicht: (2025)
Point Tracking in Surgery--The 2024 Surgical Tattoos in Infrared (STIR) Challenge
von: Schmidt, Adam, et al.
Veröffentlicht: (2025)
von: Schmidt, Adam, et al.
Veröffentlicht: (2025)
HUP-3D: A 3D multi-view synthetic dataset for assisted-egocentric hand-ultrasound pose estimation
von: Birlo, Manuel, et al.
Veröffentlicht: (2024)
von: Birlo, Manuel, et al.
Veröffentlicht: (2024)
Mismatched: Evaluating the Limits of Image Matching Approaches and Benchmarks
von: Bonilla, Sierra, et al.
Veröffentlicht: (2024)
von: Bonilla, Sierra, et al.
Veröffentlicht: (2024)
TQD-Track: Temporal Query Denoising for 3D Multi-Object Tracking
von: Ding, Shuxiao, et al.
Veröffentlicht: (2025)
von: Ding, Shuxiao, et al.
Veröffentlicht: (2025)
Future Slot Prediction for Unsupervised Object Discovery in Surgical Video
von: Liao, Guiqiu, et al.
Veröffentlicht: (2025)
von: Liao, Guiqiu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Learning from Single Timestamps: Complexity Estimation in Laparoscopic Cholecystectomy
von: Anastasiou, Dimitrios, et al.
Veröffentlicht: (2025) -
Graph Neural Networks for Surgical Scene Segmentation
von: Li, Yihan, et al.
Veröffentlicht: (2025) -
Temporal Cluster Assignment for Efficient Real-Time Video Segmentation
von: Yung, Ka-Wai, et al.
Veröffentlicht: (2025) -
Depth Augmented and FE Free 3D/2D Liver Registration for Laparoscopic Liver AR
von: Zhang, Hanyuan, et al.
Veröffentlicht: (2026) -
Warm-Started Reinforcement Learning for Iterative 3D/2D Liver Registration
von: Zhang, Hanyuan, et al.
Veröffentlicht: (2026)