Training-free Temporal Object Tracking in Surgical Videos
Fuente:
arXiv
Guardado en:
| Autores principales: | Koley, Subhadeep, Kadkhodamohammadi, Abdolrahim, Barbarisi, Santiago, Stoyanov, Danail, Luengo, Imanol |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Learning from Single Timestamps: Complexity Estimation in Laparoscopic Cholecystectomy
por: Anastasiou, Dimitrios, et al.
Publicado: (2025)
por: Anastasiou, Dimitrios, et al.
Publicado: (2025)
Graph Neural Networks for Surgical Scene Segmentation
por: Li, Yihan, et al.
Publicado: (2025)
por: Li, Yihan, et al.
Publicado: (2025)
Temporal Cluster Assignment for Efficient Real-Time Video Segmentation
por: Yung, Ka-Wai, et al.
Publicado: (2025)
por: Yung, Ka-Wai, et al.
Publicado: (2025)
Depth Augmented and FE Free 3D/2D Liver Registration for Laparoscopic Liver AR
por: Zhang, Hanyuan, et al.
Publicado: (2026)
por: Zhang, Hanyuan, et al.
Publicado: (2026)
Warm-Started Reinforcement Learning for Iterative 3D/2D Liver Registration
por: Zhang, Hanyuan, et al.
Publicado: (2026)
por: Zhang, Hanyuan, et al.
Publicado: (2026)
TemporalDoRA: Temporal PEFT for Robust Surgical Video Question Answering
por: Carlini, Luca, et al.
Publicado: (2026)
por: Carlini, Luca, et al.
Publicado: (2026)
Zero-shot Monocular Metric Depth for Endoscopic Images
por: Toussaint, Nicolas, et al.
Publicado: (2025)
por: Toussaint, Nicolas, et al.
Publicado: (2025)
SurgViVQA: Temporally-Grounded Video Question Answering for Surgical Scene Understanding
por: Drago, Mauro Orazio, et al.
Publicado: (2025)
por: Drago, Mauro Orazio, et al.
Publicado: (2025)
Think Step by Step: Chain-of-Gesture Prompting for Error Detection in Robotic Surgical Videos
por: Shao, Zhimin, et al.
Publicado: (2024)
por: Shao, Zhimin, et al.
Publicado: (2024)
Diff2DGS: Reliable Reconstruction of Occluded Surgical Scenes via 2D Gaussian Splatting
por: Song, Tianyi, et al.
Publicado: (2026)
por: Song, Tianyi, et al.
Publicado: (2026)
SHADeS: Self-supervised Monocular Depth Estimation Through Non-Lambertian Image Decomposition
por: Daher, Rema, et al.
Publicado: (2025)
por: Daher, Rema, et al.
Publicado: (2025)
SurgicalGS: Dynamic 3D Gaussian Splatting for Accurate Robotic-Assisted Surgical Scene Reconstruction
por: Chen, Jialei, et al.
Publicado: (2024)
por: Chen, Jialei, et al.
Publicado: (2024)
Confidence-aware Monocular Depth Estimation for Minimally Invasive Surgery
por: Asad, Muhammad, et al.
Publicado: (2026)
por: Asad, Muhammad, et al.
Publicado: (2026)
DreamColour: Controllable Video Colour Editing without Training
por: Utintu, Chaitat, et al.
Publicado: (2024)
por: Utintu, Chaitat, et al.
Publicado: (2024)
Tracking Everything in Robotic-Assisted Surgery
por: Zhan, Bohan, et al.
Publicado: (2024)
por: Zhan, Bohan, et al.
Publicado: (2024)
Temporally Guided Articulated Hand Pose Tracking in Surgical Videos
por: Louis, Nathan, et al.
Publicado: (2021)
por: Louis, Nathan, et al.
Publicado: (2021)
SurgAnt-ViVQA: Learning to Anticipate Surgical Events through GRU-Driven Temporal Cross-Attention
por: Dhake, Shreyas C., et al.
Publicado: (2025)
por: Dhake, Shreyas C., et al.
Publicado: (2025)
Surgical AI Copilot: Energy-Based Fourier Gradient Low-Rank Adaptation for Surgical LLM Agent Reasoning and Planning
por: Huang, Jiayuan, et al.
Publicado: (2025)
por: Huang, Jiayuan, et al.
Publicado: (2025)
Multi-Modal Monocular Endoscopic Depth and Pose Estimation with Edge-Guided Self-Supervision
por: Ju, Xinwei, et al.
Publicado: (2026)
por: Ju, Xinwei, et al.
Publicado: (2026)
CoRe-DA: Contrastive Regression for Unsupervised Domain Adaptation in Surgical Skill Assessment
por: Anastasiou, Dimitrios, et al.
Publicado: (2026)
por: Anastasiou, Dimitrios, et al.
Publicado: (2026)
SEDMamba: Enhancing Selective State Space Modelling with Bottleneck Mechanism and Fine-to-Coarse Temporal Fusion for Efficient Error Detection in Robot-Assisted Surgery
por: Xu, Jialang, et al.
Publicado: (2024)
por: Xu, Jialang, et al.
Publicado: (2024)
EndoLRMGS: Complete Endoscopic Scene Reconstruction combining Large Reconstruction Modelling and Gaussian Splatting
por: Wang, Xu, et al.
Publicado: (2025)
por: Wang, Xu, et al.
Publicado: (2025)
TSMS-SAM2: Multi-scale Temporal Sampling Augmentation and Memory-Splitting Pruning for Promptable Video Object Segmentation and Tracking in Surgical Scenarios
por: Xu, Guoping, et al.
Publicado: (2025)
por: Xu, Guoping, et al.
Publicado: (2025)
SegCol Challenge: Semantic Segmentation for Tools and Fold Edges in Colonoscopy data
por: Ju, Xinwei, et al.
Publicado: (2024)
por: Ju, Xinwei, et al.
Publicado: (2024)
High-fidelity Endoscopic Image Synthesis by Utilizing Depth-guided Neural Surfaces
por: Huang, Baoru, et al.
Publicado: (2024)
por: Huang, Baoru, et al.
Publicado: (2024)
Gaussian Pancakes: Geometrically-Regularized 3D Gaussian Splatting for Realistic Endoscopic Reconstruction
por: Bonilla, Sierra, et al.
Publicado: (2024)
por: Bonilla, Sierra, et al.
Publicado: (2024)
Multimodal Optimal Transport for Training-free Temporal Segmentation in Surgical Robotics
por: Mohamed, Omar, et al.
Publicado: (2026)
por: Mohamed, Omar, et al.
Publicado: (2026)
MOVi: Training-free Text-conditioned Multi-Object Video Generation
por: Rahman, Aimon, et al.
Publicado: (2025)
por: Rahman, Aimon, et al.
Publicado: (2025)
RGB to Hyperspectral: Spectral Reconstruction for Enhanced Surgical Imaging
por: Czempiel, Tobias, et al.
Publicado: (2024)
por: Czempiel, Tobias, et al.
Publicado: (2024)
SURGIVID: Annotation-Efficient Surgical Video Object Discovery
por: Köksal, Çağhan, et al.
Publicado: (2024)
por: Köksal, Çağhan, et al.
Publicado: (2024)
When to Trust the Answer: Question-Aligned Semantic Nearest Neighbor Entropy for Safer Surgical VQA
por: Carlini, Luca, et al.
Publicado: (2025)
por: Carlini, Luca, et al.
Publicado: (2025)
Freeview Sketching: View-Aware Fine-Grained Sketch-Based Image Retrieval
por: Sain, Aneeshan, et al.
Publicado: (2024)
por: Sain, Aneeshan, et al.
Publicado: (2024)
SurgiTrack: Fine-Grained Multi-Class Multi-Tool Tracking in Surgical Videos
por: Nwoye, Chinedu Innocent, et al.
Publicado: (2024)
por: Nwoye, Chinedu Innocent, et al.
Publicado: (2024)
Training-free Video Temporal Grounding using Large-scale Pre-trained Models
por: Zheng, Minghang, et al.
Publicado: (2024)
por: Zheng, Minghang, et al.
Publicado: (2024)
SMTrack: End-to-End Trained Spiking Neural Networks for Multi-Object Tracking in RGB Videos
por: Zhong, Pengzhi, et al.
Publicado: (2025)
por: Zhong, Pengzhi, et al.
Publicado: (2025)
Point Tracking in Surgery--The 2024 Surgical Tattoos in Infrared (STIR) Challenge
por: Schmidt, Adam, et al.
Publicado: (2025)
por: Schmidt, Adam, et al.
Publicado: (2025)
HUP-3D: A 3D multi-view synthetic dataset for assisted-egocentric hand-ultrasound pose estimation
por: Birlo, Manuel, et al.
Publicado: (2024)
por: Birlo, Manuel, et al.
Publicado: (2024)
Mismatched: Evaluating the Limits of Image Matching Approaches and Benchmarks
por: Bonilla, Sierra, et al.
Publicado: (2024)
por: Bonilla, Sierra, et al.
Publicado: (2024)
TQD-Track: Temporal Query Denoising for 3D Multi-Object Tracking
por: Ding, Shuxiao, et al.
Publicado: (2025)
por: Ding, Shuxiao, et al.
Publicado: (2025)
Future Slot Prediction for Unsupervised Object Discovery in Surgical Video
por: Liao, Guiqiu, et al.
Publicado: (2025)
por: Liao, Guiqiu, et al.
Publicado: (2025)
Ejemplares similares
-
Learning from Single Timestamps: Complexity Estimation in Laparoscopic Cholecystectomy
por: Anastasiou, Dimitrios, et al.
Publicado: (2025) -
Graph Neural Networks for Surgical Scene Segmentation
por: Li, Yihan, et al.
Publicado: (2025) -
Temporal Cluster Assignment for Efficient Real-Time Video Segmentation
por: Yung, Ka-Wai, et al.
Publicado: (2025) -
Depth Augmented and FE Free 3D/2D Liver Registration for Laparoscopic Liver AR
por: Zhang, Hanyuan, et al.
Publicado: (2026) -
Warm-Started Reinforcement Learning for Iterative 3D/2D Liver Registration
por: Zhang, Hanyuan, et al.
Publicado: (2026)