SURGIVID: Annotation-Efficient Surgical Video Object Discovery
Fuente:
arXiv
Salvato in:
| Autori principali: | Köksal, Çağhan, Ghazaei, Ghazal, Navab, Nassir |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SANGRIA: Surgical Video Scene Graph Optimization for Surgical Workflow Prediction
di: Köksal, Çağhan, et al.
Pubblicazione: (2024)
di: Köksal, Çağhan, et al.
Pubblicazione: (2024)
ProtoFlow: Interpretable and Robust Surgical Workflow Modeling with Learned Dynamic Scene Graph Prototypes
di: Holm, Felix, et al.
Pubblicazione: (2025)
di: Holm, Felix, et al.
Pubblicazione: (2025)
Towards Comprehensive Real-Time Scene Understanding in Ophthalmic Surgery through Multimodal Image Fusion
di: Rohrmoser, Nikolo, et al.
Pubblicazione: (2026)
di: Rohrmoser, Nikolo, et al.
Pubblicazione: (2026)
CAT-SG: A Large Dynamic Scene Graph Dataset for Fine-Grained Understanding of Cataract Surgery
di: Holm, Felix, et al.
Pubblicazione: (2025)
di: Holm, Felix, et al.
Pubblicazione: (2025)
SurGrID: Controllable Surgical Simulation via Scene Graph to Image Diffusion
di: Frisch, Yannik, et al.
Pubblicazione: (2025)
di: Frisch, Yannik, et al.
Pubblicazione: (2025)
Watch and Learn: Leveraging Expert Knowledge and Language for Surgical Video Understanding
di: Gastager, David, et al.
Pubblicazione: (2025)
di: Gastager, David, et al.
Pubblicazione: (2025)
HieraSurg: Hierarchy-Aware Diffusion Model for Surgical Video Generation
di: Biagini, Diego, et al.
Pubblicazione: (2025)
di: Biagini, Diego, et al.
Pubblicazione: (2025)
SG2VID: Scene Graphs Enable Fine-Grained Control for Video Synthesis
di: Sivakumar, Ssharvien Kumar, et al.
Pubblicazione: (2025)
di: Sivakumar, Ssharvien Kumar, et al.
Pubblicazione: (2025)
Procedure-Aware Surgical Video-language Pretraining with Hierarchical Knowledge Augmentation
di: Yuan, Kun, et al.
Pubblicazione: (2024)
di: Yuan, Kun, et al.
Pubblicazione: (2024)
HecVL: Hierarchical Video-Language Pretraining for Zero-shot Surgical Phase Recognition
di: Yuan, Kun, et al.
Pubblicazione: (2024)
di: Yuan, Kun, et al.
Pubblicazione: (2024)
SurgOnAir: Hierarchy-Aware Real-Time Surgical Video Commentary
di: He, Jingyi, et al.
Pubblicazione: (2026)
di: He, Jingyi, et al.
Pubblicazione: (2026)
CliPPER: Contextual Video-Language Pretraining on Long-form Intraoperative Surgical Procedures for Event Recognition
di: Stilz, Florian, et al.
Pubblicazione: (2026)
di: Stilz, Florian, et al.
Pubblicazione: (2026)
Mitigating Biases in Surgical Operating Rooms with Geometry
di: Wang, Tony Danjun, et al.
Pubblicazione: (2025)
di: Wang, Tony Danjun, et al.
Pubblicazione: (2025)
Efficient Remote Sensing Change Detection with Change State Space Models
di: Ghazaei, Elman, et al.
Pubblicazione: (2025)
di: Ghazaei, Elman, et al.
Pubblicazione: (2025)
Object Pose Transformer: Unifying Unseen Object Pose Estimation
di: Li, Weihang, et al.
Pubblicazione: (2026)
di: Li, Weihang, et al.
Pubblicazione: (2026)
Future Slot Prediction for Unsupervised Object Discovery in Surgical Video
di: Liao, Guiqiu, et al.
Pubblicazione: (2025)
di: Liao, Guiqiu, et al.
Pubblicazione: (2025)
BridgeSplat: Bidirectionally Coupled CT and Non-Rigid Gaussian Splatting for Deformable Intraoperative Surgical Navigation
di: Fehrentz, Maximilian, et al.
Pubblicazione: (2025)
di: Fehrentz, Maximilian, et al.
Pubblicazione: (2025)
Text-driven Adaptation of Foundation Models for Few-shot Surgical Workflow Analysis
di: Chen, Tingxuan, et al.
Pubblicazione: (2025)
di: Chen, Tingxuan, et al.
Pubblicazione: (2025)
Learning Multi-modal Representations by Watching Hundreds of Surgical Video Lectures
di: Yuan, Kun, et al.
Pubblicazione: (2023)
di: Yuan, Kun, et al.
Pubblicazione: (2023)
Temporal Differential Fields for 4D Motion Modeling via Image-to-Video Synthesis
di: You, Xin, et al.
Pubblicazione: (2025)
di: You, Xin, et al.
Pubblicazione: (2025)
Text-conditioned State Space Model For Domain-generalized Change Detection Visual Question Answering
di: Ghazaei, Elman, et al.
Pubblicazione: (2025)
di: Ghazaei, Elman, et al.
Pubblicazione: (2025)
EgoExOR: An Ego-Exo-Centric Operating Room Dataset for Surgical Activity Understanding
di: Özsoy, Ege, et al.
Pubblicazione: (2025)
di: Özsoy, Ege, et al.
Pubblicazione: (2025)
Neural Semantic Map-Learning for Autonomous Vehicles
di: Herb, Markus, et al.
Pubblicazione: (2024)
di: Herb, Markus, et al.
Pubblicazione: (2024)
Beyond Role-Based Surgical Domain Modeling: Generalizable Re-Identification in the Operating Room
di: Wang, Tony Danjun, et al.
Pubblicazione: (2025)
di: Wang, Tony Danjun, et al.
Pubblicazione: (2025)
Advancing Surgical VQA with Scene Graph Knowledge
di: Yuan, Kun, et al.
Pubblicazione: (2023)
di: Yuan, Kun, et al.
Pubblicazione: (2023)
MatchU: Matching Unseen Objects for 6D Pose Estimation from RGB-D Images
di: Huang, Junwen, et al.
Pubblicazione: (2024)
di: Huang, Junwen, et al.
Pubblicazione: (2024)
Next-generation Surgical Navigation: Marker-less Multi-view 6DoF Pose Estimation of Surgical Instruments
di: Hein, Jonas, et al.
Pubblicazione: (2023)
di: Hein, Jonas, et al.
Pubblicazione: (2023)
Deep Spectral Methods for Unsupervised Ultrasound Image Interpretation
di: Tmenova, Oleksandra, et al.
Pubblicazione: (2024)
di: Tmenova, Oleksandra, et al.
Pubblicazione: (2024)
ORacle: Large Vision-Language Models for Knowledge-Guided Holistic OR Domain Modeling
di: Özsoy, Ege, et al.
Pubblicazione: (2024)
di: Özsoy, Ege, et al.
Pubblicazione: (2024)
ESCAPE: Equivariant Shape Completion via Anchor Point Encoding
di: Bekci, Burak, et al.
Pubblicazione: (2024)
di: Bekci, Burak, et al.
Pubblicazione: (2024)
A Taxonomy and Library for Visualizing Learned Features in Convolutional Neural Networks
di: Grün, Felix, et al.
Pubblicazione: (2016)
di: Grün, Felix, et al.
Pubblicazione: (2016)
Hybrid Functional Maps for Crease-Aware Non-Isometric Shape Matching
di: Bastian, Lennart, et al.
Pubblicazione: (2023)
di: Bastian, Lennart, et al.
Pubblicazione: (2023)
Language-Guided Open-World Anomaly Segmentation
di: Reichard, Klara, et al.
Pubblicazione: (2025)
di: Reichard, Klara, et al.
Pubblicazione: (2025)
Forecasting Continuous Non-Conservative Dynamical Systems in SO(3)
di: Bastian, Lennart, et al.
Pubblicazione: (2025)
di: Bastian, Lennart, et al.
Pubblicazione: (2025)
MS-CLR: Multi-Skeleton Contrastive Learning for Human Action Recognition
di: Kiray, Mert, et al.
Pubblicazione: (2025)
di: Kiray, Mert, et al.
Pubblicazione: (2025)
RayPose: Ray Bundling Diffusion for Template Views in Unseen 6D Object Pose Estimation
di: Huang, Junwen, et al.
Pubblicazione: (2025)
di: Huang, Junwen, et al.
Pubblicazione: (2025)
How Far Are Surgeons from Surgical World Models? A Pilot Study on Zero-shot Surgical Video Generation with Expert Assessment
di: Chen, Zhen, et al.
Pubblicazione: (2025)
di: Chen, Zhen, et al.
Pubblicazione: (2025)
GCE-Pose: Global Context Enhancement for Category-level Object Pose Estimation
di: Li, Weihang, et al.
Pubblicazione: (2025)
di: Li, Weihang, et al.
Pubblicazione: (2025)
Slot-BERT: Self-supervised Object Discovery in Surgical Video
di: Liao, Guiqiu, et al.
Pubblicazione: (2025)
di: Liao, Guiqiu, et al.
Pubblicazione: (2025)
SurgMotion: A Video-Native Foundation Model for Universal Understanding of Surgical Videos
di: Wu, Jinlin, et al.
Pubblicazione: (2026)
di: Wu, Jinlin, et al.
Pubblicazione: (2026)
Documenti analoghi
-
SANGRIA: Surgical Video Scene Graph Optimization for Surgical Workflow Prediction
di: Köksal, Çağhan, et al.
Pubblicazione: (2024) -
ProtoFlow: Interpretable and Robust Surgical Workflow Modeling with Learned Dynamic Scene Graph Prototypes
di: Holm, Felix, et al.
Pubblicazione: (2025) -
Towards Comprehensive Real-Time Scene Understanding in Ophthalmic Surgery through Multimodal Image Fusion
di: Rohrmoser, Nikolo, et al.
Pubblicazione: (2026) -
CAT-SG: A Large Dynamic Scene Graph Dataset for Fine-Grained Understanding of Cataract Surgery
di: Holm, Felix, et al.
Pubblicazione: (2025) -
SurGrID: Controllable Surgical Simulation via Scene Graph to Image Diffusion
di: Frisch, Yannik, et al.
Pubblicazione: (2025)