Decoding the Surgical Scene: A Scoping Review of Scene Graphs in Surgery
Fuente:
arXiv
Guardado en:
| Autores principales: | Henriques, Angelo, Hoxha, Korab, Zapp, Daniel, Issa, Peter C., Navab, Nassir, Nasseri, M. Ali |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
por: Raoufi, Behnam, et al.
Publicado: (2025)
por: Raoufi, Behnam, et al.
Publicado: (2025)
From eye to AI: studying rodent social behavior in the era of machine Learning
por: Chindemi, Giuseppe, et al.
Publicado: (2025)
por: Chindemi, Giuseppe, et al.
Publicado: (2025)
vS-Graphs: Tightly Coupling Visual SLAM and 3D Scene Graphs Exploiting Hierarchical Scene Understanding
por: Tourani, Ali, et al.
Publicado: (2025)
por: Tourani, Ali, et al.
Publicado: (2025)
A Reverse Causal Framework to Mitigate Spurious Correlations for Debiasing Scene Graph Generation
por: Sun, Shuzhou, et al.
Publicado: (2025)
por: Sun, Shuzhou, et al.
Publicado: (2025)
FlowIBR: Leveraging Pre-Training for Efficient Neural Image-Based Rendering of Dynamic Scenes
por: Büsching, Marcel, et al.
Publicado: (2023)
por: Büsching, Marcel, et al.
Publicado: (2023)
SimWorld: A Unified Benchmark for Simulator-Conditioned Scene Generation via World Model
por: Li, Xinqing, et al.
Publicado: (2025)
por: Li, Xinqing, et al.
Publicado: (2025)
DualPrompt-MedCap: A Dual-Prompt Enhanced Approach for Medical Image Captioning
por: Zhao, Yining, et al.
Publicado: (2025)
por: Zhao, Yining, et al.
Publicado: (2025)
SurgicalMamba: Dual-Path SSD with State Regramming for Online Surgical Phase Recognition
por: Oh, Sukju, et al.
Publicado: (2026)
por: Oh, Sukju, et al.
Publicado: (2026)
VLM-VPI: A Vision-Language Reasoning Framework for Improving Automated Vehicle-Pedestrian Interactions
por: Pu, Qingwen, et al.
Publicado: (2026)
por: Pu, Qingwen, et al.
Publicado: (2026)
Hierarchical Image-Guided 3D Point Cloud Segmentation in Industrial Scenes via Multi-View Bayesian Fusion
por: Zhu, Yu, et al.
Publicado: (2025)
por: Zhu, Yu, et al.
Publicado: (2025)
Semi supervised GAN for smart microscopy, fast and data efficient cell cycle classification
por: Manick, Rajeev, et al.
Publicado: (2026)
por: Manick, Rajeev, et al.
Publicado: (2026)
Lifelong Learning in Vision-Language Models: Enhanced EWC with Cross-Modal Knowledge Retention
por: Durrani, Hamza Ahmed, et al.
Publicado: (2026)
por: Durrani, Hamza Ahmed, et al.
Publicado: (2026)
4D Synchronized Fields: Motion-Language Gaussian Splatting for Temporal Scene Understanding
por: Barhdadi, Mohamed Rayan, et al.
Publicado: (2026)
por: Barhdadi, Mohamed Rayan, et al.
Publicado: (2026)
Pedestrian Detection in Low-Light Conditions: A Comprehensive Survey
por: Ghari, Bahareh, et al.
Publicado: (2024)
por: Ghari, Bahareh, et al.
Publicado: (2024)
Towards Accurate and Efficient Waste Image Classification: A Hybrid Deep Learning and Machine Learning Approach
por: Nguyen, Ngoc-Bao-Quang, et al.
Publicado: (2025)
por: Nguyen, Ngoc-Bao-Quang, et al.
Publicado: (2025)
DRIFT open dataset: A drone-derived intelligence for traffic analysis in urban environment
por: Lee, Hyejin, et al.
Publicado: (2025)
por: Lee, Hyejin, et al.
Publicado: (2025)
Scene Detection Policies and Keyframe Extraction Strategies for Large-Scale Video Analysis
por: Korolkov, Vasilii
Publicado: (2025)
por: Korolkov, Vasilii
Publicado: (2025)
OmniAcc: Personalized Accessibility Assistant Using Generative AI
por: Karki, Siddhant, et al.
Publicado: (2025)
por: Karki, Siddhant, et al.
Publicado: (2025)
ProtoFlow: Interpretable and Robust Surgical Workflow Modeling with Learned Dynamic Scene Graph Prototypes
por: Holm, Felix, et al.
Publicado: (2025)
por: Holm, Felix, et al.
Publicado: (2025)
StemVLA:An Open-Source Vision-Language-Action Model with Future 3D Spatial Geometry Knowledge and 4D Historical Representation
por: Xiao, Jiasong, et al.
Publicado: (2026)
por: Xiao, Jiasong, et al.
Publicado: (2026)
Introspection in Learned Semantic Scene Graph Localisation
por: Bissessur, Manshika Charvi, et al.
Publicado: (2025)
por: Bissessur, Manshika Charvi, et al.
Publicado: (2025)
TAG-Head: Time-Aligned Graph Head for Plug-and-Play Fine-grained Action Recognition
por: Hassan, Imtiaz Ul, et al.
Publicado: (2026)
por: Hassan, Imtiaz Ul, et al.
Publicado: (2026)
Context in object detection: a systematic literature review
por: Jamali, Mahtab, et al.
Publicado: (2025)
por: Jamali, Mahtab, et al.
Publicado: (2025)
Mask-Conditioned Voxel Diffusion for Joint Geometry and Color Inpainting
por: Sumuk, Aarya
Publicado: (2026)
por: Sumuk, Aarya
Publicado: (2026)
PhysVideoGenerator: Towards Physically Aware Video Generation via Latent Physics Guidance
por: Satish, Siddarth Nilol Kundur, et al.
Publicado: (2026)
por: Satish, Siddarth Nilol Kundur, et al.
Publicado: (2026)
IMASHRIMP: Automatic White Shrimp (Penaeus vannamei) Biometrical Analysis from Laboratory Images Using Computer Vision and Deep Learning
por: González, Abiam Remache, et al.
Publicado: (2025)
por: González, Abiam Remache, et al.
Publicado: (2025)
OCC-MLLM-CoT-Alpha: Towards Multi-stage Occlusion Recognition Based on Large Language Models via 3D-Aware Supervision and Chain-of-Thoughts Guidance
por: Wang, Chaoyi, et al.
Publicado: (2025)
por: Wang, Chaoyi, et al.
Publicado: (2025)
NOAH: Benchmarking Narrative Prior driven Hallucination and Omission in Video Large Language Models
por: Lee, Kyuho, et al.
Publicado: (2025)
por: Lee, Kyuho, et al.
Publicado: (2025)
Reducing Object Hallucination in LVLMs via Emphasizing Image-negative Tokens
por: Shen, Meng, et al.
Publicado: (2026)
por: Shen, Meng, et al.
Publicado: (2026)
Towards a Generalizable Fusion Architecture for Multimodal Object Detection
por: Berjawi, Jad, et al.
Publicado: (2025)
por: Berjawi, Jad, et al.
Publicado: (2025)
A Simple Baseline for Streaming Video Understanding
por: Shen, Yujiao, et al.
Publicado: (2026)
por: Shen, Yujiao, et al.
Publicado: (2026)
EmoVerse: A MLLMs-Driven Emotion Representation Dataset for Interpretable Visual Emotion Analysis
por: Guo, Yijie, et al.
Publicado: (2025)
por: Guo, Yijie, et al.
Publicado: (2025)
SCA-Net: Spatial-Contextual Aggregation Network for Enhanced Small Building and Road Change Detection
por: Gholibeigi, Emad, et al.
Publicado: (2026)
por: Gholibeigi, Emad, et al.
Publicado: (2026)
Synthetic-Child: An AIGC-Based Synthetic Data Pipeline for Privacy-Preserving Child Posture Estimation
por: Zeng, Taowen
Publicado: (2026)
por: Zeng, Taowen
Publicado: (2026)
Exploring Surround-View Fisheye Camera 3D Object Detection
por: Li, Changcai, et al.
Publicado: (2025)
por: Li, Changcai, et al.
Publicado: (2025)
A Vision-Language Model for Focal Liver Lesion Classification
por: Jian, Song, et al.
Publicado: (2025)
por: Jian, Song, et al.
Publicado: (2025)
Car Object Counting and Position Estimation via Extension of the CLIP-EBC Framework
por: Jung, Seoik, et al.
Publicado: (2025)
por: Jung, Seoik, et al.
Publicado: (2025)
Pixel-Level Pavement Distress Assessment Using Instance Segmentation
por: Dewick, Logan, et al.
Publicado: (2026)
por: Dewick, Logan, et al.
Publicado: (2026)
Vi-SAFE: A Spatial-Temporal Framework for Efficient Violence Detection in Public Surveillance
por: Chang, Ligang, et al.
Publicado: (2025)
por: Chang, Ligang, et al.
Publicado: (2025)
Action Anticipation from SoccerNet Football Video Broadcasts
por: Dalal, Mohamad, et al.
Publicado: (2025)
por: Dalal, Mohamad, et al.
Publicado: (2025)
Ejemplares similares
-
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
por: Raoufi, Behnam, et al.
Publicado: (2025) -
From eye to AI: studying rodent social behavior in the era of machine Learning
por: Chindemi, Giuseppe, et al.
Publicado: (2025) -
vS-Graphs: Tightly Coupling Visual SLAM and 3D Scene Graphs Exploiting Hierarchical Scene Understanding
por: Tourani, Ali, et al.
Publicado: (2025) -
A Reverse Causal Framework to Mitigate Spurious Correlations for Debiasing Scene Graph Generation
por: Sun, Shuzhou, et al.
Publicado: (2025) -
FlowIBR: Leveraging Pre-Training for Efficient Neural Image-Based Rendering of Dynamic Scenes
por: Büsching, Marcel, et al.
Publicado: (2023)