Specialized Foundation Models for Intelligent Operating Rooms
Fuente:
arXiv
Saved in:
| Main Authors: | Özsoy, Ege, Pellegrini, Chantal, Bani-Harouni, David, Yuan, Kun, Keicher, Matthias, Navab, Nassir |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ORacle: Large Vision-Language Models for Knowledge-Guided Holistic OR Domain Modeling
by: Özsoy, Ege, et al.
Published: (2024)
by: Özsoy, Ege, et al.
Published: (2024)
MM-OR: A Large Multimodal Operating Room Dataset for Semantic Understanding of High-Intensity Surgical Environments
by: Özsoy, Ege, et al.
Published: (2025)
by: Özsoy, Ege, et al.
Published: (2025)
RaDialog: A Large Vision-Language Model for Radiology Report Generation and Conversational Assistance
by: Pellegrini, Chantal, et al.
Published: (2023)
by: Pellegrini, Chantal, et al.
Published: (2023)
Prototype-Based Knowledge Guidance for Fine-Grained Structured Radiology Reporting
by: Pellegrini, Chantal, et al.
Published: (2026)
by: Pellegrini, Chantal, et al.
Published: (2026)
EHR2Path: Scalable Modeling of Longitudinal Patient Pathways from Multimodal Electronic Health Records
by: Pellegrini, Chantal, et al.
Published: (2025)
by: Pellegrini, Chantal, et al.
Published: (2025)
Language Agents for Hypothesis-driven Clinical Decision Making with Reinforcement Learning
by: Bani-Harouni, David, et al.
Published: (2025)
by: Bani-Harouni, David, et al.
Published: (2025)
EgoExOR: An Ego-Exo-Centric Operating Room Dataset for Surgical Activity Understanding
by: Özsoy, Ege, et al.
Published: (2025)
by: Özsoy, Ege, et al.
Published: (2025)
PanORama: Multiview Consistent Panoptic Segmentation in Operating Rooms
by: Gürbüz, Tuna, et al.
Published: (2026)
by: Gürbüz, Tuna, et al.
Published: (2026)
Rewarding Doubt: A Reinforcement Learning Approach to Calibrated Confidence Expression of Large Language Models
by: Bani-Harouni, David, et al.
Published: (2025)
by: Bani-Harouni, David, et al.
Published: (2025)
Location-Free Scene Graph Generation
by: Özsoy, Ege, et al.
Published: (2023)
by: Özsoy, Ege, et al.
Published: (2023)
MAGDA: Multi-agent guideline-driven diagnostic assistance
by: Bani-Harouni, David, et al.
Published: (2024)
by: Bani-Harouni, David, et al.
Published: (2024)
TrackOR: Towards Personalized Intelligent Operating Rooms Through Robust Tracking
by: Wang, Tony Danjun, et al.
Published: (2025)
by: Wang, Tony Danjun, et al.
Published: (2025)
Mitigating Biases in Surgical Operating Rooms with Geometry
by: Wang, Tony Danjun, et al.
Published: (2025)
by: Wang, Tony Danjun, et al.
Published: (2025)
TopoOR: A Unified Topological Scene Representation for the Operating Room
by: Wang, Tony Danjun, et al.
Published: (2026)
by: Wang, Tony Danjun, et al.
Published: (2026)
Text-driven Adaptation of Foundation Models for Few-shot Surgical Workflow Analysis
by: Chen, Tingxuan, et al.
Published: (2025)
by: Chen, Tingxuan, et al.
Published: (2025)
Beyond Role-Based Surgical Domain Modeling: Generalizable Re-Identification in the Operating Room
by: Wang, Tony Danjun, et al.
Published: (2025)
by: Wang, Tony Danjun, et al.
Published: (2025)
LapFM: A Laparoscopic Segmentation Foundation Model via Hierarchical Concept Evolving Pre-training
by: Xu, Qing, et al.
Published: (2025)
by: Xu, Qing, et al.
Published: (2025)
Recognizing Surgical Phases Anywhere: Few-Shot Test-time Adaptation and Task-graph Guided Refinement
by: Yuan, Kun, et al.
Published: (2025)
by: Yuan, Kun, et al.
Published: (2025)
HecVL: Hierarchical Video-Language Pretraining for Zero-shot Surgical Phase Recognition
by: Yuan, Kun, et al.
Published: (2024)
by: Yuan, Kun, et al.
Published: (2024)
Procedure-Aware Surgical Video-language Pretraining with Hierarchical Knowledge Augmentation
by: Yuan, Kun, et al.
Published: (2024)
by: Yuan, Kun, et al.
Published: (2024)
Counterfactual Explanations for Medical Image Classification and Regression using Diffusion Autoencoder
by: Atad, Matan, et al.
Published: (2024)
by: Atad, Matan, et al.
Published: (2024)
Learning Diagnostic Reasoning for Decision Support in Toxicology
by: Oberländer, Nico, et al.
Published: (2026)
by: Oberländer, Nico, et al.
Published: (2026)
HieraSurg: Hierarchy-Aware Diffusion Model for Surgical Video Generation
by: Biagini, Diego, et al.
Published: (2025)
by: Biagini, Diego, et al.
Published: (2025)
Geometry OR Tracker: Universal Geometric Operating Room Tracking
by: Shao, Yihua, et al.
Published: (2026)
by: Shao, Yihua, et al.
Published: (2026)
SurgOnAir: Hierarchy-Aware Real-Time Surgical Video Commentary
by: He, Jingyi, et al.
Published: (2026)
by: He, Jingyi, et al.
Published: (2026)
Calibrated Confidence Expression for Radiology Report Generation
by: Bani-Harouni, David, et al.
Published: (2026)
by: Bani-Harouni, David, et al.
Published: (2026)
SURGIVID: Annotation-Efficient Surgical Video Object Discovery
by: Köksal, Çağhan, et al.
Published: (2024)
by: Köksal, Çağhan, et al.
Published: (2024)
From Linear Probing to Joint-Weighted Token Hierarchy: A Foundation Model Bridging Global and Cellular Representations in Biomarker Detection
by: Liu, Jingsong, et al.
Published: (2025)
by: Liu, Jingsong, et al.
Published: (2025)
Neural Semantic Map-Learning for Autonomous Vehicles
by: Herb, Markus, et al.
Published: (2024)
by: Herb, Markus, et al.
Published: (2024)
RaNeuS: Ray-adaptive Neural Surface Reconstruction
by: Wang, Yida, et al.
Published: (2024)
by: Wang, Yida, et al.
Published: (2024)
Language-Guided Open-World Anomaly Segmentation
by: Reichard, Klara, et al.
Published: (2025)
by: Reichard, Klara, et al.
Published: (2025)
Forecasting Continuous Non-Conservative Dynamical Systems in SO(3)
by: Bastian, Lennart, et al.
Published: (2025)
by: Bastian, Lennart, et al.
Published: (2025)
MS-CLR: Multi-Skeleton Contrastive Learning for Human Action Recognition
by: Kiray, Mert, et al.
Published: (2025)
by: Kiray, Mert, et al.
Published: (2025)
A Taxonomy and Library for Visualizing Learned Features in Convolutional Neural Networks
by: Grün, Felix, et al.
Published: (2016)
by: Grün, Felix, et al.
Published: (2016)
Deep Spectral Methods for Unsupervised Ultrasound Image Interpretation
by: Tmenova, Oleksandra, et al.
Published: (2024)
by: Tmenova, Oleksandra, et al.
Published: (2024)
Hybrid Functional Maps for Crease-Aware Non-Isometric Shape Matching
by: Bastian, Lennart, et al.
Published: (2023)
by: Bastian, Lennart, et al.
Published: (2023)
Towards Comprehensive Real-Time Scene Understanding in Ophthalmic Surgery through Multimodal Image Fusion
by: Rohrmoser, Nikolo, et al.
Published: (2026)
by: Rohrmoser, Nikolo, et al.
Published: (2026)
ESCAPE: Equivariant Shape Completion via Anchor Point Encoding
by: Bekci, Burak, et al.
Published: (2024)
by: Bekci, Burak, et al.
Published: (2024)
Temporal Differential Fields for 4D Motion Modeling via Image-to-Video Synthesis
by: You, Xin, et al.
Published: (2025)
by: You, Xin, et al.
Published: (2025)
Visual Autoregressive Modelling for Monocular Depth Estimation
by: El-Ghoussani, Amir, et al.
Published: (2025)
by: El-Ghoussani, Amir, et al.
Published: (2025)
Similar Items
-
ORacle: Large Vision-Language Models for Knowledge-Guided Holistic OR Domain Modeling
by: Özsoy, Ege, et al.
Published: (2024) -
MM-OR: A Large Multimodal Operating Room Dataset for Semantic Understanding of High-Intensity Surgical Environments
by: Özsoy, Ege, et al.
Published: (2025) -
RaDialog: A Large Vision-Language Model for Radiology Report Generation and Conversational Assistance
by: Pellegrini, Chantal, et al.
Published: (2023) -
Prototype-Based Knowledge Guidance for Fine-Grained Structured Radiology Reporting
by: Pellegrini, Chantal, et al.
Published: (2026) -
EHR2Path: Scalable Modeling of Longitudinal Patient Pathways from Multimodal Electronic Health Records
by: Pellegrini, Chantal, et al.
Published: (2025)