Towards Controllable Video Synthesis of Routine and Rare OR Events
Fuente:
arXiv
Saved in:
| Main Authors: | Schneider, Dominik, Seenivasan, Lalithkumar, Rapuri, Sampath, Anil, Vishalroshan, Maksutova, Aiza, Shen, Yiqing, Mangulabnan, Jan Emily, Ding, Hao, Porras, Jose L., Ishii, Masaru, Unberath, Mathias |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AffordTissue: Dense Affordance Prediction for Tool-Action Specific Tissue Interaction
by: Maksutova, Aiza, et al.
Published: (2026)
by: Maksutova, Aiza, et al.
Published: (2026)
Online Reasoning Video Segmentation with Just-in-Time Digital Twins
by: Shen, Yiqing, et al.
Published: (2025)
by: Shen, Yiqing, et al.
Published: (2025)
SAW: Toward a Surgical Action World Model via Controllable and Scalable Video Generation
by: Rapuri, Sampath, et al.
Published: (2026)
by: Rapuri, Sampath, et al.
Published: (2026)
Counterfactual World Models via Digital Twin-conditioned Video Diffusion
by: Shen, Yiqing, et al.
Published: (2025)
by: Shen, Yiqing, et al.
Published: (2025)
Investigating a Policy-Based Formulation for Endoscopic Camera Pose Recovery
by: Mangulabnan, Jan Emily, et al.
Published: (2026)
by: Mangulabnan, Jan Emily, et al.
Published: (2026)
Performance and Non-adversarial Robustness of the Segment Anything Model 2 in Surgical Video Segmentation
by: Shen, Yiqing, et al.
Published: (2024)
by: Shen, Yiqing, et al.
Published: (2024)
Position: Foundation Models Need Digital Twin Representations
by: Shen, Yiqing, et al.
Published: (2025)
by: Shen, Yiqing, et al.
Published: (2025)
Promptable Counterfactual Diffusion Model for Unified Brain Tumor Segmentation and Generation with MRIs
by: Shen, Yiqing, et al.
Published: (2024)
by: Shen, Yiqing, et al.
Published: (2024)
DualVision ArthroNav: Investigating Opportunities to Enhance Localization and Reconstruction in Image-based Arthroscopy Navigation via External Cameras
by: Shu, Hongchao, et al.
Published: (2025)
by: Shu, Hongchao, et al.
Published: (2025)
Operating Room Workflow Analysis via Reasoning Segmentation over Digital Twins
by: Shen, Yiqing, et al.
Published: (2025)
by: Shen, Yiqing, et al.
Published: (2025)
FastSAM-3DSlicer: A 3D-Slicer Extension for 3D Volumetric Segment Anything Model with Uncertainty Quantification
by: Shen, Yiqing, et al.
Published: (2024)
by: Shen, Yiqing, et al.
Published: (2024)
Beyond Rigid AI: Towards Natural Human-Machine Symbiosis for Interoperative Surgical Assistance
by: Seenivasan, Lalithkumar, et al.
Published: (2025)
by: Seenivasan, Lalithkumar, et al.
Published: (2025)
Fast Reasoning Segmentation for Images and Videos
by: Shen, Yiqing, et al.
Published: (2025)
by: Shen, Yiqing, et al.
Published: (2025)
FastSAM3D: An Efficient Segment Anything Model for 3D Volumetric Medical Images
by: Shen, Yiqing, et al.
Published: (2024)
by: Shen, Yiqing, et al.
Published: (2024)
An Intrinsically Explainable Approach to Detecting Vertebral Compression Fractures in CT Scans via Neurosymbolic Modeling
by: Inigo, Blanca, et al.
Published: (2024)
by: Inigo, Blanca, et al.
Published: (2024)
TwinOR: Photorealistic Digital Twins of Dynamic Operating Rooms for Embodied AI Research
by: Zhang, Han, et al.
Published: (2025)
by: Zhang, Han, et al.
Published: (2025)
Explainable AI for Automated User-specific Feedback in Surgical Skill Acquisition
by: Gomez, Catalina, et al.
Published: (2025)
by: Gomez, Catalina, et al.
Published: (2025)
Did you just see that? Arbitrary view synthesis for egocentric replay of operating room workflows from ambient sensors
by: Zhang, Han, et al.
Published: (2025)
by: Zhang, Han, et al.
Published: (2025)
An Endoscopic Chisel: Intraoperative Imaging Carves 3D Anatomical Models
by: Mangulabnan, Jan Emily, et al.
Published: (2024)
by: Mangulabnan, Jan Emily, et al.
Published: (2024)
Constructing and Interpreting Digital Twin Representations for Visual Reasoning via Reinforcement Learning
by: Shen, Yiqing, et al.
Published: (2025)
by: Shen, Yiqing, et al.
Published: (2025)
Text-Driven Reasoning Video Editing via Reinforcement Learning on Digital Twin Representations
by: Shen, Yiqing, et al.
Published: (2025)
by: Shen, Yiqing, et al.
Published: (2025)
Investigating Robot Control Policy Learning for Autonomous X-ray-guided Spine Procedures
by: Klitzner, Florence, et al.
Published: (2025)
by: Klitzner, Florence, et al.
Published: (2025)
Humanoid Robots as First Assistants in Endoscopic Surgery
by: Cho, Sue Min, et al.
Published: (2026)
by: Cho, Sue Min, et al.
Published: (2026)
Reasoning Text-to-Video Retrieval via Digital Twin Video Representations and Large Language Models
by: Shen, Yiqing, et al.
Published: (2025)
by: Shen, Yiqing, et al.
Published: (2025)
M$^4$oE: A Foundation Model for Medical Multimodal Image Segmentation with Mixture of Experts
by: Jiang, Yufeng, et al.
Published: (2024)
by: Jiang, Yufeng, et al.
Published: (2024)
Neural Finite-State Machines for Surgical Phase Recognition
by: Ding, Hao, et al.
Published: (2024)
by: Ding, Hao, et al.
Published: (2024)
Seamless Augmented Reality Integration in Arthroscopy: A Pipeline for Articular Reconstruction and Guidance
by: Shu, Hongchao, et al.
Published: (2024)
by: Shu, Hongchao, et al.
Published: (2024)
Temporally-Constrained Video Reasoning Segmentation and Automated Benchmark Construction
by: Shen, Yiqing, et al.
Published: (2025)
by: Shen, Yiqing, et al.
Published: (2025)
Coupled Video Frame Interpolation and Encoding with Hybrid Event Cameras for Low-Power High-Framerate Video
by: Takahashi, Hidekazu, et al.
Published: (2025)
by: Takahashi, Hidekazu, et al.
Published: (2025)
Privacy-Preserving Operating Room Workflow Analysis using Digital Twins
by: Perez, Alejandra, et al.
Published: (2025)
by: Perez, Alejandra, et al.
Published: (2025)
AI-Based Detection of Temporal Changes in MR-Linac Images Acquired During Routine Prostate Radiotherapy
by: Park, Seungbin, et al.
Published: (2026)
by: Park, Seungbin, et al.
Published: (2026)
Constrained Natural Language Action Planning for Resilient Embodied Systems
by: Byrd, Grayson, et al.
Published: (2025)
by: Byrd, Grayson, et al.
Published: (2025)
Memorizing SAM: 3D Medical Segment Anything Model with Memorizing Transformer
by: Shao, Xinyuan, et al.
Published: (2024)
by: Shao, Xinyuan, et al.
Published: (2024)
A Neural Enhancement Post-Processor with a Dynamic AV1 Encoder Configuration Strategy for CLIC 2024
by: Ramsook, Darren, et al.
Published: (2024)
by: Ramsook, Darren, et al.
Published: (2024)
Synthetic Video Enhances Physical Fidelity in Video Synthesis
by: Zhao, Qi, et al.
Published: (2025)
by: Zhao, Qi, et al.
Published: (2025)
An Efficient Quality Metric for Video Frame Interpolation Based on Motion-Field Divergence
by: Daly, Conall, et al.
Published: (2025)
by: Daly, Conall, et al.
Published: (2025)
Evaluating Deep Learning-based Melanoma Classification using Immunohistochemistry and Routine Histology: A Three Center Study
by: Wies, Christoph, et al.
Published: (2023)
by: Wies, Christoph, et al.
Published: (2023)
LiteVPNet: A Lightweight Network for Video Encoding Control in Quality-Critical Applications
by: Vibhoothi, Vibhoothi, et al.
Published: (2025)
by: Vibhoothi, Vibhoothi, et al.
Published: (2025)
Inferring Clinically Relevant Molecular Subtypes of Pancreatic Cancer from Routine Histopathology Using Deep Learning
by: Akbar, Abdul Rehman, et al.
Published: (2026)
by: Akbar, Abdul Rehman, et al.
Published: (2026)
EvEnhancer: Empowering Effectiveness, Efficiency and Generalizability for Continuous Space-Time Video Super-Resolution with Events
by: Wei, Shuoyan, et al.
Published: (2025)
by: Wei, Shuoyan, et al.
Published: (2025)
Similar Items
-
AffordTissue: Dense Affordance Prediction for Tool-Action Specific Tissue Interaction
by: Maksutova, Aiza, et al.
Published: (2026) -
Online Reasoning Video Segmentation with Just-in-Time Digital Twins
by: Shen, Yiqing, et al.
Published: (2025) -
SAW: Toward a Surgical Action World Model via Controllable and Scalable Video Generation
by: Rapuri, Sampath, et al.
Published: (2026) -
Counterfactual World Models via Digital Twin-conditioned Video Diffusion
by: Shen, Yiqing, et al.
Published: (2025) -
Investigating a Policy-Based Formulation for Endoscopic Camera Pose Recovery
by: Mangulabnan, Jan Emily, et al.
Published: (2026)