Ophora: A Large-Scale Data-Driven Text-Guided Ophthalmic Surgical Video Generation Model
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Wei, Hu, Ming, Wang, Guoan, Liu, Lihao, Zhou, Kaijing, Ning, Junzhi, Guo, Xin, Ge, Zongyuan, Gu, Lixu, He, Junjun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
F^2TTA: Free-Form Test-Time Adaptation on Cross-Domain Medical Image Classification via Image-Level Disentangled Prompt Tuning
by: Li, Wei, et al.
Published: (2025)
by: Li, Wei, et al.
Published: (2025)
Open World MRI Reconstruction with Bias-Calibrated Adaptation
by: Liu, Jiyao, et al.
Published: (2026)
by: Liu, Jiyao, et al.
Published: (2026)
RetinaLogos: Fine-Grained Synthesis of High-Resolution Retinal Images Through Captions
by: Ning, Junzhi, et al.
Published: (2025)
by: Ning, Junzhi, et al.
Published: (2025)
Unified Medical Image Tokenizer for Autoregressive Synthesis and Understanding
by: Ma, Chenglong, et al.
Published: (2025)
by: Ma, Chenglong, et al.
Published: (2025)
Multi-modal MRI Translation via Evidential Regression and Distribution Calibration
by: Liu, Jiyao, et al.
Published: (2024)
by: Liu, Jiyao, et al.
Published: (2024)
Retrieval-Guided Photovoltaic Inventory Estimation from Satellite Imagery for Distribution Grid Planning
by: Guo, Muhao, et al.
Published: (2026)
by: Guo, Muhao, et al.
Published: (2026)
Ophthalmic Biomarker Detection: Highlights from the IEEE Video and Image Processing Cup 2023 Student Competition
by: AlRegib, Ghassan, et al.
Published: (2024)
by: AlRegib, Ghassan, et al.
Published: (2024)
Diffusion Model Driven Test-Time Image Adaptation for Robust Skin Lesion Classification
by: Hu, Ming, et al.
Published: (2024)
by: Hu, Ming, et al.
Published: (2024)
Comprehensive Generative Replay for Task-Incremental Segmentation with Concurrent Appearance and Semantic Forgetting
by: Li, Wei, et al.
Published: (2024)
by: Li, Wei, et al.
Published: (2024)
Non-Interrupting Rail Track Geometry Measurement System Using UAV and LiDAR
by: Qiu, Lihao, et al.
Published: (2024)
by: Qiu, Lihao, et al.
Published: (2024)
Adapting SAM for Surgical Instrument Tracking and Segmentation in Endoscopic Submucosal Dissection Videos
by: Yu, Jieming, et al.
Published: (2024)
by: Yu, Jieming, et al.
Published: (2024)
Interpretable Few-Shot Retinal Disease Diagnosis with Concept-Guided Prompting of Vision-Language Models
by: Mehta, Deval, et al.
Published: (2025)
by: Mehta, Deval, et al.
Published: (2025)
Performance and Non-adversarial Robustness of the Segment Anything Model 2 in Surgical Video Segmentation
by: Shen, Yiqing, et al.
Published: (2024)
by: Shen, Yiqing, et al.
Published: (2024)
ADAgent: LLM Agent for Alzheimer's Disease Analysis with Collaborative Coordinator
by: Hou, Wenlong, et al.
Published: (2025)
by: Hou, Wenlong, et al.
Published: (2025)
A-Eval: A Benchmark for Cross-Dataset Evaluation of Abdominal Multi-Organ Segmentation
by: Huang, Ziyan, et al.
Published: (2023)
by: Huang, Ziyan, et al.
Published: (2023)
OmniMedVQA: A New Large-Scale Comprehensive Evaluation Benchmark for Medical LVLM
by: Hu, Yutao, et al.
Published: (2024)
by: Hu, Yutao, et al.
Published: (2024)
Fundus2Video: Cross-Modal Angiography Video Generation from Static Fundus Photography with Clinical Knowledge Guidance
by: Zhang, Weiyi, et al.
Published: (2024)
by: Zhang, Weiyi, et al.
Published: (2024)
ColonAdapter: Geometry Estimation Through Foundation Model Adaptation for Colonoscopy
by: Jiang, Zhiyi, et al.
Published: (2025)
by: Jiang, Zhiyi, et al.
Published: (2025)
Prompted Contextual Transformer for Incomplete-View CT Reconstruction
by: Ma, Chenglong, et al.
Published: (2023)
by: Ma, Chenglong, et al.
Published: (2023)
VRS-UIE: Value-Driven Reordering Scanning for Underwater Image Enhancement
by: Jiang, Kui, et al.
Published: (2025)
by: Jiang, Kui, et al.
Published: (2025)
PET Image Denoising via Text-Guided Diffusion: Integrating Anatomical Priors through Text Prompts
by: Yu, Boxiao, et al.
Published: (2025)
by: Yu, Boxiao, et al.
Published: (2025)
RDDM: A Residual-Driven Drifting Model for High-Fidelity Low-Dose CT Denoising
by: Wang, Jianxu, et al.
Published: (2026)
by: Wang, Jianxu, et al.
Published: (2026)
CGCCE-Net:Change-Guided Cross Correlation Enhancement Network for Remote Sensing Building Change Detection
by: Wang, ChengMing
Published: (2025)
by: Wang, ChengMing
Published: (2025)
VISTA: A Benchmark for Real-Time Video Streaming under Network Impairments in Surgical Teleoperation
by: Deng, Zexin, et al.
Published: (2026)
by: Deng, Zexin, et al.
Published: (2026)
Weakly Supervised YOLO Network for Surgical Instrument Localization in Endoscopic Videos
by: Wei, Rongfeng, et al.
Published: (2023)
by: Wei, Rongfeng, et al.
Published: (2023)
Unsupervised Cardiac Video Translation Via Motion Feature Guided Diffusion Model
by: Deb, Swakshar, et al.
Published: (2025)
by: Deb, Swakshar, et al.
Published: (2025)
A Multi-Scale Spatial-Temporal Network for Wireless Video Transmission
by: Zhou, Xinyi, et al.
Published: (2024)
by: Zhou, Xinyi, et al.
Published: (2024)
Attention-Guided Fair AI Modeling for Skin Cancer Diagnosis
by: Zhu, Mingcheng, et al.
Published: (2025)
by: Zhu, Mingcheng, et al.
Published: (2025)
Video Coding with Cross-Component Sample Offset
by: Gao, Han, et al.
Published: (2024)
by: Gao, Han, et al.
Published: (2024)
MSWAL: 3D Multi-class Segmentation of Whole Abdominal Lesions Dataset
by: Wu, Zhaodong, et al.
Published: (2025)
by: Wu, Zhaodong, et al.
Published: (2025)
SASVi -- Segment Any Surgical Video
by: Sivakumar, Ssharvien Kumar, et al.
Published: (2025)
by: Sivakumar, Ssharvien Kumar, et al.
Published: (2025)
Towards Interpretable Counterfactual Generation via Multimodal Autoregression
by: Ma, Chenglong, et al.
Published: (2025)
by: Ma, Chenglong, et al.
Published: (2025)
I$^2$VC: A Unified Framework for Intra- & Inter-frame Video Compression
by: Liu, Meiqin, et al.
Published: (2024)
by: Liu, Meiqin, et al.
Published: (2024)
Video-rate gigapixel ptychography via space-time neural field representations
by: Wang, Ruihai, et al.
Published: (2025)
by: Wang, Ruihai, et al.
Published: (2025)
Poisson Flow Joint Model for Multiphase contrast-enhanced CT
by: Ge, Rongjun, et al.
Published: (2025)
by: Ge, Rongjun, et al.
Published: (2025)
On Quantizing Neural Representation for Variable-Rate Video Coding
by: Shi, Junqi, et al.
Published: (2025)
by: Shi, Junqi, et al.
Published: (2025)
A Video Coding Method Based on Neural Network for CLIC2024
by: Li, Zhengang, et al.
Published: (2024)
by: Li, Zhengang, et al.
Published: (2024)
Position Dependent Prediction Combination For Intra-Frame Video Coding
by: Said, Amir, et al.
Published: (2025)
by: Said, Amir, et al.
Published: (2025)
High-Quality and Large-Scale Image Downscaling for Modern Display Devices
by: Mitra, Suvrojit, et al.
Published: (2025)
by: Mitra, Suvrojit, et al.
Published: (2025)
PupiNet: Seamless OCT-OCTA Interconversion Through Wavelet-Driven and Multi-Scale Attention Mechanisms
by: Tian, Renzhi, et al.
Published: (2025)
by: Tian, Renzhi, et al.
Published: (2025)
Similar Items
-
F^2TTA: Free-Form Test-Time Adaptation on Cross-Domain Medical Image Classification via Image-Level Disentangled Prompt Tuning
by: Li, Wei, et al.
Published: (2025) -
Open World MRI Reconstruction with Bias-Calibrated Adaptation
by: Liu, Jiyao, et al.
Published: (2026) -
RetinaLogos: Fine-Grained Synthesis of High-Resolution Retinal Images Through Captions
by: Ning, Junzhi, et al.
Published: (2025) -
Unified Medical Image Tokenizer for Autoregressive Synthesis and Understanding
by: Ma, Chenglong, et al.
Published: (2025) -
Multi-modal MRI Translation via Evidential Regression and Distribution Calibration
by: Liu, Jiyao, et al.
Published: (2024)