Endora: Video Generation Models as Endoscopy Simulators
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Chenxin, Liu, Hengyu, Liu, Yifan, Feng, Brandon Y., Li, Wuyang, Liu, Xinyu, Chen, Zhen, Shao, Jing, Yuan, Yixuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LGS: A Light-weight 4D Gaussian Splatting for Efficient Surgical Scene Reconstruction
von: Liu, Hengyu, et al.
Veröffentlicht: (2024)
von: Liu, Hengyu, et al.
Veröffentlicht: (2024)
U-KAN Makes Strong Backbone for Medical Image Segmentation and Generation
von: Li, Chenxin, et al.
Veröffentlicht: (2024)
von: Li, Chenxin, et al.
Veröffentlicht: (2024)
GaussianStego: A Generalizable Stenography Pipeline for Generative 3D Gaussians Splatting
von: Li, Chenxin, et al.
Veröffentlicht: (2024)
von: Li, Chenxin, et al.
Veröffentlicht: (2024)
Harnessing Lightweight Transformer with Contextual Synergic Enhancement for Efficient 3D Medical Image Segmentation
von: Liu, Xinyu, et al.
Veröffentlicht: (2026)
von: Liu, Xinyu, et al.
Veröffentlicht: (2026)
MetaScope: Optics-Driven Neural Network for Ultra-Micro Metalens Endoscopy
von: Li, Wuyang, et al.
Veröffentlicht: (2025)
von: Li, Wuyang, et al.
Veröffentlicht: (2025)
EndoSparse: Real-Time Sparse View Synthesis of Endoscopic Scenes using Gaussian Splatting
von: Li, Chenxin, et al.
Veröffentlicht: (2024)
von: Li, Chenxin, et al.
Veröffentlicht: (2024)
GTP-4o: Modality-prompted Heterogeneous Graph Learning for Omni-modal Biomedical Representation
von: Li, Chenxin, et al.
Veröffentlicht: (2024)
von: Li, Chenxin, et al.
Veröffentlicht: (2024)
Track Any Anomalous Object: A Granular Video Anomaly Detection Pipeline
von: Huang, Yuzhi, et al.
Veröffentlicht: (2025)
von: Huang, Yuzhi, et al.
Veröffentlicht: (2025)
DiffRect: Latent Diffusion Label Rectification for Semi-supervised Medical Image Segmentation
von: Liu, Xinyu, et al.
Veröffentlicht: (2024)
von: Liu, Xinyu, et al.
Veröffentlicht: (2024)
EndoGen: Conditional Autoregressive Endoscopic Video Generation
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
EndoGaussian: Real-time Gaussian Splatting for Dynamic Endoscopic Scene Reconstruction
von: Liu, Yifan, et al.
Veröffentlicht: (2024)
von: Liu, Yifan, et al.
Veröffentlicht: (2024)
ConcealGS: Concealing Invisible Copyright Information in 3D Gaussian Splatting
von: Yang, Yifeng, et al.
Veröffentlicht: (2025)
von: Yang, Yifeng, et al.
Veröffentlicht: (2025)
X-GRM: Large Gaussian Reconstruction Model for Sparse-view X-rays to Computed Tomography
von: Liu, Yifan, et al.
Veröffentlicht: (2025)
von: Liu, Yifan, et al.
Veröffentlicht: (2025)
When 3D Partial Points Meets SAM: Tooth Point Cloud Segmentation with Sparse Labels
von: Liu, Yifan, et al.
Veröffentlicht: (2024)
von: Liu, Yifan, et al.
Veröffentlicht: (2024)
Bridge the Gap Between Visual and Linguistic Comprehension for Generalized Zero-shot Semantic Segmentation
von: Guo, Xiaoqing, et al.
Veröffentlicht: (2025)
von: Guo, Xiaoqing, et al.
Veröffentlicht: (2025)
MonoSplat: Generalizable 3D Gaussian Splatting from Monocular Depth Foundation Models
von: Liu, Yifan, et al.
Veröffentlicht: (2025)
von: Liu, Yifan, et al.
Veröffentlicht: (2025)
EndoBench: A Comprehensive Evaluation of Multi-Modal Large Language Models for Endoscopy Analysis
von: Liu, Shengyuan, et al.
Veröffentlicht: (2025)
von: Liu, Shengyuan, et al.
Veröffentlicht: (2025)
IR3D-Bench: Evaluating Vision-Language Model Scene Understanding as Agentic Inverse Rendering
von: Liu, Parker, et al.
Veröffentlicht: (2025)
von: Liu, Parker, et al.
Veröffentlicht: (2025)
GeoT: Geometry-guided Instance-dependent Transition Matrix for Semi-supervised Tooth Point Cloud Segmentation
von: Yu, Weihao, et al.
Veröffentlicht: (2025)
von: Yu, Weihao, et al.
Veröffentlicht: (2025)
FlexGS: Train Once, Deploy Everywhere with Many-in-One Flexible 3D Gaussian Splatting
von: Liu, Hengyu, et al.
Veröffentlicht: (2025)
von: Liu, Hengyu, et al.
Veröffentlicht: (2025)
ID-Crafter: VLM-Grounded Online RL for Compositional Multi-Subject Video Generation
von: Pan, Panwang, et al.
Veröffentlicht: (2025)
von: Pan, Panwang, et al.
Veröffentlicht: (2025)
UN-SAM: Universal Prompt-Free Segmentation for Generalized Nuclei Images
von: Chen, Zhen, et al.
Veröffentlicht: (2024)
von: Chen, Zhen, et al.
Veröffentlicht: (2024)
Anchored Video Generation: Decoupling Scene Construction and Temporal Synthesis in Text-to-Video Diffusion Models
von: Hassan, Mariam, et al.
Veröffentlicht: (2025)
von: Hassan, Mariam, et al.
Veröffentlicht: (2025)
LiveWorld: Simulating Out-of-Sight Dynamics in Generative Video World Models
von: Duan, Zicheng, et al.
Veröffentlicht: (2026)
von: Duan, Zicheng, et al.
Veröffentlicht: (2026)
Customize-A-Video: One-Shot Motion Customization of Text-to-Video Diffusion Models
von: Ren, Yixuan, et al.
Veröffentlicht: (2024)
von: Ren, Yixuan, et al.
Veröffentlicht: (2024)
Learning to Adapt Foundation Model DINOv2 for Capsule Endoscopy Diagnosis
von: Zhang, Bowen, et al.
Veröffentlicht: (2024)
von: Zhang, Bowen, et al.
Veröffentlicht: (2024)
WorldSimBench: Towards Video Generation Models as World Simulators
von: Qin, Yiran, et al.
Veröffentlicht: (2024)
von: Qin, Yiran, et al.
Veröffentlicht: (2024)
Proprio: Latent Self-Scoring and Inference-Time Refinement for Physically Plausible Video Generation
von: Hassan, Mariam, et al.
Veröffentlicht: (2026)
von: Hassan, Mariam, et al.
Veröffentlicht: (2026)
Stable Video Infinity: Infinite-Length Video Generation with Error Recycling
von: Li, Wuyang, et al.
Veröffentlicht: (2025)
von: Li, Wuyang, et al.
Veröffentlicht: (2025)
Learnable Patchmatch and Self-Teaching for Multi-Frame Depth Estimation in Monocular Endoscopy
von: Shao, Shuwei, et al.
Veröffentlicht: (2022)
von: Shao, Shuwei, et al.
Veröffentlicht: (2022)
Foundation Model for Endoscopy Video Analysis via Large-scale Self-supervised Pre-train
von: Wang, Zhao, et al.
Veröffentlicht: (2023)
von: Wang, Zhao, et al.
Veröffentlicht: (2023)
MedVSR: Medical Video Super-Resolution with Cross State-Space Propagation
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
von: Liu, Xinyu, et al.
Veröffentlicht: (2025)
First Frame Is the Place to Go for Video Content Customization
von: Chen, Jingxi, et al.
Veröffentlicht: (2025)
von: Chen, Jingxi, et al.
Veröffentlicht: (2025)
Divide-then-Diagnose: Weaving Clinician-Inspired Contexts for Ultra-Long Capsule Endoscopy Videos
von: Liu, Bowen, et al.
Veröffentlicht: (2026)
von: Liu, Bowen, et al.
Veröffentlicht: (2026)
Editable Scene Simulation for Autonomous Driving via Collaborative LLM-Agents
von: Wei, Yuxi, et al.
Veröffentlicht: (2024)
von: Wei, Yuxi, et al.
Veröffentlicht: (2024)
WildSmoke: Ready-to-Use Dynamic 3D Smoke Assets from a Single Video in the Wild
von: Liu, Yuqiu, et al.
Veröffentlicht: (2025)
von: Liu, Yuqiu, et al.
Veröffentlicht: (2025)
EndoOmni: Zero-Shot Cross-Dataset Depth Estimation in Endoscopy by Robust Self-Learning from Noisy Labels
von: Tian, Qingyao, et al.
Veröffentlicht: (2024)
von: Tian, Qingyao, et al.
Veröffentlicht: (2024)
X$^{2}$-Gaussian: 4D Radiative Gaussian Splatting for Continuous-time Tomographic Reconstruction
von: Yu, Weihao, et al.
Veröffentlicht: (2025)
von: Yu, Weihao, et al.
Veröffentlicht: (2025)
Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation
von: Meng, Fanqing, et al.
Veröffentlicht: (2024)
von: Meng, Fanqing, et al.
Veröffentlicht: (2024)
Dynamic Gaussian Scene Reconstruction from Unsynchronized Videos
von: Xu, Zhixin, et al.
Veröffentlicht: (2025)
von: Xu, Zhixin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LGS: A Light-weight 4D Gaussian Splatting for Efficient Surgical Scene Reconstruction
von: Liu, Hengyu, et al.
Veröffentlicht: (2024) -
U-KAN Makes Strong Backbone for Medical Image Segmentation and Generation
von: Li, Chenxin, et al.
Veröffentlicht: (2024) -
GaussianStego: A Generalizable Stenography Pipeline for Generative 3D Gaussians Splatting
von: Li, Chenxin, et al.
Veröffentlicht: (2024) -
Harnessing Lightweight Transformer with Contextual Synergic Enhancement for Efficient 3D Medical Image Segmentation
von: Liu, Xinyu, et al.
Veröffentlicht: (2026) -
MetaScope: Optics-Driven Neural Network for Ultra-Micro Metalens Endoscopy
von: Li, Wuyang, et al.
Veröffentlicht: (2025)