EndoUFM: Utilizing Foundation Models for Monocular depth estimation of endoscopic images
Fuente:
arXiv
Saved in:
| Main Authors: | Yao, Xinning, Liu, Bo, Li, Bojian, Wang, Jingjing, Yue, Jinghua, Zhou, Fugen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Advancing Depth Anything Model for Unsupervised Monocular Depth Estimation in Endoscopy
by: Li, Bojian, et al.
Published: (2024)
by: Li, Bojian, et al.
Published: (2024)
Prior-Guided DETR for Ultrasound Nodule Detection
by: Wang, Jingjing, et al.
Published: (2026)
by: Wang, Jingjing, et al.
Published: (2026)
Nodule-DETR: A Novel DETR Architecture with Frequency-Channel Attention for Ultrasound Thyroid Nodule Detection
by: Wang, Jingjing, et al.
Published: (2026)
by: Wang, Jingjing, et al.
Published: (2026)
RT-SRTS: Angle-Agnostic Real-Time Simultaneous 3D Reconstruction and Tumor Segmentation from Single X-Ray Projection
by: Zhu, Miao, et al.
Published: (2023)
by: Zhu, Miao, et al.
Published: (2023)
Empirical curvelet based Fully Convolutional Network for supervised texture image segmentation
by: Huang, Yuan, et al.
Published: (2024)
by: Huang, Yuan, et al.
Published: (2024)
EndoDepthL: Lightweight Endoscopic Monocular Depth Estimation with CNN-Transformer
by: Li, Yangke
Published: (2023)
by: Li, Yangke
Published: (2023)
EndoMetric: Near-Light Monocular Metric Scale Estimation in Endoscopy
by: Iranzo, Raúl, et al.
Published: (2024)
by: Iranzo, Raúl, et al.
Published: (2024)
Semantic-CC: Boosting Remote Sensing Image Change Captioning via Foundational Knowledge and Semantic Guidance
by: Zhu, Yongshuo, et al.
Published: (2024)
by: Zhu, Yongshuo, et al.
Published: (2024)
EndoStreamDepth: Temporally Consistent Monocular Depth Estimation for Endoscopic Video Streams
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
UFM: Unified Feature Matching Pre-training with Multi-Modal Image Assistants
by: Di, Yide, et al.
Published: (2025)
by: Di, Yide, et al.
Published: (2025)
Endo-4DGS: Endoscopic Monocular Scene Reconstruction with 4D Gaussian Splatting
by: Huang, Yiming, et al.
Published: (2024)
by: Huang, Yiming, et al.
Published: (2024)
EndoMamba: An Efficient Foundation Model for Endoscopic Videos via Hierarchical Pre-training
by: Tian, Qingyao, et al.
Published: (2025)
by: Tian, Qingyao, et al.
Published: (2025)
Endo3R: Unified Online Reconstruction from Dynamic Monocular Endoscopic Video
by: Guo, Jiaxin, et al.
Published: (2025)
by: Guo, Jiaxin, et al.
Published: (2025)
EndoDINO: A Foundation Model for GI Endoscopy
by: Dermyer, Patrick, et al.
Published: (2025)
by: Dermyer, Patrick, et al.
Published: (2025)
UFM: A Simple Path towards Unified Dense Correspondence with Flow
by: Zhang, Yuchen, et al.
Published: (2025)
by: Zhang, Yuchen, et al.
Published: (2025)
EndoGMDE: Generalizable Monocular Depth Estimation with Mixture of Low-Rank Experts for Diverse Endoscopic Scenes
by: Shao, Liangjing, et al.
Published: (2025)
by: Shao, Liangjing, et al.
Published: (2025)
Monocular absolute depth estimation from endoscopy via domain-invariant feature learning and latent consistency
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
M^3: Dense Matching Meets Multi-View Foundation Models for Monocular Gaussian Splatting SLAM
by: Ren, Kerui, et al.
Published: (2026)
by: Ren, Kerui, et al.
Published: (2026)
MonoSplat: Generalizable 3D Gaussian Splatting from Monocular Depth Foundation Models
by: Liu, Yifan, et al.
Published: (2025)
by: Liu, Yifan, et al.
Published: (2025)
Distilling Monocular Foundation Model for Fine-grained Depth Completion
by: Liang, Yingping, et al.
Published: (2025)
by: Liang, Yingping, et al.
Published: (2025)
Endo-FASt3r: Endoscopic Foundation model Adaptation for Structure from motion
by: Zeinoddin, Mona Sheikh, et al.
Published: (2025)
by: Zeinoddin, Mona Sheikh, et al.
Published: (2025)
Semantic-CD: Remote Sensing Image Semantic Change Detection towards Open-vocabulary Setting
by: Zhu, Yongshuo, et al.
Published: (2025)
by: Zhu, Yongshuo, et al.
Published: (2025)
Colonoscopy polyp detection with massive endoscopic images
by: Yu, Jialin, et al.
Published: (2022)
by: Yu, Jialin, et al.
Published: (2022)
Endo-SemiS: Towards Robust Semi-Supervised Image Segmentation for Endoscopic Video
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
EndoMUST: Monocular Depth Estimation for Robotic Endoscopy via End-to-end Multi-step Self-supervised Training
by: Shao, Liangjing, et al.
Published: (2025)
by: Shao, Liangjing, et al.
Published: (2025)
Bridging Geometric and Semantic Foundation Models for Generalized Monocular Depth Estimation
by: Ma, Sanggyun, et al.
Published: (2025)
by: Ma, Sanggyun, et al.
Published: (2025)
DepthMaster: Taming Diffusion Models for Monocular Depth Estimation
by: Song, Ziyang, et al.
Published: (2025)
by: Song, Ziyang, et al.
Published: (2025)
IQ-LUT: interpolated and quantized LUT for efficient image super-resolution
by: Zhang, Yuxuan, et al.
Published: (2026)
by: Zhang, Yuxuan, et al.
Published: (2026)
EndoChat: Grounded Multimodal Large Language Model for Endoscopic Surgery
by: Wang, Guankun, et al.
Published: (2025)
by: Wang, Guankun, et al.
Published: (2025)
Tackling domain generalization for out-of-distribution endoscopic imaging
by: Teevno, Mansoor Ali, et al.
Published: (2024)
by: Teevno, Mansoor Ali, et al.
Published: (2024)
Self-Consistent Model-based Adaptation for Visual Reinforcement Learning
by: Zhou, Xinning, et al.
Published: (2025)
by: Zhou, Xinning, et al.
Published: (2025)
Multi-needle Localization for Pelvic Seed Implant Brachytherapy based on Tip-handle Detection and Matching
by: Xiao, Zhuo, et al.
Published: (2025)
by: Xiao, Zhuo, et al.
Published: (2025)
ArtHOI: Taming Foundation Models for Monocular 4D Reconstruction of Hand-Articulated-Object Interactions
by: Wang, Zikai, et al.
Published: (2026)
by: Wang, Zikai, et al.
Published: (2026)
EndoGaussian: Real-time Gaussian Splatting for Dynamic Endoscopic Scene Reconstruction
by: Liu, Yifan, et al.
Published: (2024)
by: Liu, Yifan, et al.
Published: (2024)
EndoCoT: Scaling Endogenous Chain-of-Thought Reasoning in Diffusion Models
by: Dai, Xuanlang, et al.
Published: (2026)
by: Dai, Xuanlang, et al.
Published: (2026)
Glossy Object Reconstruction with Cost-effective Polarized Acquisition
by: Wu, Bojian, et al.
Published: (2025)
by: Wu, Bojian, et al.
Published: (2025)
Monocular Biomechanical Tracking of Fingers with Inverse Kinematics to Foundation Models
by: Cotton, R. James, et al.
Published: (2026)
by: Cotton, R. James, et al.
Published: (2026)
EndoSfM3D: Learning to 3D Reconstruct Any Endoscopic Surgery Scene using Self-supervised Foundation Model
by: Zhang, Changhao, et al.
Published: (2025)
by: Zhang, Changhao, et al.
Published: (2025)
EndoVGGT: GNN-Enhanced Depth Estimation for Surgical 3D Reconstruction
by: Fan, Falong, et al.
Published: (2026)
by: Fan, Falong, et al.
Published: (2026)
VFMM3D: Releasing the Potential of Image by Vision Foundation Model for Monocular 3D Object Detection
by: Ding, Bonan, et al.
Published: (2024)
by: Ding, Bonan, et al.
Published: (2024)
Similar Items
-
Advancing Depth Anything Model for Unsupervised Monocular Depth Estimation in Endoscopy
by: Li, Bojian, et al.
Published: (2024) -
Prior-Guided DETR for Ultrasound Nodule Detection
by: Wang, Jingjing, et al.
Published: (2026) -
Nodule-DETR: A Novel DETR Architecture with Frequency-Channel Attention for Ultrasound Thyroid Nodule Detection
by: Wang, Jingjing, et al.
Published: (2026) -
RT-SRTS: Angle-Agnostic Real-Time Simultaneous 3D Reconstruction and Tumor Segmentation from Single X-Ray Projection
by: Zhu, Miao, et al.
Published: (2023) -
Empirical curvelet based Fully Convolutional Network for supervised texture image segmentation
by: Huang, Yuan, et al.
Published: (2024)