ColonCrafter: A Depth Estimation Model for Colonoscopy Videos Using Diffusion Priors
Fuente:
arXiv
Guardado en:
| Autores principales: | Hardy, Romain, Berzin, Tyler, Rajpurkar, Pranav |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
GeometryCrafter: Consistent Geometry Estimation for Open-world Videos with Diffusion Priors
por: Xu, Tian-Xing, et al.
Publicado: (2025)
por: Xu, Tian-Xing, et al.
Publicado: (2025)
ECoDepth: Effective Conditioning of Diffusion Models for Monocular Depth Estimation
por: Patni, Suraj, et al.
Publicado: (2024)
por: Patni, Suraj, et al.
Publicado: (2024)
InverseCrafter: Efficient Video ReCapture as a Latent Domain Inverse Problem
por: Hong, Yeobin, et al.
Publicado: (2025)
por: Hong, Yeobin, et al.
Publicado: (2025)
WorDepth: Variational Language Prior for Monocular Depth Estimation
por: Zeng, Ziyao, et al.
Publicado: (2024)
por: Zeng, Ziyao, et al.
Publicado: (2024)
Residual Prior Diffusion: A Probabilistic Framework Integrating Coarse Latent Priors with Diffusion Models
por: Kutsuna, Takuro
Publicado: (2025)
por: Kutsuna, Takuro
Publicado: (2025)
Control-A-Video: Controllable Text-to-Video Diffusion Models with Motion Prior and Reward Feedback Learning
por: Chen, Weifeng, et al.
Publicado: (2023)
por: Chen, Weifeng, et al.
Publicado: (2023)
Solving Video Inverse Problems Using Image Diffusion Models
por: Kwon, Taesung, et al.
Publicado: (2024)
por: Kwon, Taesung, et al.
Publicado: (2024)
DepthCrafter: Generating Consistent Long Depth Sequences for Open-world Videos
por: Hu, Wenbo, et al.
Publicado: (2024)
por: Hu, Wenbo, et al.
Publicado: (2024)
DepthPilot: From Controllability to Interpretability in Colonoscopy Video Generation
por: Fu, Junhu, et al.
Publicado: (2026)
por: Fu, Junhu, et al.
Publicado: (2026)
Evaluating Contextual Intelligence in Recyclability: A Comprehensive Study of Image-Based Reasoning Systems
por: Park, Eliot, et al.
Publicado: (2025)
por: Park, Eliot, et al.
Publicado: (2025)
MotionCrafter: Dense Geometry and Motion Reconstruction with a 4D VAE
por: Zhu, Ruijie, et al.
Publicado: (2026)
por: Zhu, Ruijie, et al.
Publicado: (2026)
A Survey on Video Diffusion Models
por: Xing, Zhen, et al.
Publicado: (2023)
por: Xing, Zhen, et al.
Publicado: (2023)
A Critical Synthesis of Uncertainty Quantification and Foundation Models in Monocular Depth Estimation
por: Landgraf, Steven, et al.
Publicado: (2025)
por: Landgraf, Steven, et al.
Publicado: (2025)
VideoGPA: Distilling Geometry Priors for 3D-Consistent Video Generation
por: Du, Hongyang, et al.
Publicado: (2026)
por: Du, Hongyang, et al.
Publicado: (2026)
Learning to Play Video Games with Intuitive Physics Priors
por: Jaiswal, Abhishek, et al.
Publicado: (2024)
por: Jaiswal, Abhishek, et al.
Publicado: (2024)
TurboDiffusion: Accelerating Video Diffusion Models by 100-200 Times
por: Zhang, Jintao, et al.
Publicado: (2025)
por: Zhang, Jintao, et al.
Publicado: (2025)
Warped Diffusion: Solving Video Inverse Problems with Image Diffusion Models
por: Daras, Giannis, et al.
Publicado: (2024)
por: Daras, Giannis, et al.
Publicado: (2024)
CCIS-Diff: A Generative Model with Stable Diffusion Prior for Controlled Colonoscopy Image Synthesis
por: Xie, Yifan, et al.
Publicado: (2024)
por: Xie, Yifan, et al.
Publicado: (2024)
TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Models
por: YU, Mark, et al.
Publicado: (2025)
por: YU, Mark, et al.
Publicado: (2025)
Beyond Generative Priors: Minority Sampling with JEPA-Guided Diffusion
por: Park, Sol, et al.
Publicado: (2026)
por: Park, Sol, et al.
Publicado: (2026)
ViBiDSampler: Enhancing Video Interpolation Using Bidirectional Diffusion Sampler
por: Yang, Serin, et al.
Publicado: (2024)
por: Yang, Serin, et al.
Publicado: (2024)
VideoPDE: Unified Generative PDE Solving via Video Inpainting Diffusion Models
por: Li, Edward, et al.
Publicado: (2025)
por: Li, Edward, et al.
Publicado: (2025)
Correlation of Object Detection Performance with Visual Saliency and Depth Estimation
por: Bartolo, Matthias, et al.
Publicado: (2024)
por: Bartolo, Matthias, et al.
Publicado: (2024)
ReXSonoVQA: A Video QA Benchmark for Procedure-Centric Ultrasound Understanding
por: Wang, Xucheng, et al.
Publicado: (2026)
por: Wang, Xucheng, et al.
Publicado: (2026)
Towards Effective Usage of Human-Centric Priors in Diffusion Models for Text-based Human Image Generation
por: Wang, Junyan, et al.
Publicado: (2024)
por: Wang, Junyan, et al.
Publicado: (2024)
Learning to Generate Rigid Body Interactions with Video Diffusion Models
por: Romero, David, et al.
Publicado: (2025)
por: Romero, David, et al.
Publicado: (2025)
Accelerating Video Inverse Problem Solvers with Autoregressive Diffusion Models
por: Kwon, Taesung, et al.
Publicado: (2026)
por: Kwon, Taesung, et al.
Publicado: (2026)
Contextualized Diffusion Models for Text-Guided Image and Video Generation
por: Yang, Ling, et al.
Publicado: (2024)
por: Yang, Ling, et al.
Publicado: (2024)
Track4Gen: Teaching Video Diffusion Models to Track Points Improves Video Generation
por: Jeong, Hyeonho, et al.
Publicado: (2024)
por: Jeong, Hyeonho, et al.
Publicado: (2024)
VideoGuide: Improving Video Diffusion Models without Training Through a Teacher's Guide
por: Lee, Dohun, et al.
Publicado: (2024)
por: Lee, Dohun, et al.
Publicado: (2024)
Jasmine: Harnessing Diffusion Prior for Self-supervised Depth Estimation
por: Wang, Jiyuan, et al.
Publicado: (2025)
por: Wang, Jiyuan, et al.
Publicado: (2025)
Extending Foundational Monocular Depth Estimators to Fisheye Cameras with Calibration Tokens
por: Gangopadhyay, Suchisrit, et al.
Publicado: (2025)
por: Gangopadhyay, Suchisrit, et al.
Publicado: (2025)
DMin: Scalable Training Data Influence Estimation for Diffusion Models
por: Lin, Huawei, et al.
Publicado: (2024)
por: Lin, Huawei, et al.
Publicado: (2024)
Diagnosing and Improving Diffusion Models by Estimating the Optimal Loss Value
por: Xu, Yixian, et al.
Publicado: (2025)
por: Xu, Yixian, et al.
Publicado: (2025)
GaitCrafter: Diffusion Model for Biometric Preserving Gait Synthesis
por: Mitra, Sirshapan, et al.
Publicado: (2025)
por: Mitra, Sirshapan, et al.
Publicado: (2025)
Ground-A-Video: Zero-shot Grounded Video Editing using Text-to-image Diffusion Models
por: Jeong, Hyeonho, et al.
Publicado: (2023)
por: Jeong, Hyeonho, et al.
Publicado: (2023)
Spectral Motion Alignment for Video Motion Transfer using Diffusion Models
por: Park, Geon Yeong, et al.
Publicado: (2024)
por: Park, Geon Yeong, et al.
Publicado: (2024)
SemSegDepth: A Combined Model for Semantic Segmentation and Depth Completion
por: Lagos, Juan Pablo, et al.
Publicado: (2022)
por: Lagos, Juan Pablo, et al.
Publicado: (2022)
Programmatic Video Prediction Using Large Language Models
por: Tang, Hao, et al.
Publicado: (2025)
por: Tang, Hao, et al.
Publicado: (2025)
VideoDPO: Omni-Preference Alignment for Video Diffusion Generation
por: Liu, Runtao, et al.
Publicado: (2024)
por: Liu, Runtao, et al.
Publicado: (2024)
Ejemplares similares
-
GeometryCrafter: Consistent Geometry Estimation for Open-world Videos with Diffusion Priors
por: Xu, Tian-Xing, et al.
Publicado: (2025) -
ECoDepth: Effective Conditioning of Diffusion Models for Monocular Depth Estimation
por: Patni, Suraj, et al.
Publicado: (2024) -
InverseCrafter: Efficient Video ReCapture as a Latent Domain Inverse Problem
por: Hong, Yeobin, et al.
Publicado: (2025) -
WorDepth: Variational Language Prior for Monocular Depth Estimation
por: Zeng, Ziyao, et al.
Publicado: (2024) -
Residual Prior Diffusion: A Probabilistic Framework Integrating Coarse Latent Priors with Diffusion Models
por: Kutsuna, Takuro
Publicado: (2025)