Learning Temporally Consistent Video Depth from Video Diffusion Priors
Fuente:
arXiv
Saved in:
| Main Authors: | Shao, Jiahao, Yang, Yuanbo, Zhou, Hongyu, Zhang, Youmin, Shen, Yujun, Guizilini, Vitor, Wang, Yue, Poggi, Matteo, Liao, Yiyi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Prometheus: 3D-Aware Latent Diffusion Models for Feed-Forward Text-to-3D Scene Generation
by: Yang, Yuanbo, et al.
Published: (2024)
by: Yang, Yuanbo, et al.
Published: (2024)
The Constant Eye: Benchmarking and Bridging Appearance Robustness in Autonomous Driving
by: Wang, Jiabao, et al.
Published: (2026)
by: Wang, Jiabao, et al.
Published: (2026)
ReRoPE: Repurposing RoPE for Relative Camera Control
by: Li, Chunyang, et al.
Published: (2026)
by: Li, Chunyang, et al.
Published: (2026)
NormalCrafter: Learning Temporally Consistent Normals from Video Diffusion Priors
by: Bin, Yanrui, et al.
Published: (2025)
by: Bin, Yanrui, et al.
Published: (2025)
ScenDi: 3D-to-2D Scene Diffusion Cascades for Urban Generation
by: Guo, Hanlei, et al.
Published: (2026)
by: Guo, Hanlei, et al.
Published: (2026)
Learning 3D-Aware GANs from Unposed Images with Template Feature Field
by: Chen, Xinya, et al.
Published: (2024)
by: Chen, Xinya, et al.
Published: (2024)
Denoise to Track: Harnessing Video Diffusion Priors for Robust Correspondence
by: Yuan, Tianyu, et al.
Published: (2025)
by: Yuan, Tianyu, et al.
Published: (2025)
Depth-Guided Metric-Aware Temporal Consistency for Monocular Video Human Mesh Recovery
by: Cen, Jiaxin, et al.
Published: (2026)
by: Cen, Jiaxin, et al.
Published: (2026)
Towards Realistic Scene Generation with LiDAR Diffusion Models
by: Ran, Haoxi, et al.
Published: (2024)
by: Ran, Haoxi, et al.
Published: (2024)
GRIN: Zero-Shot Metric Depth with Pixel-Level Diffusion
by: Guizilini, Vitor, et al.
Published: (2024)
by: Guizilini, Vitor, et al.
Published: (2024)
AnchorDream: Repurposing Video Diffusion for Embodiment-Aware Robot Data Synthesis
by: Ye, Junjie, et al.
Published: (2025)
by: Ye, Junjie, et al.
Published: (2025)
Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models
by: Jin, Yudong, et al.
Published: (2025)
by: Jin, Yudong, et al.
Published: (2025)
Spatial Degradation-Aware and Temporal Consistent Diffusion Model for Compressed Video Super-Resolution
by: An, Hongyu, et al.
Published: (2025)
by: An, Hongyu, et al.
Published: (2025)
DepthSync: Diffusion Guidance-Based Depth Synchronization for Scale- and Geometry-Consistent Video Depth Estimation
by: Dong, Yue-Jiang, et al.
Published: (2025)
by: Dong, Yue-Jiang, et al.
Published: (2025)
CoDeF: Content Deformation Fields for Temporally Consistent Video Processing
by: Ouyang, Hao, et al.
Published: (2023)
by: Ouyang, Hao, et al.
Published: (2023)
Diffusion Models for Monocular Depth Estimation: Overcoming Challenging Conditions
by: Tosi, Fabio, et al.
Published: (2024)
by: Tosi, Fabio, et al.
Published: (2024)
GemDepth: Geometry-Embedded Features for 3D-Consistent Video Depth
by: Liu, Yuecheng, et al.
Published: (2026)
by: Liu, Yuecheng, et al.
Published: (2026)
DVD: Deterministic Video Depth Estimation with Generative Priors
by: Zhang, Hongfei, et al.
Published: (2026)
by: Zhang, Hongfei, et al.
Published: (2026)
STATIC : Surface Temporal Affine for TIme Consistency in Video Monocular Depth Estimation
by: Yang, Sunghun, et al.
Published: (2024)
by: Yang, Sunghun, et al.
Published: (2024)
Novel View Extrapolation with Video Diffusion Priors
by: Liu, Kunhao, et al.
Published: (2024)
by: Liu, Kunhao, et al.
Published: (2024)
Towards Depth Foundation Model: Recent Trends in Vision-Based Depth Estimation
by: Xu, Zhen, et al.
Published: (2025)
by: Xu, Zhen, et al.
Published: (2025)
Zero-Shot Novel View and Depth Synthesis with Multi-View Geometric Diffusion
by: Guizilini, Vitor, et al.
Published: (2025)
by: Guizilini, Vitor, et al.
Published: (2025)
EndoStreamDepth: Temporally Consistent Monocular Depth Estimation for Endoscopic Video Streams
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
Online Video Depth Anything: Temporally-Consistent Depth Prediction with Low Memory Consumption
by: Feiden, Johann-Friedrich, et al.
Published: (2025)
by: Feiden, Johann-Friedrich, et al.
Published: (2025)
Temporal-Consistent Video Restoration with Pre-trained Diffusion Models
by: Wang, Hengkang, et al.
Published: (2025)
by: Wang, Hengkang, et al.
Published: (2025)
Edit Temporal-Consistent Videos with Image Diffusion Model
by: Wang, Yuanzhi, et al.
Published: (2023)
by: Wang, Yuanzhi, et al.
Published: (2023)
HS-SLAM: Hybrid Representation with Structural Supervision for Improved Dense SLAM
by: Gong, Ziren, et al.
Published: (2025)
by: Gong, Ziren, et al.
Published: (2025)
FlowSeek: Optical Flow Made Easier with Depth Foundation Models and Motion Bases
by: Poggi, Matteo, et al.
Published: (2025)
by: Poggi, Matteo, et al.
Published: (2025)
Pose-Free Omnidirectional Gaussian Splatting for 360-Degree Videos with Consistent Depth Priors
by: Zhuang, Chuanqing, et al.
Published: (2026)
by: Zhuang, Chuanqing, et al.
Published: (2026)
TRIP: Temporal Residual Learning with Image Noise Prior for Image-to-Video Diffusion Models
by: Zhang, Zhongwei, et al.
Published: (2024)
by: Zhang, Zhongwei, et al.
Published: (2024)
JVID: Joint Video-Image Diffusion for Visual-Quality and Temporal-Consistency in Video Generation
by: Reynaud, Hadrien, et al.
Published: (2024)
by: Reynaud, Hadrien, et al.
Published: (2024)
Natural Human Motion Recovery by Aligning High-Order Temporal Dynamics from Monocular Videos
by: Wei, Dingkun, et al.
Published: (2026)
by: Wei, Dingkun, et al.
Published: (2026)
Video Depth Anything: Consistent Depth Estimation for Super-Long Videos
by: Chen, Sili, et al.
Published: (2025)
by: Chen, Sili, et al.
Published: (2025)
InFusion: Inpainting 3D Gaussians via Learning Depth Completion from Diffusion Prior
by: Liu, Zhiheng, et al.
Published: (2024)
by: Liu, Zhiheng, et al.
Published: (2024)
Generative Neural Video Compression via Video Diffusion Prior
by: Mao, Qi, et al.
Published: (2025)
by: Mao, Qi, et al.
Published: (2025)
VideoGPA: Distilling Geometry Priors for 3D-Consistent Video Generation
by: Du, Hongyang, et al.
Published: (2026)
by: Du, Hongyang, et al.
Published: (2026)
DVFace: Spatio-Temporal Dual-Prior Diffusion for Video Face Restoration
by: Chen, Zheng, et al.
Published: (2026)
by: Chen, Zheng, et al.
Published: (2026)
Vivid4D: Improving 4D Reconstruction from Monocular Video by Video Inpainting
by: Huang, Jiaxin, et al.
Published: (2025)
by: Huang, Jiaxin, et al.
Published: (2025)
HUGS: Holistic Urban 3D Scene Understanding via Gaussian Splatting
by: Zhou, Hongyu, et al.
Published: (2024)
by: Zhou, Hongyu, et al.
Published: (2024)
FreeFix: Boosting 3D Gaussian Splatting via Fine-Tuning-Free Diffusion Models
by: Zhou, Hongyu, et al.
Published: (2026)
by: Zhou, Hongyu, et al.
Published: (2026)
Similar Items
-
Prometheus: 3D-Aware Latent Diffusion Models for Feed-Forward Text-to-3D Scene Generation
by: Yang, Yuanbo, et al.
Published: (2024) -
The Constant Eye: Benchmarking and Bridging Appearance Robustness in Autonomous Driving
by: Wang, Jiabao, et al.
Published: (2026) -
ReRoPE: Repurposing RoPE for Relative Camera Control
by: Li, Chunyang, et al.
Published: (2026) -
NormalCrafter: Learning Temporally Consistent Normals from Video Diffusion Priors
by: Bin, Yanrui, et al.
Published: (2025) -
ScenDi: 3D-to-2D Scene Diffusion Cascades for Urban Generation
by: Guo, Hanlei, et al.
Published: (2026)