Align3R: Aligned Monocular Depth Estimation for Dynamic Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Lu, Jiahao, Huang, Tianyu, Li, Peng, Dou, Zhiyang, Lin, Cheng, Cui, Zhiming, Dong, Zhen, Yeung, Sai-Kit, Wang, Wenping, Liu, Yuan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TrackingWorld: World-centric Monocular 3D Tracking of Almost All Pixels
by: Lu, Jiahao, et al.
Published: (2025)
by: Lu, Jiahao, et al.
Published: (2025)
Diffusion as Shader: 3D-aware Video Diffusion for Versatile Video Generation Control
by: Gu, Zekai, et al.
Published: (2025)
by: Gu, Zekai, et al.
Published: (2025)
Photo-SLAM: Real-time Simultaneous Localization and Photorealistic Mapping for Monocular, Stereo, and RGB-D Cameras
by: Huang, Huajian, et al.
Published: (2023)
by: Huang, Huajian, et al.
Published: (2023)
MoDGS: Dynamic Gaussian Splatting from Casually-captured Monocular Videos with Depth Priors
by: Liu, Qingming, et al.
Published: (2024)
by: Liu, Qingming, et al.
Published: (2024)
360DVO: Deep Visual Odometry for Monocular 360-Degree Camera
by: Guo, Xiaopeng, et al.
Published: (2026)
by: Guo, Xiaopeng, et al.
Published: (2026)
FreeReg: Image-to-Point Cloud Registration Leveraging Pretrained Diffusion Models and Monocular Depth Estimators
by: Wang, Haiping, et al.
Published: (2023)
by: Wang, Haiping, et al.
Published: (2023)
Endo3R: Unified Online Reconstruction from Dynamic Monocular Endoscopic Video
by: Guo, Jiaxin, et al.
Published: (2025)
by: Guo, Jiaxin, et al.
Published: (2025)
HuPrior3R: Incorporating Human Priors for Better 3D Dynamic Reconstruction from Monocular Videos
by: Xiong, Weitao, et al.
Published: (2025)
by: Xiong, Weitao, et al.
Published: (2025)
Mining Supervision for Dynamic Regions in Self-Supervised Monocular Depth Estimation
by: Nguyen, Hoang Chuong, et al.
Published: (2024)
by: Nguyen, Hoang Chuong, et al.
Published: (2024)
Deep Neighbor Layer Aggregation for Lightweight Self-Supervised Monocular Depth Estimation
by: Boya, Wang, et al.
Published: (2023)
by: Boya, Wang, et al.
Published: (2023)
EMID: An Emotional Aligned Dataset in Audio-Visual Modality
by: Zou, Jialing, et al.
Published: (2023)
by: Zou, Jialing, et al.
Published: (2023)
Real-time Monocular Depth Estimation on Embedded Systems
by: Feng, Cheng, et al.
Published: (2023)
by: Feng, Cheng, et al.
Published: (2023)
360VOTS: Visual Object Tracking and Segmentation in Omnidirectional Videos
by: Xu, Yinzhe, et al.
Published: (2024)
by: Xu, Yinzhe, et al.
Published: (2024)
Natural Human Motion Recovery by Aligning High-Order Temporal Dynamics from Monocular Videos
by: Wei, Dingkun, et al.
Published: (2026)
by: Wei, Dingkun, et al.
Published: (2026)
EndoStreamDepth: Temporally Consistent Monocular Depth Estimation for Endoscopic Video Streams
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
Revisiting NLI: Towards Cost-Effective and Human-Aligned Metrics for Evaluating LLMs in Question Answering
by: Balamurali, Sai Shridhar, et al.
Published: (2025)
by: Balamurali, Sai Shridhar, et al.
Published: (2025)
CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos
by: Zhao, Chengfeng, et al.
Published: (2026)
by: Zhao, Chengfeng, et al.
Published: (2026)
Focusable Monocular Depth Estimation
by: Du, Yuxin, et al.
Published: (2026)
by: Du, Yuxin, et al.
Published: (2026)
Vision-Language Embodiment for Monocular Depth Estimation
by: Zhang, Jinchang, et al.
Published: (2025)
by: Zhang, Jinchang, et al.
Published: (2025)
OmniGS: Fast Radiance Field Reconstruction using Omnidirectional Gaussian Splatting
by: Li, Longwei, et al.
Published: (2024)
by: Li, Longwei, et al.
Published: (2024)
Dynamic Realms: 4D Content Analysis, Recovery and Generation with Geometric, Topological and Physical Priors
by: Dou, Zhiyang
Published: (2024)
by: Dou, Zhiyang
Published: (2024)
AlignPxtr: Aligning Predicted Behavior Distributions for Bias-Free Video Recommendations
by: Lin, Chengzhi, et al.
Published: (2025)
by: Lin, Chengzhi, et al.
Published: (2025)
Physical 3D Adversarial Attacks against Monocular Depth Estimation in Autonomous Driving
by: Zheng, Junhao, et al.
Published: (2024)
by: Zheng, Junhao, et al.
Published: (2024)
PMPNet: Pixel Movement Prediction Network for Monocular Depth Estimation in Dynamic Scenes
by: Peng, Kebin, et al.
Published: (2024)
by: Peng, Kebin, et al.
Published: (2024)
The Third Monocular Depth Estimation Challenge
by: Spencer, Jaime, et al.
Published: (2024)
by: Spencer, Jaime, et al.
Published: (2024)
WonderHuman: Hallucinating Unseen Parts in Dynamic 3D Human Reconstruction
by: Wang, Zilong, et al.
Published: (2025)
by: Wang, Zilong, et al.
Published: (2025)
Geometry-Constrained Monocular Scale Estimation Using Semantic Segmentation for Dynamic Scenes
by: Zhang, Hui, et al.
Published: (2025)
by: Zhang, Hui, et al.
Published: (2025)
DepthDark: Robust Monocular Depth Estimation for Low-Light Environments
by: Zeng, Longjian, et al.
Published: (2025)
by: Zeng, Longjian, et al.
Published: (2025)
StableDPT: Temporal Stable Monocular Video Depth Estimation
by: Sobko, Ivan, et al.
Published: (2026)
by: Sobko, Ivan, et al.
Published: (2026)
MAMo: Leveraging Memory and Attention for Monocular Video Depth Estimation
by: Yasarla, Rajeev, et al.
Published: (2023)
by: Yasarla, Rajeev, et al.
Published: (2023)
PPEA-Depth: Progressive Parameter-Efficient Adaptation for Self-Supervised Monocular Depth Estimation
by: Dong, Yue-Jiang, et al.
Published: (2023)
by: Dong, Yue-Jiang, et al.
Published: (2023)
VPO: Aligning Text-to-Video Generation Models with Prompt Optimization
by: Cheng, Jiale, et al.
Published: (2025)
by: Cheng, Jiale, et al.
Published: (2025)
FA-Depth: Toward Fast and Accurate Self-supervised Monocular Depth Estimation
by: Wang, Fei, et al.
Published: (2024)
by: Wang, Fei, et al.
Published: (2024)
Aligning Step-by-Step Instructional Diagrams to Video Demonstrations
by: Zhang, Jiahao, et al.
Published: (2023)
by: Zhang, Jiahao, et al.
Published: (2023)
DepthMaster: Taming Diffusion Models for Monocular Depth Estimation
by: Song, Ziyang, et al.
Published: (2025)
by: Song, Ziyang, et al.
Published: (2025)
WorDepth: Variational Language Prior for Monocular Depth Estimation
by: Zeng, Ziyao, et al.
Published: (2024)
by: Zeng, Ziyao, et al.
Published: (2024)
Robust Reasoning via Dynamic Token Selection for Distribution-Aligned Self-Distillation
by: Zhang, Ruiqi, et al.
Published: (2026)
by: Zhang, Ruiqi, et al.
Published: (2026)
Distill Any Depth: Distillation Creates a Stronger Monocular Depth Estimator
by: He, Xiankang, et al.
Published: (2025)
by: He, Xiankang, et al.
Published: (2025)
Align before Adapt: Leveraging Entity-to-Region Alignments for Generalizable Video Action Recognition
by: Chen, Yifei, et al.
Published: (2023)
by: Chen, Yifei, et al.
Published: (2023)
VIMD: Monocular Visual-Inertial Motion and Depth Estimation
by: Katragadda, Saimouli, et al.
Published: (2025)
by: Katragadda, Saimouli, et al.
Published: (2025)
Similar Items
-
TrackingWorld: World-centric Monocular 3D Tracking of Almost All Pixels
by: Lu, Jiahao, et al.
Published: (2025) -
Diffusion as Shader: 3D-aware Video Diffusion for Versatile Video Generation Control
by: Gu, Zekai, et al.
Published: (2025) -
Photo-SLAM: Real-time Simultaneous Localization and Photorealistic Mapping for Monocular, Stereo, and RGB-D Cameras
by: Huang, Huajian, et al.
Published: (2023) -
MoDGS: Dynamic Gaussian Splatting from Casually-captured Monocular Videos with Depth Priors
by: Liu, Qingming, et al.
Published: (2024) -
360DVO: Deep Visual Odometry for Monocular 360-Degree Camera
by: Guo, Xiaopeng, et al.
Published: (2026)