The Midas Touch for Metric Depth
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Yu, Guo, Zizhan, Xiong, Zuyi, Zhang, Haoran, Feng, Yi, Zhao, Hongbo, Wang, Hanli, Fan, Rui |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Discriminately Treating Motion Components Evolves Joint Depth and Ego-Motion Learning
by: Zhang, Mengtan, et al.
Published: (2025)
by: Zhang, Mengtan, et al.
Published: (2025)
An Instance-Centric Panoptic Occupancy Prediction Benchmark for Autonomous Driving
by: Feng, Yi, et al.
Published: (2026)
by: Feng, Yi, et al.
Published: (2026)
SCIPaD: Incorporating Spatial Clues into Unsupervised Pose-Depth Joint Learning
by: Feng, Yi, et al.
Published: (2024)
by: Feng, Yi, et al.
Published: (2024)
Rebenchmarking Unsupervised Monocular 3D Occupancy Prediction
by: Guo, Zizhan, et al.
Published: (2026)
by: Guo, Zizhan, et al.
Published: (2026)
A Birotation Solution for Relative Pose Problems
by: Zhao, Hongbo, et al.
Published: (2025)
by: Zhao, Hongbo, et al.
Published: (2025)
Integrating Disparity Confidence Estimation into Relative Depth Prior-Guided Unsupervised Stereo Matching
by: Liu, Chuang-Wei, et al.
Published: (2025)
by: Liu, Chuang-Wei, et al.
Published: (2025)
DCPI-Depth: Explicitly Infusing Dense Correspondence Prior to Unsupervised Monocular Depth Estimation
by: Zhang, Mengtan, et al.
Published: (2024)
by: Zhang, Mengtan, et al.
Published: (2024)
Promoting CNNs with Cross-Architecture Knowledge Distillation for Efficient Monocular Depth Estimation
by: Zheng, Zhimeng, et al.
Published: (2024)
by: Zheng, Zhimeng, et al.
Published: (2024)
Two-Stream Interactive Joint Learning of Scene Parsing and Geometric Vision Tasks
by: Tang, Guanfeng, et al.
Published: (2026)
by: Tang, Guanfeng, et al.
Published: (2026)
Unsupervised Collaborative Domain Adaptation for Driving Scene Parsing
by: Fan, Jiahe, et al.
Published: (2026)
by: Fan, Jiahe, et al.
Published: (2026)
DepthLM: Metric Depth From Vision Language Models
by: Cai, Zhipeng, et al.
Published: (2025)
by: Cai, Zhipeng, et al.
Published: (2025)
DeepSight: Bridging Depth Maps and Language with a Depth-Driven Multimodal Model
by: Yang, Hao, et al.
Published: (2026)
by: Yang, Hao, et al.
Published: (2026)
ScaleDepth: Decomposing Metric Depth Estimation into Scale Prediction and Relative Depth Estimation
by: Zhu, Ruijie, et al.
Published: (2024)
by: Zhu, Ruijie, et al.
Published: (2024)
HybridDepth: Robust Metric Depth Fusion by Leveraging Depth from Focus and Single-Image Priors
by: Ganj, Ashkan, et al.
Published: (2024)
by: Ganj, Ashkan, et al.
Published: (2024)
RaCalNet: Radar Calibration Network for Sparse-Supervised Metric Depth Estimation
by: Qin, Xingrui, et al.
Published: (2025)
by: Qin, Xingrui, et al.
Published: (2025)
Unlocking Dense Metric Depth Estimation in VLMs
by: Yu, Hanxun, et al.
Published: (2026)
by: Yu, Hanxun, et al.
Published: (2026)
UniDepth: Universal Monocular Metric Depth Estimation
by: Piccinelli, Luigi, et al.
Published: (2024)
by: Piccinelli, Luigi, et al.
Published: (2024)
Survey on Monocular Metric Depth Estimation
by: Zhang, Jiuling
Published: (2025)
by: Zhang, Jiuling
Published: (2025)
MetricDepth: Enhancing Monocular Depth Estimation with Deep Metric Learning
by: Liu, Chunpu, et al.
Published: (2024)
by: Liu, Chunpu, et al.
Published: (2024)
Depth-Guided Metric-Aware Temporal Consistency for Monocular Video Human Mesh Recovery
by: Cen, Jiaxin, et al.
Published: (2026)
by: Cen, Jiaxin, et al.
Published: (2026)
SM4Depth: Seamless Monocular Metric Depth Estimation across Multiple Cameras and Scenes by One Model
by: Liu, Yihao, et al.
Published: (2024)
by: Liu, Yihao, et al.
Published: (2024)
Dive Deeper into Rectifying Homography for Stereo Camera Online Self-Calibration
by: Zhao, Hongbo, et al.
Published: (2023)
by: Zhao, Hongbo, et al.
Published: (2023)
Monocular One-Shot Metric-Depth Alignment for RGB-Based Robot Grasping
by: Guo, Teng, et al.
Published: (2025)
by: Guo, Teng, et al.
Published: (2025)
RadarCam-Depth: Radar-Camera Fusion for Depth Estimation with Learned Metric Scale
by: Li, Han, et al.
Published: (2024)
by: Li, Han, et al.
Published: (2024)
Metric-Solver: Sliding Anchored Metric Depth Estimation from a Single Image
by: Wen, Tao, et al.
Published: (2025)
by: Wen, Tao, et al.
Published: (2025)
Prompting Depth Anything for 4K Resolution Accurate Metric Depth Estimation
by: Lin, Haotong, et al.
Published: (2024)
by: Lin, Haotong, et al.
Published: (2024)
M${^2}$Depth: Self-supervised Two-Frame Multi-camera Metric Depth Estimation
by: Zou, Yingshuang, et al.
Published: (2024)
by: Zou, Yingshuang, et al.
Published: (2024)
Perfecting Depth: Uncertainty-Aware Enhancement of Metric Depth
by: Jun, Jinyoung, et al.
Published: (2025)
by: Jun, Jinyoung, et al.
Published: (2025)
UniDAC: Universal Metric Depth Estimation for Any Camera
by: Ganesan, Girish Chandar, et al.
Published: (2026)
by: Ganesan, Girish Chandar, et al.
Published: (2026)
SNE-RoadSegV2: Advancing Heterogeneous Feature Fusion and Fallibility Awareness for Freespace Detection
by: Feng, Yi, et al.
Published: (2024)
by: Feng, Yi, et al.
Published: (2024)
MetricAnything: Scaling Metric Depth Pretraining with Noisy Heterogeneous Sources
by: Ma, Baorui, et al.
Published: (2026)
by: Ma, Baorui, et al.
Published: (2026)
Language as Prior, Vision as Calibration: Metric Scale Recovery for Monocular Depth Estimation
by: Zhan, Mingxia, et al.
Published: (2026)
by: Zhan, Mingxia, et al.
Published: (2026)
Touch2Shape: Touch-Conditioned 3D Diffusion for Shape Exploration and Reconstruction
by: Wang, Yuanbo, et al.
Published: (2025)
by: Wang, Yuanbo, et al.
Published: (2025)
SharpDepth: Sharpening Metric Depth Predictions Using Diffusion Distillation
by: Pham, Duc-Hai, et al.
Published: (2024)
by: Pham, Duc-Hai, et al.
Published: (2024)
Self-supervised Event-based Monocular Depth Estimation using Cross-modal Consistency
by: Zhu, Junyu, et al.
Published: (2024)
by: Zhu, Junyu, et al.
Published: (2024)
RA-Touch: Retrieval-Augmented Touch Understanding with Enriched Visual Data
by: Cho, Yoorhim, et al.
Published: (2025)
by: Cho, Yoorhim, et al.
Published: (2025)
TouchMap-OR: Multi-View 3D Mapping of Hand-Surface Contacts
by: Ktistakis, Sophokles, et al.
Published: (2026)
by: Ktistakis, Sophokles, et al.
Published: (2026)
Metric3Dv2: A Versatile Monocular Geometric Foundation Model for Zero-shot Metric Depth and Surface Normal Estimation
by: Hu, Mu, et al.
Published: (2024)
by: Hu, Mu, et al.
Published: (2024)
Boosting Monocular Metric Depth Estimation via Bokeh Rendering
by: Zhang, Hangwei, et al.
Published: (2025)
by: Zhang, Hangwei, et al.
Published: (2025)
Weakly-Supervised Referring Video Object Segmentation through Text Supervision
by: Shi, Miaojing, et al.
Published: (2026)
by: Shi, Miaojing, et al.
Published: (2026)
Similar Items
-
Discriminately Treating Motion Components Evolves Joint Depth and Ego-Motion Learning
by: Zhang, Mengtan, et al.
Published: (2025) -
An Instance-Centric Panoptic Occupancy Prediction Benchmark for Autonomous Driving
by: Feng, Yi, et al.
Published: (2026) -
SCIPaD: Incorporating Spatial Clues into Unsupervised Pose-Depth Joint Learning
by: Feng, Yi, et al.
Published: (2024) -
Rebenchmarking Unsupervised Monocular 3D Occupancy Prediction
by: Guo, Zizhan, et al.
Published: (2026) -
A Birotation Solution for Relative Pose Problems
by: Zhao, Hongbo, et al.
Published: (2025)