Bridging Geometric and Semantic Foundation Models for Generalized Monocular Depth Estimation
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Sanggyun, Choi, Wonjoon, Park, Jihun, Kim, Jaeyeul, Lee, Seunghun, Seo, Jiwan, Im, Sunghoon |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CAVIS: Context-Aware Video Instance Segmentation
by: Lee, Seunghun, et al.
Published: (2024)
by: Lee, Seunghun, et al.
Published: (2024)
CVA: Context-aware Video-text Alignment for Video Temporal Grounding
by: Moon, Sungho, et al.
Published: (2026)
by: Moon, Sungho, et al.
Published: (2026)
Intrinsic Image Decomposition for Robust Self-supervised Monocular Depth Estimation on Reflective Surfaces
by: Choi, Wonhyeok, et al.
Published: (2025)
by: Choi, Wonhyeok, et al.
Published: (2025)
A Training-Free Style-aligned Image Generation with Scale-wise Autoregressive Model
by: Park, Jihun, et al.
Published: (2025)
by: Park, Jihun, et al.
Published: (2025)
Scale-invariant and View-relational Representation Learning for Full Surround Monocular Depth
by: Hwang, Kyumin, et al.
Published: (2025)
by: Hwang, Kyumin, et al.
Published: (2025)
A Training-Free Style-Personalization via SVD-Based Feature Decomposition
by: Lee, Kyoungmin, et al.
Published: (2025)
by: Lee, Kyoungmin, et al.
Published: (2025)
Style-Editor: Text-driven object-centric style editing
by: Park, Jihun, et al.
Published: (2024)
by: Park, Jihun, et al.
Published: (2024)
Rethinking LiDAR Domain Generalization: Single Source as Multiple Density Domains
by: Kim, Jaeyeul, et al.
Published: (2023)
by: Kim, Jaeyeul, et al.
Published: (2023)
Depth-discriminative Metric Learning for Monocular 3D Object Detection
by: Choi, Wonhyeok, et al.
Published: (2024)
by: Choi, Wonhyeok, et al.
Published: (2024)
Infinite-Story: A Training-Free Consistent Text-to-Image Generation
by: Park, Jihun, et al.
Published: (2025)
by: Park, Jihun, et al.
Published: (2025)
Flow4D: Leveraging 4D Voxel Network for LiDAR Scene Flow Estimation
by: Kim, Jaeyeul, et al.
Published: (2024)
by: Kim, Jaeyeul, et al.
Published: (2024)
Self-supervised Monocular Depth Estimation Robust to Reflective Surface Leveraged by Triplet Mining
by: Choi, Wonhyeok, et al.
Published: (2025)
by: Choi, Wonhyeok, et al.
Published: (2025)
Latest Object Memory Management for Temporally Consistent Video Instance Segmentation
by: Lee, Seunghun, et al.
Published: (2025)
by: Lee, Seunghun, et al.
Published: (2025)
ProDepth: Boosting Self-Supervised Multi-Frame Monocular Depth with Probabilistic Fusion
by: Woo, Sungmin, et al.
Published: (2024)
by: Woo, Sungmin, et al.
Published: (2024)
Temporal Grounding as a Learning Signal for Referring Video Object Segmentation
by: Lee, Seunghun, et al.
Published: (2025)
by: Lee, Seunghun, et al.
Published: (2025)
Monocular Depth Estimation and Segmentation for Transparent Object with Iterative Semantic and Geometric Fusion
by: Liu, Jiangyuan, et al.
Published: (2025)
by: Liu, Jiangyuan, et al.
Published: (2025)
TIE-KD: Teacher-Independent and Explainable Knowledge Distillation for Monocular Depth Estimation
by: Choi, Sangwon, et al.
Published: (2024)
by: Choi, Sangwon, et al.
Published: (2024)
Multi-task Learning for Real-time Autonomous Driving Leveraging Task-adaptive Attention Generator
by: Choi, Wonhyeok, et al.
Published: (2024)
by: Choi, Wonhyeok, et al.
Published: (2024)
Enhanced Scale-aware Depth Estimation for Monocular Endoscopic Scenes with Geometric Modeling
by: Wei, Ruofeng, et al.
Published: (2024)
by: Wei, Ruofeng, et al.
Published: (2024)
Adversarial Manhole: Challenging Monocular Depth Estimation and Semantic Segmentation Models with Patch Attack
by: Suryanto, Naufal, et al.
Published: (2024)
by: Suryanto, Naufal, et al.
Published: (2024)
Extending Foundational Monocular Depth Estimators to Fisheye Cameras with Calibration Tokens
by: Gangopadhyay, Suchisrit, et al.
Published: (2025)
by: Gangopadhyay, Suchisrit, et al.
Published: (2025)
Stereo-Matching Knowledge Distilled Monocular Depth Estimation Filtered by Multiple Disparity Consistency
by: Ka, Woonghyun, et al.
Published: (2024)
by: Ka, Woonghyun, et al.
Published: (2024)
PTC-Depth: Pose-Refined Monocular Depth Estimation with Temporal Consistency
by: Han, Leezy, et al.
Published: (2026)
by: Han, Leezy, et al.
Published: (2026)
UM-Depth : Uncertainty Masked Self-Supervised Monocular Depth Estimation with Visual Odometry
by: Um, Tae-Wook, et al.
Published: (2025)
by: Um, Tae-Wook, et al.
Published: (2025)
SPACE-CLIP: Spatial Perception via Adaptive CLIP Embeddings for Monocular Depth Estimation
by: Cho, Taewan, et al.
Published: (2026)
by: Cho, Taewan, et al.
Published: (2026)
Metric3Dv2: A Versatile Monocular Geometric Foundation Model for Zero-shot Metric Depth and Surface Normal Estimation
by: Hu, Mu, et al.
Published: (2024)
by: Hu, Mu, et al.
Published: (2024)
On the Viability of Monocular Depth Pre-training for Semantic Segmentation
by: Lao, Dong, et al.
Published: (2022)
by: Lao, Dong, et al.
Published: (2022)
CompoDistill: Attention Distillation for Compositional Reasoning in Multimodal LLMs
by: Kim, Jiwan, et al.
Published: (2025)
by: Kim, Jiwan, et al.
Published: (2025)
StarryGazer: Leveraging Monocular Depth Estimation Models for Domain-Agnostic Single Depth Image Completion
by: Hong, Sangmin, et al.
Published: (2025)
by: Hong, Sangmin, et al.
Published: (2025)
The Third Monocular Depth Estimation Challenge
by: Spencer, Jaime, et al.
Published: (2024)
by: Spencer, Jaime, et al.
Published: (2024)
DepthMaster: Taming Diffusion Models for Monocular Depth Estimation
by: Song, Ziyang, et al.
Published: (2025)
by: Song, Ziyang, et al.
Published: (2025)
Leveraging Stable Diffusion for Monocular Depth Estimation via Image Semantic Encoding
by: Xia, Jingming, et al.
Published: (2025)
by: Xia, Jingming, et al.
Published: (2025)
Rate-Adaptive Quantization: A Multi-Rate Codebook Adaptation for Vector Quantization-based Generative Models
by: Seo, Jiwan, et al.
Published: (2024)
by: Seo, Jiwan, et al.
Published: (2024)
Visual Autoregressive Modelling for Monocular Depth Estimation
by: El-Ghoussani, Amir, et al.
Published: (2025)
by: El-Ghoussani, Amir, et al.
Published: (2025)
DepthFM: Fast Monocular Depth Estimation with Flow Matching
by: Gui, Ming, et al.
Published: (2024)
by: Gui, Ming, et al.
Published: (2024)
VLF-MSC: Vision-Language Feature-Based Multimodal Semantic Communication System
by: Ahn, Gwangyeon, et al.
Published: (2025)
by: Ahn, Gwangyeon, et al.
Published: (2025)
RPG360: Robust 360 Depth Estimation with Perspective Foundation Models and Graph Optimization
by: Jung, Dongki, et al.
Published: (2025)
by: Jung, Dongki, et al.
Published: (2025)
Distilling Monocular Foundation Model for Fine-grained Depth Completion
by: Liang, Yingping, et al.
Published: (2025)
by: Liang, Yingping, et al.
Published: (2025)
Multi-task Geometric Estimation of Depth and Surface Normal from Monocular 360° Images
by: Huang, Kun, et al.
Published: (2024)
by: Huang, Kun, et al.
Published: (2024)
The Fourth Monocular Depth Estimation Challenge
by: Obukhov, Anton, et al.
Published: (2025)
by: Obukhov, Anton, et al.
Published: (2025)
Similar Items
-
CAVIS: Context-Aware Video Instance Segmentation
by: Lee, Seunghun, et al.
Published: (2024) -
CVA: Context-aware Video-text Alignment for Video Temporal Grounding
by: Moon, Sungho, et al.
Published: (2026) -
Intrinsic Image Decomposition for Robust Self-supervised Monocular Depth Estimation on Reflective Surfaces
by: Choi, Wonhyeok, et al.
Published: (2025) -
A Training-Free Style-aligned Image Generation with Scale-wise Autoregressive Model
by: Park, Jihun, et al.
Published: (2025) -
Scale-invariant and View-relational Representation Learning for Full Surround Monocular Depth
by: Hwang, Kyumin, et al.
Published: (2025)