RSA: Resolving Scale Ambiguities in Monocular Depth Estimators through Language Descriptions
Fuente:
arXiv
Salvato in:
| Autori principali: | Zeng, Ziyao, Wu, Yangchao, Park, Hyoungseob, Wang, Daniel, Yang, Fengyu, Soatto, Stefano, Lao, Dong, Hong, Byung-Woo, Wong, Alex |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
WorDepth: Variational Language Prior for Monocular Depth Estimation
di: Zeng, Ziyao, et al.
Pubblicazione: (2024)
di: Zeng, Ziyao, et al.
Pubblicazione: (2024)
AugUndo: Scaling Up Augmentations for Monocular Depth Completion and Estimation
di: Wu, Yangchao, et al.
Pubblicazione: (2023)
di: Wu, Yangchao, et al.
Pubblicazione: (2023)
On the Viability of Monocular Depth Pre-training for Semantic Segmentation
di: Lao, Dong, et al.
Pubblicazione: (2022)
di: Lao, Dong, et al.
Pubblicazione: (2022)
Iris: Integrating Language into Diffusion-based Monocular Depth Estimation
di: Zeng, Ziyao, et al.
Pubblicazione: (2024)
di: Zeng, Ziyao, et al.
Pubblicazione: (2024)
Sub-token ViT Embedding via Stochastic Resonance Transformers
di: Lao, Dong, et al.
Pubblicazione: (2023)
di: Lao, Dong, et al.
Pubblicazione: (2023)
ProtoDepth: Unsupervised Continual Depth Completion with Prototypes
di: Rim, Patrick, et al.
Pubblicazione: (2025)
di: Rim, Patrick, et al.
Pubblicazione: (2025)
Test-Time Defense Against Adversarial Attacks via Stochastic Resonance of Latent Ensembles
di: Lao, Dong, et al.
Pubblicazione: (2025)
di: Lao, Dong, et al.
Pubblicazione: (2025)
Extending Foundational Monocular Depth Estimators to Fisheye Cameras with Calibration Tokens
di: Gangopadhyay, Suchisrit, et al.
Pubblicazione: (2025)
di: Gangopadhyay, Suchisrit, et al.
Pubblicazione: (2025)
ETA: Energy-based Test-time Adaptation for Depth Completion
di: Chung, Younjoon, et al.
Pubblicazione: (2025)
di: Chung, Younjoon, et al.
Pubblicazione: (2025)
STree: Speculative Tree Decoding for Hybrid State-Space Models
di: Wu, Yangchao, et al.
Pubblicazione: (2025)
di: Wu, Yangchao, et al.
Pubblicazione: (2025)
Test-Time Adaptation for Depth Completion
di: Park, Hyoungseob, et al.
Pubblicazione: (2024)
di: Park, Hyoungseob, et al.
Pubblicazione: (2024)
Diffeomorphic Template Registration for Atmospheric Turbulence Mitigation
di: Lao, Dong, et al.
Pubblicazione: (2024)
di: Lao, Dong, et al.
Pubblicazione: (2024)
Radar-Guided Polynomial Fitting for Metric Depth Estimation
di: Rim, Patrick, et al.
Pubblicazione: (2025)
di: Rim, Patrick, et al.
Pubblicazione: (2025)
UnCLe: Benchmarking Unsupervised Continual Learning for Depth Completion
di: Chen, Xien, et al.
Pubblicazione: (2024)
di: Chen, Xien, et al.
Pubblicazione: (2024)
Binding Touch to Everything: Learning Unified Multimodal Tactile Representations
di: Yang, Fengyu, et al.
Pubblicazione: (2024)
di: Yang, Fengyu, et al.
Pubblicazione: (2024)
NeuroBind: Towards Unified Multimodal Representations for Neural Signals
di: Yang, Fengyu, et al.
Pubblicazione: (2024)
di: Yang, Fengyu, et al.
Pubblicazione: (2024)
Progressive Test Time Energy Adaptation for Medical Image Segmentation
di: Zhang, Xiaoran, et al.
Pubblicazione: (2025)
di: Zhang, Xiaoran, et al.
Pubblicazione: (2025)
X as Supervision: Contending with Depth Ambiguity in Unsupervised Monocular 3D Pose Estimation
di: Yang, Yuchen, et al.
Pubblicazione: (2024)
di: Yang, Yuchen, et al.
Pubblicazione: (2024)
Scale-Invariant Monocular Depth Estimation via SSI Depth
di: Miangoleh, S. Mahdi H., et al.
Pubblicazione: (2024)
di: Miangoleh, S. Mahdi H., et al.
Pubblicazione: (2024)
Divided Attention: Unsupervised Multi-Object Discovery with Contextually Separated Slots
di: Lao, Dong, et al.
Pubblicazione: (2023)
di: Lao, Dong, et al.
Pubblicazione: (2023)
Language-Based Depth Hints for Monocular Depth Estimation
di: Auty, Dylan, et al.
Pubblicazione: (2024)
di: Auty, Dylan, et al.
Pubblicazione: (2024)
Cycles of Thought: Measuring LLM Confidence through Stable Explanations
di: Becker, Evan, et al.
Pubblicazione: (2024)
di: Becker, Evan, et al.
Pubblicazione: (2024)
Follow-Up Differential Descriptions: Language Models Resolve Ambiguities for Image Classification
di: Esfandiarpoor, Reza, et al.
Pubblicazione: (2023)
di: Esfandiarpoor, Reza, et al.
Pubblicazione: (2023)
MultiDepth: Multi-Sample Priors for Refining Monocular Metric Depth Estimations in Indoor Scenes
di: Byun, Sanghyun, et al.
Pubblicazione: (2024)
di: Byun, Sanghyun, et al.
Pubblicazione: (2024)
Language as Prior, Vision as Calibration: Metric Scale Recovery for Monocular Depth Estimation
di: Zhan, Mingxia, et al.
Pubblicazione: (2026)
di: Zhan, Mingxia, et al.
Pubblicazione: (2026)
Vision-Language Embodiment for Monocular Depth Estimation
di: Zhang, Jinchang, et al.
Pubblicazione: (2025)
di: Zhang, Jinchang, et al.
Pubblicazione: (2025)
All-day Depth Completion
di: Ezhov, Vadim, et al.
Pubblicazione: (2024)
di: Ezhov, Vadim, et al.
Pubblicazione: (2024)
The Fourth Monocular Depth Estimation Challenge
di: Obukhov, Anton, et al.
Pubblicazione: (2025)
di: Obukhov, Anton, et al.
Pubblicazione: (2025)
Robust Monocular Depth Estimation under Challenging Conditions
di: Gasperini, Stefano, et al.
Pubblicazione: (2023)
di: Gasperini, Stefano, et al.
Pubblicazione: (2023)
Demo-Pose: Depth-Monocular Modality Fusion For Object Pose Estimation
di: Agarwal, Rachit, et al.
Pubblicazione: (2026)
di: Agarwal, Rachit, et al.
Pubblicazione: (2026)
Adaptive Depth-converted-Scale Convolution for Self-supervised Monocular Depth Estimation
di: Gao, Yanbo, et al.
Pubblicazione: (2026)
di: Gao, Yanbo, et al.
Pubblicazione: (2026)
TR2M: Transferring Monocular Relative Depth to Metric Depth with Language Descriptions and Dual-Level Scale-Oriented Contrast
di: Cui, Beilei, et al.
Pubblicazione: (2025)
di: Cui, Beilei, et al.
Pubblicazione: (2025)
DepthDark: Robust Monocular Depth Estimation for Low-Light Environments
di: Zeng, Longjian, et al.
Pubblicazione: (2025)
di: Zeng, Longjian, et al.
Pubblicazione: (2025)
Focusable Monocular Depth Estimation
di: Du, Yuxin, et al.
Pubblicazione: (2026)
di: Du, Yuxin, et al.
Pubblicazione: (2026)
TREND: Unsupervised 3D Representation Learning via Temporal Forecasting for LiDAR Perception
di: Chen, Runjian, et al.
Pubblicazione: (2024)
di: Chen, Runjian, et al.
Pubblicazione: (2024)
ProDepth: Boosting Self-Supervised Multi-Frame Monocular Depth with Probabilistic Fusion
di: Woo, Sungmin, et al.
Pubblicazione: (2024)
di: Woo, Sungmin, et al.
Pubblicazione: (2024)
UniDepth: Universal Monocular Metric Depth Estimation
di: Piccinelli, Luigi, et al.
Pubblicazione: (2024)
di: Piccinelli, Luigi, et al.
Pubblicazione: (2024)
Relative Pose Estimation through Affine Corrections of Monocular Depth Priors
di: Yu, Yifan, et al.
Pubblicazione: (2025)
di: Yu, Yifan, et al.
Pubblicazione: (2025)
Rethinking Transparent Object Grasping: Depth Completion with Monocular Depth Estimation and Instance Mask
di: Cheng, Yaofeng, et al.
Pubblicazione: (2025)
di: Cheng, Yaofeng, et al.
Pubblicazione: (2025)
Depth AnyEvent: A Cross-Modal Distillation Paradigm for Event-Based Monocular Depth Estimation
di: Bartolomei, Luca, et al.
Pubblicazione: (2025)
di: Bartolomei, Luca, et al.
Pubblicazione: (2025)
Documenti analoghi
-
WorDepth: Variational Language Prior for Monocular Depth Estimation
di: Zeng, Ziyao, et al.
Pubblicazione: (2024) -
AugUndo: Scaling Up Augmentations for Monocular Depth Completion and Estimation
di: Wu, Yangchao, et al.
Pubblicazione: (2023) -
On the Viability of Monocular Depth Pre-training for Semantic Segmentation
di: Lao, Dong, et al.
Pubblicazione: (2022) -
Iris: Integrating Language into Diffusion-based Monocular Depth Estimation
di: Zeng, Ziyao, et al.
Pubblicazione: (2024) -
Sub-token ViT Embedding via Stochastic Resonance Transformers
di: Lao, Dong, et al.
Pubblicazione: (2023)