WorDepth: Variational Language Prior for Monocular Depth Estimation
Fuente:
arXiv
Salvato in:
| Autori principali: | Zeng, Ziyao, Wang, Daniel, Yang, Fengyu, Park, Hyoungseob, Wu, Yangchao, Soatto, Stefano, Hong, Byung-Woo, Lao, Dong, Wong, Alex |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
RSA: Resolving Scale Ambiguities in Monocular Depth Estimators through Language Descriptions
di: Zeng, Ziyao, et al.
Pubblicazione: (2024)
di: Zeng, Ziyao, et al.
Pubblicazione: (2024)
Iris: Integrating Language into Diffusion-based Monocular Depth Estimation
di: Zeng, Ziyao, et al.
Pubblicazione: (2024)
di: Zeng, Ziyao, et al.
Pubblicazione: (2024)
AugUndo: Scaling Up Augmentations for Monocular Depth Completion and Estimation
di: Wu, Yangchao, et al.
Pubblicazione: (2023)
di: Wu, Yangchao, et al.
Pubblicazione: (2023)
On the Viability of Monocular Depth Pre-training for Semantic Segmentation
di: Lao, Dong, et al.
Pubblicazione: (2022)
di: Lao, Dong, et al.
Pubblicazione: (2022)
ProtoDepth: Unsupervised Continual Depth Completion with Prototypes
di: Rim, Patrick, et al.
Pubblicazione: (2025)
di: Rim, Patrick, et al.
Pubblicazione: (2025)
ETA: Energy-based Test-time Adaptation for Depth Completion
di: Chung, Younjoon, et al.
Pubblicazione: (2025)
di: Chung, Younjoon, et al.
Pubblicazione: (2025)
Extending Foundational Monocular Depth Estimators to Fisheye Cameras with Calibration Tokens
di: Gangopadhyay, Suchisrit, et al.
Pubblicazione: (2025)
di: Gangopadhyay, Suchisrit, et al.
Pubblicazione: (2025)
Test-Time Adaptation for Depth Completion
di: Park, Hyoungseob, et al.
Pubblicazione: (2024)
di: Park, Hyoungseob, et al.
Pubblicazione: (2024)
Nutrition Estimation for Dietary Management: A Transformer Approach with Depth Sensing
di: Kwan, Zhengyi, et al.
Pubblicazione: (2024)
di: Kwan, Zhengyi, et al.
Pubblicazione: (2024)
Sub-token ViT Embedding via Stochastic Resonance Transformers
di: Lao, Dong, et al.
Pubblicazione: (2023)
di: Lao, Dong, et al.
Pubblicazione: (2023)
Radar-Guided Polynomial Fitting for Metric Depth Estimation
di: Rim, Patrick, et al.
Pubblicazione: (2025)
di: Rim, Patrick, et al.
Pubblicazione: (2025)
Extending Depth of Field for Varifocal Multiview Images
di: Li, Zhilong, et al.
Pubblicazione: (2024)
di: Li, Zhilong, et al.
Pubblicazione: (2024)
DepthGait: Multi-Scale Cross-Level Feature Fusion of RGB-Derived Depth and Silhouette Sequences for Robust Gait Recognition
di: Li, Xinzhu, et al.
Pubblicazione: (2025)
di: Li, Xinzhu, et al.
Pubblicazione: (2025)
Single Image Dehazing Using Scene Depth Ordering
di: Ling, Pengyang, et al.
Pubblicazione: (2024)
di: Ling, Pengyang, et al.
Pubblicazione: (2024)
Depth and Image Fusion for Road Obstacle Detection Using Stereo Camera
di: Perezyabov, Oleg, et al.
Pubblicazione: (2025)
di: Perezyabov, Oleg, et al.
Pubblicazione: (2025)
A Sleep Monitoring System Based on Audio, Video and Depth Information
di: Chen, Lyn Chao-ling, et al.
Pubblicazione: (2025)
di: Chen, Lyn Chao-ling, et al.
Pubblicazione: (2025)
Estimating Indoor Scene Depth Maps from Ultrasonic Echoes
di: Honma, Junpei, et al.
Pubblicazione: (2024)
di: Honma, Junpei, et al.
Pubblicazione: (2024)
Test-Time Defense Against Adversarial Attacks via Stochastic Resonance of Latent Ensembles
di: Lao, Dong, et al.
Pubblicazione: (2025)
di: Lao, Dong, et al.
Pubblicazione: (2025)
MultiDepth: Multi-Sample Priors for Refining Monocular Metric Depth Estimations in Indoor Scenes
di: Byun, Sanghyun, et al.
Pubblicazione: (2024)
di: Byun, Sanghyun, et al.
Pubblicazione: (2024)
Doctor Sun: A Bilingual Multimodal Large Language Model for Biomedical AI
di: Xue, Dong, et al.
Pubblicazione: (2025)
di: Xue, Dong, et al.
Pubblicazione: (2025)
Perceptual Depth Quality Assessment of Stereoscopic Omnidirectional Images
di: Zhou, Wei, et al.
Pubblicazione: (2024)
di: Zhou, Wei, et al.
Pubblicazione: (2024)
Diffeomorphic Template Registration for Atmospheric Turbulence Mitigation
di: Lao, Dong, et al.
Pubblicazione: (2024)
di: Lao, Dong, et al.
Pubblicazione: (2024)
Enhancing Multimodal Entity and Relation Extraction with Variational Information Bottleneck
di: Cui, Shiyao, et al.
Pubblicazione: (2023)
di: Cui, Shiyao, et al.
Pubblicazione: (2023)
IBMEA: Exploring Variational Information Bottleneck for Multi-modal Entity Alignment
di: Su, Taoyu, et al.
Pubblicazione: (2024)
di: Su, Taoyu, et al.
Pubblicazione: (2024)
CMATH: Cross-Modality Augmented Transformer with Hierarchical Variational Distillation for Multimodal Emotion Recognition in Conversation
di: Zhu, Xiaofei, et al.
Pubblicazione: (2024)
di: Zhu, Xiaofei, et al.
Pubblicazione: (2024)
UnCLe: Benchmarking Unsupervised Continual Learning for Depth Completion
di: Chen, Xien, et al.
Pubblicazione: (2024)
di: Chen, Xien, et al.
Pubblicazione: (2024)
Why Multi-Interest Fairness Matters: Hypergraph Contrastive Multi-Interest Learning for Fair Conversational Recommender System
di: Zheng, Yongsen, et al.
Pubblicazione: (2025)
di: Zheng, Yongsen, et al.
Pubblicazione: (2025)
Creatively Upscaling Images with Global-Regional Priors
di: Qian, Yurui, et al.
Pubblicazione: (2025)
di: Qian, Yurui, et al.
Pubblicazione: (2025)
NAIMA: Semantics Aware RGB Guided Depth Super-Resolution
di: Nasir, Tayyab, et al.
Pubblicazione: (2026)
di: Nasir, Tayyab, et al.
Pubblicazione: (2026)
Holistic Visual-Textual Sentiment Analysis with Prior Models
di: Chen, Junyu, et al.
Pubblicazione: (2022)
di: Chen, Junyu, et al.
Pubblicazione: (2022)
Natural Language Induced Adversarial Images
di: Zhu, Xiaopei, et al.
Pubblicazione: (2024)
di: Zhu, Xiaopei, et al.
Pubblicazione: (2024)
Self-similarity Prior Distillation for Unsupervised Remote Physiological Measurement
di: Zhang, Xinyu, et al.
Pubblicazione: (2023)
di: Zhang, Xinyu, et al.
Pubblicazione: (2023)
Mono3DVG-EnSD: Enhanced Spatial-aware and Dimension-decoupled Text Encoding for Monocular 3D Visual Grounding
di: Li, Yuzhen, et al.
Pubblicazione: (2025)
di: Li, Yuzhen, et al.
Pubblicazione: (2025)
Enhancing DETRs Variants through Improved Content Query and Similar Query Aggregation
di: Zhang, Yingying, et al.
Pubblicazione: (2024)
di: Zhang, Yingying, et al.
Pubblicazione: (2024)
StableDub: Taming Diffusion Prior for Generalized and Efficient Visual Dubbing
di: Chen, Liyang, et al.
Pubblicazione: (2025)
di: Chen, Liyang, et al.
Pubblicazione: (2025)
Disparity-based Stereo Image Compression with Aligned Cross-View Priors
di: Zhai, Yongqi, et al.
Pubblicazione: (2022)
di: Zhai, Yongqi, et al.
Pubblicazione: (2022)
Efficient Object-centric Representation Learning with Pre-trained Geometric Prior
di: Khac, Phúc H. Le, et al.
Pubblicazione: (2024)
di: Khac, Phúc H. Le, et al.
Pubblicazione: (2024)
Enhancing Interactive Image Retrieval With Query Rewriting Using Large Language Models and Vision Language Models
di: Zhu, Hongyi, et al.
Pubblicazione: (2024)
di: Zhu, Hongyi, et al.
Pubblicazione: (2024)
Towards Point Cloud Compression for Machine Perception: A Simple and Strong Baseline by Learning the Octree Depth Level Predictor
di: Liu, Lei, et al.
Pubblicazione: (2024)
di: Liu, Lei, et al.
Pubblicazione: (2024)
CoPRS: Learning Positional Prior from Chain-of-Thought for Reasoning Segmentation
di: Lu, Zhenyu, et al.
Pubblicazione: (2025)
di: Lu, Zhenyu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
RSA: Resolving Scale Ambiguities in Monocular Depth Estimators through Language Descriptions
di: Zeng, Ziyao, et al.
Pubblicazione: (2024) -
Iris: Integrating Language into Diffusion-based Monocular Depth Estimation
di: Zeng, Ziyao, et al.
Pubblicazione: (2024) -
AugUndo: Scaling Up Augmentations for Monocular Depth Completion and Estimation
di: Wu, Yangchao, et al.
Pubblicazione: (2023) -
On the Viability of Monocular Depth Pre-training for Semantic Segmentation
di: Lao, Dong, et al.
Pubblicazione: (2022) -
ProtoDepth: Unsupervised Continual Depth Completion with Prototypes
di: Rim, Patrick, et al.
Pubblicazione: (2025)