Hybrid-grained Feature Aggregation with Coarse-to-fine Language Guidance for Self-supervised Monocular Depth Estimation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Wenyao, Liu, Hongsi, Li, Bohan, He, Jiawei, Qi, Zekun, Wang, Yunnan, Zhao, Shengyang, Yu, Xinqiang, Zeng, Wenjun, Jin, Xin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge
von: Zhang, Wenyao, et al.
Veröffentlicht: (2025)
von: Zhang, Wenyao, et al.
Veröffentlicht: (2025)
ReWorld: Multi-Dimensional Reward Modeling for Embodied World Models
von: Peng, Baorui, et al.
Veröffentlicht: (2026)
von: Peng, Baorui, et al.
Veröffentlicht: (2026)
Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining
von: Zhang, Wenyao, et al.
Veröffentlicht: (2026)
von: Zhang, Wenyao, et al.
Veröffentlicht: (2026)
Self-supervised Monocular Depth Estimation with Large Kernel Attention
von: Xiang, Xuezhi, et al.
Veröffentlicht: (2024)
von: Xiang, Xuezhi, et al.
Veröffentlicht: (2024)
FA-Depth: Toward Fast and Accurate Self-supervised Monocular Depth Estimation
von: Wang, Fei, et al.
Veröffentlicht: (2024)
von: Wang, Fei, et al.
Veröffentlicht: (2024)
Adaptive Depth-converted-Scale Convolution for Self-supervised Monocular Depth Estimation
von: Gao, Yanbo, et al.
Veröffentlicht: (2026)
von: Gao, Yanbo, et al.
Veröffentlicht: (2026)
Scene Graph Disentanglement and Composition for Generalizable Complex Image Generation
von: Wang, Yunnan, et al.
Veröffentlicht: (2024)
von: Wang, Yunnan, et al.
Veröffentlicht: (2024)
Adaptive Discrete Disparity Volume for Self-supervised Monocular Depth Estimation
von: Ren, Jianwei
Veröffentlicht: (2024)
von: Ren, Jianwei
Veröffentlicht: (2024)
SPIdepth: Strengthened Pose Information for Self-supervised Monocular Depth Estimation
von: Lavreniuk, Mykola
Veröffentlicht: (2024)
von: Lavreniuk, Mykola
Veröffentlicht: (2024)
BaseBoostDepth: Exploiting Larger Baselines For Self-supervised Monocular Depth Estimation
von: Saunders, Kieran, et al.
Veröffentlicht: (2024)
von: Saunders, Kieran, et al.
Veröffentlicht: (2024)
BoRe-Depth: Self-supervised Monocular Depth Estimation with Boundary Refinement for Embedded Systems
von: Liu, Chang, et al.
Veröffentlicht: (2025)
von: Liu, Chang, et al.
Veröffentlicht: (2025)
$\mathrm{F^2Depth}$: Self-supervised Indoor Monocular Depth Estimation via Optical Flow Consistency and Feature Map Synthesis
von: Guo, Xiaotong, et al.
Veröffentlicht: (2024)
von: Guo, Xiaotong, et al.
Veröffentlicht: (2024)
OmniSpatial: Towards Comprehensive Spatial Reasoning Benchmark for Vision Language Models
von: Jia, Mengdi, et al.
Veröffentlicht: (2025)
von: Jia, Mengdi, et al.
Veröffentlicht: (2025)
Unsupervised Monocular Depth Estimation Based on Hierarchical Feature-Guided Diffusion
von: Liu, Runze, et al.
Veröffentlicht: (2024)
von: Liu, Runze, et al.
Veröffentlicht: (2024)
Deep Neighbor Layer Aggregation for Lightweight Self-Supervised Monocular Depth Estimation
von: Boya, Wang, et al.
Veröffentlicht: (2023)
von: Boya, Wang, et al.
Veröffentlicht: (2023)
Self-supervised Event-based Monocular Depth Estimation using Cross-modal Consistency
von: Zhu, Junyu, et al.
Veröffentlicht: (2024)
von: Zhu, Junyu, et al.
Veröffentlicht: (2024)
Self-supervised Adversarial Training of Monocular Depth Estimation against Physical-World Attacks
von: Cheng, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Cheng, Zhiyuan, et al.
Veröffentlicht: (2024)
Self-supervised Monocular Depth Estimation on Water Scenes via Specular Reflection Prior
von: Lu, Zhengyang, et al.
Veröffentlicht: (2024)
von: Lu, Zhengyang, et al.
Veröffentlicht: (2024)
Intrinsic Image Decomposition for Robust Self-supervised Monocular Depth Estimation on Reflective Surfaces
von: Choi, Wonhyeok, et al.
Veröffentlicht: (2025)
von: Choi, Wonhyeok, et al.
Veröffentlicht: (2025)
SHADeS: Self-supervised Monocular Depth Estimation Through Non-Lambertian Image Decomposition
von: Daher, Rema, et al.
Veröffentlicht: (2025)
von: Daher, Rema, et al.
Veröffentlicht: (2025)
DASP: Self-supervised Nighttime Monocular Depth Estimation with Domain Adaptation of Spatiotemporal Priors
von: Huang, Yiheng, et al.
Veröffentlicht: (2025)
von: Huang, Yiheng, et al.
Veröffentlicht: (2025)
Self-supervised Monocular Depth Estimation Robust to Reflective Surface Leveraged by Triplet Mining
von: Choi, Wonhyeok, et al.
Veröffentlicht: (2025)
von: Choi, Wonhyeok, et al.
Veröffentlicht: (2025)
Hierarchical Temporal Context Learning for Camera-based Semantic Scene Completion
von: Li, Bohan, et al.
Veröffentlicht: (2024)
von: Li, Bohan, et al.
Veröffentlicht: (2024)
Plane2Depth: Hierarchical Adaptive Plane Guidance for Monocular Depth Estimation
von: Liu, Li, et al.
Veröffentlicht: (2024)
von: Liu, Li, et al.
Veröffentlicht: (2024)
Self-supervised Pretraining and Finetuning for Monocular Depth and Visual Odometry
von: Chidlovskii, Boris, et al.
Veröffentlicht: (2024)
von: Chidlovskii, Boris, et al.
Veröffentlicht: (2024)
AVS-Net: Audio-Visual Scale Net for Self-supervised Monocular Metric Depth Estimation
von: Liu, Xiaohu, et al.
Veröffentlicht: (2024)
von: Liu, Xiaohu, et al.
Veröffentlicht: (2024)
Improving Domain Generalization in Self-supervised Monocular Depth Estimation via Stabilized Adversarial Training
von: Yao, Yuanqi, et al.
Veröffentlicht: (2024)
von: Yao, Yuanqi, et al.
Veröffentlicht: (2024)
Index-Assisted Stratified Sampling for Online Aggregation
von: Yu, Yunnan, et al.
Veröffentlicht: (2026)
von: Yu, Yunnan, et al.
Veröffentlicht: (2026)
Graph-based Unsupervised Disentangled Representation Learning via Multimodal Large Language Models
von: Xie, Baao, et al.
Veröffentlicht: (2024)
von: Xie, Baao, et al.
Veröffentlicht: (2024)
Bridging Stereo Geometry and BEV Representation with Reliable Mutual Interaction for Semantic Scene Completion
von: Li, Bohan, et al.
Veröffentlicht: (2023)
von: Li, Bohan, et al.
Veröffentlicht: (2023)
Improving Depth Gradient Continuity in Transformers: A Comparative Study on Monocular Depth Estimation with CNN
von: Yao, Jiawei, et al.
Veröffentlicht: (2023)
von: Yao, Jiawei, et al.
Veröffentlicht: (2023)
Closed-Loop Unsupervised Representation Disentanglement with $β$-VAE Distillation and Diffusion Probabilistic Feedback
von: Jin, Xin, et al.
Veröffentlicht: (2024)
von: Jin, Xin, et al.
Veröffentlicht: (2024)
DexVLG: Dexterous Vision-Language-Grasp Model at Scale
von: He, Jiawei, et al.
Veröffentlicht: (2025)
von: He, Jiawei, et al.
Veröffentlicht: (2025)
Efficient Hybrid CNN-GNN Architecture for Monocular Depth Estimation
von: Narayan, Ishan
Veröffentlicht: (2026)
von: Narayan, Ishan
Veröffentlicht: (2026)
Multiple Prior Representation Learning for Self-Supervised Monocular Depth Estimation via Hybrid Transformer
von: Sun, Guodong, et al.
Veröffentlicht: (2024)
von: Sun, Guodong, et al.
Veröffentlicht: (2024)
Reasoning in Space via Grounding in the World
von: Chen, Yiming, et al.
Veröffentlicht: (2025)
von: Chen, Yiming, et al.
Veröffentlicht: (2025)
NimbleD: Enhancing Self-supervised Monocular Depth Estimation with Pseudo-labels and Large-scale Video Pre-training
von: Luginov, Albert, et al.
Veröffentlicht: (2024)
von: Luginov, Albert, et al.
Veröffentlicht: (2024)
Mono-ViFI: A Unified Learning Framework for Self-supervised Single- and Multi-frame Monocular Depth Estimation
von: Liu, Jinfeng, et al.
Veröffentlicht: (2024)
von: Liu, Jinfeng, et al.
Veröffentlicht: (2024)
WS-SfMLearner: Self-supervised Monocular Depth and Ego-motion Estimation on Surgical Videos with Unknown Camera Parameters
von: Lou, Ange, et al.
Veröffentlicht: (2023)
von: Lou, Ange, et al.
Veröffentlicht: (2023)
EndoMUST: Monocular Depth Estimation for Robotic Endoscopy via End-to-end Multi-step Self-supervised Training
von: Shao, Liangjing, et al.
Veröffentlicht: (2025)
von: Shao, Liangjing, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge
von: Zhang, Wenyao, et al.
Veröffentlicht: (2025) -
ReWorld: Multi-Dimensional Reward Modeling for Embodied World Models
von: Peng, Baorui, et al.
Veröffentlicht: (2026) -
Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining
von: Zhang, Wenyao, et al.
Veröffentlicht: (2026) -
Self-supervised Monocular Depth Estimation with Large Kernel Attention
von: Xiang, Xuezhi, et al.
Veröffentlicht: (2024) -
FA-Depth: Toward Fast and Accurate Self-supervised Monocular Depth Estimation
von: Wang, Fei, et al.
Veröffentlicht: (2024)