A Systematic Literature Review on Deep Learning-based Depth Estimation in Computer Vision
Fuente:
arXiv
Salvato in:
| Autori principali: | Rohan, Ali, Hasan, Md Junayed, Petrovski, Andrei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Systematic Literature Review of Computer Vision Applications in Robotized Wire Harness Assembly
di: Wang, Hao, et al.
Pubblicazione: (2023)
di: Wang, Hao, et al.
Pubblicazione: (2023)
FuzzRisk: Online Collision Risk Estimation for Autonomous Vehicles based on Depth-Aware Object Detection via Fuzzy Inference
di: Liao, Brian Hsuan-Cheng, et al.
Pubblicazione: (2024)
di: Liao, Brian Hsuan-Cheng, et al.
Pubblicazione: (2024)
GST-VLA: Structured Gaussian Spatial Tokens for 3D Depth-Aware Vision-Language-Action Models
di: Sarowar, Md Selim, et al.
Pubblicazione: (2026)
di: Sarowar, Md Selim, et al.
Pubblicazione: (2026)
An Embedded Real-time Object Alert System for Visually Impaired: A Monocular Depth Estimation based Approach through Computer Vision
di: Anjom, Jareen, et al.
Pubblicazione: (2025)
di: Anjom, Jareen, et al.
Pubblicazione: (2025)
Agricultural Object Detection with You Look Only Once (YOLO) Algorithm: A Bibliometric and Systematic Literature Review
di: Badgujar, Chetan M, et al.
Pubblicazione: (2024)
di: Badgujar, Chetan M, et al.
Pubblicazione: (2024)
Depth Any Camera: Zero-Shot Metric Depth Estimation from Any Camera
di: Guo, Yuliang, et al.
Pubblicazione: (2025)
di: Guo, Yuliang, et al.
Pubblicazione: (2025)
GVDepth: Zero-Shot Monocular Depth Estimation for Ground Vehicles based on Probabilistic Cue Fusion
di: Koledić, Karlo, et al.
Pubblicazione: (2024)
di: Koledić, Karlo, et al.
Pubblicazione: (2024)
EGSA-PT:Edge-Guided Spatial Attention with Progressive Training for Monocular Depth Estimation and Segmentation of Transparent Objects
di: Omotara, Gbenga, et al.
Pubblicazione: (2025)
di: Omotara, Gbenga, et al.
Pubblicazione: (2025)
Bridging Spectral-wise and Multi-spectral Depth Estimation via Geometry-guided Contrastive Learning
di: Shin, Ukcheol, et al.
Pubblicazione: (2025)
di: Shin, Ukcheol, et al.
Pubblicazione: (2025)
Helvipad: A Real-World Dataset for Omnidirectional Stereo Depth Estimation
di: Zayene, Mehdi, et al.
Pubblicazione: (2024)
di: Zayene, Mehdi, et al.
Pubblicazione: (2024)
Towards Sharper Object Boundaries in Self-Supervised Depth Estimation
di: Cecille, Aurélien, et al.
Pubblicazione: (2025)
di: Cecille, Aurélien, et al.
Pubblicazione: (2025)
Adaptive Discrete Disparity Volume for Self-supervised Monocular Depth Estimation
di: Ren, Jianwei
Pubblicazione: (2024)
di: Ren, Jianwei
Pubblicazione: (2024)
Learning Point Cloud Representations with Pose Continuity for Depth-Based Category-Level 6D Object Pose Estimation
di: Li, Zhujun, et al.
Pubblicazione: (2025)
di: Li, Zhujun, et al.
Pubblicazione: (2025)
DepthVision: Enabling Robust Vision-Language Models with GAN-Based LiDAR-to-RGB Synthesis for Autonomous Driving
di: Kirchner, Sven, et al.
Pubblicazione: (2025)
di: Kirchner, Sven, et al.
Pubblicazione: (2025)
DeepKalPose: An Enhanced Deep-Learning Kalman Filter for Temporally Consistent Monocular Vehicle Pose Estimation
di: Di Bella, Leandro, et al.
Pubblicazione: (2024)
di: Di Bella, Leandro, et al.
Pubblicazione: (2024)
ShapeICP: Iterative Category-level Object Pose and Shape Estimation from Depth
di: Zhang, Yihao, et al.
Pubblicazione: (2024)
di: Zhang, Yihao, et al.
Pubblicazione: (2024)
{S\textsuperscript{2}M\textsuperscript{2}}: Scalable Stereo Matching Model for Reliable Depth Estimation
di: Min, Junhong, et al.
Pubblicazione: (2025)
di: Min, Junhong, et al.
Pubblicazione: (2025)
RoboFlamingo-Plus: Fusion of Depth and RGB Perception with Vision-Language Models for Enhanced Robotic Manipulation
di: Wang, Sheng
Pubblicazione: (2025)
di: Wang, Sheng
Pubblicazione: (2025)
Deep Learning Innovations for Underwater Waste Detection: An In-Depth Analysis
di: Walia, Jaskaran Singh, et al.
Pubblicazione: (2024)
di: Walia, Jaskaran Singh, et al.
Pubblicazione: (2024)
MM3DGS SLAM: Multi-modal 3D Gaussian Splatting for SLAM Using Vision, Depth, and Inertial Measurements
di: Sun, Lisong C., et al.
Pubblicazione: (2024)
di: Sun, Lisong C., et al.
Pubblicazione: (2024)
AutoVDC: Automated Vision Data Cleaning Using Vision-Language Models
di: Vasa, Santosh, et al.
Pubblicazione: (2025)
di: Vasa, Santosh, et al.
Pubblicazione: (2025)
4th Workshop on Maritime Computer Vision (MaCVi): Challenge Overview
di: Kiefer, Benjamin, et al.
Pubblicazione: (2026)
di: Kiefer, Benjamin, et al.
Pubblicazione: (2026)
A Deep Learning-based Pest Insect Monitoring System for Ultra-low Power Pocket-sized Drones
di: Crupi, Luca, et al.
Pubblicazione: (2024)
di: Crupi, Luca, et al.
Pubblicazione: (2024)
MetricGold: Leveraging Text-To-Image Latent Diffusion Models for Metric Depth Estimation
di: Shah, Ansh, et al.
Pubblicazione: (2024)
di: Shah, Ansh, et al.
Pubblicazione: (2024)
Overview of Computer Vision Techniques in Robotized Wire Harness Assembly: Current State and Future Opportunities
di: Wang, Hao, et al.
Pubblicazione: (2023)
di: Wang, Hao, et al.
Pubblicazione: (2023)
Depth-aware Fusion Method based on Image and 4D Radar Spectrum for 3D Object Detection
di: Sun, Yue, et al.
Pubblicazione: (2025)
di: Sun, Yue, et al.
Pubblicazione: (2025)
V$^2$-SfMLearner: Learning Monocular Depth and Ego-motion for Multimodal Wireless Capsule Endoscopy
di: Bai, Long, et al.
Pubblicazione: (2024)
di: Bai, Long, et al.
Pubblicazione: (2024)
Vision-based Manipulation of Transparent Plastic Bags in Industrial Setups
di: Adetunji, F., et al.
Pubblicazione: (2024)
di: Adetunji, F., et al.
Pubblicazione: (2024)
Zero-Shot Peg Insertion: Identifying Mating Holes and Estimating SE(2) Poses with Vision-Language Models
di: Yajima, Masaru, et al.
Pubblicazione: (2025)
di: Yajima, Masaru, et al.
Pubblicazione: (2025)
A Large Vision-Language Model based Environment Perception System for Visually Impaired People
di: Chen, Zezhou, et al.
Pubblicazione: (2025)
di: Chen, Zezhou, et al.
Pubblicazione: (2025)
UNIC: Learning Unified Multimodal Extrinsic Contact Estimation
di: Xu, Zhengtong, et al.
Pubblicazione: (2026)
di: Xu, Zhengtong, et al.
Pubblicazione: (2026)
Computer Vision for Multimedia Geolocation in Human Trafficking Investigation: A Systematic Literature Review
di: Bamigbade, Opeyemi, et al.
Pubblicazione: (2024)
di: Bamigbade, Opeyemi, et al.
Pubblicazione: (2024)
MapDream: Task-Driven Map Learning for Vision-Language Navigation
di: Lian, Guoxin, et al.
Pubblicazione: (2026)
di: Lian, Guoxin, et al.
Pubblicazione: (2026)
Robotic Grasping of Harvested Tomato Trusses Using Vision and Online Learning
di: Bent, Luuk van den, et al.
Pubblicazione: (2023)
di: Bent, Luuk van den, et al.
Pubblicazione: (2023)
Reinforcement Learning-Based Monocular Vision Approach for Autonomous UAV Landing
di: Houichime, Tarik, et al.
Pubblicazione: (2025)
di: Houichime, Tarik, et al.
Pubblicazione: (2025)
EvtSlowTV -- A Large and Diverse Dataset for Event-Based Depth Estimation
di: Macaulay, Sadiq Layi, et al.
Pubblicazione: (2025)
di: Macaulay, Sadiq Layi, et al.
Pubblicazione: (2025)
VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving
di: Huang, Zilin, et al.
Pubblicazione: (2024)
di: Huang, Zilin, et al.
Pubblicazione: (2024)
LoopVLA: Learning Sufficiency in Recurrent Refinement for Vision-Language-Action Models
di: Shen, Boyang, et al.
Pubblicazione: (2026)
di: Shen, Boyang, et al.
Pubblicazione: (2026)
Multimodal Object Detection using Depth and Image Data for Manufacturing Parts
di: Mahjourian, Nazanin, et al.
Pubblicazione: (2024)
di: Mahjourian, Nazanin, et al.
Pubblicazione: (2024)
STAR: A Foundation Model-driven Framework for Robust Task Planning and Failure Recovery in Robotic Systems
di: Sakib, Md Sadman, et al.
Pubblicazione: (2025)
di: Sakib, Md Sadman, et al.
Pubblicazione: (2025)
Documenti analoghi
-
A Systematic Literature Review of Computer Vision Applications in Robotized Wire Harness Assembly
di: Wang, Hao, et al.
Pubblicazione: (2023) -
FuzzRisk: Online Collision Risk Estimation for Autonomous Vehicles based on Depth-Aware Object Detection via Fuzzy Inference
di: Liao, Brian Hsuan-Cheng, et al.
Pubblicazione: (2024) -
GST-VLA: Structured Gaussian Spatial Tokens for 3D Depth-Aware Vision-Language-Action Models
di: Sarowar, Md Selim, et al.
Pubblicazione: (2026) -
An Embedded Real-time Object Alert System for Visually Impaired: A Monocular Depth Estimation based Approach through Computer Vision
di: Anjom, Jareen, et al.
Pubblicazione: (2025) -
Agricultural Object Detection with You Look Only Once (YOLO) Algorithm: A Bibliometric and Systematic Literature Review
di: Badgujar, Chetan M, et al.
Pubblicazione: (2024)