Deep Learning-based Depth Estimation Methods from Monocular Image and Videos: A Comprehensive Survey
Fuente:
arXiv
Salvato in:
| Autori principali: | Rajapaksha, Uchitha, Sohel, Ferdous, Laga, Hamid, Diepeveen, Dean, Bennamoun, Mohammed |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
di: Raoufi, Behnam, et al.
Pubblicazione: (2025)
di: Raoufi, Behnam, et al.
Pubblicazione: (2025)
DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification
di: Ho, Darryl, et al.
Pubblicazione: (2025)
di: Ho, Darryl, et al.
Pubblicazione: (2025)
Adapting SAM with Dynamic Similarity Graphs for Few-Shot Parameter-Efficient Small Dense Object Detection: A Case Study of Chickpea Pods in Field Conditions
di: Jiang, Xintong, et al.
Pubblicazione: (2025)
di: Jiang, Xintong, et al.
Pubblicazione: (2025)
CG-HOI: Contact-Guided 3D Human-Object Interaction Generation
di: Diller, Christian, et al.
Pubblicazione: (2023)
di: Diller, Christian, et al.
Pubblicazione: (2023)
Sign language recognition based on deep learning and low-cost handcrafted descriptors
di: Carneiro, Alvaro Leandro Cavalcante, et al.
Pubblicazione: (2024)
di: Carneiro, Alvaro Leandro Cavalcante, et al.
Pubblicazione: (2024)
Neuromorphic Monocular Depth Estimation with Uncertainty Modeling
di: Bergkvist, Viktor, et al.
Pubblicazione: (2026)
di: Bergkvist, Viktor, et al.
Pubblicazione: (2026)
Capacity Constraint Analysis Using Object Detection for Smart Manufacturing
di: Ahmad, Hafiz Mughees, et al.
Pubblicazione: (2024)
di: Ahmad, Hafiz Mughees, et al.
Pubblicazione: (2024)
SH17: A Dataset for Human Safety and Personal Protective Equipment Detection in Manufacturing Industry
di: Ahmad, Hafiz Mughees, et al.
Pubblicazione: (2024)
di: Ahmad, Hafiz Mughees, et al.
Pubblicazione: (2024)
Detecting AI-Generated Videos with Spiking Neural Networks
di: Jang, Minsuk, et al.
Pubblicazione: (2026)
di: Jang, Minsuk, et al.
Pubblicazione: (2026)
TAG-Head: Time-Aligned Graph Head for Plug-and-Play Fine-grained Action Recognition
di: Hassan, Imtiaz Ul, et al.
Pubblicazione: (2026)
di: Hassan, Imtiaz Ul, et al.
Pubblicazione: (2026)
High-Frequency Semantics and Geometric Priors for End-to-End Detection Transformers in Challenging UAV Imagery
di: Peng, Hongxing, et al.
Pubblicazione: (2025)
di: Peng, Hongxing, et al.
Pubblicazione: (2025)
PhysicsNeRF: Physics-Guided 3D Reconstruction from Sparse Views
di: Barhdadi, Mohamed Rayan, et al.
Pubblicazione: (2025)
di: Barhdadi, Mohamed Rayan, et al.
Pubblicazione: (2025)
FutureHuman3D: Forecasting Complex Long-Term 3D Human Behavior from Video Observations
di: Diller, Christian, et al.
Pubblicazione: (2022)
di: Diller, Christian, et al.
Pubblicazione: (2022)
Systematic Comparison of Projection Methods for Monocular 3D Human Pose Estimation on Fisheye Images
di: Käs, Stephanie, et al.
Pubblicazione: (2025)
di: Käs, Stephanie, et al.
Pubblicazione: (2025)
Geo2Sound: A Scalable Geo-Aligned Framework for Soundscape Generation from Satellite Imagery
di: Wu, Kunlin, et al.
Pubblicazione: (2026)
di: Wu, Kunlin, et al.
Pubblicazione: (2026)
Decoupling Vision and Language: Codebook Anchored Visual Adaptation
di: Wu, Jason, et al.
Pubblicazione: (2026)
di: Wu, Jason, et al.
Pubblicazione: (2026)
Scene Detection Policies and Keyframe Extraction Strategies for Large-Scale Video Analysis
di: Korolkov, Vasilii
Pubblicazione: (2025)
di: Korolkov, Vasilii
Pubblicazione: (2025)
SpectralCA: Bi-Directional Cross-Attention for Next-Generation UAV Hyperspectral Vision
di: Brovko, D. V.
Pubblicazione: (2025)
di: Brovko, D. V.
Pubblicazione: (2025)
Beyond still images: Temporal features and input variance resilience
di: Fadaei, Amir Hosein, et al.
Pubblicazione: (2023)
di: Fadaei, Amir Hosein, et al.
Pubblicazione: (2023)
FlowDet: Overcoming Perspective and Scale Challenges in Real-Time End-to-End Traffic Detection
di: Wang, Zixing, et al.
Pubblicazione: (2025)
di: Wang, Zixing, et al.
Pubblicazione: (2025)
UGOD: Uncertainty-Guided Differentiable Opacity and Soft Dropout for Enhanced Sparse-View 3DGS
di: Guo, Zhihao, et al.
Pubblicazione: (2025)
di: Guo, Zhihao, et al.
Pubblicazione: (2025)
CARScenes: Semantic VLM Dataset for Safe Autonomous Driving
di: He, Yuankai, et al.
Pubblicazione: (2025)
di: He, Yuankai, et al.
Pubblicazione: (2025)
Pedestrian Detection in Low-Light Conditions: A Comprehensive Survey
di: Ghari, Bahareh, et al.
Pubblicazione: (2024)
di: Ghari, Bahareh, et al.
Pubblicazione: (2024)
CCVA-FL: Cross-Client Variations Adaptive Federated Learning for Medical Imaging
di: Gupta, Sunny, et al.
Pubblicazione: (2024)
di: Gupta, Sunny, et al.
Pubblicazione: (2024)
Mistake Attribution: Fine-Grained Mistake Understanding in Egocentric Videos
di: Li, Yayuan, et al.
Pubblicazione: (2025)
di: Li, Yayuan, et al.
Pubblicazione: (2025)
THIRDEYE: Cue-Aware Monocular Depth Estimation via Brain-Inspired Multi-Stage Fusion
di: Ioan, Calin Teodor
Pubblicazione: (2025)
di: Ioan, Calin Teodor
Pubblicazione: (2025)
WaveMix: A Resource-efficient Neural Network for Image Analysis
di: Jeevan, Pranav, et al.
Pubblicazione: (2022)
di: Jeevan, Pranav, et al.
Pubblicazione: (2022)
UnCageNet: Tracking and Pose Estimation of Caged Animal
di: Dutta, Sayak, et al.
Pubblicazione: (2025)
di: Dutta, Sayak, et al.
Pubblicazione: (2025)
Taming the Tail: Leveraging Asymmetric Loss and Pade Approximation to Overcome Medical Image Long-Tailed Class Imbalance
di: Kashyap, Pankhi, et al.
Pubblicazione: (2024)
di: Kashyap, Pankhi, et al.
Pubblicazione: (2024)
Joint Learning of Depth, Pose, and Local Radiance Field for Large Scale Monocular 3D Reconstruction
di: Syed, Shahram Najam, et al.
Pubblicazione: (2025)
di: Syed, Shahram Najam, et al.
Pubblicazione: (2025)
Motion-Guided Semantic Alignment with Negative Prompts for Zero-Shot Video Action Recognition
di: Wang, Yiming, et al.
Pubblicazione: (2026)
di: Wang, Yiming, et al.
Pubblicazione: (2026)
Lifelong Learning in Vision-Language Models: Enhanced EWC with Cross-Modal Knowledge Retention
di: Durrani, Hamza Ahmed, et al.
Pubblicazione: (2026)
di: Durrani, Hamza Ahmed, et al.
Pubblicazione: (2026)
Interpreting Structured Perturbations in Image Protection Methods for Diffusion Models
di: Martin, Michael R., et al.
Pubblicazione: (2025)
di: Martin, Michael R., et al.
Pubblicazione: (2025)
Foreground Focus: Enhancing Coherence and Fidelity in Camouflaged Image Generation
di: Chen, Pei-Chi, et al.
Pubblicazione: (2025)
di: Chen, Pei-Chi, et al.
Pubblicazione: (2025)
Privacy-Preserving Structureless Visual Localization via Image Obfuscation
di: Panek, Vojtech, et al.
Pubblicazione: (2026)
di: Panek, Vojtech, et al.
Pubblicazione: (2026)
Semi supervised GAN for smart microscopy, fast and data efficient cell cycle classification
di: Manick, Rajeev, et al.
Pubblicazione: (2026)
di: Manick, Rajeev, et al.
Pubblicazione: (2026)
Visible Iris Area as a Quality Metric for Reliable Iris Recognition Under Pupil Dilation and Eyelid Occlusion
di: Pessaud, Jack, et al.
Pubblicazione: (2025)
di: Pessaud, Jack, et al.
Pubblicazione: (2025)
Normalizing Flow-Based Metric for Image Generation
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
di: Jeevan, Pranav, et al.
Pubblicazione: (2024)
Image-Based Leopard Seal Recognition: Approaches and Challenges in Current Automated Systems
di: Salazar, Jorge Yero, et al.
Pubblicazione: (2024)
di: Salazar, Jorge Yero, et al.
Pubblicazione: (2024)
Selection, Not Fusion: Radar-Modulated State Space Models for Radar-Camera Depth Estimation
di: Hou, Zhangcheng, et al.
Pubblicazione: (2026)
di: Hou, Zhangcheng, et al.
Pubblicazione: (2026)
Documenti analoghi
-
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
di: Raoufi, Behnam, et al.
Pubblicazione: (2025) -
DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification
di: Ho, Darryl, et al.
Pubblicazione: (2025) -
Adapting SAM with Dynamic Similarity Graphs for Few-Shot Parameter-Efficient Small Dense Object Detection: A Case Study of Chickpea Pods in Field Conditions
di: Jiang, Xintong, et al.
Pubblicazione: (2025) -
CG-HOI: Contact-Guided 3D Human-Object Interaction Generation
di: Diller, Christian, et al.
Pubblicazione: (2023) -
Sign language recognition based on deep learning and low-cost handcrafted descriptors
di: Carneiro, Alvaro Leandro Cavalcante, et al.
Pubblicazione: (2024)