D4D: An RGBD diffusion model to boost monocular depth estimation
Fuente:
arXiv
Guardado en:
| Autores principales: | Papa, L., Russo, P., Amerini, I. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
METER: a mobile vision transformer architecture for monocular depth estimation
por: Papa, L., et al.
Publicado: (2024)
por: Papa, L., et al.
Publicado: (2024)
A survey on efficient vision transformers: algorithms, techniques, and performance benchmarking
por: Papa, Lorenzo, et al.
Publicado: (2023)
por: Papa, Lorenzo, et al.
Publicado: (2023)
DepthFake: a depth-based strategy for detecting Deepfake videos
por: Maiano, Luca, et al.
Publicado: (2022)
por: Maiano, Luca, et al.
Publicado: (2022)
Shedding Light on Depth: Explainability Assessment in Monocular Depth Estimation
por: Cirillo, Lorenzo, et al.
Publicado: (2025)
por: Cirillo, Lorenzo, et al.
Publicado: (2025)
FMPose3D: monocular 3D pose estimation via flow matching
por: Wang, Ti, et al.
Publicado: (2026)
por: Wang, Ti, et al.
Publicado: (2026)
VidCLearn: A Continual Learning Approach for Text-to-Video Generation
por: Zanchetta, Luca, et al.
Publicado: (2025)
por: Zanchetta, Luca, et al.
Publicado: (2025)
EC-Depth: Exploring the consistency of self-supervised monocular depth estimation in challenging scenes
por: Song, Ziyang, et al.
Publicado: (2023)
por: Song, Ziyang, et al.
Publicado: (2023)
Extended monocular 3D imaging
por: Shen, Zicheng, et al.
Publicado: (2025)
por: Shen, Zicheng, et al.
Publicado: (2025)
On the impact of key design aspects in simulated Hybrid Quantum Neural Networks for Earth Observation
por: Papa, Lorenzo, et al.
Publicado: (2024)
por: Papa, Lorenzo, et al.
Publicado: (2024)
Spiking monocular event based 6D pose estimation for space application
por: Courtois, Jonathan, et al.
Publicado: (2025)
por: Courtois, Jonathan, et al.
Publicado: (2025)
FitDiff: Robust monocular 3D facial shape and reflectance estimation using Diffusion Models
por: Galanakis, Stathis, et al.
Publicado: (2023)
por: Galanakis, Stathis, et al.
Publicado: (2023)
3D Human Mesh Estimation from Single View RGBD
por: Suat, Ozhan, et al.
Publicado: (2025)
por: Suat, Ozhan, et al.
Publicado: (2025)
Diffusion Models for Earth Observation Use-cases: from cloud removal to urban change detection
por: Sanguigni, Fulvio, et al.
Publicado: (2023)
por: Sanguigni, Fulvio, et al.
Publicado: (2023)
R3ST: A Synthetic 3D Dataset With Realistic Trajectories
por: Teglia, Simone, et al.
Publicado: (2025)
por: Teglia, Simone, et al.
Publicado: (2025)
Generalizing monocular colonoscopy image depth estimation by uncertainty-based global and local fusion network
por: Du, Sijia, et al.
Publicado: (2024)
por: Du, Sijia, et al.
Publicado: (2024)
KBody: Towards general, robust, and aligned monocular whole-body estimation
por: Zioulis, Nikolaos, et al.
Publicado: (2023)
por: Zioulis, Nikolaos, et al.
Publicado: (2023)
RGBD-Glue: General Feature Combination for Robust RGB-D Point Cloud Registration
por: Chen, Congjia, et al.
Publicado: (2024)
por: Chen, Congjia, et al.
Publicado: (2024)
COSMU: Complete 3D human shape from monocular unconstrained images
por: Pesavento, Marco, et al.
Publicado: (2024)
por: Pesavento, Marco, et al.
Publicado: (2024)
RGBD GS-ICP SLAM
por: Ha, Seongbo, et al.
Publicado: (2024)
por: Ha, Seongbo, et al.
Publicado: (2024)
RGBD Objects in the Wild: Scaling Real-World 3D Object Learning from RGB-D Videos
por: Xia, Hongchi, et al.
Publicado: (2024)
por: Xia, Hongchi, et al.
Publicado: (2024)
Markerless Stride Length estimation in Athletic using Pose Estimation with monocular vision
por: Skorupski, Patryk, et al.
Publicado: (2025)
por: Skorupski, Patryk, et al.
Publicado: (2025)
Implicit Event-RGBD Neural SLAM
por: Qu, Delin, et al.
Publicado: (2023)
por: Qu, Delin, et al.
Publicado: (2023)
VTGaussian-SLAM: RGBD SLAM for Large Scale Scenes with Splatting View-Tied 3D Gaussians
por: Hu, Pengchong, et al.
Publicado: (2025)
por: Hu, Pengchong, et al.
Publicado: (2025)
WildLIFT: Lifting monocular drone video to 3D for species-agnostic wildlife monitoring
por: Shukla, Vandita, et al.
Publicado: (2026)
por: Shukla, Vandita, et al.
Publicado: (2026)
Enhancing Ground-to-Aerial Image Matching for Visual Misinformation Detection Using Semantic Segmentation
por: Mule, Emanuele, et al.
Publicado: (2025)
por: Mule, Emanuele, et al.
Publicado: (2025)
Geometry Depth Consistency in RGBD Relative Pose Estimation
por: Kumar, Sourav, et al.
Publicado: (2024)
por: Kumar, Sourav, et al.
Publicado: (2024)
Osmosis: RGBD Diffusion Prior for Underwater Image Restoration
por: Nathan, Opher Bar, et al.
Publicado: (2024)
por: Nathan, Opher Bar, et al.
Publicado: (2024)
DFormer: Rethinking RGBD Representation Learning for Semantic Segmentation
por: Yin, Bowen, et al.
Publicado: (2023)
por: Yin, Bowen, et al.
Publicado: (2023)
Forest canopy height estimation from satellite RGB imagery using large-scale airborne LiDAR-derived training data and monocular depth estimation
por: Lai, Yongkang, et al.
Publicado: (2026)
por: Lai, Yongkang, et al.
Publicado: (2026)
DFormerv2: Geometry Self-Attention for RGBD Semantic Segmentation
por: Yin, Bo-Wen, et al.
Publicado: (2025)
por: Yin, Bo-Wen, et al.
Publicado: (2025)
MonoVisual3DFilter: 3D tomatoes' localisation with monocular cameras using histogram filters
por: Magalhães, Sandro Costa, et al.
Publicado: (2023)
por: Magalhães, Sandro Costa, et al.
Publicado: (2023)
DRSM: efficient neural 4d decomposition for dynamic reconstruction in stationary monocular cameras
por: Xie, Weixing, et al.
Publicado: (2024)
por: Xie, Weixing, et al.
Publicado: (2024)
Z-SASLM: Zero-Shot Style-Aligned SLI Blending Latent Manipulation
por: Borgi, Alessio, et al.
Publicado: (2025)
por: Borgi, Alessio, et al.
Publicado: (2025)
Enhancing Abnormality Identification: Robust Out-of-Distribution Strategies for Deepfake Detection
por: Maiano, Luca, et al.
Publicado: (2025)
por: Maiano, Luca, et al.
Publicado: (2025)
LADLE-MM: Limited Annotation based Detector with Learned Ensembles for Multimodal Misinformation
por: Cardullo, Daniele, et al.
Publicado: (2025)
por: Cardullo, Daniele, et al.
Publicado: (2025)
Digging into contrastive learning for robust depth estimation with diffusion models
por: Wang, Jiyuan, et al.
Publicado: (2024)
por: Wang, Jiyuan, et al.
Publicado: (2024)
Self-localization on a 3D map by fusing global and local features from a monocular camera
por: Kikuchi, Satoshi, et al.
Publicado: (2025)
por: Kikuchi, Satoshi, et al.
Publicado: (2025)
PoCo: Point Context Cluster for RGBD Indoor Place Recognition
por: Liang, Jing, et al.
Publicado: (2024)
por: Liang, Jing, et al.
Publicado: (2024)
Stroke3D: Lifting 2D strokes into rigged 3D model via latent diffusion models
por: Zhao, Ruisi, et al.
Publicado: (2026)
por: Zhao, Ruisi, et al.
Publicado: (2026)
SimpleDepthPose: Fast and Reliable Human Pose Estimation with RGBD-Images
por: Bermuth, Daniel, et al.
Publicado: (2025)
por: Bermuth, Daniel, et al.
Publicado: (2025)
Ejemplares similares
-
METER: a mobile vision transformer architecture for monocular depth estimation
por: Papa, L., et al.
Publicado: (2024) -
A survey on efficient vision transformers: algorithms, techniques, and performance benchmarking
por: Papa, Lorenzo, et al.
Publicado: (2023) -
DepthFake: a depth-based strategy for detecting Deepfake videos
por: Maiano, Luca, et al.
Publicado: (2022) -
Shedding Light on Depth: Explainability Assessment in Monocular Depth Estimation
por: Cirillo, Lorenzo, et al.
Publicado: (2025) -
FMPose3D: monocular 3D pose estimation via flow matching
por: Wang, Ti, et al.
Publicado: (2026)