MonoDETR: Depth-guided Transformer for Monocular 3D Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Renrui, Qiu, Han, Wang, Tai, Guo, Ziyu, Tang, Yiwen, Xu, Xuanzhuo, Cui, Ziteng, Qiao, Yu, Gao, Peng, Li, Hongsheng |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
S$^3$-MonoDETR: Supervised Shape&Scale-perceptive Deformable Transformer for Monocular 3D Object Detection
by: He, Xuan, et al.
Published: (2023)
by: He, Xuan, et al.
Published: (2023)
MonoSIM: An open source SIL framework for Ackermann Vehicular Systems with Monocular Vision
by: Rahman, Shantanu, et al.
Published: (2026)
by: Rahman, Shantanu, et al.
Published: (2026)
Monocular Depth Estimation From the Perspective of Feature Restoration: A Diffusion Enhanced Depth Restoration Approach
by: Bai, Huibin, et al.
Published: (2026)
by: Bai, Huibin, et al.
Published: (2026)
CDSE-UNet: Enhancing COVID-19 CT Image Segmentation with Canny Edge Detection and Dual-Path SENet Feature Fusion
by: Ding, Jiao, et al.
Published: (2024)
by: Ding, Jiao, et al.
Published: (2024)
Multiple Prior Representation Learning for Self-Supervised Monocular Depth Estimation via Hybrid Transformer
by: Sun, Guodong, et al.
Published: (2024)
by: Sun, Guodong, et al.
Published: (2024)
LiftFormer: Lifting and Frame Theory Based Monocular Depth Estimation Using Depth and Edge Oriented Subspace Representation
by: Li, Shuai, et al.
Published: (2026)
by: Li, Shuai, et al.
Published: (2026)
Enhanced Encoder-Decoder Architecture for Accurate Monocular Depth Estimation
by: Das, Dabbrata, et al.
Published: (2024)
by: Das, Dabbrata, et al.
Published: (2024)
Global and Local Attention-Based Transformer for Hyperspectral Image Change Detection
by: Wang, Ziyi, et al.
Published: (2024)
by: Wang, Ziyi, et al.
Published: (2024)
SPIdepth: Strengthened Pose Information for Self-supervised Monocular Depth Estimation
by: Lavreniuk, Mykola
Published: (2024)
by: Lavreniuk, Mykola
Published: (2024)
Domain-Transferred Synthetic Data Generation for Improving Monocular Depth Estimation
by: Lee, Seungyeop, et al.
Published: (2024)
by: Lee, Seungyeop, et al.
Published: (2024)
Privacy-Preserving Autoencoder for Collaborative Object Detection
by: Azizian, Bardia, et al.
Published: (2024)
by: Azizian, Bardia, et al.
Published: (2024)
Fifty Years of Object Detection and Recognition from Synthetic Aperture Radar Remote Sensing Imagery: The Road Forward
by: Zhou, Jie, et al.
Published: (2025)
by: Zhou, Jie, et al.
Published: (2025)
V2M4: 4D Mesh Animation Reconstruction from a Single Monocular Video
by: Chen, Jianqi, et al.
Published: (2025)
by: Chen, Jianqi, et al.
Published: (2025)
LiDAR Depth Map Guided Image Compression Model
by: Gnutti, Alessandro, et al.
Published: (2024)
by: Gnutti, Alessandro, et al.
Published: (2024)
RF-DETR for Robust Mitotic Figure Detection: A MIDOG 2025 Track 1 Approach
by: Giedziun, Piotr, et al.
Published: (2025)
by: Giedziun, Piotr, et al.
Published: (2025)
A foundation model for generalizable disease diagnosis in chest X-ray images
by: Xu, Lijian, et al.
Published: (2024)
by: Xu, Lijian, et al.
Published: (2024)
Dataset and Benchmark for Enhancing Critical Retained Foreign Object Detection
by: Wang, Yuli, et al.
Published: (2025)
by: Wang, Yuli, et al.
Published: (2025)
Intermediate Domain-guided Adaptation for Unsupervised Chorioallantoic Membrane Vessel Segmentation
by: Song, Pengwu, et al.
Published: (2025)
by: Song, Pengwu, et al.
Published: (2025)
Feature Compression for Cloud-Edge Multimodal 3D Object Detection
by: Tian, Chongzhen, et al.
Published: (2024)
by: Tian, Chongzhen, et al.
Published: (2024)
LUMEN: Low-light Unified Multi-stage Enhancement Network using depth-guided flash, clustering, and attention-based Transformers
by: Debnath, Bibhabasu, et al.
Published: (2026)
by: Debnath, Bibhabasu, et al.
Published: (2026)
FEFormer: Frequency-enhanced Vision Transformer for Generic Knowledge Extraction and Adaptive Feature Fusion in Volumetric Medical Image Segmentation
by: Yang, Jin, et al.
Published: (2026)
by: Yang, Jin, et al.
Published: (2026)
Exploring Autoregressive Vision Foundation Models for Image Compression
by: Phung, Huu-Tai, et al.
Published: (2025)
by: Phung, Huu-Tai, et al.
Published: (2025)
MaskCRT: Masked Conditional Residual Transformer for Learned Video Compression
by: Chen, Yi-Hsin, et al.
Published: (2023)
by: Chen, Yi-Hsin, et al.
Published: (2023)
Fine-tuned Transformer Models for Breast Cancer Detection and Classification
by: Osman, Showkat, et al.
Published: (2025)
by: Osman, Showkat, et al.
Published: (2025)
MetaFE-DE: Learning Meta Feature Embedding for Depth Estimation from Monocular Endoscopic Images
by: Lu, Dawei, et al.
Published: (2025)
by: Lu, Dawei, et al.
Published: (2025)
RoTIR: Rotation-Equivariant Network and Transformers for Fish Scale Image Registration
by: Wang, Ruixiong, et al.
Published: (2024)
by: Wang, Ruixiong, et al.
Published: (2024)
Streamlined Hybrid Annotation Framework using Scalable Codestream for Bandwidth-Restricted UAV Object Detection
by: Khoury, Karim El, et al.
Published: (2024)
by: Khoury, Karim El, et al.
Published: (2024)
ReRAW: RGB-to-RAW Image Reconstruction via Stratified Sampling for Efficient Object Detection on the Edge
by: Berdan, Radu, et al.
Published: (2025)
by: Berdan, Radu, et al.
Published: (2025)
Point Cloud Feature Coding for Object Detection over an Error-Prone Cloud-Edge Collaborative System
by: Tian, Chongzhen, et al.
Published: (2026)
by: Tian, Chongzhen, et al.
Published: (2026)
CoBEV: Elevating Roadside 3D Object Detection with Depth and Height Complementarity
by: Shi, Hao, et al.
Published: (2023)
by: Shi, Hao, et al.
Published: (2023)
AI-Driven Collaborative Satellite Object Detection for Space Sustainability
by: Hu, Peng, et al.
Published: (2025)
by: Hu, Peng, et al.
Published: (2025)
Color-Guided Flying Pixel Correction in Depth Images
by: Vasudevan, Ekamresh, et al.
Published: (2024)
by: Vasudevan, Ekamresh, et al.
Published: (2024)
Learning Optimal Linear Block Transform by Rate Distortion Minimization
by: Gnutti, Alessandro, et al.
Published: (2024)
by: Gnutti, Alessandro, et al.
Published: (2024)
Control Copy-Paste: Controllable Diffusion-Based Augmentation Method for Remote Sensing Few-Shot Object Detection
by: Liu, Yanxing, et al.
Published: (2025)
by: Liu, Yanxing, et al.
Published: (2025)
Diverse Instance Generation via Diffusion Models for Enhanced Few-Shot Object Detection in Remote Sensing Images
by: Liu, Yanxing, et al.
Published: (2025)
by: Liu, Yanxing, et al.
Published: (2025)
MH-LVC: Multi-Hypothesis Temporal Prediction for Learned Conditional Residual Video Coding
by: Phung, Huu-Tai, et al.
Published: (2025)
by: Phung, Huu-Tai, et al.
Published: (2025)
A Dual-Feature Extractor Framework for Accurate Back Depth and Spine Morphology Estimation from Monocular RGB Images
by: Wei, Yuxin, et al.
Published: (2025)
by: Wei, Yuxin, et al.
Published: (2025)
SaViD: Spectravista Aesthetic Vision Integration for Robust and Discerning 3D Object Detection in Challenging Environments
by: Dam, Tanmoy, et al.
Published: (2025)
by: Dam, Tanmoy, et al.
Published: (2025)
Depth Separable architecture for Sentinel-5P Super-Resolution
by: Ali, Hyam Omar, et al.
Published: (2025)
by: Ali, Hyam Omar, et al.
Published: (2025)
LoLiSRFlow: Joint Single Image Low-light Enhancement and Super-resolution via Cross-scale Transformer-based Conditional Flow
by: Yue, Ziyu, et al.
Published: (2024)
by: Yue, Ziyu, et al.
Published: (2024)
Similar Items
-
S$^3$-MonoDETR: Supervised Shape&Scale-perceptive Deformable Transformer for Monocular 3D Object Detection
by: He, Xuan, et al.
Published: (2023) -
MonoSIM: An open source SIL framework for Ackermann Vehicular Systems with Monocular Vision
by: Rahman, Shantanu, et al.
Published: (2026) -
Monocular Depth Estimation From the Perspective of Feature Restoration: A Diffusion Enhanced Depth Restoration Approach
by: Bai, Huibin, et al.
Published: (2026) -
CDSE-UNet: Enhancing COVID-19 CT Image Segmentation with Canny Edge Detection and Dual-Path SENet Feature Fusion
by: Ding, Jiao, et al.
Published: (2024) -
Multiple Prior Representation Learning for Self-Supervised Monocular Depth Estimation via Hybrid Transformer
by: Sun, Guodong, et al.
Published: (2024)