VFMM3D: Releasing the Potential of Image by Vision Foundation Model for Monocular 3D Object Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ding, Bonan, Xie, Jin, Nie, Jing, Cao, Jiale, Li, Xuelong, Pang, Yanwei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SSLFusion: Scale & Space Aligned Latent Fusion Model for Multimodal 3D Object Detection
von: Ding, Bonan, et al.
Veröffentlicht: (2025)
von: Ding, Bonan, et al.
Veröffentlicht: (2025)
Transformer-based stereo-aware 3D object detection from binocular images
von: Sun, Hanqing, et al.
Veröffentlicht: (2023)
von: Sun, Hanqing, et al.
Veröffentlicht: (2023)
SNNSIR: A Simple Spiking Neural Network for Stereo Image Restoration
von: Xu, Ronghua, et al.
Veröffentlicht: (2025)
von: Xu, Ronghua, et al.
Veröffentlicht: (2025)
Aerial Monocular 3D Object Detection
von: Hu, Yue, et al.
Veröffentlicht: (2022)
von: Hu, Yue, et al.
Veröffentlicht: (2022)
Open Vocabulary Monocular 3D Object Detection
von: Yao, Jin, et al.
Veröffentlicht: (2024)
von: Yao, Jin, et al.
Veröffentlicht: (2024)
Revisiting Monocular 3D Object Detection with Depth Thickness Field
von: Zhang, Qiude, et al.
Veröffentlicht: (2024)
von: Zhang, Qiude, et al.
Veröffentlicht: (2024)
Deep Intra-Image Contrastive Learning for Weakly Supervised One-Step Person Search
von: Wang, Jiabei, et al.
Veröffentlicht: (2023)
von: Wang, Jiabei, et al.
Veröffentlicht: (2023)
MonoDINO-DETR: Depth-Enhanced Monocular 3D Object Detection Using a Vision Foundation Model
von: Kim, Jihyeok, et al.
Veröffentlicht: (2025)
von: Kim, Jihyeok, et al.
Veröffentlicht: (2025)
Generalizing Monocular 3D Object Detection
von: Kumar, Abhinav
Veröffentlicht: (2025)
von: Kumar, Abhinav
Veröffentlicht: (2025)
Selective Transfer Learning of Cross-Modality Distillation for Monocular 3D Object Detection
von: Ding, Rui, et al.
Veröffentlicht: (2026)
von: Ding, Rui, et al.
Veröffentlicht: (2026)
Simultaneous Multiple Object Detection and Pose Estimation using 3D Model Infusion with Monocular Vision
von: Li, Congliang, et al.
Veröffentlicht: (2022)
von: Li, Congliang, et al.
Veröffentlicht: (2022)
FOMO-3D: Using Vision Foundation Models for Long-Tailed 3D Object Detection
von: Yang, Anqi Joyce, et al.
Veröffentlicht: (2026)
von: Yang, Anqi Joyce, et al.
Veröffentlicht: (2026)
CLIPer: Hierarchically Improving Spatial Representation of CLIP for Open-Vocabulary Semantic Segmentation
von: Sun, Lin, et al.
Veröffentlicht: (2024)
von: Sun, Lin, et al.
Veröffentlicht: (2024)
CLIP-VIS: Adapting CLIP for Open-Vocabulary Video Instance Segmentation
von: Zhu, Wenqi, et al.
Veröffentlicht: (2024)
von: Zhu, Wenqi, et al.
Veröffentlicht: (2024)
Implicit and Explicit Language Guidance for Diffusion-based Visual Perception
von: Wang, Hefeng, et al.
Veröffentlicht: (2024)
von: Wang, Hefeng, et al.
Veröffentlicht: (2024)
SED: A Simple Encoder-Decoder for Open-Vocabulary Semantic Segmentation
von: Xie, Bin, et al.
Veröffentlicht: (2023)
von: Xie, Bin, et al.
Veröffentlicht: (2023)
Object-Scene-Camera Decomposition and Recomposition for Data-Efficient Monocular 3D Object Detection
von: Kuang, Zhaonian, et al.
Veröffentlicht: (2026)
von: Kuang, Zhaonian, et al.
Veröffentlicht: (2026)
Decoupled Pseudo-labeling for Semi-Supervised Monocular 3D Object Detection
von: Zhang, Jiacheng, et al.
Veröffentlicht: (2024)
von: Zhang, Jiacheng, et al.
Veröffentlicht: (2024)
Training an Open-Vocabulary Monocular 3D Object Detection Model without 3D Data
von: Huang, Rui, et al.
Veröffentlicht: (2024)
von: Huang, Rui, et al.
Veröffentlicht: (2024)
Towards Intrinsic-Aware Monocular 3D Object Detection
von: Zhang, Zhihao, et al.
Veröffentlicht: (2026)
von: Zhang, Zhihao, et al.
Veröffentlicht: (2026)
LAM3D: Leveraging Attention for Monocular 3D Object Detection
von: Sas, Diana-Alexandra, et al.
Veröffentlicht: (2024)
von: Sas, Diana-Alexandra, et al.
Veröffentlicht: (2024)
MonoPlace3D: Learning 3D-Aware Object Placement for 3D Monocular Detection
von: Parihar, Rishubh, et al.
Veröffentlicht: (2025)
von: Parihar, Rishubh, et al.
Veröffentlicht: (2025)
3D-MOOD: Lifting 2D to 3D for Monocular Open-Set Object Detection
von: Yang, Yung-Hsu, et al.
Veröffentlicht: (2025)
von: Yang, Yung-Hsu, et al.
Veröffentlicht: (2025)
Scalable Vision-Based 3D Object Detection and Monocular Depth Estimation for Autonomous Driving
von: Liu, Yuxuan
Veröffentlicht: (2024)
von: Liu, Yuxuan
Veröffentlicht: (2024)
iSeg: An Iterative Refinement-based Framework for Training-free Segmentation
von: Sun, Lin, et al.
Veröffentlicht: (2024)
von: Sun, Lin, et al.
Veröffentlicht: (2024)
Weakly Supervised Monocular 3D Detection with a Single-View Image
von: Jiang, Xueying, et al.
Veröffentlicht: (2024)
von: Jiang, Xueying, et al.
Veröffentlicht: (2024)
MonoCD: Monocular 3D Object Detection with Complementary Depths
von: Yan, Longfei, et al.
Veröffentlicht: (2024)
von: Yan, Longfei, et al.
Veröffentlicht: (2024)
Depth-discriminative Metric Learning for Monocular 3D Object Detection
von: Choi, Wonhyeok, et al.
Veröffentlicht: (2024)
von: Choi, Wonhyeok, et al.
Veröffentlicht: (2024)
Unleashing the Power of Chain-of-Prediction for Monocular 3D Object Detection
von: Zhang, Zhihao, et al.
Veröffentlicht: (2025)
von: Zhang, Zhihao, et al.
Veröffentlicht: (2025)
Fully Test-Time Adaptation for Monocular 3D Object Detection
von: Lin, Hongbin, et al.
Veröffentlicht: (2024)
von: Lin, Hongbin, et al.
Veröffentlicht: (2024)
SPAN: Spatial-Projection Alignment for Monocular 3D Object Detection
von: Wang, Yifan, et al.
Veröffentlicht: (2025)
von: Wang, Yifan, et al.
Veröffentlicht: (2025)
MonoCLUE : Object-Aware Clustering Enhances Monocular 3D Object Detection
von: Yang, Sunghun, et al.
Veröffentlicht: (2025)
von: Yang, Sunghun, et al.
Veröffentlicht: (2025)
OBMO: One Bounding Box Multiple Objects for Monocular 3D Object Detection
von: Huang, Chenxi, et al.
Veröffentlicht: (2022)
von: Huang, Chenxi, et al.
Veröffentlicht: (2022)
Sub-Image Recapture for Multi-View 3D Reconstruction
von: Wang, Yanwei
Veröffentlicht: (2025)
von: Wang, Yanwei
Veröffentlicht: (2025)
You Only Look Bottom-Up for Monocular 3D Object Detection
von: Xiong, Kaixin, et al.
Veröffentlicht: (2024)
von: Xiong, Kaixin, et al.
Veröffentlicht: (2024)
Difficulty-Aware Label-Guided Denoising for Monocular 3D Object Detection
von: Lee, Soyul, et al.
Veröffentlicht: (2025)
von: Lee, Soyul, et al.
Veröffentlicht: (2025)
MonoSAOD: Monocular 3D Object Detection with Sparsely Annotated Label
von: Jung, Junyoung, et al.
Veröffentlicht: (2026)
von: Jung, Junyoung, et al.
Veröffentlicht: (2026)
MonoWAD: Weather-Adaptive Diffusion Model for Robust Monocular 3D Object Detection
von: Oh, Youngmin, et al.
Veröffentlicht: (2024)
von: Oh, Youngmin, et al.
Veröffentlicht: (2024)
Learning Geometry-Guided Depth via Projective Modeling for Monocular 3D Object Detection
von: Zhang, Yinmin, et al.
Veröffentlicht: (2021)
von: Zhang, Yinmin, et al.
Veröffentlicht: (2021)
RaGS: Unleashing 3D Gaussian Splatting from 4D Radar and Monocular Cues for 3D Object Detection
von: Bai, Xiaokai, et al.
Veröffentlicht: (2025)
von: Bai, Xiaokai, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SSLFusion: Scale & Space Aligned Latent Fusion Model for Multimodal 3D Object Detection
von: Ding, Bonan, et al.
Veröffentlicht: (2025) -
Transformer-based stereo-aware 3D object detection from binocular images
von: Sun, Hanqing, et al.
Veröffentlicht: (2023) -
SNNSIR: A Simple Spiking Neural Network for Stereo Image Restoration
von: Xu, Ronghua, et al.
Veröffentlicht: (2025) -
Aerial Monocular 3D Object Detection
von: Hu, Yue, et al.
Veröffentlicht: (2022) -
Open Vocabulary Monocular 3D Object Detection
von: Yao, Jin, et al.
Veröffentlicht: (2024)