Object-Scene-Camera Decomposition and Recomposition for Data-Efficient Monocular 3D Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Kuang, Zhaonian, Ding, Rui, Yang, Meng, Zheng, Xinhu, Hua, Gang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CoIn3D: Revisiting Configuration-Invariant Multi-Camera 3D Object Detection
by: Kuang, Zhaonian, et al.
Published: (2026)
by: Kuang, Zhaonian, et al.
Published: (2026)
Multi-Modal Decouple and Recouple Network for Robust 3D Object Detection
by: Ding, Rui, et al.
Published: (2026)
by: Ding, Rui, et al.
Published: (2026)
RayD3D: Distilling Depth Knowledge Along the Ray for Robust Multi-View 3D Object Detection
by: Ding, Rui, et al.
Published: (2026)
by: Ding, Rui, et al.
Published: (2026)
Synergistic Perception and Generative Recomposition: A Multi-Agent Orchestration for Expert-Level Building Inspection
by: Zhong, Hui, et al.
Published: (2026)
by: Zhong, Hui, et al.
Published: (2026)
Selective Transfer Learning of Cross-Modality Distillation for Monocular 3D Object Detection
by: Ding, Rui, et al.
Published: (2026)
by: Ding, Rui, et al.
Published: (2026)
UniT: Unified Geometry Learning with Group Autoregressive Transformer
by: Wang, Haotian, et al.
Published: (2026)
by: Wang, Haotian, et al.
Published: (2026)
IROAM: Improving Roadside Monocular 3D Object Detection Learning from Autonomous Vehicle Data Domain
by: Wang, Zhe, et al.
Published: (2025)
by: Wang, Zhe, et al.
Published: (2025)
Scalable Vision-Based 3D Object Detection and Monocular Depth Estimation for Autonomous Driving
by: Liu, Yuxuan
Published: (2024)
by: Liu, Yuxuan
Published: (2024)
Chasing Day and Night: Towards Robust and Efficient All-Day Object Detection Guided by an Event Camera
by: Cao, Jiahang, et al.
Published: (2023)
by: Cao, Jiahang, et al.
Published: (2023)
OmniPose6D: Towards Short-Term Object Pose Tracking in Dynamic Scenes from Monocular RGB
by: Lin, Yunzhi, et al.
Published: (2024)
by: Lin, Yunzhi, et al.
Published: (2024)
Challenges for Monocular 6D Object Pose Estimation in Robotics
by: Thalhammer, Stefan, et al.
Published: (2023)
by: Thalhammer, Stefan, et al.
Published: (2023)
Learning High-resolution Vector Representation from Multi-Camera Images for 3D Object Detection
by: Chen, Zhili, et al.
Published: (2024)
by: Chen, Zhili, et al.
Published: (2024)
SHOW3D: Capturing Scenes of 3D Hands and Objects in the Wild
by: Rim, Patrick, et al.
Published: (2026)
by: Rim, Patrick, et al.
Published: (2026)
MonoDiff9D: Monocular Category-Level 9D Object Pose Estimation via Diffusion Model
by: Liu, Jian, et al.
Published: (2025)
by: Liu, Jian, et al.
Published: (2025)
Label-Efficient 3D Object Detection For Road-Side Units
by: Dao, Minh-Quan, et al.
Published: (2024)
by: Dao, Minh-Quan, et al.
Published: (2024)
Incremental Joint Learning of Depth, Pose and Implicit Scene Representation on Monocular Camera in Large-scale Scenes
by: Deng, Tianchen, et al.
Published: (2024)
by: Deng, Tianchen, et al.
Published: (2024)
Leveraging Previous-Traversal Point Cloud Map Priors for Camera-Based 3D Object Detection and Tracking
by: Käppeler, Markus, et al.
Published: (2026)
by: Käppeler, Markus, et al.
Published: (2026)
Informative Object-centric Next Best View for Object-aware 3D Gaussian Splatting in Cluttered Scenes
by: Jeong, Seunghoon, et al.
Published: (2026)
by: Jeong, Seunghoon, et al.
Published: (2026)
Online 3D Scene Reconstruction Using Neural Object Priors
by: Chabal, Thomas, et al.
Published: (2025)
by: Chabal, Thomas, et al.
Published: (2025)
Perspective-Invariant 3D Object Detection
by: Liang, Ao, et al.
Published: (2025)
by: Liang, Ao, et al.
Published: (2025)
An Efficient LiDAR-Camera Fusion Network for Multi-Class 3D Dynamic Object Detection and Trajectory Prediction
by: He, Yushen, et al.
Published: (2025)
by: He, Yushen, et al.
Published: (2025)
S$^3$-MonoDETR: Supervised Shape&Scale-perceptive Deformable Transformer for Monocular 3D Object Detection
by: He, Xuan, et al.
Published: (2023)
by: He, Xuan, et al.
Published: (2023)
DynamicPose: Real-time and Robust 6D Object Pose Tracking for Fast-Moving Cameras and Objects
by: Liang, Tingbang, et al.
Published: (2025)
by: Liang, Tingbang, et al.
Published: (2025)
MSSF: A 4D Radar and Camera Fusion Framework With Multi-Stage Sampling for 3D Object Detection in Autonomous Driving
by: Liu, Hongsi, et al.
Published: (2024)
by: Liu, Hongsi, et al.
Published: (2024)
Camera Height Doesn't Change: Unsupervised Training for Metric Monocular Road-Scene Depth Estimation
by: Kinoshita, Genki, et al.
Published: (2023)
by: Kinoshita, Genki, et al.
Published: (2023)
Cross-Cluster Shifting for Efficient and Effective 3D Object Detection in Autonomous Driving
by: Chen, Zhili, et al.
Published: (2024)
by: Chen, Zhili, et al.
Published: (2024)
Benchmarking Multi-View BEV Object Detection with Mixed Pinhole and Fisheye Cameras
by: Liu, Xiangzhong, et al.
Published: (2026)
by: Liu, Xiangzhong, et al.
Published: (2026)
CRKD: Enhanced Camera-Radar Object Detection with Cross-modality Knowledge Distillation
by: Zhao, Lingjun, et al.
Published: (2024)
by: Zhao, Lingjun, et al.
Published: (2024)
Cross-Level Sensor Fusion with Object Lists via Transformer for 3D Object Detection
by: Liu, Xiangzhong, et al.
Published: (2025)
by: Liu, Xiangzhong, et al.
Published: (2025)
Robust Fusion of Object-Level V2X for Learned 3D Object Detection
by: Ostendorf, Lukas, et al.
Published: (2026)
by: Ostendorf, Lukas, et al.
Published: (2026)
Sparse3DTrack: Monocular 3D Object Tracking Using Sparse Supervision
by: Gosala, Nikhil, et al.
Published: (2026)
by: Gosala, Nikhil, et al.
Published: (2026)
Self-Supervised Monocular 4D Scene Reconstruction for Egocentric Videos
by: Yuan, Chengbo, et al.
Published: (2024)
by: Yuan, Chengbo, et al.
Published: (2024)
Robot See Robot Do: Imitating Articulated Object Manipulation with Monocular 4D Reconstruction
by: Kerr, Justin, et al.
Published: (2024)
by: Kerr, Justin, et al.
Published: (2024)
DeformGS: Scene Flow in Highly Deformable Scenes for Deformable Object Manipulation
by: Duisterhof, Bardienus P., et al.
Published: (2023)
by: Duisterhof, Bardienus P., et al.
Published: (2023)
MoD-SLAM: Monocular Dense Mapping for Unbounded 3D Scene Reconstruction
by: Zhou, Heng, et al.
Published: (2024)
by: Zhou, Heng, et al.
Published: (2024)
RockTrack: A 3D Robust Multi-Camera-Ken Multi-Object Tracking Framework
by: Li, Xiaoyu, et al.
Published: (2024)
by: Li, Xiaoyu, et al.
Published: (2024)
Improved Scene Landmark Detection for Camera Localization
by: Do, Tien, et al.
Published: (2024)
by: Do, Tien, et al.
Published: (2024)
Rethink 3D Object Detection from Physical World
by: Tanaka, Satoshi, et al.
Published: (2025)
by: Tanaka, Satoshi, et al.
Published: (2025)
FOMO-3D: Using Vision Foundation Models for Long-Tailed 3D Object Detection
by: Yang, Anqi Joyce, et al.
Published: (2026)
by: Yang, Anqi Joyce, et al.
Published: (2026)
SToRe3D: Sparse Token Relevance in ViTs for Efficient Multi-View 3D Object Detection
by: Papais, Sandro, et al.
Published: (2026)
by: Papais, Sandro, et al.
Published: (2026)
Similar Items
-
CoIn3D: Revisiting Configuration-Invariant Multi-Camera 3D Object Detection
by: Kuang, Zhaonian, et al.
Published: (2026) -
Multi-Modal Decouple and Recouple Network for Robust 3D Object Detection
by: Ding, Rui, et al.
Published: (2026) -
RayD3D: Distilling Depth Knowledge Along the Ray for Robust Multi-View 3D Object Detection
by: Ding, Rui, et al.
Published: (2026) -
Synergistic Perception and Generative Recomposition: A Multi-Agent Orchestration for Expert-Level Building Inspection
by: Zhong, Hui, et al.
Published: (2026) -
Selective Transfer Learning of Cross-Modality Distillation for Monocular 3D Object Detection
by: Ding, Rui, et al.
Published: (2026)