MoCA3D: Monocular 3D Bounding Box Prediction in the Image Plane
Fuente:
arXiv
Saved in:
| Main Authors: | Jeon, Changwoo, Upadhyay, Rishi, Kadambi, Achuta |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MoCA: Mixture-of-Components Attention for Scalable Compositional 3D Generation
by: Li, Zhiqi, et al.
Published: (2025)
by: Li, Zhiqi, et al.
Published: (2025)
SparseGS: Real-Time 360° Sparse View Synthesis using Gaussian Splatting
by: Xiong, Haolin, et al.
Published: (2023)
by: Xiong, Haolin, et al.
Published: (2023)
SpatialStack: Layered Geometry-Language Fusion for 3D VLM Spatial Reasoning
by: Zhang, Jian, et al.
Published: (2026)
by: Zhang, Jian, et al.
Published: (2026)
MoCA: Identity-Preserving Text-to-Video Generation via Mixture of Cross Attention
by: Xie, Qi, et al.
Published: (2025)
by: Xie, Qi, et al.
Published: (2025)
OBMO: One Bounding Box Multiple Objects for Monocular 3D Object Detection
by: Huang, Chenxi, et al.
Published: (2022)
by: Huang, Chenxi, et al.
Published: (2022)
Solutions to Deepfakes: Can Camera Hardware, Cryptography, and Deep Learning Verify Real Images?
by: Vilesov, Alexander, et al.
Published: (2024)
by: Vilesov, Alexander, et al.
Published: (2024)
MoCA-Video: Motion-Aware Concept Alignment for Consistent Video Editing
by: Zhang, Tong, et al.
Published: (2025)
by: Zhang, Tong, et al.
Published: (2025)
Feature4X: Bridging Any Monocular Video to 4D Agentic AI with Versatile Gaussian Feature Fields
by: Zhou, Shijie, et al.
Published: (2025)
by: Zhou, Shijie, et al.
Published: (2025)
Feature 3DGS: Supercharging 3D Gaussian Splatting to Enable Distilled Feature Fields
by: Zhou, Shijie, et al.
Published: (2023)
by: Zhou, Shijie, et al.
Published: (2023)
Large Spatial Model: End-to-end Unposed Images to Semantic 3D
by: Fan, Zhiwen, et al.
Published: (2024)
by: Fan, Zhiwen, et al.
Published: (2024)
All-day Depth Completion
by: Ezhov, Vadim, et al.
Published: (2024)
by: Ezhov, Vadim, et al.
Published: (2024)
InstantRestore: Single-Step Personalized Face Restoration with Shared-Image Attention
by: Zhang, Howard, et al.
Published: (2024)
by: Zhang, Howard, et al.
Published: (2024)
OpenBox: Annotate Any Bounding Boxes in 3D
by: Lee, In-Jae, et al.
Published: (2025)
by: Lee, In-Jae, et al.
Published: (2025)
MonoPlane: Exploiting Monocular Geometric Cues for Generalizable 3D Plane Reconstruction
by: Zhao, Wang, et al.
Published: (2024)
by: Zhao, Wang, et al.
Published: (2024)
WorldBench: Disambiguating Physics for Diagnostic Evaluation of World Models
by: Upadhyay, Rishi, et al.
Published: (2026)
by: Upadhyay, Rishi, et al.
Published: (2026)
DreamScene360: Unconstrained Text-to-3D Scene Generation with Panoramic Gaussian Splatting
by: Zhou, Shijie, et al.
Published: (2024)
by: Zhou, Shijie, et al.
Published: (2024)
WeatherProof: Leveraging Language Guidance for Semantic Segmentation in Adverse Weather
by: Gella, Blake, et al.
Published: (2024)
by: Gella, Blake, et al.
Published: (2024)
CA-W3D: Leveraging Context-Aware Knowledge for Weakly Supervised Monocular 3D Detection
by: Liu, Chupeng, et al.
Published: (2025)
by: Liu, Chupeng, et al.
Published: (2025)
Rebenchmarking Unsupervised Monocular 3D Occupancy Prediction
by: Guo, Zizhan, et al.
Published: (2026)
by: Guo, Zizhan, et al.
Published: (2026)
Unsupervised Monocular 3D Keypoint Discovery from Multi-View Diffusion Priors
by: Jeon, Subin, et al.
Published: (2025)
by: Jeon, Subin, et al.
Published: (2025)
MoD-SLAM: Monocular Dense Mapping for Unbounded 3D Scene Reconstruction
by: Zhou, Heng, et al.
Published: (2024)
by: Zhou, Heng, et al.
Published: (2024)
Boxer: Robust Lifting of Open-World 2D Bounding Boxes to 3D
by: DeTone, Daniel, et al.
Published: (2026)
by: DeTone, Daniel, et al.
Published: (2026)
CVCP-Fusion: On Implicit Depth Estimation for 3D Bounding Box Prediction
by: Gupta, Pranav, et al.
Published: (2024)
by: Gupta, Pranav, et al.
Published: (2024)
MorphoSim: An Interactive, Controllable, and Editable Language-guided 4D World Simulator
by: He, Xuehai, et al.
Published: (2025)
by: He, Xuehai, et al.
Published: (2025)
Thermal Imaging and Radar for Remote Sleep Monitoring of Breathing and Apnea
by: Del Regno, Kai, et al.
Published: (2024)
by: Del Regno, Kai, et al.
Published: (2024)
MoGA: 3D Generative Avatar Prior for Monocular Gaussian Avatar Reconstruction
by: Dong, Zijian, et al.
Published: (2025)
by: Dong, Zijian, et al.
Published: (2025)
BoxSplitGen: A Generative Model for 3D Part Bounding Boxes in Varying Granularity
by: Koo, Juil, et al.
Published: (2026)
by: Koo, Juil, et al.
Published: (2026)
Unleashing the Power of Chain-of-Prediction for Monocular 3D Object Detection
by: Zhang, Zhihao, et al.
Published: (2025)
by: Zhang, Zhihao, et al.
Published: (2025)
Harnessing Uncertainty-aware Bounding Boxes for Unsupervised 3D Object Detection
by: Zhang, Ruiyang, et al.
Published: (2024)
by: Zhang, Ruiyang, et al.
Published: (2024)
MoBGS: Motion Deblurring Dynamic 3D Gaussian Splatting for Blurry Monocular Video
by: Bui, Minh-Quan Viet, et al.
Published: (2025)
by: Bui, Minh-Quan Viet, et al.
Published: (2025)
4K4DGen: Panoramic 4D Generation at 4K Resolution
by: Li, Renjie, et al.
Published: (2024)
by: Li, Renjie, et al.
Published: (2024)
VFMM3D: Releasing the Potential of Image by Vision Foundation Model for Monocular 3D Object Detection
by: Ding, Bonan, et al.
Published: (2024)
by: Ding, Bonan, et al.
Published: (2024)
Rethinking the Encoding and Annotating of 3D Bounding Box: Corner-Aware 3D Object Detection from Point Clouds
by: Meng, Qinghao, et al.
Published: (2025)
by: Meng, Qinghao, et al.
Published: (2025)
MoCapAnything: Unified 3D Motion Capture for Arbitrary Skeletons from Monocular Videos
by: Gong, Kehong, et al.
Published: (2025)
by: Gong, Kehong, et al.
Published: (2025)
ContactGen: Contact-Guided Interactive 3D Human Generation for Partners
by: Gu, Dongjun, et al.
Published: (2024)
by: Gu, Dongjun, et al.
Published: (2024)
UniK3D: Universal Camera Monocular 3D Estimation
by: Piccinelli, Luigi, et al.
Published: (2025)
by: Piccinelli, Luigi, et al.
Published: (2025)
OAHuman: Occlusion-Aware 3D Human Reconstruction from Monocular Images
by: Yang, Yuanwang, et al.
Published: (2026)
by: Yang, Yuanwang, et al.
Published: (2026)
Generalizing Monocular 3D Object Detection
by: Kumar, Abhinav
Published: (2025)
by: Kumar, Abhinav
Published: (2025)
Weakly Supervised Monocular 3D Detection with a Single-View Image
by: Jiang, Xueying, et al.
Published: (2024)
by: Jiang, Xueying, et al.
Published: (2024)
MonoPlace3D: Learning 3D-Aware Object Placement for 3D Monocular Detection
by: Parihar, Rishubh, et al.
Published: (2025)
by: Parihar, Rishubh, et al.
Published: (2025)
Similar Items
-
MoCA: Mixture-of-Components Attention for Scalable Compositional 3D Generation
by: Li, Zhiqi, et al.
Published: (2025) -
SparseGS: Real-Time 360° Sparse View Synthesis using Gaussian Splatting
by: Xiong, Haolin, et al.
Published: (2023) -
SpatialStack: Layered Geometry-Language Fusion for 3D VLM Spatial Reasoning
by: Zhang, Jian, et al.
Published: (2026) -
MoCA: Identity-Preserving Text-to-Video Generation via Mixture of Cross Attention
by: Xie, Qi, et al.
Published: (2025) -
OBMO: One Bounding Box Multiple Objects for Monocular 3D Object Detection
by: Huang, Chenxi, et al.
Published: (2022)