3D-MOOD: Lifting 2D to 3D for Monocular Open-Set Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Yung-Hsu, Piccinelli, Luigi, Segu, Mattia, Li, Siyuan, Huang, Rui, Fu, Yuqian, Pollefeys, Marc, Blum, Hermann, Bauer, Zuria |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
UniK3D: Universal Camera Monocular 3D Estimation
by: Piccinelli, Luigi, et al.
Published: (2025)
by: Piccinelli, Luigi, et al.
Published: (2025)
Samba: Synchronized Set-of-Sequences Modeling for Multiple Object Tracking
by: Segu, Mattia, et al.
Published: (2024)
by: Segu, Mattia, et al.
Published: (2024)
Lost & Found: Tracking Changes from Egocentric Observations in 3D Dynamic Scene Graphs
by: Behrens, Tjark, et al.
Published: (2024)
by: Behrens, Tjark, et al.
Published: (2024)
LeAD-M3D: Leveraging Asymmetric Distillation for Real-Time Monocular 3D Detection
by: Meier, Johannes, et al.
Published: (2025)
by: Meier, Johannes, et al.
Published: (2025)
UniDepth: Universal Monocular Metric Depth Estimation
by: Piccinelli, Luigi, et al.
Published: (2024)
by: Piccinelli, Luigi, et al.
Published: (2024)
Articulated 3D Scene Graphs for Open-World Mobile Manipulation
by: Büchner, Martin, et al.
Published: (2026)
by: Büchner, Martin, et al.
Published: (2026)
UniDepthV2: Universal Monocular Metric Depth Estimation Made Simpler
by: Piccinelli, Luigi, et al.
Published: (2025)
by: Piccinelli, Luigi, et al.
Published: (2025)
FunFact: Building Probabilistic Functional 3D Scene Graphs via Factor-Graph Reasoning
by: Fu, Zhengyu, et al.
Published: (2026)
by: Fu, Zhengyu, et al.
Published: (2026)
Spot-Compose: A Framework for Open-Vocabulary Object Retrieval and Drawer Manipulation in Point Clouds
by: Lemke, Oliver, et al.
Published: (2024)
by: Lemke, Oliver, et al.
Published: (2024)
SLAck: Semantic, Location, and Appearance Aware Open-Vocabulary Tracking
by: Li, Siyuan, et al.
Published: (2024)
by: Li, Siyuan, et al.
Published: (2024)
SpotLight: Robotic Scene Understanding through Interaction and Affordance Detection
by: Engelbracht, Tim, et al.
Published: (2024)
by: Engelbracht, Tim, et al.
Published: (2024)
OpenDAS: Open-Vocabulary Domain Adaptation for 2D and 3D Segmentation
by: Yilmaz, Gonca, et al.
Published: (2024)
by: Yilmaz, Gonca, et al.
Published: (2024)
Walker: Self-supervised Multiple Object Tracking by Walking on Temporal Appearance Graphs
by: Segu, Mattia, et al.
Published: (2024)
by: Segu, Mattia, et al.
Published: (2024)
Video Perception Models for 3D Scene Synthesis
by: Huang, Rui, et al.
Published: (2025)
by: Huang, Rui, et al.
Published: (2025)
Open Vocabulary Monocular 3D Object Detection
by: Yao, Jin, et al.
Published: (2024)
by: Yao, Jin, et al.
Published: (2024)
Geometry‐Guided Transformer for Monocular 3D Object Detection
by: Man Zhang, et al.
Published: (2025)
by: Man Zhang, et al.
Published: (2025)
Training an Open-Vocabulary Monocular 3D Object Detection Model without 3D Data
by: Huang, Rui, et al.
Published: (2024)
by: Huang, Rui, et al.
Published: (2024)
MaRINeR: Enhancing Novel Views by Matching Rendered Images with Nearby References
by: Bösiger, Lukas, et al.
Published: (2024)
by: Bösiger, Lukas, et al.
Published: (2024)
ARKit LabelMaker: A New Scale for Indoor 3D Scene Understanding
by: Ji, Guangda, et al.
Published: (2024)
by: Ji, Guangda, et al.
Published: (2024)
CR3DT: Camera-RADAR Fusion for 3D Detection and Tracking
by: Baumann, Nicolas, et al.
Published: (2024)
by: Baumann, Nicolas, et al.
Published: (2024)
Matching Anything by Segmenting Anything
by: Li, Siyuan, et al.
Published: (2024)
by: Li, Siyuan, et al.
Published: (2024)
Spot-On: A Mixed Reality Interface for Multi-Robot Cooperation
by: Engelbracht, Tim, et al.
Published: (2025)
by: Engelbracht, Tim, et al.
Published: (2025)
Hoi! - A Multimodal Dataset for Force-Grounded, Cross-View Articulated Manipulation
by: Engelbracht, Tim, et al.
Published: (2025)
by: Engelbracht, Tim, et al.
Published: (2025)
Loop Closure from Two Views: Revisiting PGO for Scalable Trajectory Estimation through Monocular Priors
by: Lim, Tian Yi, et al.
Published: (2025)
by: Lim, Tian Yi, et al.
Published: (2025)
Generalizing Monocular 3D Object Detection
by: Kumar, Abhinav
Published: (2025)
by: Kumar, Abhinav
Published: (2025)
Aerial Monocular 3D Object Detection
by: Hu, Yue, et al.
Published: (2022)
by: Hu, Yue, et al.
Published: (2022)
HoloSpot: Intuitive Object Manipulation via Mixed Reality Drag-and-Drop
by: Garcia, Pablo Soler, et al.
Published: (2024)
by: Garcia, Pablo Soler, et al.
Published: (2024)
Memory Over Maps: 3D Object Localization Without Reconstruction
by: Zhou, Rui, et al.
Published: (2026)
by: Zhou, Rui, et al.
Published: (2026)
Object-X: Learning to Reconstruct Multi-Modal 3D Object Representations
by: Di Lorenzo, Gaia, et al.
Published: (2025)
by: Di Lorenzo, Gaia, et al.
Published: (2025)
LAM3D: Leveraging Attention for Monocular 3D Object Detection
by: Sas, Diana-Alexandra, et al.
Published: (2024)
by: Sas, Diana-Alexandra, et al.
Published: (2024)
Know Your Neighbors: Improving Single-View Reconstruction via Spatial Vision-Language Reasoning
by: Li, Rui, et al.
Published: (2024)
by: Li, Rui, et al.
Published: (2024)
Mocap-2-to-3: Multi-view Lifting for Monocular Motion Recovery with 2D Pretraining
by: Wang, Zhumei, et al.
Published: (2025)
by: Wang, Zhumei, et al.
Published: (2025)
OpenNeRF: Open Set 3D Neural Scene Segmentation with Pixel-Wise Features and Rendered Novel Views
by: Engelmann, Francis, et al.
Published: (2024)
by: Engelmann, Francis, et al.
Published: (2024)
NeRFmentation: NeRF-based Augmentation for Monocular Depth Estimation
by: Feldmann, Casimir, et al.
Published: (2024)
by: Feldmann, Casimir, et al.
Published: (2024)
Selective Transfer Learning of Cross-Modality Distillation for Monocular 3D Object Detection
by: Ding, Rui, et al.
Published: (2026)
by: Ding, Rui, et al.
Published: (2026)
Object-Scene-Camera Decomposition and Recomposition for Data-Efficient Monocular 3D Object Detection
by: Kuang, Zhaonian, et al.
Published: (2026)
by: Kuang, Zhaonian, et al.
Published: (2026)
MonoPlace3D: Learning 3D-Aware Object Placement for 3D Monocular Detection
by: Parihar, Rishubh, et al.
Published: (2025)
by: Parihar, Rishubh, et al.
Published: (2025)
OpenFrontier: General Navigation with Visual-Language Grounded Frontiers
by: Padilla-Cerdio, Esteban, et al.
Published: (2026)
by: Padilla-Cerdio, Esteban, et al.
Published: (2026)
OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations
by: Hsu, Peng-Hao, et al.
Published: (2025)
by: Hsu, Peng-Hao, et al.
Published: (2025)
Towards Intrinsic-Aware Monocular 3D Object Detection
by: Zhang, Zhihao, et al.
Published: (2026)
by: Zhang, Zhihao, et al.
Published: (2026)
Similar Items
-
UniK3D: Universal Camera Monocular 3D Estimation
by: Piccinelli, Luigi, et al.
Published: (2025) -
Samba: Synchronized Set-of-Sequences Modeling for Multiple Object Tracking
by: Segu, Mattia, et al.
Published: (2024) -
Lost & Found: Tracking Changes from Egocentric Observations in 3D Dynamic Scene Graphs
by: Behrens, Tjark, et al.
Published: (2024) -
LeAD-M3D: Leveraging Asymmetric Distillation for Real-Time Monocular 3D Detection
by: Meier, Johannes, et al.
Published: (2025) -
UniDepth: Universal Monocular Metric Depth Estimation
by: Piccinelli, Luigi, et al.
Published: (2024)