Leveraging Automatic CAD Annotations for Supervised Learning in 3D Scene Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Rao, Yuchen, Ainetter, Stefan, Stekovic, Sinisa, Lepetit, Vincent, Fraundorfer, Friedrich |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PyTorchGeoNodes: Enabling Differentiable Shape Programs for 3D Shape Reconstruction
by: Stekovic, Sinisa, et al.
Published: (2024)
by: Stekovic, Sinisa, et al.
Published: (2024)
GuideFlow3D: Optimization-Guided Rectified Flow For Appearance Transfer
by: Sarkar, Sayan Deb, et al.
Published: (2025)
by: Sarkar, Sayan Deb, et al.
Published: (2025)
SAda-Net: A Self-Supervised Adaptive Stereo Estimation CNN For Remote Sensing Image Data
by: Hirner, Dominik, et al.
Published: (2024)
by: Hirner, Dominik, et al.
Published: (2024)
Scene Generation at Absolute Scale: Utilizing Semantic and Geometric Guidance From Text for Accurate and Interpretable 3D Indoor Scene Generation
by: Ainetter, Stefan, et al.
Published: (2026)
by: Ainetter, Stefan, et al.
Published: (2026)
ICG-MVSNet: Learning Intra-view and Cross-view Relationships for Guidance in Multi-View Stereo
by: Hu, Yuxi, et al.
Published: (2025)
by: Hu, Yuxi, et al.
Published: (2025)
Paparazzo: Active Mapping of Moving 3D Objects
by: Allegro, Davide, et al.
Published: (2026)
by: Allegro, Davide, et al.
Published: (2026)
DreamAnywhere: Object-Centric Panoramic 3D Scene Generation
by: Dominici, Edoardo Alberto, et al.
Published: (2025)
by: Dominici, Edoardo Alberto, et al.
Published: (2025)
Gaussian Frosting: Editable Complex Radiance Fields with Real-Time Rendering
by: Guédon, Antoine, et al.
Published: (2024)
by: Guédon, Antoine, et al.
Published: (2024)
TerraSky3D: Multi-View Reconstructions of European Landmarks in 4K
by: D'Urso, Mattia, et al.
Published: (2026)
by: D'Urso, Mattia, et al.
Published: (2026)
Articulation in Prime: Primitive-Based Articulated Object Understanding from a Single Casual Video
by: Artykov, Arslan, et al.
Published: (2026)
by: Artykov, Arslan, et al.
Published: (2026)
EVLoc: Event-based Visual Localization in LiDAR Maps via Event-Depth Registration
by: Chen, Kuangyi, et al.
Published: (2025)
by: Chen, Kuangyi, et al.
Published: (2025)
Mask as Supervision: Leveraging Unified Mask Information for Unsupervised 3D Pose Estimation
by: Yang, Yuchen, et al.
Published: (2023)
by: Yang, Yuchen, et al.
Published: (2023)
Masked Scene Modeling: Narrowing the Gap Between Supervised and Self-Supervised Learning in 3D Scene Understanding
by: Hermosilla, Pedro, et al.
Published: (2025)
by: Hermosilla, Pedro, et al.
Published: (2025)
Alligat0R: Pre-Training Through Co-Visibility Segmentation for Relative Camera Pose Regression
by: Loiseau, Thibaut, et al.
Published: (2025)
by: Loiseau, Thibaut, et al.
Published: (2025)
Correspondences of the Third Kind: Camera Pose Estimation from Object Reflection
by: Yamashita, Kohei, et al.
Published: (2023)
by: Yamashita, Kohei, et al.
Published: (2023)
DiffCAD: Weakly-Supervised Probabilistic CAD Model Retrieval and Alignment from an RGB Image
by: Gao, Daoyi, et al.
Published: (2023)
by: Gao, Daoyi, et al.
Published: (2023)
NextBestPath: Efficient 3D Mapping of Unseen Environments
by: Li, Shiyao, et al.
Published: (2025)
by: Li, Shiyao, et al.
Published: (2025)
3DRS: MLLMs Need 3D-Aware Representation Supervision for Scene Understanding
by: Huang, Xiaohu, et al.
Published: (2025)
by: Huang, Xiaohu, et al.
Published: (2025)
LLaVA$^3$: Representing 3D Scenes like a Cubist Painter to Boost 3D Scene Understanding of VLMs
by: Petit, Doriand, et al.
Published: (2025)
by: Petit, Doriand, et al.
Published: (2025)
DC-Scene: Data-Centric Learning for 3D Scene Understanding
by: Huang, Ting, et al.
Published: (2025)
by: Huang, Ting, et al.
Published: (2025)
${C}^{3}$-GS: Learning Context-aware, Cross-dimension, Cross-scale Feature for Generalizable Gaussian Splatting
by: Hu, Yuxi, et al.
Published: (2025)
by: Hu, Yuxi, et al.
Published: (2025)
LEAR: Learning Edge-Aware Representations for Event-to-LiDAR Localization
by: Chen, Kuangyi, et al.
Published: (2026)
by: Chen, Kuangyi, et al.
Published: (2026)
Towards Foundation Models for 3D Scene Understanding: Instance-Aware Self-Supervised Learning for Point Clouds
by: Yang, Bin, et al.
Published: (2026)
by: Yang, Bin, et al.
Published: (2026)
DenseScan: Advancing 3D Scene Understanding with 2D Dense Annotation
by: Wang, Zirui, et al.
Published: (2025)
by: Wang, Zirui, et al.
Published: (2025)
Leveraging 2D-VLM for Label-Free 3D Segmentation in Large-Scale Outdoor Scene Understanding
by: Nishimura, Toshihiko, et al.
Published: (2026)
by: Nishimura, Toshihiko, et al.
Published: (2026)
ShelfGaussian: Shelf-Supervised Open-Vocabulary Gaussian-based 3D Scene Understanding
by: Zhao, Lingjun, et al.
Published: (2025)
by: Zhao, Lingjun, et al.
Published: (2025)
UrbanCAD: Towards Highly Controllable and Photorealistic 3D Vehicles for Urban Scene Simulation
by: Lu, Yichong, et al.
Published: (2024)
by: Lu, Yichong, et al.
Published: (2024)
Leveraging VLM-Based Pipelines to Annotate 3D Objects
by: Kabra, Rishabh, et al.
Published: (2023)
by: Kabra, Rishabh, et al.
Published: (2023)
BOP-Distrib: Revisiting 6D Pose Estimation Benchmarks for Better Evaluation under Visual Ambiguities
by: Meden, Boris, et al.
Published: (2024)
by: Meden, Boris, et al.
Published: (2024)
GigaPose: Fast and Robust Novel Object Pose Estimation via One Correspondence
by: Nguyen, Van Nguyen, et al.
Published: (2023)
by: Nguyen, Van Nguyen, et al.
Published: (2023)
SceneGPT: A Language Model for 3D Scene Understanding
by: Chandhok, Shivam
Published: (2024)
by: Chandhok, Shivam
Published: (2024)
Scene-R1: Video-Grounded Large Language Models for 3D Scene Reasoning without 3D Annotations
by: Yuan, Zhihao, et al.
Published: (2025)
by: Yuan, Zhihao, et al.
Published: (2025)
sim2art: Accurate Articulated Object Modeling from a Single Video using Synthetic Training Data Only
by: Artykov, Arslan, et al.
Published: (2025)
by: Artykov, Arslan, et al.
Published: (2025)
GenCAD-Self-Repairing: Feasibility Enhancement for 3D CAD Generation
by: Tsuji, Chikaha, et al.
Published: (2025)
by: Tsuji, Chikaha, et al.
Published: (2025)
MAGICIAN: Efficient Long-Term Planning with Imagined Gaussians for Active Mapping
by: Li, Shiyao, et al.
Published: (2026)
by: Li, Shiyao, et al.
Published: (2026)
Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning
by: Fu, Rao, et al.
Published: (2024)
by: Fu, Rao, et al.
Published: (2024)
CAD-Llama: Leveraging Large Language Models for Computer-Aided Design Parametric 3D Model Generation
by: Li, Jiahao, et al.
Published: (2025)
by: Li, Jiahao, et al.
Published: (2025)
Open-Vocabulary vs Supervised Learning Methods for Post-Disaster Visual Scene Understanding
by: Michailidou, Anna, et al.
Published: (2026)
by: Michailidou, Anna, et al.
Published: (2026)
R3DS: Reality-linked 3D Scenes for Panoramic Scene Understanding
by: Wu, Qirui, et al.
Published: (2024)
by: Wu, Qirui, et al.
Published: (2024)
Enhancing Generalizability of Representation Learning for Data-Efficient 3D Scene Understanding
by: Wang, Yunsong, et al.
Published: (2024)
by: Wang, Yunsong, et al.
Published: (2024)
Similar Items
-
PyTorchGeoNodes: Enabling Differentiable Shape Programs for 3D Shape Reconstruction
by: Stekovic, Sinisa, et al.
Published: (2024) -
GuideFlow3D: Optimization-Guided Rectified Flow For Appearance Transfer
by: Sarkar, Sayan Deb, et al.
Published: (2025) -
SAda-Net: A Self-Supervised Adaptive Stereo Estimation CNN For Remote Sensing Image Data
by: Hirner, Dominik, et al.
Published: (2024) -
Scene Generation at Absolute Scale: Utilizing Semantic and Geometric Guidance From Text for Accurate and Interpretable 3D Indoor Scene Generation
by: Ainetter, Stefan, et al.
Published: (2026) -
ICG-MVSNet: Learning Intra-view and Cross-view Relationships for Guidance in Multi-View Stereo
by: Hu, Yuxi, et al.
Published: (2025)