Leveraging Automatic CAD Annotations for Supervised Learning in 3D Scene Understanding
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Rao, Yuchen, Ainetter, Stefan, Stekovic, Sinisa, Lepetit, Vincent, Fraundorfer, Friedrich |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
PyTorchGeoNodes: Enabling Differentiable Shape Programs for 3D Shape Reconstruction
par: Stekovic, Sinisa, et autres
Publié: (2024)
par: Stekovic, Sinisa, et autres
Publié: (2024)
GuideFlow3D: Optimization-Guided Rectified Flow For Appearance Transfer
par: Sarkar, Sayan Deb, et autres
Publié: (2025)
par: Sarkar, Sayan Deb, et autres
Publié: (2025)
SAda-Net: A Self-Supervised Adaptive Stereo Estimation CNN For Remote Sensing Image Data
par: Hirner, Dominik, et autres
Publié: (2024)
par: Hirner, Dominik, et autres
Publié: (2024)
Scene Generation at Absolute Scale: Utilizing Semantic and Geometric Guidance From Text for Accurate and Interpretable 3D Indoor Scene Generation
par: Ainetter, Stefan, et autres
Publié: (2026)
par: Ainetter, Stefan, et autres
Publié: (2026)
ICG-MVSNet: Learning Intra-view and Cross-view Relationships for Guidance in Multi-View Stereo
par: Hu, Yuxi, et autres
Publié: (2025)
par: Hu, Yuxi, et autres
Publié: (2025)
Paparazzo: Active Mapping of Moving 3D Objects
par: Allegro, Davide, et autres
Publié: (2026)
par: Allegro, Davide, et autres
Publié: (2026)
DreamAnywhere: Object-Centric Panoramic 3D Scene Generation
par: Dominici, Edoardo Alberto, et autres
Publié: (2025)
par: Dominici, Edoardo Alberto, et autres
Publié: (2025)
Gaussian Frosting: Editable Complex Radiance Fields with Real-Time Rendering
par: Guédon, Antoine, et autres
Publié: (2024)
par: Guédon, Antoine, et autres
Publié: (2024)
TerraSky3D: Multi-View Reconstructions of European Landmarks in 4K
par: D'Urso, Mattia, et autres
Publié: (2026)
par: D'Urso, Mattia, et autres
Publié: (2026)
Articulation in Prime: Primitive-Based Articulated Object Understanding from a Single Casual Video
par: Artykov, Arslan, et autres
Publié: (2026)
par: Artykov, Arslan, et autres
Publié: (2026)
EVLoc: Event-based Visual Localization in LiDAR Maps via Event-Depth Registration
par: Chen, Kuangyi, et autres
Publié: (2025)
par: Chen, Kuangyi, et autres
Publié: (2025)
Mask as Supervision: Leveraging Unified Mask Information for Unsupervised 3D Pose Estimation
par: Yang, Yuchen, et autres
Publié: (2023)
par: Yang, Yuchen, et autres
Publié: (2023)
Masked Scene Modeling: Narrowing the Gap Between Supervised and Self-Supervised Learning in 3D Scene Understanding
par: Hermosilla, Pedro, et autres
Publié: (2025)
par: Hermosilla, Pedro, et autres
Publié: (2025)
Alligat0R: Pre-Training Through Co-Visibility Segmentation for Relative Camera Pose Regression
par: Loiseau, Thibaut, et autres
Publié: (2025)
par: Loiseau, Thibaut, et autres
Publié: (2025)
Correspondences of the Third Kind: Camera Pose Estimation from Object Reflection
par: Yamashita, Kohei, et autres
Publié: (2023)
par: Yamashita, Kohei, et autres
Publié: (2023)
DiffCAD: Weakly-Supervised Probabilistic CAD Model Retrieval and Alignment from an RGB Image
par: Gao, Daoyi, et autres
Publié: (2023)
par: Gao, Daoyi, et autres
Publié: (2023)
NextBestPath: Efficient 3D Mapping of Unseen Environments
par: Li, Shiyao, et autres
Publié: (2025)
par: Li, Shiyao, et autres
Publié: (2025)
3DRS: MLLMs Need 3D-Aware Representation Supervision for Scene Understanding
par: Huang, Xiaohu, et autres
Publié: (2025)
par: Huang, Xiaohu, et autres
Publié: (2025)
LLaVA$^3$: Representing 3D Scenes like a Cubist Painter to Boost 3D Scene Understanding of VLMs
par: Petit, Doriand, et autres
Publié: (2025)
par: Petit, Doriand, et autres
Publié: (2025)
DC-Scene: Data-Centric Learning for 3D Scene Understanding
par: Huang, Ting, et autres
Publié: (2025)
par: Huang, Ting, et autres
Publié: (2025)
${C}^{3}$-GS: Learning Context-aware, Cross-dimension, Cross-scale Feature for Generalizable Gaussian Splatting
par: Hu, Yuxi, et autres
Publié: (2025)
par: Hu, Yuxi, et autres
Publié: (2025)
LEAR: Learning Edge-Aware Representations for Event-to-LiDAR Localization
par: Chen, Kuangyi, et autres
Publié: (2026)
par: Chen, Kuangyi, et autres
Publié: (2026)
Towards Foundation Models for 3D Scene Understanding: Instance-Aware Self-Supervised Learning for Point Clouds
par: Yang, Bin, et autres
Publié: (2026)
par: Yang, Bin, et autres
Publié: (2026)
DenseScan: Advancing 3D Scene Understanding with 2D Dense Annotation
par: Wang, Zirui, et autres
Publié: (2025)
par: Wang, Zirui, et autres
Publié: (2025)
Leveraging 2D-VLM for Label-Free 3D Segmentation in Large-Scale Outdoor Scene Understanding
par: Nishimura, Toshihiko, et autres
Publié: (2026)
par: Nishimura, Toshihiko, et autres
Publié: (2026)
ShelfGaussian: Shelf-Supervised Open-Vocabulary Gaussian-based 3D Scene Understanding
par: Zhao, Lingjun, et autres
Publié: (2025)
par: Zhao, Lingjun, et autres
Publié: (2025)
UrbanCAD: Towards Highly Controllable and Photorealistic 3D Vehicles for Urban Scene Simulation
par: Lu, Yichong, et autres
Publié: (2024)
par: Lu, Yichong, et autres
Publié: (2024)
Leveraging VLM-Based Pipelines to Annotate 3D Objects
par: Kabra, Rishabh, et autres
Publié: (2023)
par: Kabra, Rishabh, et autres
Publié: (2023)
BOP-Distrib: Revisiting 6D Pose Estimation Benchmarks for Better Evaluation under Visual Ambiguities
par: Meden, Boris, et autres
Publié: (2024)
par: Meden, Boris, et autres
Publié: (2024)
GigaPose: Fast and Robust Novel Object Pose Estimation via One Correspondence
par: Nguyen, Van Nguyen, et autres
Publié: (2023)
par: Nguyen, Van Nguyen, et autres
Publié: (2023)
SceneGPT: A Language Model for 3D Scene Understanding
par: Chandhok, Shivam
Publié: (2024)
par: Chandhok, Shivam
Publié: (2024)
Scene-R1: Video-Grounded Large Language Models for 3D Scene Reasoning without 3D Annotations
par: Yuan, Zhihao, et autres
Publié: (2025)
par: Yuan, Zhihao, et autres
Publié: (2025)
sim2art: Accurate Articulated Object Modeling from a Single Video using Synthetic Training Data Only
par: Artykov, Arslan, et autres
Publié: (2025)
par: Artykov, Arslan, et autres
Publié: (2025)
GenCAD-Self-Repairing: Feasibility Enhancement for 3D CAD Generation
par: Tsuji, Chikaha, et autres
Publié: (2025)
par: Tsuji, Chikaha, et autres
Publié: (2025)
MAGICIAN: Efficient Long-Term Planning with Imagined Gaussians for Active Mapping
par: Li, Shiyao, et autres
Publié: (2026)
par: Li, Shiyao, et autres
Publié: (2026)
Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning
par: Fu, Rao, et autres
Publié: (2024)
par: Fu, Rao, et autres
Publié: (2024)
CAD-Llama: Leveraging Large Language Models for Computer-Aided Design Parametric 3D Model Generation
par: Li, Jiahao, et autres
Publié: (2025)
par: Li, Jiahao, et autres
Publié: (2025)
Open-Vocabulary vs Supervised Learning Methods for Post-Disaster Visual Scene Understanding
par: Michailidou, Anna, et autres
Publié: (2026)
par: Michailidou, Anna, et autres
Publié: (2026)
R3DS: Reality-linked 3D Scenes for Panoramic Scene Understanding
par: Wu, Qirui, et autres
Publié: (2024)
par: Wu, Qirui, et autres
Publié: (2024)
Enhancing Generalizability of Representation Learning for Data-Efficient 3D Scene Understanding
par: Wang, Yunsong, et autres
Publié: (2024)
par: Wang, Yunsong, et autres
Publié: (2024)
Documents similaires
-
PyTorchGeoNodes: Enabling Differentiable Shape Programs for 3D Shape Reconstruction
par: Stekovic, Sinisa, et autres
Publié: (2024) -
GuideFlow3D: Optimization-Guided Rectified Flow For Appearance Transfer
par: Sarkar, Sayan Deb, et autres
Publié: (2025) -
SAda-Net: A Self-Supervised Adaptive Stereo Estimation CNN For Remote Sensing Image Data
par: Hirner, Dominik, et autres
Publié: (2024) -
Scene Generation at Absolute Scale: Utilizing Semantic and Geometric Guidance From Text for Accurate and Interpretable 3D Indoor Scene Generation
par: Ainetter, Stefan, et autres
Publié: (2026) -
ICG-MVSNet: Learning Intra-view and Cross-view Relationships for Guidance in Multi-View Stereo
par: Hu, Yuxi, et autres
Publié: (2025)