OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model
Fuente:
arXiv
Salvato in:
| Autore principale: | Zhang, Kunshen |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model
di: Huang, Kuan-Chih, et al.
Pubblicazione: (2024)
di: Huang, Kuan-Chih, et al.
Pubblicazione: (2024)
DINO in the Room: Leveraging 2D Foundation Models for 3D Segmentation
di: Knaebel, Karim, et al.
Pubblicazione: (2025)
di: Knaebel, Karim, et al.
Pubblicazione: (2025)
OVS-DINO: Open-Vocabulary Segmentation via Structure-Aligned SAM-DINO with Language Guidance
di: Zeng, Haoxi, et al.
Pubblicazione: (2026)
di: Zeng, Haoxi, et al.
Pubblicazione: (2026)
Reasoning3D -- Grounding and Reasoning in 3D: Fine-Grained Zero-Shot Open-Vocabulary 3D Reasoning Part Segmentation via Large Vision-Language Models
di: Chen, Tianrun, et al.
Pubblicazione: (2024)
di: Chen, Tianrun, et al.
Pubblicazione: (2024)
Open3DIS: Open-Vocabulary 3D Instance Segmentation with 2D Mask Guidance
di: Nguyen, Phuc D. A., et al.
Pubblicazione: (2023)
di: Nguyen, Phuc D. A., et al.
Pubblicazione: (2023)
MLLM-For3D: Adapting Multimodal Large Language Model for 3D Reasoning Segmentation
di: Huang, Jiaxin, et al.
Pubblicazione: (2025)
di: Huang, Jiaxin, et al.
Pubblicazione: (2025)
XMask3D: Cross-modal Mask Reasoning for Open Vocabulary 3D Semantic Segmentation
di: Wang, Ziyi, et al.
Pubblicazione: (2024)
di: Wang, Ziyi, et al.
Pubblicazione: (2024)
SegDINO3D: 3D Instance Segmentation Empowered by Both Image-Level and Object-Level 2D Features
di: Qu, Jinyuan, et al.
Pubblicazione: (2025)
di: Qu, Jinyuan, et al.
Pubblicazione: (2025)
DI-MaskDINO: A Joint Object Detection and Instance Segmentation Model
di: Nan, Zhixiong, et al.
Pubblicazione: (2024)
di: Nan, Zhixiong, et al.
Pubblicazione: (2024)
RefMask3D: Language-Guided Transformer for 3D Referring Segmentation
di: He, Shuting, et al.
Pubblicazione: (2024)
di: He, Shuting, et al.
Pubblicazione: (2024)
MaskClustering: View Consensus based Mask Graph Clustering for Open-Vocabulary 3D Instance Segmentation
di: Yan, Mi, et al.
Pubblicazione: (2024)
di: Yan, Mi, et al.
Pubblicazione: (2024)
SegDINO: An Efficient Design for Medical and Natural Image Segmentation with DINO-V3
di: Yang, Sicheng, et al.
Pubblicazione: (2025)
di: Yang, Sicheng, et al.
Pubblicazione: (2025)
DINO Soars: DINOv3 for Open-Vocabulary Semantic Segmentation of Remote Sensing Imagery
di: Faulkenberry, Ryan, et al.
Pubblicazione: (2026)
di: Faulkenberry, Ryan, et al.
Pubblicazione: (2026)
DINO Eats CLIP: Adapting Beyond Knowns for Open-set 3D Object Retrieval
di: He, Xinwei, et al.
Pubblicazione: (2026)
di: He, Xinwei, et al.
Pubblicazione: (2026)
Open3D-VQA: A Benchmark for Comprehensive Spatial Reasoning with Multimodal Large Language Model in Open Space
di: Zhang, Weichen, et al.
Pubblicazione: (2025)
di: Zhang, Weichen, et al.
Pubblicazione: (2025)
MV3DIS: Multi-View Mask Matching via 3D Guides for Zero-Shot 3D Instance Segmentation
di: Zhao, Yibo, et al.
Pubblicazione: (2026)
di: Zhao, Yibo, et al.
Pubblicazione: (2026)
3D Open-Vocabulary Panoptic Segmentation with 2D-3D Vision-Language Distillation
di: Xiao, Zihao, et al.
Pubblicazione: (2024)
di: Xiao, Zihao, et al.
Pubblicazione: (2024)
DINO-SLAM: DINO-informed RGB-D SLAM for Neural Implicit and Explicit Representations
di: Gong, Ziren, et al.
Pubblicazione: (2025)
di: Gong, Ziren, et al.
Pubblicazione: (2025)
Mosaic3D: Foundation Dataset and Model for Open-Vocabulary 3D Segmentation
di: Lee, Junha, et al.
Pubblicazione: (2025)
di: Lee, Junha, et al.
Pubblicazione: (2025)
Effective Feature Learning for 3D Medical Registration via Domain-Specialized DINO Pretraining
di: Kats, Eytan, et al.
Pubblicazione: (2026)
di: Kats, Eytan, et al.
Pubblicazione: (2026)
SGS-3D: High-Fidelity 3D Instance Segmentation via Reliable Semantic Mask Splitting and Growing
di: Wang, Chaolei, et al.
Pubblicazione: (2025)
di: Wang, Chaolei, et al.
Pubblicazione: (2025)
PGOV3D: Open-Vocabulary 3D Semantic Segmentation with Partial-to-Global Curriculum
di: Zhang, Shiqi, et al.
Pubblicazione: (2025)
di: Zhang, Shiqi, et al.
Pubblicazione: (2025)
Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
di: Liu, Shilong, et al.
Pubblicazione: (2023)
di: Liu, Shilong, et al.
Pubblicazione: (2023)
COS3D: Collaborative Open-Vocabulary 3D Segmentation
di: Zhu, Runsong, et al.
Pubblicazione: (2025)
di: Zhu, Runsong, et al.
Pubblicazione: (2025)
OVSeg3R: Learn Open-vocabulary Instance Segmentation from 2D via 3D Reconstruction
di: Li, Hongyang, et al.
Pubblicazione: (2025)
di: Li, Hongyang, et al.
Pubblicazione: (2025)
Search3D: Hierarchical Open-Vocabulary 3D Segmentation
di: Takmaz, Ayca, et al.
Pubblicazione: (2024)
di: Takmaz, Ayca, et al.
Pubblicazione: (2024)
Jigsaw3D: Disentangled 3D Style Transfer via Patch Shuffling and Masking
di: Ye, Yuteng, et al.
Pubblicazione: (2025)
di: Ye, Yuteng, et al.
Pubblicazione: (2025)
Point Linguist Model: Segment Any Object via Bridged Large 3D-Language Model
di: Huang, Zhuoxu, et al.
Pubblicazione: (2025)
di: Huang, Zhuoxu, et al.
Pubblicazione: (2025)
OpenIns3D: Snap and Lookup for 3D Open-vocabulary Instance Segmentation
di: Huang, Zhening, et al.
Pubblicazione: (2023)
di: Huang, Zhening, et al.
Pubblicazione: (2023)
DINO-RotateMatch: A Rotation-Aware Deep Framework for Robust Image Matching in Large-Scale 3D Reconstruction
di: Zhang, Kaichen, et al.
Pubblicazione: (2025)
di: Zhang, Kaichen, et al.
Pubblicazione: (2025)
DecoDINO: 3D Human-Scene Contact Prediction with Semantic Classification
di: Bierling, Lukas, et al.
Pubblicazione: (2025)
di: Bierling, Lukas, et al.
Pubblicazione: (2025)
Empowering DINO Representations for Underwater Instance Segmentation via Aligner and Prompter
di: Chen, Zhiyang, et al.
Pubblicazione: (2025)
di: Chen, Zhiyang, et al.
Pubblicazione: (2025)
InterMask: 3D Human Interaction Generation via Collaborative Masked Modeling
di: Javed, Muhammad Gohar, et al.
Pubblicazione: (2024)
di: Javed, Muhammad Gohar, et al.
Pubblicazione: (2024)
Any3DIS: Class-Agnostic 3D Instance Segmentation by 2D Mask Tracking
di: Nguyen, Phuc, et al.
Pubblicazione: (2024)
di: Nguyen, Phuc, et al.
Pubblicazione: (2024)
Segment then Splat: Unified 3D Open-Vocabulary Segmentation via Gaussian Splatting
di: Lu, Yiren, et al.
Pubblicazione: (2025)
di: Lu, Yiren, et al.
Pubblicazione: (2025)
Weakly Supervised 3D Open-vocabulary Segmentation
di: Liu, Kunhao, et al.
Pubblicazione: (2023)
di: Liu, Kunhao, et al.
Pubblicazione: (2023)
Scene-R1: Video-Grounded Large Language Models for 3D Scene Reasoning without 3D Annotations
di: Yuan, Zhihao, et al.
Pubblicazione: (2025)
di: Yuan, Zhihao, et al.
Pubblicazione: (2025)
OpenTrack3D: Towards Accurate and Generalizable Open-Vocabulary 3D Instance Segmentation
di: Zhou, Zhishan, et al.
Pubblicazione: (2025)
di: Zhou, Zhishan, et al.
Pubblicazione: (2025)
OpenSplat3D: Open-Vocabulary 3D Instance Segmentation using Gaussian Splatting
di: Piekenbrinck, Jens, et al.
Pubblicazione: (2025)
di: Piekenbrinck, Jens, et al.
Pubblicazione: (2025)
Open-YOLO 3D: Towards Fast and Accurate Open-Vocabulary 3D Instance Segmentation
di: Boudjoghra, Mohamed El Amine, et al.
Pubblicazione: (2024)
di: Boudjoghra, Mohamed El Amine, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model
di: Huang, Kuan-Chih, et al.
Pubblicazione: (2024) -
DINO in the Room: Leveraging 2D Foundation Models for 3D Segmentation
di: Knaebel, Karim, et al.
Pubblicazione: (2025) -
OVS-DINO: Open-Vocabulary Segmentation via Structure-Aligned SAM-DINO with Language Guidance
di: Zeng, Haoxi, et al.
Pubblicazione: (2026) -
Reasoning3D -- Grounding and Reasoning in 3D: Fine-Grained Zero-Shot Open-Vocabulary 3D Reasoning Part Segmentation via Large Vision-Language Models
di: Chen, Tianrun, et al.
Pubblicazione: (2024) -
Open3DIS: Open-Vocabulary 3D Instance Segmentation with 2D Mask Guidance
di: Nguyen, Phuc D. A., et al.
Pubblicazione: (2023)