CoSMo3D: Open-World Promptable 3D Semantic Part Segmentation through LLM-Guided Canonical Spatial Modeling
Fuente:
arXiv
Saved in:
| Main Authors: | Jin, Li, Chen, Weikai, Wang, Yujie, Yin, Yingda, Hu, Zeyu, Zhang, Runze, Luo, Keyang, Qian, Shengju, Wang, Xin, Qin, Xueying |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CoSMo: A Multimodal Transformer for Page Stream Segmentation in Comic Books
by: Ortega, Marc Serra, et al.
Published: (2025)
by: Ortega, Marc Serra, et al.
Published: (2025)
CanoVerse: 3D Object Scalable Canonicalization and Dataset for Generation and Pose
by: Jin, Li, et al.
Published: (2026)
by: Jin, Li, et al.
Published: (2026)
CoSMo: a Framework to Instantiate Conditioned Process Simulation Models
by: Oyamada, Rafael S., et al.
Published: (2023)
by: Oyamada, Rafael S., et al.
Published: (2023)
CaliTex: Geometry-Calibrated Attention for View-Coherent 3D Texture Generation
by: Liu, Chenyu, et al.
Published: (2025)
by: Liu, Chenyu, et al.
Published: (2025)
ReVSeg: Incentivizing the Reasoning Chain for Video Segmentation with Reinforcement Learning
by: Li, Yifan, et al.
Published: (2025)
by: Li, Yifan, et al.
Published: (2025)
MuMA: 3D PBR Texturing via Multi-Channel Multi-View Generation and Agentic Post-Processing
by: Zhu, Lingting, et al.
Published: (2025)
by: Zhu, Lingting, et al.
Published: (2025)
SAP: Segment Any 4K Panorama
by: Jiang, Lutao, et al.
Published: (2026)
by: Jiang, Lutao, et al.
Published: (2026)
MAR-3D: Progressive Masked Auto-regressor for High-Resolution 3D Generation
by: Chen, Jinnan, et al.
Published: (2025)
by: Chen, Jinnan, et al.
Published: (2025)
PartSAM: A Scalable Promptable Part Segmentation Model Trained on Native 3D Data
by: Zhu, Zhe, et al.
Published: (2025)
by: Zhu, Zhe, et al.
Published: (2025)
nnInteractive: Redefining 3D Promptable Segmentation
by: Isensee, Fabian, et al.
Published: (2025)
by: Isensee, Fabian, et al.
Published: (2025)
TagCLIP: Improving Discrimination Ability of Open-Vocabulary Semantic Segmentation
by: Li, Jingyao, et al.
Published: (2023)
by: Li, Jingyao, et al.
Published: (2023)
WildDet3D: Scaling Promptable 3D Detection in the Wild
by: Huang, Weikai, et al.
Published: (2026)
by: Huang, Weikai, et al.
Published: (2026)
Open-Vocabulary Semantic Part Segmentation of 3D Human
by: Suzuki, Keito, et al.
Published: (2025)
by: Suzuki, Keito, et al.
Published: (2025)
Opportunistic Promptable Segmentation: Leveraging Routine Radiological Annotations to Guide 3D CT Lesion Segmentation
by: Church, Samuel, et al.
Published: (2026)
by: Church, Samuel, et al.
Published: (2026)
Large Material Gaussian Model for Relightable 3D Generation
by: Ye, Jingrui, et al.
Published: (2025)
by: Ye, Jingrui, et al.
Published: (2025)
Point-SAM: Promptable 3D Segmentation Model for Point Clouds
by: Zhou, Yuchen, et al.
Published: (2024)
by: Zhou, Yuchen, et al.
Published: (2024)
CLIP-Guided SAM: Parameter-Efficient Semantic Conditioning for Promptable Segmentation
by: Jalilian, Shayan, et al.
Published: (2026)
by: Jalilian, Shayan, et al.
Published: (2026)
Skill-Evolving Grounded Reasoning for Free-Text Promptable 3D Medical Image Segmentation
by: Zhang, Tongrui, et al.
Published: (2026)
by: Zhang, Tongrui, et al.
Published: (2026)
GeoGuide: Hierarchical Geometric Guidance for Open-Vocabulary 3D Semantic Segmentation
by: Tao, Xujing, et al.
Published: (2026)
by: Tao, Xujing, et al.
Published: (2026)
SAI3D: Segment Any Instance in 3D Scenes
by: Yin, Yingda, et al.
Published: (2023)
by: Yin, Yingda, et al.
Published: (2023)
SurgTPGS: Semantic 3D Surgical Scene Understanding with Text Promptable Gaussian Splatting
by: Huang, Yiming, et al.
Published: (2025)
by: Huang, Yiming, et al.
Published: (2025)
XMask3D: Cross-modal Mask Reasoning for Open Vocabulary 3D Semantic Segmentation
by: Wang, Ziyi, et al.
Published: (2024)
by: Wang, Ziyi, et al.
Published: (2024)
3x2: 3D Object Part Segmentation by 2D Semantic Correspondences
by: Thai, Anh, et al.
Published: (2024)
by: Thai, Anh, et al.
Published: (2024)
LumiTex: Towards High-Fidelity PBR Texture Generation with Illumination Context
by: Bao, Jingzhi, et al.
Published: (2025)
by: Bao, Jingzhi, et al.
Published: (2025)
CUS3D :CLIP-based Unsupervised 3D Segmentation via Object-level Denoise
by: Yu, Fuyang, et al.
Published: (2024)
by: Yu, Fuyang, et al.
Published: (2024)
VoxTell: Free-Text Promptable Universal 3D Medical Image Segmentation
by: Rokuss, Maximilian, et al.
Published: (2025)
by: Rokuss, Maximilian, et al.
Published: (2025)
3D CoCa: Contrastive Learners are 3D Captioners
by: Huang, Ting, et al.
Published: (2025)
by: Huang, Ting, et al.
Published: (2025)
D3S2: Diffusion-Guided Dataset Distillation for Semantic Segmentation
by: Zheng, Wenjie, et al.
Published: (2026)
by: Zheng, Wenjie, et al.
Published: (2026)
Light-SQ: Structure-aware Shape Abstraction with Superquadrics for Generated Meshes
by: Wang, Yuhan, et al.
Published: (2025)
by: Wang, Yuhan, et al.
Published: (2025)
SAM2Point: Segment Any 3D as Videos in Zero-shot and Promptable Manners
by: Guo, Ziyu, et al.
Published: (2024)
by: Guo, Ziyu, et al.
Published: (2024)
CoCo-SAM3: Harnessing Concept Conflict in Open-Vocabulary Semantic Segmentation
by: Chen, Yanhui, et al.
Published: (2026)
by: Chen, Yanhui, et al.
Published: (2026)
SAM 3D Animal: Promptable Animal 3D Reconstruction from Images in the Wild
by: Hu, Xuyi, et al.
Published: (2026)
by: Hu, Xuyi, et al.
Published: (2026)
PGOV3D: Open-Vocabulary 3D Semantic Segmentation with Partial-to-Global Curriculum
by: Zhang, Shiqi, et al.
Published: (2025)
by: Zhang, Shiqi, et al.
Published: (2025)
OpenUrban3D: Annotation-Free Open-Vocabulary Semantic Segmentation of Large-Scale Urban Point Clouds
by: Wang, Chongyu, et al.
Published: (2025)
by: Wang, Chongyu, et al.
Published: (2025)
PDF: A Probability-Driven Framework for Open World 3D Point Cloud Semantic Segmentation
by: Xu, Jinfeng, et al.
Published: (2024)
by: Xu, Jinfeng, et al.
Published: (2024)
SegEarth-OV3: Exploring SAM 3 for Open-Vocabulary Semantic Segmentation in Remote Sensing Images
by: Li, Kaiyu, et al.
Published: (2025)
by: Li, Kaiyu, et al.
Published: (2025)
OpenTrack3D: Towards Accurate and Generalizable Open-Vocabulary 3D Instance Segmentation
by: Zhou, Zhishan, et al.
Published: (2025)
by: Zhou, Zhishan, et al.
Published: (2025)
3D Gaussian Splatting with Deferred Reflection
by: Ye, Keyang, et al.
Published: (2024)
by: Ye, Keyang, et al.
Published: (2024)
Multimodal 3D Reasoning Segmentation with Complex Scenes
by: Jiang, Xueying, et al.
Published: (2024)
by: Jiang, Xueying, et al.
Published: (2024)
Unifying 3D Vision-Language Understanding via Promptable Queries
by: Zhu, Ziyu, et al.
Published: (2024)
by: Zhu, Ziyu, et al.
Published: (2024)
Similar Items
-
CoSMo: A Multimodal Transformer for Page Stream Segmentation in Comic Books
by: Ortega, Marc Serra, et al.
Published: (2025) -
CanoVerse: 3D Object Scalable Canonicalization and Dataset for Generation and Pose
by: Jin, Li, et al.
Published: (2026) -
CoSMo: a Framework to Instantiate Conditioned Process Simulation Models
by: Oyamada, Rafael S., et al.
Published: (2023) -
CaliTex: Geometry-Calibrated Attention for View-Coherent 3D Texture Generation
by: Liu, Chenyu, et al.
Published: (2025) -
ReVSeg: Incentivizing the Reasoning Chain for Video Segmentation with Reinforcement Learning
by: Li, Yifan, et al.
Published: (2025)