Unlocking 3D Affordance Segmentation with 2D Semantic Knowledge
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Yu, Peng, Zelin, Wen, Changsong, Yang, Xiaokang, Shen, Wei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MedSeg-R: Reasoning Segmentation in Medical Images with Multimodal Large Language Models
von: Huang, Yu, et al.
Veröffentlicht: (2025)
von: Huang, Yu, et al.
Veröffentlicht: (2025)
Learning 2D Invariant Affordance Knowledge for 3D Affordance Grounding
von: Gao, Xianqiang, et al.
Veröffentlicht: (2024)
von: Gao, Xianqiang, et al.
Veröffentlicht: (2024)
Tackling View-Dependent Semantics in 3D Language Gaussian Splatting
von: Cen, Jiazhong, et al.
Veröffentlicht: (2025)
von: Cen, Jiazhong, et al.
Veröffentlicht: (2025)
Tendency-driven Mutual Exclusivity for Weakly Supervised Incremental Semantic Segmentation
von: Si, Chongjie, et al.
Veröffentlicht: (2024)
von: Si, Chongjie, et al.
Veröffentlicht: (2024)
HyperET: Efficient Training in Hyperbolic Space for Multi-modal Large Language Models
von: Peng, Zelin, et al.
Veröffentlicht: (2025)
von: Peng, Zelin, et al.
Veröffentlicht: (2025)
Parameter-efficient Fine-tuning in Hyperspherical Space for Open-vocabulary Semantic Segmentation
von: Peng, Zelin, et al.
Veröffentlicht: (2024)
von: Peng, Zelin, et al.
Veröffentlicht: (2024)
NEARL-CLIP: Interacted Query Adaptation with Orthogonal Regularization for Medical Vision-Language Understanding
von: Peng, Zelin, et al.
Veröffentlicht: (2025)
von: Peng, Zelin, et al.
Veröffentlicht: (2025)
AffordanceSAM: Segment Anything Once More in Affordance Grounding
von: Jiang, Dengyang, et al.
Veröffentlicht: (2025)
von: Jiang, Dengyang, et al.
Veröffentlicht: (2025)
3DAffordSplat: Efficient Affordance Reasoning with 3D Gaussians
von: Wei, Zeming, et al.
Veröffentlicht: (2025)
von: Wei, Zeming, et al.
Veröffentlicht: (2025)
Gradient-Driven 3D Segmentation and Affordance Transfer in Gaussian Splatting Using 2D Masks
von: Joseph, Joji, et al.
Veröffentlicht: (2024)
von: Joseph, Joji, et al.
Veröffentlicht: (2024)
FOLK: Fast Open-Vocabulary 3D Instance Segmentation via Label-guided Knowledge Distillation
von: Wu, Hongrui, et al.
Veröffentlicht: (2025)
von: Wu, Hongrui, et al.
Veröffentlicht: (2025)
HumanCrafter: Synergizing Generalizable Human Reconstruction and Semantic 3D Segmentation
von: Pan, Panwang, et al.
Veröffentlicht: (2025)
von: Pan, Panwang, et al.
Veröffentlicht: (2025)
Adaptive Margin Contrastive Learning for Ambiguity-aware 3D Semantic Segmentation
von: Chen, Yang, et al.
Veröffentlicht: (2025)
von: Chen, Yang, et al.
Veröffentlicht: (2025)
TCATSeg: A Tooth Center-Wise Attention Network for 3D Dental Model Semantic Segmentation
von: He, Qiang, et al.
Veröffentlicht: (2026)
von: He, Qiang, et al.
Veröffentlicht: (2026)
MFSeg: Efficient Multi-frame 3D Semantic Segmentation
von: Huang, Chengjie, et al.
Veröffentlicht: (2025)
von: Huang, Chengjie, et al.
Veröffentlicht: (2025)
PGOV3D: Open-Vocabulary 3D Semantic Segmentation with Partial-to-Global Curriculum
von: Zhang, Shiqi, et al.
Veröffentlicht: (2025)
von: Zhang, Shiqi, et al.
Veröffentlicht: (2025)
GEAL: Generalizable 3D Affordance Learning with Cross-Modal Consistency
von: Lu, Dongyue, et al.
Veröffentlicht: (2024)
von: Lu, Dongyue, et al.
Veröffentlicht: (2024)
CutS3D: Cutting Semantics in 3D for 2D Unsupervised Instance Segmentation
von: Sick, Leon, et al.
Veröffentlicht: (2024)
von: Sick, Leon, et al.
Veröffentlicht: (2024)
A Unified Framework with Multimodal Fine-tuning for Remote Sensing Semantic Segmentation
von: Ma, Xianping, et al.
Veröffentlicht: (2024)
von: Ma, Xianping, et al.
Veröffentlicht: (2024)
RS3Mamba: Visual State Space Model for Remote Sensing Images Semantic Segmentation
von: Ma, Xianping, et al.
Veröffentlicht: (2024)
von: Ma, Xianping, et al.
Veröffentlicht: (2024)
Grounding 3D Scene Affordance From Egocentric Interactions
von: Liu, Cuiyu, et al.
Veröffentlicht: (2024)
von: Liu, Cuiyu, et al.
Veröffentlicht: (2024)
Segment Any 3D Gaussians
von: Cen, Jiazhong, et al.
Veröffentlicht: (2023)
von: Cen, Jiazhong, et al.
Veröffentlicht: (2023)
Task-Aware 3D Affordance Segmentation via 2D Guidance and Geometric Refinement
von: He, Lian, et al.
Veröffentlicht: (2025)
von: He, Lian, et al.
Veröffentlicht: (2025)
Part-Aware Open-Vocabulary 3D Affordance Grounding via Prototypical Semantic and Geometric Alignment
von: Gou, Dongqiang, et al.
Veröffentlicht: (2026)
von: Gou, Dongqiang, et al.
Veröffentlicht: (2026)
3x2: 3D Object Part Segmentation by 2D Semantic Correspondences
von: Thai, Anh, et al.
Veröffentlicht: (2024)
von: Thai, Anh, et al.
Veröffentlicht: (2024)
Affostruction: 3D Affordance Grounding with Generative Reconstruction
von: Park, Chunghyun, et al.
Veröffentlicht: (2026)
von: Park, Chunghyun, et al.
Veröffentlicht: (2026)
3D-AffordanceLLM: Harnessing Large Language Models for Open-Vocabulary Affordance Detection in 3D Worlds
von: Chu, Hengshuo, et al.
Veröffentlicht: (2025)
von: Chu, Hengshuo, et al.
Veröffentlicht: (2025)
AdaCo: Overcoming Visual Foundation Model Noise in 3D Semantic Segmentation via Adaptive Label Correction
von: Zou, Pufan, et al.
Veröffentlicht: (2024)
von: Zou, Pufan, et al.
Veröffentlicht: (2024)
SPHERE: Semantic-PHysical Engaged REpresentation for 3D Semantic Scene Completion
von: Yang, Zhiwen, et al.
Veröffentlicht: (2025)
von: Yang, Zhiwen, et al.
Veröffentlicht: (2025)
Domain Adaptation-Based Crossmodal Knowledge Distillation for 3D Semantic Segmentation
von: Kang, Jialiang, et al.
Veröffentlicht: (2025)
von: Kang, Jialiang, et al.
Veröffentlicht: (2025)
EGFormer: Towards Efficient and Generalizable Multimodal Semantic Segmentation
von: Zhang, Zelin, et al.
Veröffentlicht: (2025)
von: Zhang, Zelin, et al.
Veröffentlicht: (2025)
VAGNet: Grounding 3D Affordance from Human-Object Interactions in Videos
von: Mao, Aihua, et al.
Veröffentlicht: (2026)
von: Mao, Aihua, et al.
Veröffentlicht: (2026)
MMRad-22K: A Structured Multimodal Evidence Dataset for Chest X-ray Report Generation
von: Zhao, Yichen, et al.
Veröffentlicht: (2026)
von: Zhao, Yichen, et al.
Veröffentlicht: (2026)
3D Weakly Supervised Semantic Segmentation with 2D Vision-Language Guidance
von: Xu, Xiaoxu, et al.
Veröffentlicht: (2024)
von: Xu, Xiaoxu, et al.
Veröffentlicht: (2024)
Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances
von: Wang, Qirui, et al.
Veröffentlicht: (2026)
von: Wang, Qirui, et al.
Veröffentlicht: (2026)
BFANet: Revisiting 3D Semantic Segmentation with Boundary Feature Analysis
von: Zhao, Weiguang, et al.
Veröffentlicht: (2025)
von: Zhao, Weiguang, et al.
Veröffentlicht: (2025)
Affordance-Guided Diffusion Prior for 3D Hand Reconstruction
von: Suzuki, Naru, et al.
Veröffentlicht: (2025)
von: Suzuki, Naru, et al.
Veröffentlicht: (2025)
Segment Anything in 3D with Radiance Fields
von: Cen, Jiazhong, et al.
Veröffentlicht: (2023)
von: Cen, Jiazhong, et al.
Veröffentlicht: (2023)
3D-PointZshotS: Geometry-Aware 3D Point Cloud Zero-Shot Semantic Segmentation Narrowing the Visual-Semantic Gap
von: Yang, Minmin, et al.
Veröffentlicht: (2025)
von: Yang, Minmin, et al.
Veröffentlicht: (2025)
SceneTeract: Agentic Functional Affordances and VLM Grounding in 3D Scenes
von: Maillard, Léopold, et al.
Veröffentlicht: (2026)
von: Maillard, Léopold, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MedSeg-R: Reasoning Segmentation in Medical Images with Multimodal Large Language Models
von: Huang, Yu, et al.
Veröffentlicht: (2025) -
Learning 2D Invariant Affordance Knowledge for 3D Affordance Grounding
von: Gao, Xianqiang, et al.
Veröffentlicht: (2024) -
Tackling View-Dependent Semantics in 3D Language Gaussian Splatting
von: Cen, Jiazhong, et al.
Veröffentlicht: (2025) -
Tendency-driven Mutual Exclusivity for Weakly Supervised Incremental Semantic Segmentation
von: Si, Chongjie, et al.
Veröffentlicht: (2024) -
HyperET: Efficient Training in Hyperbolic Space for Multi-modal Large Language Models
von: Peng, Zelin, et al.
Veröffentlicht: (2025)