PC-CrossDiff: Point-Cluster Dual-Level Cross-Modal Differential Attention for Unified 3D Referring and Segmentation
Fuente:
arXiv
Saved in:
| Main Authors: | Tan, Wenbin, Lin, Jiawen, Wang, Fangyong, Xie, Yuan, Xie, Yong, Zhang, Yachao, Qu, Yanyun |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Target Refocusing via Attention Redistribution for Open-Vocabulary Semantic Segmentation: An Explainability Perspective
by: Li, Jiahao, et al.
Published: (2025)
by: Li, Jiahao, et al.
Published: (2025)
Novel Category Discovery with X-Agent Attention for Open-Vocabulary Semantic Segmentation
by: Li, Jiahao, et al.
Published: (2025)
by: Li, Jiahao, et al.
Published: (2025)
Direct Segmentation without Logits Optimization for Training-Free Open-Vocabulary Semantic Segmentation
by: Li, Jiahao, et al.
Published: (2026)
by: Li, Jiahao, et al.
Published: (2026)
Fusion-then-Distillation: Toward Cross-modal Positive Distillation for Domain Adaptive 3D Semantic Segmentation
by: Wu, Yao, et al.
Published: (2024)
by: Wu, Yao, et al.
Published: (2024)
CrossDiff: Diffusion Probabilistic Model With Cross-conditional Encoder-Decoder for Crack Segmentation
by: Shi, Xianglong, et al.
Published: (2025)
by: Shi, Xianglong, et al.
Published: (2025)
SeqVLM: Proposal-Guided Multi-View Sequences Reasoning via VLM for Zero-Shot 3D Visual Grounding
by: Lin, Jiawen, et al.
Published: (2025)
by: Lin, Jiawen, et al.
Published: (2025)
CrossDiff: Exploring Self-Supervised Representation of Pansharpening via Cross-Predictive Diffusion Model
by: Xing, Yinghui, et al.
Published: (2024)
by: Xing, Yinghui, et al.
Published: (2024)
Beyond the Label Itself: Latent Labels Enhance Semi-supervised Point Cloud Panoptic Segmentation
by: Chen, Yujun, et al.
Published: (2023)
by: Chen, Yujun, et al.
Published: (2023)
Dual-Level Cross-Modal Contrastive Clustering
by: Zhang, Haixin, et al.
Published: (2024)
by: Zhang, Haixin, et al.
Published: (2024)
Camera-Aware Cross-View Alignment for Referring 3D Gaussian Splatting Segmentation
by: Tao, Yuwen, et al.
Published: (2025)
by: Tao, Yuwen, et al.
Published: (2025)
Exploring the Untouched Sweeps for Conflict-Aware 3D Segmentation Pretraining
by: Sun, Tianfang, et al.
Published: (2024)
by: Sun, Tianfang, et al.
Published: (2024)
Learning Commonality, Divergence and Variety for Unsupervised Visible-Infrared Person Re-identification
by: Shi, Jiangming, et al.
Published: (2024)
by: Shi, Jiangming, et al.
Published: (2024)
Beyond Semantics: Uncovering the Physics of Fakes via Universal Physical Descriptors for Cross-Modal Synthetic Detection
by: Qiu, Mei, et al.
Published: (2026)
by: Qiu, Mei, et al.
Published: (2026)
Progressive Prompt-Guided Cross-Modal Reasoning for Referring Image Segmentation
by: Li, Jiachen, et al.
Published: (2026)
by: Li, Jiachen, et al.
Published: (2026)
Channel Attention-Guided Cross-Modal Knowledge Distillation for Referring Image Segmentation
by: Yang, Chen
Published: (2026)
by: Yang, Chen
Published: (2026)
DyKen-Hyena: Dynamic Kernel Generation via Cross-Modal Attention for Multimodal Intent Recognition
by: Wang, Yifei, et al.
Published: (2025)
by: Wang, Yifei, et al.
Published: (2025)
Continual-NExT: A Unified Comprehension And Generation Continual Learning Framework
by: Qiao, Jingyang, et al.
Published: (2026)
by: Qiao, Jingyang, et al.
Published: (2026)
Robust Pseudo-label Learning with Neighbor Relation for Unsupervised Visible-Infrared Person Re-Identification
by: Yin, Xiangbo, et al.
Published: (2024)
by: Yin, Xiangbo, et al.
Published: (2024)
Multi-Memory Matching for Unsupervised Visible-Infrared Person Re-Identification
by: Shi, Jiangming, et al.
Published: (2024)
by: Shi, Jiangming, et al.
Published: (2024)
DiffX: Guide Your Layout to Cross-Modal Generative Modeling
by: Wang, Zeyu, et al.
Published: (2024)
by: Wang, Zeyu, et al.
Published: (2024)
Modality Selection and Skill Segmentation via Cross-Modality Attention
by: Jiang, Jiawei, et al.
Published: (2025)
by: Jiang, Jiawei, et al.
Published: (2025)
Mutual Information Guided Optimal Transport for Unsupervised Visible-Infrared Person Re-identification
by: Zhang, Zhizhong, et al.
Published: (2024)
by: Zhang, Zhizhong, et al.
Published: (2024)
DEYOLO: Dual-Feature-Enhancement YOLO for Cross-Modality Object Detection
by: Chen, Yishuo, et al.
Published: (2024)
by: Chen, Yishuo, et al.
Published: (2024)
MSCT: Differential Cross-Modal Attention for Deepfake Detection
by: Wei, Fangda, et al.
Published: (2026)
by: Wei, Fangda, et al.
Published: (2026)
A Unified Attention U-Net Framework for Cross-Modality Tumor Segmentation in MRI and CT
by: Rai, Nishan, et al.
Published: (2026)
by: Rai, Nishan, et al.
Published: (2026)
Dual Cross-Attention for Medical Image Segmentation
by: Ates, Gorkem Can, et al.
Published: (2023)
by: Ates, Gorkem Can, et al.
Published: (2023)
Referring Video Object Segmentation with Cross-Modality Proxy Queries
by: Sun, Baoli, et al.
Published: (2025)
by: Sun, Baoli, et al.
Published: (2025)
sleep2vec: Unified Cross-Modal Alignment for Heterogeneous Nocturnal Biosignals
by: Yuan, Weixuan, et al.
Published: (2026)
by: Yuan, Weixuan, et al.
Published: (2026)
Uni-RCM: Unified Reference-guided Cross-modal Mapping for Multi-Class Anomaly Detection
by: Wu, Yangchen, et al.
Published: (2026)
by: Wu, Yangchen, et al.
Published: (2026)
Data-free Distillation with Degradation-prompt Diffusion for Multi-weather Image Restoration
by: Wang, Pei, et al.
Published: (2024)
by: Wang, Pei, et al.
Published: (2024)
Training-Free Anomaly Generation via Dual-Attention Enhancement in Diffusion Model
by: Zuo, Zuo, et al.
Published: (2025)
by: Zuo, Zuo, et al.
Published: (2025)
PointDC:Unsupervised Semantic Segmentation of 3D Point Clouds via Cross-modal Distillation and Super-Voxel Clustering
by: Chen, Zisheng, et al.
Published: (2023)
by: Chen, Zisheng, et al.
Published: (2023)
Stability and Synchronization of a Fractional‐Order Unified System with Complex Variables
by: Yanyun Xie, et al.
Published: (2024)
by: Yanyun Xie, et al.
Published: (2024)
Cross-Modal Attention Network with Dual Graph Learning in Multimodal Recommendation
by: Dai, Ji, et al.
Published: (2026)
by: Dai, Ji, et al.
Published: (2026)
DiffCrossGait: Trajectory-Level Alignment for 2D-3D Cross-Modal Gait Recognition via Latent Diffusion
by: Lu, Zhiyang, et al.
Published: (2026)
by: Lu, Zhiyang, et al.
Published: (2026)
Cross-Modal Bidirectional Interaction Model for Referring Remote Sensing Image Segmentation
by: Dong, Zhe, et al.
Published: (2024)
by: Dong, Zhe, et al.
Published: (2024)
CCSD: Cross-Modal Compositional Self-Distillation for Robust Brain Tumor Segmentation with Missing Modalities
by: Xie, Dongqing, et al.
Published: (2025)
by: Xie, Dongqing, et al.
Published: (2025)
Boosting SAM for Cross-Domain Few-Shot Segmentation via Conditional Point Sparsification
by: Nie, Jiahao, et al.
Published: (2026)
by: Nie, Jiahao, et al.
Published: (2026)
TUNI: Real-time RGB-T Semantic Segmentation with Unified Multi-Modal Feature Extraction and Cross-Modal Feature Fusion
by: Guo, Xiaodong, et al.
Published: (2025)
by: Guo, Xiaodong, et al.
Published: (2025)
COTR: Compact Occupancy TRansformer for Vision-based 3D Occupancy Prediction
by: Ma, Qihang, et al.
Published: (2023)
by: Ma, Qihang, et al.
Published: (2023)
Similar Items
-
Target Refocusing via Attention Redistribution for Open-Vocabulary Semantic Segmentation: An Explainability Perspective
by: Li, Jiahao, et al.
Published: (2025) -
Novel Category Discovery with X-Agent Attention for Open-Vocabulary Semantic Segmentation
by: Li, Jiahao, et al.
Published: (2025) -
Direct Segmentation without Logits Optimization for Training-Free Open-Vocabulary Semantic Segmentation
by: Li, Jiahao, et al.
Published: (2026) -
Fusion-then-Distillation: Toward Cross-modal Positive Distillation for Domain Adaptive 3D Semantic Segmentation
by: Wu, Yao, et al.
Published: (2024) -
CrossDiff: Diffusion Probabilistic Model With Cross-conditional Encoder-Decoder for Crack Segmentation
by: Shi, Xianglong, et al.
Published: (2025)