COM3D: Leveraging Cross-View Correspondence and Cross-Modal Mining for 3D Retrieval
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Hao, LI, Ruochong, Wang, Hao, Xiong, Hui |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SCA3D: Enhancing Cross-modal 3D Retrieval via 3D Shape and Caption Paired Data Augmentation
by: Ren, Junlong, et al.
Published: (2025)
by: Ren, Junlong, et al.
Published: (2025)
3DBonsai: Structure-Aware Bonsai Modeling Using Conditioned 3D Gaussian Splatting
by: Wu, Hao, et al.
Published: (2025)
by: Wu, Hao, et al.
Published: (2025)
CorrespondentDream: Enhancing 3D Fidelity of Text-to-3D using Cross-View Correspondences
by: Kim, Seungwook, et al.
Published: (2024)
by: Kim, Seungwook, et al.
Published: (2024)
C3Po: Cross-View Cross-Modality Correspondence by Pointmap Prediction
by: Huang, Kuan Wei, et al.
Published: (2025)
by: Huang, Kuan Wei, et al.
Published: (2025)
Enhanced Cross-modal 3D Retrieval via Tri-modal Reconstruction
by: Ren, Junlong, et al.
Published: (2025)
by: Ren, Junlong, et al.
Published: (2025)
Leveraging RGB-D Data with Cross-Modal Context Mining for Glass Surface Detection
by: Lin, Jiaying, et al.
Published: (2022)
by: Lin, Jiaying, et al.
Published: (2022)
Leveraging Modality Tags for Enhanced Cross-Modal Video Retrieval
by: Fragomeni, Adriano, et al.
Published: (2025)
by: Fragomeni, Adriano, et al.
Published: (2025)
CrossOver: 3D Scene Cross-Modal Alignment
by: Sarkar, Sayan Deb, et al.
Published: (2025)
by: Sarkar, Sayan Deb, et al.
Published: (2025)
CL2CM: Improving Cross-Lingual Cross-Modal Retrieval via Cross-Lingual Knowledge Transfer
by: Wang, Yabing, et al.
Published: (2023)
by: Wang, Yabing, et al.
Published: (2023)
Boosting Instance Awareness via Cross-View Correlation with 4D Radar and Camera for 3D Object Detection
by: Bai, Xiaokai, et al.
Published: (2026)
by: Bai, Xiaokai, et al.
Published: (2026)
DuoSpaceNet: Leveraging Both Bird's-Eye-View and Perspective View Representations for 3D Object Detection
by: Huang, Zhe, et al.
Published: (2024)
by: Huang, Zhe, et al.
Published: (2024)
3x2: 3D Object Part Segmentation by 2D Semantic Correspondences
by: Thai, Anh, et al.
Published: (2024)
by: Thai, Anh, et al.
Published: (2024)
No Calibration, No Depth, No Problem: Cross-Sensor View Synthesis with 3D Consistency
by: Wu, Cho-Ying, et al.
Published: (2026)
by: Wu, Cho-Ying, et al.
Published: (2026)
FusionBERT: Multi-View Image-3D Retrieval via Cross-Attention Visual Fusion and Normal-Aware 3D Encoder
by: Li, Wei, et al.
Published: (2026)
by: Li, Wei, et al.
Published: (2026)
Towards Cross-View Point Correspondence in Vision-Language Models
by: Wang, Yipu, et al.
Published: (2025)
by: Wang, Yipu, et al.
Published: (2025)
SP3D: Boosting Sparsely-Supervised 3D Object Detection via Accurate Cross-Modal Semantic Prompts
by: Zhao, Shijia, et al.
Published: (2025)
by: Zhao, Shijia, et al.
Published: (2025)
CVFusion: Cross-View Fusion of 4D Radar and Camera for 3D Object Detection
by: Zhong, Hanzhi, et al.
Published: (2025)
by: Zhong, Hanzhi, et al.
Published: (2025)
GEAL: Generalizable 3D Affordance Learning with Cross-Modal Consistency
by: Lu, Dongyue, et al.
Published: (2024)
by: Lu, Dongyue, et al.
Published: (2024)
Cross-View Completion Models are Zero-shot Correspondence Estimators
by: An, Honggyu, et al.
Published: (2024)
by: An, Honggyu, et al.
Published: (2024)
WonderFree: Enhancing Novel View Quality and Cross-View Consistency for 3D Scene Exploration
by: Ni, Chaojun, et al.
Published: (2025)
by: Ni, Chaojun, et al.
Published: (2025)
UPose3D: Uncertainty-Aware 3D Human Pose Estimation with Cross-View and Temporal Cues
by: Davoodnia, Vandad, et al.
Published: (2024)
by: Davoodnia, Vandad, et al.
Published: (2024)
3D Multi-View Stylization with Pose-Free Correspondences Matching for Robust 3D Geometry Preservation
by: Bose, Shirsha
Published: (2026)
by: Bose, Shirsha
Published: (2026)
Hierarchical Neural Semantic Representation for 3D Semantic Correspondence
by: Du, Keyu, et al.
Published: (2025)
by: Du, Keyu, et al.
Published: (2025)
CrossJEPA: Cross-Modal Joint-Embedding Predictive Architecture for Efficient 3D Representation Learning from 2D Images
by: Perera, Avishka, et al.
Published: (2025)
by: Perera, Avishka, et al.
Published: (2025)
Hierarchical Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
by: Zhao, Youjun, et al.
Published: (2025)
by: Zhao, Youjun, et al.
Published: (2025)
Learning Cross-View Object Correspondence via Cycle-Consistent Mask Prediction
by: Yan, Shannan, et al.
Published: (2026)
by: Yan, Shannan, et al.
Published: (2026)
iSegMan: Interactive Segment-and-Manipulate 3D Gaussians
by: Zhao, Yian, et al.
Published: (2025)
by: Zhao, Yian, et al.
Published: (2025)
Cross-Domain Few-Shot Segmentation via Iterative Support-Query Correspondence Mining
by: Nie, Jiahao, et al.
Published: (2024)
by: Nie, Jiahao, et al.
Published: (2024)
PC-CrossDiff: Point-Cluster Dual-Level Cross-Modal Differential Attention for Unified 3D Referring and Segmentation
by: Tan, Wenbin, et al.
Published: (2026)
by: Tan, Wenbin, et al.
Published: (2026)
3D Gaussian Inpainting with Depth-Guided Cross-View Consistency
by: Huang, Sheng-Yu, et al.
Published: (2025)
by: Huang, Sheng-Yu, et al.
Published: (2025)
3D-Aware Multi-Task Learning with Cross-View Correlations for Dense Scene Understanding
by: Wang, Xiaoye, et al.
Published: (2025)
by: Wang, Xiaoye, et al.
Published: (2025)
Leveraging Single-View Images for Unsupervised 3D Point Cloud Completion
by: Wu, Lintai, et al.
Published: (2022)
by: Wu, Lintai, et al.
Published: (2022)
Cross-Modal Attention Guided Unlearning in Vision-Language Models
by: Bhaila, Karuna, et al.
Published: (2025)
by: Bhaila, Karuna, et al.
Published: (2025)
FPL+: Filtered Pseudo Label-based Unsupervised Cross-Modality Adaptation for 3D Medical Image Segmentation
by: Wu, Jianghao, et al.
Published: (2024)
by: Wu, Jianghao, et al.
Published: (2024)
CT-MVSNet: Efficient Multi-View Stereo with Cross-scale Transformer
by: Wang, Sicheng, et al.
Published: (2023)
by: Wang, Sicheng, et al.
Published: (2023)
MSAM: Multi-Semantic Adaptive Mining for Cross-Modal Drone Video-Text Retrieval
by: Huang, Jinghao, et al.
Published: (2025)
by: Huang, Jinghao, et al.
Published: (2025)
Cross-View Multi-Modal Segmentation @ Ego-Exo4D Challenges 2025
by: Fu, Yuqian, et al.
Published: (2025)
by: Fu, Yuqian, et al.
Published: (2025)
Multimodal LLM Enhanced Cross-lingual Cross-modal Retrieval
by: Wang, Yabing, et al.
Published: (2024)
by: Wang, Yabing, et al.
Published: (2024)
Camera-Aware Cross-View Alignment for Referring 3D Gaussian Splatting Segmentation
by: Tao, Yuwen, et al.
Published: (2025)
by: Tao, Yuwen, et al.
Published: (2025)
Cross-Temporal 3D Gaussian Splatting for Sparse-View Guided Scene Update
by: An, Zeyuan, et al.
Published: (2025)
by: An, Zeyuan, et al.
Published: (2025)
Similar Items
-
SCA3D: Enhancing Cross-modal 3D Retrieval via 3D Shape and Caption Paired Data Augmentation
by: Ren, Junlong, et al.
Published: (2025) -
3DBonsai: Structure-Aware Bonsai Modeling Using Conditioned 3D Gaussian Splatting
by: Wu, Hao, et al.
Published: (2025) -
CorrespondentDream: Enhancing 3D Fidelity of Text-to-3D using Cross-View Correspondences
by: Kim, Seungwook, et al.
Published: (2024) -
C3Po: Cross-View Cross-Modality Correspondence by Pointmap Prediction
by: Huang, Kuan Wei, et al.
Published: (2025) -
Enhanced Cross-modal 3D Retrieval via Tri-modal Reconstruction
by: Ren, Junlong, et al.
Published: (2025)