Benefit from Reference: Retrieval-Augmented Cross-modal Point Cloud Completion
Fuente:
arXiv
Saved in:
| Main Authors: | Hou, Hongye, Zhan, Liu, Yang, Yang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mask-aware Text-to-Image Retrieval: Referring Expression Segmentation Meets Cross-modal Retrieval
by: Shen, Li-Cheng, et al.
Published: (2025)
by: Shen, Li-Cheng, et al.
Published: (2025)
MAGE:A Multi-stage Avatar Generator with Sparse Observations
by: Du, Fangyu, et al.
Published: (2025)
by: Du, Fangyu, et al.
Published: (2025)
PointCG: Self-supervised Point Cloud Learning via Joint Completion and Generation
by: Liu, Yun, et al.
Published: (2024)
by: Liu, Yun, et al.
Published: (2024)
RAAP: Retrieval-Augmented Affordance Prediction with Cross-Image Action Alignment
by: Zhuang, Qiyuan, et al.
Published: (2026)
by: Zhuang, Qiyuan, et al.
Published: (2026)
Enhancing Performance of Point Cloud Completion Networks with Consistency Loss
by: Wijaya, Kevin Tirta, et al.
Published: (2024)
by: Wijaya, Kevin Tirta, et al.
Published: (2024)
Unsupervised Point Cloud Completion through Unbalanced Optimal Transport
by: Lee, Taekyung, et al.
Published: (2024)
by: Lee, Taekyung, et al.
Published: (2024)
Position-aware Guided Point Cloud Completion with CLIP Model
by: Zhou, Feng, et al.
Published: (2024)
by: Zhou, Feng, et al.
Published: (2024)
Explicitly Guided Information Interaction Network for Cross-modal Point Cloud Completion
by: Xu, Hang, et al.
Published: (2024)
by: Xu, Hang, et al.
Published: (2024)
DC-PCN: Point Cloud Completion Network with Dual-Codebook Guided Quantization
by: Wu, Qiuxia, et al.
Published: (2025)
by: Wu, Qiuxia, et al.
Published: (2025)
Masked Contrastive Reconstruction for Cross-modal Medical Image-Report Retrieval
by: Wei, Zeqiang, et al.
Published: (2023)
by: Wei, Zeqiang, et al.
Published: (2023)
ViewSAM: Learning View-aware Cross-modal Semantics for Weakly Supervised Cross-view Referring Multi-Object Tracking
by: Ge, Jiawei, et al.
Published: (2026)
by: Ge, Jiawei, et al.
Published: (2026)
Cross-modal RAG: Sub-dimensional Text-to-Image Retrieval-Augmented Generation
by: Zhu, Mengdan, et al.
Published: (2025)
by: Zhu, Mengdan, et al.
Published: (2025)
Generate Point Clouds with Multiscale Details from Graph-Represented Structures
by: Yang, Ximing, et al.
Published: (2021)
by: Yang, Ximing, et al.
Published: (2021)
Self-Supervised Cross-Modal Learning for Image-to-Point Cloud Registration
by: Wang, Xingmei, et al.
Published: (2025)
by: Wang, Xingmei, et al.
Published: (2025)
PointSSC: A Cooperative Vehicle-Infrastructure Point Cloud Benchmark for Semantic Scene Completion
by: Yan, Yuxiang, et al.
Published: (2023)
by: Yan, Yuxiang, et al.
Published: (2023)
Hierarchical Multi-modal Transformer for Cross-modal Long Document Classification
by: Liu, Tengfei, et al.
Published: (2024)
by: Liu, Tengfei, et al.
Published: (2024)
Knowledge Completes the Vision: A Multimodal Entity-aware Retrieval-Augmented Generation Framework for News Image Captioning
by: You, Xiaoxing, et al.
Published: (2025)
by: You, Xiaoxing, et al.
Published: (2025)
mRAG: Elucidating the Design Space of Multi-modal Retrieval-Augmented Generation
by: Hu, Chan-Wei, et al.
Published: (2025)
by: Hu, Chan-Wei, et al.
Published: (2025)
CMHANet: A Cross-Modal Hybrid Attention Network for Point Cloud Registration
by: Zhang, Dongxu, et al.
Published: (2026)
by: Zhang, Dongxu, et al.
Published: (2026)
Rethinking Multimodal Point Cloud Completion: A Completion-by-Correction Perspective
by: Luo, Wang, et al.
Published: (2025)
by: Luo, Wang, et al.
Published: (2025)
PointCloud-Text Matching: Benchmark Datasets and a Baseline
by: Feng, Yanglin, et al.
Published: (2024)
by: Feng, Yanglin, et al.
Published: (2024)
Efficient Point Cloud Processing with High-Dimensional Positional Encoding and Non-Local MLPs
by: Zou, Yanmei, et al.
Published: (2026)
by: Zou, Yanmei, et al.
Published: (2026)
Referring Remote Sensing Image Segmentation with Cross-view Semantics Interaction Network
by: Yang, Jiaxing, et al.
Published: (2025)
by: Yang, Jiaxing, et al.
Published: (2025)
TMCIR: Token Merge Benefits Composed Image Retrieval
by: Wang, Chaoyang, et al.
Published: (2025)
by: Wang, Chaoyang, et al.
Published: (2025)
Fusion-Mamba for Cross-modality Object Detection
by: Dong, Wenhao, et al.
Published: (2024)
by: Dong, Wenhao, et al.
Published: (2024)
Taming a Retrieval Framework to Read Images in Humanlike Manner for Augmenting Generation of MLLMs
by: Xi, Suyang, et al.
Published: (2025)
by: Xi, Suyang, et al.
Published: (2025)
SceneRAG: Scene-level Retrieval-Augmented Generation for Video Understanding
by: Zeng, Nianbo, et al.
Published: (2025)
by: Zeng, Nianbo, et al.
Published: (2025)
Hierarchical Direction Perception via Atomic Dot-Product Operators for Rotation-Invariant Point Clouds Learning
by: Hu, Chenyu, et al.
Published: (2025)
by: Hu, Chenyu, et al.
Published: (2025)
3D Single-object Tracking in Point Clouds with High Temporal Variation
by: Wu, Qiao, et al.
Published: (2024)
by: Wu, Qiao, et al.
Published: (2024)
Img2Loc: Revisiting Image Geolocalization using Multi-modality Foundation Models and Image-based Retrieval-Augmented Generation
by: Zhou, Zhongliang, et al.
Published: (2024)
by: Zhou, Zhongliang, et al.
Published: (2024)
MRG: A Multi-Robot Manufacturing Digital Scene Generation Method Using Multi-Instance Point Cloud Registration
by: Han, Songjie, et al.
Published: (2025)
by: Han, Songjie, et al.
Published: (2025)
A Text-Image Fusion Method with Data Augmentation Capabilities for Referring Medical Image Segmentation
by: Chai, Shurong, et al.
Published: (2025)
by: Chai, Shurong, et al.
Published: (2025)
LINR-PCGC: Lossless Implicit Neural Representations for Point Cloud Geometry Compression
by: Huang, Wenjie, et al.
Published: (2025)
by: Huang, Wenjie, et al.
Published: (2025)
Complementarity-driven Representation Learning for Multi-modal Knowledge Graph Completion
by: Li, Lijian
Published: (2025)
by: Li, Lijian
Published: (2025)
Points-to-3D: Structure-Aware 3D Generation with Point Cloud Priors
by: Xia, Jiatong, et al.
Published: (2026)
by: Xia, Jiatong, et al.
Published: (2026)
ResAgent: Entropy-based Prior Point Discovery and Visual Reasoning for Referring Expression Segmentation
by: Wang, Yihao, et al.
Published: (2026)
by: Wang, Yihao, et al.
Published: (2026)
FreeEdit: Mask-free Reference-based Image Editing with Multi-modal Instruction
by: He, Runze, et al.
Published: (2024)
by: He, Runze, et al.
Published: (2024)
Voxel-based Point Cloud Geometry Compression with Space-to-Channel Context
by: Liu, Bojun, et al.
Published: (2025)
by: Liu, Bojun, et al.
Published: (2025)
Variational Adapter for Cross-modal Similarity Representation
by: Wei, WenZhang, et al.
Published: (2026)
by: Wei, WenZhang, et al.
Published: (2026)
Enhancing Sampling Protocol for Point Cloud Classification Against Corruptions
by: Li, Chongshou, et al.
Published: (2024)
by: Li, Chongshou, et al.
Published: (2024)
Similar Items
-
Mask-aware Text-to-Image Retrieval: Referring Expression Segmentation Meets Cross-modal Retrieval
by: Shen, Li-Cheng, et al.
Published: (2025) -
MAGE:A Multi-stage Avatar Generator with Sparse Observations
by: Du, Fangyu, et al.
Published: (2025) -
PointCG: Self-supervised Point Cloud Learning via Joint Completion and Generation
by: Liu, Yun, et al.
Published: (2024) -
RAAP: Retrieval-Augmented Affordance Prediction with Cross-Image Action Alignment
by: Zhuang, Qiyuan, et al.
Published: (2026) -
Enhancing Performance of Point Cloud Completion Networks with Consistency Loss
by: Wijaya, Kevin Tirta, et al.
Published: (2024)