Enhanced Cross-modal 3D Retrieval via Tri-modal Reconstruction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ren, Junlong, Wang, Hao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SCA3D: Enhancing Cross-modal 3D Retrieval via 3D Shape and Caption Paired Data Augmentation
von: Ren, Junlong, et al.
Veröffentlicht: (2025)
von: Ren, Junlong, et al.
Veröffentlicht: (2025)
Multimodal LLM Enhanced Cross-lingual Cross-modal Retrieval
von: Wang, Yabing, et al.
Veröffentlicht: (2024)
von: Wang, Yabing, et al.
Veröffentlicht: (2024)
Prototype-based Aleatoric Uncertainty Quantification for Cross-modal Retrieval
von: Li, Hao, et al.
Veröffentlicht: (2023)
von: Li, Hao, et al.
Veröffentlicht: (2023)
Masked Contrastive Reconstruction for Cross-modal Medical Image-Report Retrieval
von: Wei, Zeqiang, et al.
Veröffentlicht: (2023)
von: Wei, Zeqiang, et al.
Veröffentlicht: (2023)
Dynamic Adapter with Semantics Disentangling for Cross-lingual Cross-modal Retrieval
von: Cai, Rui, et al.
Veröffentlicht: (2024)
von: Cai, Rui, et al.
Veröffentlicht: (2024)
SeCG: Semantic-Enhanced 3D Visual Grounding via Cross-modal Graph Attention
von: Xiao, Feng, et al.
Veröffentlicht: (2024)
von: Xiao, Feng, et al.
Veröffentlicht: (2024)
CrossTracker: Robust Multi-modal 3D Multi-Object Tracking via Cross Correction
von: Gu, Lipeng, et al.
Veröffentlicht: (2024)
von: Gu, Lipeng, et al.
Veröffentlicht: (2024)
Lightweight Contrastive Distilled Hashing for Online Cross-modal Retrieval
von: Li, Jiaxing, et al.
Veröffentlicht: (2025)
von: Li, Jiaxing, et al.
Veröffentlicht: (2025)
Cross-modal Full-mode Fine-grained Alignment for Text-to-Image Person Retrieval
von: Yin, Hao, et al.
Veröffentlicht: (2025)
von: Yin, Hao, et al.
Veröffentlicht: (2025)
Deep Reversible Consistency Learning for Cross-modal Retrieval
von: Pu, Ruitao, et al.
Veröffentlicht: (2025)
von: Pu, Ruitao, et al.
Veröffentlicht: (2025)
Towards Cross-modal Retrieval in Chinese Cultural Heritage Documents: Dataset and Solution
von: Yuan, Junyi, et al.
Veröffentlicht: (2025)
von: Yuan, Junyi, et al.
Veröffentlicht: (2025)
WaMo: Wavelet-Enhanced Multi-Frequency Trajectory Analysis for Fine-Grained Text-Motion Retrieval
von: Ren, Junlong, et al.
Veröffentlicht: (2025)
von: Ren, Junlong, et al.
Veröffentlicht: (2025)
SCPNet: Unsupervised Cross-modal Homography Estimation via Intra-modal Self-supervised Learning
von: Zhang, Runmin, et al.
Veröffentlicht: (2024)
von: Zhang, Runmin, et al.
Veröffentlicht: (2024)
Mitigating Cross-modal Representation Bias for Multicultural Image-to-Recipe Retrieval
von: Wang, Qing, et al.
Veröffentlicht: (2025)
von: Wang, Qing, et al.
Veröffentlicht: (2025)
Cross-modal Prompting for Balanced Incomplete Multi-modal Emotion Recognition
von: He, Wen-Jue, et al.
Veröffentlicht: (2025)
von: He, Wen-Jue, et al.
Veröffentlicht: (2025)
AsyncBEV: Cross-modal Flow Alignment in Asynchronous 3D Object Detection
von: Wang, Shiming, et al.
Veröffentlicht: (2026)
von: Wang, Shiming, et al.
Veröffentlicht: (2026)
Enhancing Scientific Figure Captioning Through Cross-modal Learning
von: Rojas, Mateo Alejandro, et al.
Veröffentlicht: (2024)
von: Rojas, Mateo Alejandro, et al.
Veröffentlicht: (2024)
Enhanced Partially Relevant Video Retrieval through Inter- and Intra-Sample Analysis with Coherence Prediction
von: Ren, Junlong, et al.
Veröffentlicht: (2025)
von: Ren, Junlong, et al.
Veröffentlicht: (2025)
Enhancing Incomplete Multi-modal Brain Tumor Segmentation with Intra-modal Asymmetry and Inter-modal Dependency
von: Liu, Weide, et al.
Veröffentlicht: (2024)
von: Liu, Weide, et al.
Veröffentlicht: (2024)
Tri-modal Confluence with Temporal Dynamics for Scene Graph Generation in Operating Rooms
von: Guo, Diandian, et al.
Veröffentlicht: (2024)
von: Guo, Diandian, et al.
Veröffentlicht: (2024)
Fine-grained Context and Multi-modal Alignment for Freehand 3D Ultrasound Reconstruction
von: Yan, Zhongnuo, et al.
Veröffentlicht: (2024)
von: Yan, Zhongnuo, et al.
Veröffentlicht: (2024)
FLEX-CLIP: Feature-Level GEneration Network Enhanced CLIP for X-shot Cross-modal Retrieval
von: Xie, Jingyou, et al.
Veröffentlicht: (2024)
von: Xie, Jingyou, et al.
Veröffentlicht: (2024)
Mentor3AD: Feature Reconstruction-based 3D Anomaly Detection via Multi-modality Mentor Learning
von: Liang, Hanzhe
Veröffentlicht: (2025)
von: Liang, Hanzhe
Veröffentlicht: (2025)
Multi-modal Semantic Understanding with Contrastive Cross-modal Feature Alignment
von: Zhang, Ming, et al.
Veröffentlicht: (2024)
von: Zhang, Ming, et al.
Veröffentlicht: (2024)
Multi-modal Reference Learning for Fine-grained Text-to-Image Retrieval
von: Ma, Zehong, et al.
Veröffentlicht: (2025)
von: Ma, Zehong, et al.
Veröffentlicht: (2025)
Hierarchical Cross-Attention Network for Virtual Try-On
von: Tang, Hao, et al.
Veröffentlicht: (2024)
von: Tang, Hao, et al.
Veröffentlicht: (2024)
Hierarchical Cross-modal Prompt Learning for Vision-Language Models
von: Zheng, Hao, et al.
Veröffentlicht: (2025)
von: Zheng, Hao, et al.
Veröffentlicht: (2025)
Cross-modal learning for plankton recognition
von: Kareinen, Joona, et al.
Veröffentlicht: (2026)
von: Kareinen, Joona, et al.
Veröffentlicht: (2026)
OccGen: Generative Multi-modal 3D Occupancy Prediction for Autonomous Driving
von: Wang, Guoqing, et al.
Veröffentlicht: (2024)
von: Wang, Guoqing, et al.
Veröffentlicht: (2024)
CLIP Multi-modal Hashing for Multimedia Retrieval
von: Zhu, Jian, et al.
Veröffentlicht: (2024)
von: Zhu, Jian, et al.
Veröffentlicht: (2024)
Unsupervised Spike Depth Estimation via Cross-modality Cross-domain Knowledge Transfer
von: Liu, Jiaming, et al.
Veröffentlicht: (2022)
von: Liu, Jiaming, et al.
Veröffentlicht: (2022)
T2ICount: Enhancing Cross-modal Understanding for Zero-Shot Counting
von: Qian, Yifei, et al.
Veröffentlicht: (2025)
von: Qian, Yifei, et al.
Veröffentlicht: (2025)
Reliable Cross-modal Alignment via Prototype Iterative Construction
von: Ma, Xiang, et al.
Veröffentlicht: (2025)
von: Ma, Xiang, et al.
Veröffentlicht: (2025)
Multi-modal Generation via Cross-Modal In-Context Learning
von: Kumar, Amandeep, et al.
Veröffentlicht: (2024)
von: Kumar, Amandeep, et al.
Veröffentlicht: (2024)
CDPR: Cross-modal Diffusion with Polarization for Reliable Monocular Depth Estimation
von: Yu, Rongjia, et al.
Veröffentlicht: (2026)
von: Yu, Rongjia, et al.
Veröffentlicht: (2026)
Eliminating Cross-modal Conflicts in BEV Space for LiDAR-Camera 3D Object Detection
von: Fu, Jiahui, et al.
Veröffentlicht: (2024)
von: Fu, Jiahui, et al.
Veröffentlicht: (2024)
A 3D Cross-modal Keypoint Descriptor for MR-US Matching and Registration
von: Morozov, Daniil, et al.
Veröffentlicht: (2025)
von: Morozov, Daniil, et al.
Veröffentlicht: (2025)
CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models
von: Huang, Junming, et al.
Veröffentlicht: (2026)
von: Huang, Junming, et al.
Veröffentlicht: (2026)
TRRG: Towards Truthful Radiology Report Generation With Cross-modal Disease Clue Enhanced Large Language Model
von: Wang, Yuhao, et al.
Veröffentlicht: (2024)
von: Wang, Yuhao, et al.
Veröffentlicht: (2024)
3DMIT: 3D Multi-modal Instruction Tuning for Scene Understanding
von: Li, Zeju, et al.
Veröffentlicht: (2024)
von: Li, Zeju, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
SCA3D: Enhancing Cross-modal 3D Retrieval via 3D Shape and Caption Paired Data Augmentation
von: Ren, Junlong, et al.
Veröffentlicht: (2025) -
Multimodal LLM Enhanced Cross-lingual Cross-modal Retrieval
von: Wang, Yabing, et al.
Veröffentlicht: (2024) -
Prototype-based Aleatoric Uncertainty Quantification for Cross-modal Retrieval
von: Li, Hao, et al.
Veröffentlicht: (2023) -
Masked Contrastive Reconstruction for Cross-modal Medical Image-Report Retrieval
von: Wei, Zeqiang, et al.
Veröffentlicht: (2023) -
Dynamic Adapter with Semantics Disentangling for Cross-lingual Cross-modal Retrieval
von: Cai, Rui, et al.
Veröffentlicht: (2024)