Contrastive masked auto-encoders based self-supervised hashing for 2D image and 3D point cloud cross-modal retrieval
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Rukai, Cui, Heng, Liu, Yu, Hou, Yufeng, Xie, Yanzhao, Zhou, Ke |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hyperbolic Hierarchical Contrastive Hashing
by: Wei, Rukai, et al.
Published: (2022)
by: Wei, Rukai, et al.
Published: (2022)
A multifractal-based masked auto-encoder: an application to medical images
by: Florindo, Joao Batista, et al.
Published: (2026)
by: Florindo, Joao Batista, et al.
Published: (2026)
Enhancing image quality prediction with self-supervised visual masking
by: Çoğalan, Uğur, et al.
Published: (2023)
by: Çoğalan, Uğur, et al.
Published: (2023)
L-MAE: Longitudinal masked auto-encoder with time and severity-aware encoding for diabetic retinopathy progression prediction
by: Zeghlache, Rachid, et al.
Published: (2024)
by: Zeghlache, Rachid, et al.
Published: (2024)
Fast 3D point clouds retrieval for Large-scale 3D Place Recognition
by: Zede, Chahine-Nicolas, et al.
Published: (2025)
by: Zede, Chahine-Nicolas, et al.
Published: (2025)
Contrastive-learning of language embedding and biological features for cross modality encoding and effector prediction.
by: Peng, Yue, et al.
Published: (2025)
by: Peng, Yue, et al.
Published: (2025)
Efficient physics-informed neural networks using hash encoding
by: Huang, Xinquan, et al.
Published: (2023)
by: Huang, Xinquan, et al.
Published: (2023)
Co-distilled attention guided masked image modeling with noisy teacher for self-supervised learning on medical images
by: Jiang, Jue, et al.
Published: (2026)
by: Jiang, Jue, et al.
Published: (2026)
S-MolSearch: 3D Semi-supervised Contrastive Learning for Bioactive Molecule Search
by: Zhou, Gengmo, et al.
Published: (2024)
by: Zhou, Gengmo, et al.
Published: (2024)
ExpPoint-MAE: Better interpretability and performance for self-supervised point cloud transformers
by: Romanelis, Ioannis, et al.
Published: (2023)
by: Romanelis, Ioannis, et al.
Published: (2023)
LPRnet: A self-supervised registration network for LiDAR and photogrammetric point clouds
by: Wang, Chen, et al.
Published: (2025)
by: Wang, Chen, et al.
Published: (2025)
Animal behavioral analysis and neural encoding with transformer-based self-supervised pretraining
by: Wang, Yanchen, et al.
Published: (2025)
by: Wang, Yanchen, et al.
Published: (2025)
SAFE: a SAR Feature Extractor based on self-supervised learning and masked Siamese ViTs
by: Muzeau, Max, et al.
Published: (2024)
by: Muzeau, Max, et al.
Published: (2024)
Template masks for 4D-STEM
by: Xie, Yining, et al.
Published: (2025)
by: Xie, Yining, et al.
Published: (2025)
Hi-End-MAE: Hierarchical encoder-driven masked autoencoders are stronger vision learners for medical image segmentation
by: Tang, Fenghe, et al.
Published: (2025)
by: Tang, Fenghe, et al.
Published: (2025)
An overview of neural architectures for self-supervised audio representation learning from masked spectrograms
by: Yadav, Sarthak, et al.
Published: (2025)
by: Yadav, Sarthak, et al.
Published: (2025)
Exploring bat song syllable representations in self-supervised audio encoders
by: Kloots, Marianne de Heer, et al.
Published: (2024)
by: Kloots, Marianne de Heer, et al.
Published: (2024)
Exploring scalable medical image encoders beyond text supervision
by: Pérez-García, Fernando, et al.
Published: (2024)
by: Pérez-García, Fernando, et al.
Published: (2024)
Exploiting scattering-based point spread functions for snapshot 5D and modality-switchable lensless imaging
by: Zheng, Ze, et al.
Published: (2025)
by: Zheng, Ze, et al.
Published: (2025)
MCAD: Multi-teacher Cross-modal Alignment Distillation for efficient image-text retrieval
by: Lei, Youbo, et al.
Published: (2023)
by: Lei, Youbo, et al.
Published: (2023)
Optimized encoding point distributions for efficient single-point imaging
by: Bschorr, Fabian, et al.
Published: (2026)
by: Bschorr, Fabian, et al.
Published: (2026)
Rigid point cloud registration based on correspondence cloud for image‐to‐patient registration in image‐guided surgery
by: Zhihao Li, et al.
Published: (2024)
by: Zhihao Li, et al.
Published: (2024)
PointVoxelFormer -- Reviving point cloud networks for 3D medical imaging
by: Heinrich, Mattias Paul
Published: (2024)
by: Heinrich, Mattias Paul
Published: (2024)
A generalizable 3D framework and model for self-supervised learning in medical imaging
by: Xu, Tony, et al.
Published: (2025)
by: Xu, Tony, et al.
Published: (2025)
Super-resolution of biomedical volumes with 2D supervision
by: Jiang, Cheng, et al.
Published: (2024)
by: Jiang, Cheng, et al.
Published: (2024)
Making medical vision-language models think causally across modalities with retrieval-augmented cross-modal reasoning
by: Yang, Weiqin, et al.
Published: (2026)
by: Yang, Weiqin, et al.
Published: (2026)
MC2SleepNet: Multi-modal Cross-masking with Contrastive Learning for Sleep Stage Classification
by: Na, Younghoon, et al.
Published: (2025)
by: Na, Younghoon, et al.
Published: (2025)
DenserRadar: A 4D millimeter-wave radar point cloud detector based on dense LiDAR point clouds
by: Han, Zeyu, et al.
Published: (2024)
by: Han, Zeyu, et al.
Published: (2024)
Cayley hashing with cookies
by: Shpilrain, Vladimir, et al.
Published: (2024)
by: Shpilrain, Vladimir, et al.
Published: (2024)
Multi-encoder nnU-Net outperforms transformer models with self-supervised pretraining
by: Otaghsara, Seyedeh Sahar Taheri, et al.
Published: (2025)
by: Otaghsara, Seyedeh Sahar Taheri, et al.
Published: (2025)
CAGS: Open-Vocabulary 3D Scene Understanding with Context-Aware Gaussian Splatting
by: Sun, Wei, et al.
Published: (2025)
by: Sun, Wei, et al.
Published: (2025)
Self-supervised Learning-based Reconstruction of High-resolution 4D Light Fields
by: Lei, Jianxin, et al.
Published: (2024)
by: Lei, Jianxin, et al.
Published: (2024)
Multi-modal Semantic Understanding with Contrastive Cross-modal Feature Alignment
by: Zhang, Ming, et al.
Published: (2024)
by: Zhang, Ming, et al.
Published: (2024)
3D point cloud of the Hundsalm ice cave
by: Pfeiffer, Jan, et al.
Published: (2022)
by: Pfeiffer, Jan, et al.
Published: (2022)
EdgeFormer: local patch-based edge detection transformer on point clouds
by: Xie, Yifei, et al.
Published: (2026)
by: Xie, Yifei, et al.
Published: (2026)
Compile-once block encodings for masked similarity-transformed effective Hamiltonians
by: Peng, Bo, et al.
Published: (2026)
by: Peng, Bo, et al.
Published: (2026)
Scaling up masked audio encoder learning for general audio classification
by: Dinkel, Heinrich, et al.
Published: (2024)
by: Dinkel, Heinrich, et al.
Published: (2024)
Fast and accurate sparse-view CBCT reconstruction using meta-learned neural attenuation field and hash-encoding regularization
by: Shin, Heejun, et al.
Published: (2023)
by: Shin, Heejun, et al.
Published: (2023)
Probing self-attention in self-supervised speech models for cross-linguistic differences
by: Gopinath, Sai, et al.
Published: (2024)
by: Gopinath, Sai, et al.
Published: (2024)
SlingBAG Pro: Accelerating point cloud-based iterative reconstruction for 3D photoacoustic imaging with arbitrary array geometries
by: Li, Shuang, et al.
Published: (2026)
by: Li, Shuang, et al.
Published: (2026)
Similar Items
-
Hyperbolic Hierarchical Contrastive Hashing
by: Wei, Rukai, et al.
Published: (2022) -
A multifractal-based masked auto-encoder: an application to medical images
by: Florindo, Joao Batista, et al.
Published: (2026) -
Enhancing image quality prediction with self-supervised visual masking
by: Çoğalan, Uğur, et al.
Published: (2023) -
L-MAE: Longitudinal masked auto-encoder with time and severity-aware encoding for diabetic retinopathy progression prediction
by: Zeghlache, Rachid, et al.
Published: (2024) -
Fast 3D point clouds retrieval for Large-scale 3D Place Recognition
by: Zede, Chahine-Nicolas, et al.
Published: (2025)