CLIP Multi-modal Hashing for Multimedia Retrieval
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhu, Jian, Sheng, Mingkai, Huang, Zhangmin, Chang, Jingfei, Jiang, Jinling, Long, Jian, Luo, Cheng, Liu, Lei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adaptive Confidence Multi-View Hashing for Multimedia Retrieval
von: Zhu, Jian, et al.
Veröffentlicht: (2023)
von: Zhu, Jian, et al.
Veröffentlicht: (2023)
Trusted Mamba Contrastive Network for Multi-View Clustering
von: Zhu, Jian, et al.
Veröffentlicht: (2024)
von: Zhu, Jian, et al.
Veröffentlicht: (2024)
Lightweight Contrastive Distilled Hashing for Online Cross-modal Retrieval
von: Li, Jiaxing, et al.
Veröffentlicht: (2025)
von: Li, Jiaxing, et al.
Veröffentlicht: (2025)
Distribution-Consistency-Guided Multi-modal Hashing
von: Liu, Jin-Yu, et al.
Veröffentlicht: (2024)
von: Liu, Jin-Yu, et al.
Veröffentlicht: (2024)
M$^3$amba: CLIP-driven Mamba Model for Multi-modal Remote Sensing Classification
von: Cao, Mingxiang, et al.
Veröffentlicht: (2025)
von: Cao, Mingxiang, et al.
Veröffentlicht: (2025)
From CLIP to DINO: Visual Encoders Shout in Multi-modal Large Language Models
von: Jiang, Dongsheng, et al.
Veröffentlicht: (2023)
von: Jiang, Dongsheng, et al.
Veröffentlicht: (2023)
InterCLIP-MEP: Interactive CLIP and Memory-Enhanced Predictor for Multi-modal Sarcasm Detection
von: Chen, Junjie, et al.
Veröffentlicht: (2024)
von: Chen, Junjie, et al.
Veröffentlicht: (2024)
FLEX-CLIP: Feature-Level GEneration Network Enhanced CLIP for X-shot Cross-modal Retrieval
von: Xie, Jingyou, et al.
Veröffentlicht: (2024)
von: Xie, Jingyou, et al.
Veröffentlicht: (2024)
Harmonizing and Merging Source Models for CLIP-based Domain Generalization
von: Ding, Yuhe, et al.
Veröffentlicht: (2025)
von: Ding, Yuhe, et al.
Veröffentlicht: (2025)
Deep Hashing with Semantic Hash Centers for Image Retrieval
von: Chen, Li, et al.
Veröffentlicht: (2025)
von: Chen, Li, et al.
Veröffentlicht: (2025)
HybridHash: Hybrid Convolutional and Self-Attention Deep Hashing for Image Retrieval
von: He, Chao, et al.
Veröffentlicht: (2024)
von: He, Chao, et al.
Veröffentlicht: (2024)
Multiple Code Hashing for Efficient Image Retrieval
von: Li, Ming-Wei, et al.
Veröffentlicht: (2020)
von: Li, Ming-Wei, et al.
Veröffentlicht: (2020)
Multi-Perspective Subimage CLIP with Keyword Guidance for Remote Sensing Image-Text Retrieval
von: Li, Yifan, et al.
Veröffentlicht: (2026)
von: Li, Yifan, et al.
Veröffentlicht: (2026)
NeuroHash: A Hyperdimensional Neuro-Symbolic Framework for Spatially-Aware Image Hashing and Retrieval
von: Yun, Sanggeon, et al.
Veröffentlicht: (2024)
von: Yun, Sanggeon, et al.
Veröffentlicht: (2024)
MambaHash: Visual State Space Deep Hashing Model for Large-Scale Image Retrieval
von: He, Chao, et al.
Veröffentlicht: (2025)
von: He, Chao, et al.
Veröffentlicht: (2025)
Language Augmentation in CLIP for Improved Anatomy Detection on Multi-modal Medical Images
von: Kakkar, Mansi, et al.
Veröffentlicht: (2024)
von: Kakkar, Mansi, et al.
Veröffentlicht: (2024)
Multi-view Distillation based on Multi-modal Fusion for Few-shot Action Recognition(CLIP-$\mathrm{M^2}$DF)
von: Guo, Fei, et al.
Veröffentlicht: (2024)
von: Guo, Fei, et al.
Veröffentlicht: (2024)
CLIP Adaptation by Intra-modal Overlap Reduction
von: Kravets, Alexey, et al.
Veröffentlicht: (2024)
von: Kravets, Alexey, et al.
Veröffentlicht: (2024)
Towards Cross-modal Retrieval in Chinese Cultural Heritage Documents: Dataset and Solution
von: Yuan, Junyi, et al.
Veröffentlicht: (2025)
von: Yuan, Junyi, et al.
Veröffentlicht: (2025)
Generative Diffusion Contrastive Network for Multi-View Clustering
von: Zhu, Jian, et al.
Veröffentlicht: (2025)
von: Zhu, Jian, et al.
Veröffentlicht: (2025)
CIBR: Cross-modal Information Bottleneck Regularization for Robust CLIP Generalization
von: Ji, Yingrui, et al.
Veröffentlicht: (2025)
von: Ji, Yingrui, et al.
Veröffentlicht: (2025)
THCRL: Trusted Hierarchical Contrastive Representation Learning for Multi-View Clustering
von: Zhu, Jian
Veröffentlicht: (2025)
von: Zhu, Jian
Veröffentlicht: (2025)
Building a Multi-modal Spatiotemporal Expert for Zero-shot Action Recognition with CLIP
von: Yu, Yating, et al.
Veröffentlicht: (2024)
von: Yu, Yating, et al.
Veröffentlicht: (2024)
MFCLIP: Multi-modal Fine-grained CLIP for Generalizable Diffusion Face Forgery Detection
von: Zhang, Yaning, et al.
Veröffentlicht: (2024)
von: Zhang, Yaning, et al.
Veröffentlicht: (2024)
CT-CLIP: A Multi-modal Fusion Framework for Robust Apple Leaf Disease Recognition in Complex Environments
von: Liu, Lemin, et al.
Veröffentlicht: (2025)
von: Liu, Lemin, et al.
Veröffentlicht: (2025)
Composed Multi-modal Retrieval: A Survey of Approaches and Applications
von: Zhang, Kun, et al.
Veröffentlicht: (2025)
von: Zhang, Kun, et al.
Veröffentlicht: (2025)
MaskedCLIP: Bridging the Masked and CLIP Space for Semi-Supervised Medical Vision-Language Pre-training
von: Zhu, Lei, et al.
Veröffentlicht: (2025)
von: Zhu, Lei, et al.
Veröffentlicht: (2025)
GET: Unlocking the Multi-modal Potential of CLIP for Generalized Category Discovery
von: Wang, Enguang, et al.
Veröffentlicht: (2024)
von: Wang, Enguang, et al.
Veröffentlicht: (2024)
Image-to-Brain Signal Generation for Visual Prosthesis with CLIP Guided Multimodal Diffusion Models
von: Xu, Ganxi, et al.
Veröffentlicht: (2025)
von: Xu, Ganxi, et al.
Veröffentlicht: (2025)
Rethinking Prior Information Generation with CLIP for Few-Shot Segmentation
von: Wang, Jin, et al.
Veröffentlicht: (2024)
von: Wang, Jin, et al.
Veröffentlicht: (2024)
ConceptHash: Interpretable Fine-Grained Hashing via Concept Discovery
von: Ng, Kam Woh, et al.
Veröffentlicht: (2024)
von: Ng, Kam Woh, et al.
Veröffentlicht: (2024)
UniM$^2$AE: Multi-modal Masked Autoencoders with Unified 3D Representation for 3D Perception in Autonomous Driving
von: Zou, Jian, et al.
Veröffentlicht: (2023)
von: Zou, Jian, et al.
Veröffentlicht: (2023)
Frame-Difference Guided Dynamic Region Perception for CLIP Adaptation in Text-Video Retrieval
von: Yu, Jiaao, et al.
Veröffentlicht: (2025)
von: Yu, Jiaao, et al.
Veröffentlicht: (2025)
DialogGen: Multi-modal Interactive Dialogue System for Multi-turn Text-to-Image Generation
von: Huang, Minbin, et al.
Veröffentlicht: (2024)
von: Huang, Minbin, et al.
Veröffentlicht: (2024)
Realistic Unsupervised CLIP Fine-tuning with Universal Entropy Optimization
von: Liang, Jian, et al.
Veröffentlicht: (2023)
von: Liang, Jian, et al.
Veröffentlicht: (2023)
GLCONet: Learning Multi-source Perception Representation for Camouflaged Object Detection
von: Sun, Yanguang, et al.
Veröffentlicht: (2024)
von: Sun, Yanguang, et al.
Veröffentlicht: (2024)
DTR: A Unified Deep Tensor Representation Framework for Multimedia Data Recovery
von: Zhou, Ting-Wei, et al.
Veröffentlicht: (2024)
von: Zhou, Ting-Wei, et al.
Veröffentlicht: (2024)
Grid4D: 4D Decomposed Hash Encoding for High-Fidelity Dynamic Gaussian Splatting
von: Xu, Jiawei, et al.
Veröffentlicht: (2024)
von: Xu, Jiawei, et al.
Veröffentlicht: (2024)
M4-SAR: A Multi-Resolution, Multi-Polarization, Multi-Scene, Multi-Source Dataset and Benchmark for optical-SAR Object Detection
von: Wang, Chao, et al.
Veröffentlicht: (2025)
von: Wang, Chao, et al.
Veröffentlicht: (2025)
NAU-QMUL: Utilizing BERT and CLIP for Multi-modal AI-Generated Image Detection
von: Guo, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Guo, Xiaoyu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Adaptive Confidence Multi-View Hashing for Multimedia Retrieval
von: Zhu, Jian, et al.
Veröffentlicht: (2023) -
Trusted Mamba Contrastive Network for Multi-View Clustering
von: Zhu, Jian, et al.
Veröffentlicht: (2024) -
Lightweight Contrastive Distilled Hashing for Online Cross-modal Retrieval
von: Li, Jiaxing, et al.
Veröffentlicht: (2025) -
Distribution-Consistency-Guided Multi-modal Hashing
von: Liu, Jin-Yu, et al.
Veröffentlicht: (2024) -
M$^3$amba: CLIP-driven Mamba Model for Multi-modal Remote Sensing Classification
von: Cao, Mingxiang, et al.
Veröffentlicht: (2025)