FLEX-CLIP: Feature-Level GEneration Network Enhanced CLIP for X-shot Cross-modal Retrieval
Fuente:
arXiv
Salvato in:
| Autori principali: | Xie, Jingyou, Kuang, Jiayi, Lin, Zhenzhou, Ouyang, Jiarui, Zhao, Zishuo, Shen, Ying |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DetailCLIP: Injecting Image Details into CLIP's Feature Space
di: Zhang, Zilun, et al.
Pubblicazione: (2022)
di: Zhang, Zilun, et al.
Pubblicazione: (2022)
CLIP Multi-modal Hashing for Multimedia Retrieval
di: Zhu, Jian, et al.
Pubblicazione: (2024)
di: Zhu, Jian, et al.
Pubblicazione: (2024)
InterCLIP-MEP: Interactive CLIP and Memory-Enhanced Predictor for Multi-modal Sarcasm Detection
di: Chen, Junjie, et al.
Pubblicazione: (2024)
di: Chen, Junjie, et al.
Pubblicazione: (2024)
Is CLIP Cross-Eyed? Revealing and Mitigating Center Bias in the CLIP Family
di: Chew, Oscar, et al.
Pubblicazione: (2026)
di: Chew, Oscar, et al.
Pubblicazione: (2026)
MadCLIP: Few-shot Medical Anomaly Detection with CLIP
di: Shiri, Mahshid, et al.
Pubblicazione: (2025)
di: Shiri, Mahshid, et al.
Pubblicazione: (2025)
CAGE: Controllable Articulation GEneration
di: Liu, Jiayi, et al.
Pubblicazione: (2023)
di: Liu, Jiayi, et al.
Pubblicazione: (2023)
AA-CLIP: Enhancing Zero-shot Anomaly Detection via Anomaly-Aware CLIP
di: Ma, Wenxin, et al.
Pubblicazione: (2025)
di: Ma, Wenxin, et al.
Pubblicazione: (2025)
GenCLIP: Generalizing CLIP Prompts for Zero-shot Anomaly Detection
di: Kim, Donghyeong, et al.
Pubblicazione: (2025)
di: Kim, Donghyeong, et al.
Pubblicazione: (2025)
DialCLIP: Empowering CLIP as Multi-Modal Dialog Retriever
di: Yin, Zhichao, et al.
Pubblicazione: (2024)
di: Yin, Zhichao, et al.
Pubblicazione: (2024)
MediCLIP: Adapting CLIP for Few-shot Medical Image Anomaly Detection
di: Zhang, Ximiao, et al.
Pubblicazione: (2024)
di: Zhang, Ximiao, et al.
Pubblicazione: (2024)
Natural Language Understanding and Inference with MLLM in Visual Question Answering: A Survey
di: Kuang, Jiayi, et al.
Pubblicazione: (2024)
di: Kuang, Jiayi, et al.
Pubblicazione: (2024)
Building a Multi-modal Spatiotemporal Expert for Zero-shot Action Recognition with CLIP
di: Yu, Yating, et al.
Pubblicazione: (2024)
di: Yu, Yating, et al.
Pubblicazione: (2024)
IsoCLIP: Decomposing CLIP Projectors for Efficient Intra-modal Alignment
di: Magistri, Simone, et al.
Pubblicazione: (2026)
di: Magistri, Simone, et al.
Pubblicazione: (2026)
CLIP is Strong Enough to Fight Back: Test-time Counterattacks towards Zero-shot Adversarial Robustness of CLIP
di: Xing, Songlong, et al.
Pubblicazione: (2025)
di: Xing, Songlong, et al.
Pubblicazione: (2025)
Distill CLIP (DCLIP): Enhancing Image-Text Retrieval via Cross-Modal Transformer Distillation
di: Csizmadia, Daniel, et al.
Pubblicazione: (2025)
di: Csizmadia, Daniel, et al.
Pubblicazione: (2025)
MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition
di: Song, Dan, et al.
Pubblicazione: (2023)
di: Song, Dan, et al.
Pubblicazione: (2023)
CIBR: Cross-modal Information Bottleneck Regularization for Robust CLIP Generalization
di: Ji, Yingrui, et al.
Pubblicazione: (2025)
di: Ji, Yingrui, et al.
Pubblicazione: (2025)
Enhancing CLIP Robustness via Cross-Modality Alignment
di: Zhu, Xingyu, et al.
Pubblicazione: (2025)
di: Zhu, Xingyu, et al.
Pubblicazione: (2025)
CLIP Behaves like a Bag-of-Words Model Cross-modally but not Uni-modally
di: Koishigarina, Darina, et al.
Pubblicazione: (2025)
di: Koishigarina, Darina, et al.
Pubblicazione: (2025)
CardiacCLIP: Video-based CLIP Adaptation for LVEF Prediction in a Few-shot Manner
di: Du, Yao, et al.
Pubblicazione: (2025)
di: Du, Yao, et al.
Pubblicazione: (2025)
Jina CLIP: Your CLIP Model Is Also Your Text Retriever
di: Koukounas, Andreas, et al.
Pubblicazione: (2024)
di: Koukounas, Andreas, et al.
Pubblicazione: (2024)
GET: Unlocking the Multi-modal Potential of CLIP for Generalized Category Discovery
di: Wang, Enguang, et al.
Pubblicazione: (2024)
di: Wang, Enguang, et al.
Pubblicazione: (2024)
CLIP Adaptation by Intra-modal Overlap Reduction
di: Kravets, Alexey, et al.
Pubblicazione: (2024)
di: Kravets, Alexey, et al.
Pubblicazione: (2024)
Adversarial Backdoor Defense in CLIP
di: Kuang, Junhao, et al.
Pubblicazione: (2024)
di: Kuang, Junhao, et al.
Pubblicazione: (2024)
CLIP-guided Prototype Modulating for Few-shot Action Recognition
di: Wang, Xiang, et al.
Pubblicazione: (2023)
di: Wang, Xiang, et al.
Pubblicazione: (2023)
CLIPS: An Enhanced CLIP Framework for Learning with Synthetic Captions
di: Liu, Yanqing, et al.
Pubblicazione: (2024)
di: Liu, Yanqing, et al.
Pubblicazione: (2024)
AMU-Tuning: Effective Logit Bias for CLIP-based Few-shot Learning
di: Tang, Yuwei, et al.
Pubblicazione: (2024)
di: Tang, Yuwei, et al.
Pubblicazione: (2024)
SuperCLIP: CLIP with Simple Classification Supervision
di: Zhao, Weiheng, et al.
Pubblicazione: (2025)
di: Zhao, Weiheng, et al.
Pubblicazione: (2025)
LowCLIP: Adapting the CLIP Model Architecture for Low-Resource Languages in Multimodal Image Retrieval Task
di: Asgarov, Ali, et al.
Pubblicazione: (2024)
di: Asgarov, Ali, et al.
Pubblicazione: (2024)
CLIP-driven Outliers Synthesis for few-shot OOD detection
di: Sun, Hao, et al.
Pubblicazione: (2024)
di: Sun, Hao, et al.
Pubblicazione: (2024)
Unlocking Patch-Level Features for CLIP-Based Class-Incremental Learning
di: Sun, Hao, et al.
Pubblicazione: (2026)
di: Sun, Hao, et al.
Pubblicazione: (2026)
CLIP-Mamba: CLIP Pretrained Mamba Models with OOD and Hessian Evaluation
di: Huang, Weiquan, et al.
Pubblicazione: (2024)
di: Huang, Weiquan, et al.
Pubblicazione: (2024)
CLIP-driven Zero-shot Learning with Ambiguous Labels
di: Fan, Jinfu, et al.
Pubblicazione: (2026)
di: Fan, Jinfu, et al.
Pubblicazione: (2026)
Explaining CLIP Zero-shot Predictions Through Concepts
di: Ozdemir, Onat, et al.
Pubblicazione: (2026)
di: Ozdemir, Onat, et al.
Pubblicazione: (2026)
SPECIAL: Zero-shot Hyperspectral Image Classification With CLIP
di: Pang, Li, et al.
Pubblicazione: (2025)
di: Pang, Li, et al.
Pubblicazione: (2025)
Demystifying CLIP Data
di: Xu, Hu, et al.
Pubblicazione: (2023)
di: Xu, Hu, et al.
Pubblicazione: (2023)
A CLIP-Powered Framework for Robust and Generalizable Data Selection
di: Yang, Suorong, et al.
Pubblicazione: (2024)
di: Yang, Suorong, et al.
Pubblicazione: (2024)
CLIP-Map: Structured Matrix Mapping for Parameter-Efficient CLIP Compression
di: Zhang, Kangjie, et al.
Pubblicazione: (2026)
di: Zhang, Kangjie, et al.
Pubblicazione: (2026)
GazeCLIP: Gaze-Guided CLIP with Adaptive-Enhanced Fine-Grained Language Prompt for Deepfake Attribution and Detection
di: Zhang, Yaning, et al.
Pubblicazione: (2026)
di: Zhang, Yaning, et al.
Pubblicazione: (2026)
CLIP-VIS: Adapting CLIP for Open-Vocabulary Video Instance Segmentation
di: Zhu, Wenqi, et al.
Pubblicazione: (2024)
di: Zhu, Wenqi, et al.
Pubblicazione: (2024)
Documenti analoghi
-
DetailCLIP: Injecting Image Details into CLIP's Feature Space
di: Zhang, Zilun, et al.
Pubblicazione: (2022) -
CLIP Multi-modal Hashing for Multimedia Retrieval
di: Zhu, Jian, et al.
Pubblicazione: (2024) -
InterCLIP-MEP: Interactive CLIP and Memory-Enhanced Predictor for Multi-modal Sarcasm Detection
di: Chen, Junjie, et al.
Pubblicazione: (2024) -
Is CLIP Cross-Eyed? Revealing and Mitigating Center Bias in the CLIP Family
di: Chew, Oscar, et al.
Pubblicazione: (2026) -
MadCLIP: Few-shot Medical Anomaly Detection with CLIP
di: Shiri, Mahshid, et al.
Pubblicazione: (2025)