InterCLIP-MEP: Interactive CLIP and Memory-Enhanced Predictor for Multi-modal Sarcasm Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Junjie, Yu, Hang, Huang, Subin, Liu, Sanmin, Zhang, Linfeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FLEX-CLIP: Feature-Level GEneration Network Enhanced CLIP for X-shot Cross-modal Retrieval
von: Xie, Jingyou, et al.
Veröffentlicht: (2024)
von: Xie, Jingyou, et al.
Veröffentlicht: (2024)
NAU-QMUL: Utilizing BERT and CLIP for Multi-modal AI-Generated Image Detection
von: Guo, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Guo, Xiaoyu, et al.
Veröffentlicht: (2026)
Is CLIP Cross-Eyed? Revealing and Mitigating Center Bias in the CLIP Family
von: Chew, Oscar, et al.
Veröffentlicht: (2026)
von: Chew, Oscar, et al.
Veröffentlicht: (2026)
VideoCLIP-XL: Advancing Long Description Understanding for Video CLIP Models
von: Wang, Jiapeng, et al.
Veröffentlicht: (2024)
von: Wang, Jiapeng, et al.
Veröffentlicht: (2024)
Demystifying CLIP Data
von: Xu, Hu, et al.
Veröffentlicht: (2023)
von: Xu, Hu, et al.
Veröffentlicht: (2023)
CultureCLIP: Empowering CLIP with Cultural Awareness through Synthetic Images and Contextualized Captions
von: Huang, Yuchen, et al.
Veröffentlicht: (2025)
von: Huang, Yuchen, et al.
Veröffentlicht: (2025)
Can CLIP Count Stars? An Empirical Study on Quantity Bias in CLIP
von: Zhang, Zeliang, et al.
Veröffentlicht: (2024)
von: Zhang, Zeliang, et al.
Veröffentlicht: (2024)
SPECS: Specificity-Enhanced CLIP-Score for Long Image Caption Evaluation
von: Chen, Xiaofu, et al.
Veröffentlicht: (2025)
von: Chen, Xiaofu, et al.
Veröffentlicht: (2025)
CLIP Multi-modal Hashing for Multimedia Retrieval
von: Zhu, Jian, et al.
Veröffentlicht: (2024)
von: Zhu, Jian, et al.
Veröffentlicht: (2024)
TripletCLIP: Improving Compositional Reasoning of CLIP via Synthetic Vision-Language Negatives
von: Patel, Maitreya, et al.
Veröffentlicht: (2024)
von: Patel, Maitreya, et al.
Veröffentlicht: (2024)
GET: Unlocking the Multi-modal Potential of CLIP for Generalized Category Discovery
von: Wang, Enguang, et al.
Veröffentlicht: (2024)
von: Wang, Enguang, et al.
Veröffentlicht: (2024)
Know "No" Better: A Data-Driven Approach for Enhancing Negation Awareness in CLIP
von: Park, Junsung, et al.
Veröffentlicht: (2025)
von: Park, Junsung, et al.
Veröffentlicht: (2025)
TiC-CLIP: Continual Training of CLIP Models
von: Garg, Saurabh, et al.
Veröffentlicht: (2023)
von: Garg, Saurabh, et al.
Veröffentlicht: (2023)
LowCLIP: Adapting the CLIP Model Architecture for Low-Resource Languages in Multimodal Image Retrieval Task
von: Asgarov, Ali, et al.
Veröffentlicht: (2024)
von: Asgarov, Ali, et al.
Veröffentlicht: (2024)
MFCLIP: Multi-modal Fine-grained CLIP for Generalizable Diffusion Face Forgery Detection
von: Zhang, Yaning, et al.
Veröffentlicht: (2024)
von: Zhang, Yaning, et al.
Veröffentlicht: (2024)
Meta CLIP 2: A Worldwide Scaling Recipe
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2025)
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2025)
LatteCLIP: Unsupervised CLIP Fine-Tuning via LMM-Synthetic Texts
von: Cao, Anh-Quan, et al.
Veröffentlicht: (2024)
von: Cao, Anh-Quan, et al.
Veröffentlicht: (2024)
CLIP-Adapter: Better Vision-Language Models with Feature Adapters
von: Gao, Peng, et al.
Veröffentlicht: (2021)
von: Gao, Peng, et al.
Veröffentlicht: (2021)
Language Augmentation in CLIP for Improved Anatomy Detection on Multi-modal Medical Images
von: Kakkar, Mansi, et al.
Veröffentlicht: (2024)
von: Kakkar, Mansi, et al.
Veröffentlicht: (2024)
Debiasing CLIP: Interpreting and Correcting Bias in Attention Heads
von: Yeo, Wei Jie, et al.
Veröffentlicht: (2025)
von: Yeo, Wei Jie, et al.
Veröffentlicht: (2025)
Generalizable Prompt Learning of CLIP: A Brief Overview
von: Cui, Fangming, et al.
Veröffentlicht: (2025)
von: Cui, Fangming, et al.
Veröffentlicht: (2025)
Seeing Sarcasm Through Different Eyes: Analyzing Multimodal Sarcasm Perception in Large Vision-Language Models
von: Chen, Junjie, et al.
Veröffentlicht: (2025)
von: Chen, Junjie, et al.
Veröffentlicht: (2025)
Explaining Caption-Image Interactions in CLIP Models with Second-Order Attributions
von: Möller, Lucas, et al.
Veröffentlicht: (2024)
von: Möller, Lucas, et al.
Veröffentlicht: (2024)
Distill CLIP (DCLIP): Enhancing Image-Text Retrieval via Cross-Modal Transformer Distillation
von: Csizmadia, Daniel, et al.
Veröffentlicht: (2025)
von: Csizmadia, Daniel, et al.
Veröffentlicht: (2025)
LLM2CLIP: Powerful Language Model Unlocks Richer Cross-Modality Representation
von: Huang, Weiquan, et al.
Veröffentlicht: (2024)
von: Huang, Weiquan, et al.
Veröffentlicht: (2024)
GroundingGPT:Language Enhanced Multi-modal Grounding Model
von: Li, Zhaowei, et al.
Veröffentlicht: (2024)
von: Li, Zhaowei, et al.
Veröffentlicht: (2024)
$λ$-ECLIPSE: Multi-Concept Personalized Text-to-Image Diffusion Models by Leveraging CLIP Latent Space
von: Patel, Maitreya, et al.
Veröffentlicht: (2024)
von: Patel, Maitreya, et al.
Veröffentlicht: (2024)
HiMo-CLIP: Modeling Semantic Hierarchy and Monotonicity in Vision-Language Alignment
von: Wu, Ruijia, et al.
Veröffentlicht: (2025)
von: Wu, Ruijia, et al.
Veröffentlicht: (2025)
Incremental Object Detection with CLIP
von: Huang, Ziyue, et al.
Veröffentlicht: (2023)
von: Huang, Ziyue, et al.
Veröffentlicht: (2023)
GazeCLIP: Gaze-Guided CLIP with Adaptive-Enhanced Fine-Grained Language Prompt for Deepfake Attribution and Detection
von: Zhang, Yaning, et al.
Veröffentlicht: (2026)
von: Zhang, Yaning, et al.
Veröffentlicht: (2026)
Taming CLIP for Fine-grained and Structured Visual Understanding of Museum Exhibits
von: Balauca, Ada-Astrid, et al.
Veröffentlicht: (2024)
von: Balauca, Ada-Astrid, et al.
Veröffentlicht: (2024)
CLIP Adaptation by Intra-modal Overlap Reduction
von: Kravets, Alexey, et al.
Veröffentlicht: (2024)
von: Kravets, Alexey, et al.
Veröffentlicht: (2024)
CLIP-KD: An Empirical Study of CLIP Model Distillation
von: Yang, Chuanguang, et al.
Veröffentlicht: (2023)
von: Yang, Chuanguang, et al.
Veröffentlicht: (2023)
GDCNet: Generative Discrepancy Comparison Network for Multimodal Sarcasm Detection
von: Zhang, Shuguang, et al.
Veröffentlicht: (2026)
von: Zhang, Shuguang, et al.
Veröffentlicht: (2026)
SuperCLIP: CLIP with Simple Classification Supervision
von: Zhao, Weiheng, et al.
Veröffentlicht: (2025)
von: Zhao, Weiheng, et al.
Veröffentlicht: (2025)
IsoCLIP: Decomposing CLIP Projectors for Efficient Intra-modal Alignment
von: Magistri, Simone, et al.
Veröffentlicht: (2026)
von: Magistri, Simone, et al.
Veröffentlicht: (2026)
MadCLIP: Few-shot Medical Anomaly Detection with CLIP
von: Shiri, Mahshid, et al.
Veröffentlicht: (2025)
von: Shiri, Mahshid, et al.
Veröffentlicht: (2025)
Enhancing CLIP Conceptual Embedding through Knowledge Distillation
von: Kao, Kuei-Chun
Veröffentlicht: (2024)
von: Kao, Kuei-Chun
Veröffentlicht: (2024)
CLIP-SVD: Efficient and Interpretable Vision-Language Adaptation via Singular Values
von: Koleilat, Taha, et al.
Veröffentlicht: (2025)
von: Koleilat, Taha, et al.
Veröffentlicht: (2025)
Building a Multi-modal Spatiotemporal Expert for Zero-shot Action Recognition with CLIP
von: Yu, Yating, et al.
Veröffentlicht: (2024)
von: Yu, Yating, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
FLEX-CLIP: Feature-Level GEneration Network Enhanced CLIP for X-shot Cross-modal Retrieval
von: Xie, Jingyou, et al.
Veröffentlicht: (2024) -
NAU-QMUL: Utilizing BERT and CLIP for Multi-modal AI-Generated Image Detection
von: Guo, Xiaoyu, et al.
Veröffentlicht: (2026) -
Is CLIP Cross-Eyed? Revealing and Mitigating Center Bias in the CLIP Family
von: Chew, Oscar, et al.
Veröffentlicht: (2026) -
VideoCLIP-XL: Advancing Long Description Understanding for Video CLIP Models
von: Wang, Jiapeng, et al.
Veröffentlicht: (2024) -
Demystifying CLIP Data
von: Xu, Hu, et al.
Veröffentlicht: (2023)