Omni-NegCLIP: Enhancing CLIP with Front-Layer Contrastive Fine-Tuning for Comprehensive Negation Understanding
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Xu, Jingqi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
TNG-CLIP:Training-Time Negation Data Generation for Negation Awareness of CLIP
von: Cai, Yuliang, et al.
Veröffentlicht: (2025)
von: Cai, Yuliang, et al.
Veröffentlicht: (2025)
LatteCLIP: Unsupervised CLIP Fine-Tuning via LMM-Synthetic Texts
von: Cao, Anh-Quan, et al.
Veröffentlicht: (2024)
von: Cao, Anh-Quan, et al.
Veröffentlicht: (2024)
DesignCLIP: Multimodal Learning with CLIP for Design Patent Understanding
von: Wang, Zhu, et al.
Veröffentlicht: (2025)
von: Wang, Zhu, et al.
Veröffentlicht: (2025)
No Captions, No Problem: Captionless 3D-CLIP Alignment with Hard Negatives via CLIP Knowledge and LLMs
von: Sbrolli, Cristian, et al.
Veröffentlicht: (2024)
von: Sbrolli, Cristian, et al.
Veröffentlicht: (2024)
microCLIP: Unsupervised CLIP Adaptation via Coarse-Fine Token Fusion for Fine-Grained Image Classification
von: Silva, Sathira, et al.
Veröffentlicht: (2025)
von: Silva, Sathira, et al.
Veröffentlicht: (2025)
CleanerCLIP: Fine-grained Counterfactual Semantic Augmentation for Backdoor Defense in Contrastive Learning
von: Xun, Yuan, et al.
Veröffentlicht: (2024)
von: Xun, Yuan, et al.
Veröffentlicht: (2024)
AA-CLIP: Enhancing Zero-shot Anomaly Detection via Anomaly-Aware CLIP
von: Ma, Wenxin, et al.
Veröffentlicht: (2025)
von: Ma, Wenxin, et al.
Veröffentlicht: (2025)
Enhancing Multimodal Understanding with CLIP-Based Image-to-Text Transformation
von: Che, Chang, et al.
Veröffentlicht: (2024)
von: Che, Chang, et al.
Veröffentlicht: (2024)
NegVQA: Can Vision Language Models Understand Negation?
von: Zhang, Yuhui, et al.
Veröffentlicht: (2025)
von: Zhang, Yuhui, et al.
Veröffentlicht: (2025)
FG-CLIP: Fine-Grained Visual and Textual Alignment
von: Xie, Chunyu, et al.
Veröffentlicht: (2025)
von: Xie, Chunyu, et al.
Veröffentlicht: (2025)
CLIP Can Understand Depth
von: Kim, Sohee, et al.
Veröffentlicht: (2024)
von: Kim, Sohee, et al.
Veröffentlicht: (2024)
Enhancing Image Retrieval : A Comprehensive Study on Photo Search using the CLIP Mode
von: Lahajal, Naresh Kumar, et al.
Veröffentlicht: (2024)
von: Lahajal, Naresh Kumar, et al.
Veröffentlicht: (2024)
HarmoCLIP: Harmonizing Global and Regional Representations in Contrastive Vision-Language Models
von: Zeng, Haoxi, et al.
Veröffentlicht: (2025)
von: Zeng, Haoxi, et al.
Veröffentlicht: (2025)
AdaptCLIP: Adapting CLIP for Universal Visual Anomaly Detection
von: Gao, Bin-Bin, et al.
Veröffentlicht: (2025)
von: Gao, Bin-Bin, et al.
Veröffentlicht: (2025)
RadCLIP: Enhancing Radiologic Image Analysis through Contrastive Language-Image Pre-training
von: Lu, Zhixiu, et al.
Veröffentlicht: (2024)
von: Lu, Zhixiu, et al.
Veröffentlicht: (2024)
MINT: Memory-Infused Prompt Tuning at Test-time for CLIP
von: Yi, Jiaming, et al.
Veröffentlicht: (2025)
von: Yi, Jiaming, et al.
Veröffentlicht: (2025)
MR-CLIP: Efficient Metadata-Guided Learning of MRI Contrast Representations
von: Avci, Mehmet Yigit, et al.
Veröffentlicht: (2025)
von: Avci, Mehmet Yigit, et al.
Veröffentlicht: (2025)
Explicit Uncertainty Modeling for Active CLIP Adaptation with Dual Prompt Tuning
von: Wang, Qian-Wei, et al.
Veröffentlicht: (2026)
von: Wang, Qian-Wei, et al.
Veröffentlicht: (2026)
InterCLIP-MEP: Interactive CLIP and Memory-Enhanced Predictor for Multi-modal Sarcasm Detection
von: Chen, Junjie, et al.
Veröffentlicht: (2024)
von: Chen, Junjie, et al.
Veröffentlicht: (2024)
VL-CLIP: Enhancing Multimodal Recommendations via Visual Grounding and LLM-Augmented CLIP Embeddings
von: Giahi, Ramin, et al.
Veröffentlicht: (2025)
von: Giahi, Ramin, et al.
Veröffentlicht: (2025)
RegionMed-CLIP: A Region-Aware Multimodal Contrastive Learning Pre-trained Model for Medical Image Understanding
von: Fang, Tianchen, et al.
Veröffentlicht: (2025)
von: Fang, Tianchen, et al.
Veröffentlicht: (2025)
CEIA: CLIP-Based Event-Image Alignment for Open-World Event-Based Understanding
von: Xu, Wenhao, et al.
Veröffentlicht: (2024)
von: Xu, Wenhao, et al.
Veröffentlicht: (2024)
BotaCLIP: Contrastive Learning for Botany-Aware Representation of Earth Observation Data
von: Cerna, Selene, et al.
Veröffentlicht: (2025)
von: Cerna, Selene, et al.
Veröffentlicht: (2025)
Mammo-CLIP: Leveraging Contrastive Language-Image Pre-training (CLIP) for Enhanced Breast Cancer Diagnosis with Multi-view Mammography
von: Chen, Xuxin, et al.
Veröffentlicht: (2024)
von: Chen, Xuxin, et al.
Veröffentlicht: (2024)
CLIP-MoE: Towards Building Mixture of Experts for CLIP with Diversified Multiplet Upcycling
von: Zhang, Jihai, et al.
Veröffentlicht: (2024)
von: Zhang, Jihai, et al.
Veröffentlicht: (2024)
CLIP-MUSED: CLIP-Guided Multi-Subject Visual Neural Information Semantic Decoding
von: Zhou, Qiongyi, et al.
Veröffentlicht: (2024)
von: Zhou, Qiongyi, et al.
Veröffentlicht: (2024)
DiffCLIP: Differential Attention Meets CLIP
von: Hammoud, Hasan Abed Al Kader, et al.
Veröffentlicht: (2025)
von: Hammoud, Hasan Abed Al Kader, et al.
Veröffentlicht: (2025)
CLIP-DQA: Blindly Evaluating Dehazed Images from Global and Local Perspectives Using CLIP
von: Zeng, Yirui, et al.
Veröffentlicht: (2025)
von: Zeng, Yirui, et al.
Veröffentlicht: (2025)
FoCLIP: A Feature-Space Misalignment Framework for CLIP-Based Image Manipulation and Detection
von: Chen, Yulin, et al.
Veröffentlicht: (2025)
von: Chen, Yulin, et al.
Veröffentlicht: (2025)
CountCLIP -- [Re] Teaching CLIP to Count to Ten
von: Mestha, Harshvardhan, et al.
Veröffentlicht: (2024)
von: Mestha, Harshvardhan, et al.
Veröffentlicht: (2024)
Robust CLIP: Unsupervised Adversarial Fine-Tuning of Vision Embeddings for Robust Large Vision-Language Models
von: Schlarmann, Christian, et al.
Veröffentlicht: (2024)
von: Schlarmann, Christian, et al.
Veröffentlicht: (2024)
FB-CLIP: Fine-Grained Zero-Shot Anomaly Detection with Foreground-Background Disentanglement
von: Hu, Ming, et al.
Veröffentlicht: (2026)
von: Hu, Ming, et al.
Veröffentlicht: (2026)
TokLIP: Marry Visual Tokens to CLIP for Multimodal Comprehension and Generation
von: Lin, Haokun, et al.
Veröffentlicht: (2025)
von: Lin, Haokun, et al.
Veröffentlicht: (2025)
Enhancing Compositional Reasoning in CLIP via Reconstruction and Alignment of Text Descriptions
von: Kwon, Jihoon, et al.
Veröffentlicht: (2025)
von: Kwon, Jihoon, et al.
Veröffentlicht: (2025)
AdaNeg: Adaptive Negative Proxy Guided OOD Detection with Vision-Language Models
von: Zhang, Yabin, et al.
Veröffentlicht: (2024)
von: Zhang, Yabin, et al.
Veröffentlicht: (2024)
Fine-tuning CLIP Text Encoders with Two-step Paraphrasing
von: Kim, Hyunjae, et al.
Veröffentlicht: (2024)
von: Kim, Hyunjae, et al.
Veröffentlicht: (2024)
FineLIP: Extending CLIP's Reach via Fine-Grained Alignment with Longer Text Inputs
von: Asokan, Mothilal, et al.
Veröffentlicht: (2025)
von: Asokan, Mothilal, et al.
Veröffentlicht: (2025)
CLIP-Guided Unsupervised Semantic-Aware Exposure Correction
von: Wu, Puzhen, et al.
Veröffentlicht: (2026)
von: Wu, Puzhen, et al.
Veröffentlicht: (2026)
IDEA: Image Description Enhanced CLIP-Adapter
von: Ye, Zhipeng, et al.
Veröffentlicht: (2025)
von: Ye, Zhipeng, et al.
Veröffentlicht: (2025)
DIST-CLIP: Arbitrary Metadata and Image Guided MRI Harmonization via Disentangled Anatomy-Contrast Representations
von: Avci, Mehmet Yigit, et al.
Veröffentlicht: (2025)
von: Avci, Mehmet Yigit, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
TNG-CLIP:Training-Time Negation Data Generation for Negation Awareness of CLIP
von: Cai, Yuliang, et al.
Veröffentlicht: (2025) -
LatteCLIP: Unsupervised CLIP Fine-Tuning via LMM-Synthetic Texts
von: Cao, Anh-Quan, et al.
Veröffentlicht: (2024) -
DesignCLIP: Multimodal Learning with CLIP for Design Patent Understanding
von: Wang, Zhu, et al.
Veröffentlicht: (2025) -
No Captions, No Problem: Captionless 3D-CLIP Alignment with Hard Negatives via CLIP Knowledge and LLMs
von: Sbrolli, Cristian, et al.
Veröffentlicht: (2024) -
microCLIP: Unsupervised CLIP Adaptation via Coarse-Fine Token Fusion for Fine-Grained Image Classification
von: Silva, Sathira, et al.
Veröffentlicht: (2025)