CLIP Adaptation by Intra-modal Overlap Reduction
Fuente:
arXiv
Saved in:
| Main Authors: | Kravets, Alexey, Namboodiri, Vinay |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Rethinking Few Shot CLIP Benchmarks: A Critical Analysis in the Inductive Setting
by: Kravets, Alexey, et al.
Published: (2025)
by: Kravets, Alexey, et al.
Published: (2025)
Zero-Shot Class Unlearning in CLIP with Synthetic Samples
by: Kravets, A., et al.
Published: (2024)
by: Kravets, A., et al.
Published: (2024)
Interpretability Transfer from Language to Vision via Sparse Autoencoders
by: Kravets, Alexey, et al.
Published: (2026)
by: Kravets, Alexey, et al.
Published: (2026)
TalkLoRA: Low-Rank Adaptation for Speech-Driven Animation
by: Saunders, Jack, et al.
Published: (2024)
by: Saunders, Jack, et al.
Published: (2024)
MedFocusCLIP : Improving few shot classification in medical datasets using pixel wise attention
by: Arora, Aadya, et al.
Published: (2025)
by: Arora, Aadya, et al.
Published: (2025)
StyleYourSmile: Cross-Domain Face Retargeting Without Paired Multi-Style Data
by: Dey, Avirup, et al.
Published: (2025)
by: Dey, Avirup, et al.
Published: (2025)
IsoCLIP: Decomposing CLIP Projectors for Efficient Intra-modal Alignment
by: Magistri, Simone, et al.
Published: (2026)
by: Magistri, Simone, et al.
Published: (2026)
Dubbing for Everyone: Data-Efficient Visual Dubbing using Neural Rendering Priors
by: Saunders, Jack, et al.
Published: (2024)
by: Saunders, Jack, et al.
Published: (2024)
Trusting Semantic Segmentation Networks
by: Some, Samik, et al.
Published: (2024)
by: Some, Samik, et al.
Published: (2024)
Can Unsupervised Segmentation Reduce Annotation Costs for Video Semantic Segmentation?
by: Some, Samik, et al.
Published: (2026)
by: Some, Samik, et al.
Published: (2026)
Self-supervised Representation Learning for Cell Event Recognition through Time Arrow Prediction
by: Chen, Cangxiong, et al.
Published: (2024)
by: Chen, Cangxiong, et al.
Published: (2024)
Determinantal Point Process as an alternative to NMS
by: Some, Samik, et al.
Published: (2020)
by: Some, Samik, et al.
Published: (2020)
RAW: Robust Avatar Watermarking -- Benchmarking and Baseline
by: Parry, Jack, et al.
Published: (2026)
by: Parry, Jack, et al.
Published: (2026)
Reevaluating the Intra-Modal Misalignment Hypothesis in CLIP
by: Herzog, Jonas, et al.
Published: (2026)
by: Herzog, Jonas, et al.
Published: (2026)
EIDT-V: Exploiting Intersections in Diffusion Trajectories for Model-Agnostic, Zero-Shot, Training-Free Text-to-Video Generation
by: Jagpal, Diljeet, et al.
Published: (2025)
by: Jagpal, Diljeet, et al.
Published: (2025)
Cross the Gap: Exposing the Intra-modal Misalignment in CLIP via Modality Inversion
by: Mistretta, Marco, et al.
Published: (2025)
by: Mistretta, Marco, et al.
Published: (2025)
CLIP4Sketch: Enhancing Sketch to Mugshot Matching through Dataset Augmentation using Diffusion Models
by: Jain, Kushal Kumar, et al.
Published: (2024)
by: Jain, Kushal Kumar, et al.
Published: (2024)
CLIP Multi-modal Hashing for Multimedia Retrieval
by: Zhu, Jian, et al.
Published: (2024)
by: Zhu, Jian, et al.
Published: (2024)
PS-StyleGAN: Illustrative Portrait Sketching using Attention-Based Style Adaptation
by: Jain, Kushal Kumar, et al.
Published: (2024)
by: Jain, Kushal Kumar, et al.
Published: (2024)
Enhancing Incomplete Multi-modal Brain Tumor Segmentation with Intra-modal Asymmetry and Inter-modal Dependency
by: Liu, Weide, et al.
Published: (2024)
by: Liu, Weide, et al.
Published: (2024)
RISSOLE: Parameter-efficient Diffusion Models via Block-wise Generation and Retrieval-Guidance
by: Mukherjee, Avideep, et al.
Published: (2024)
by: Mukherjee, Avideep, et al.
Published: (2024)
SCPNet: Unsupervised Cross-modal Homography Estimation via Intra-modal Self-supervised Learning
by: Zhang, Runmin, et al.
Published: (2024)
by: Zhang, Runmin, et al.
Published: (2024)
CLIP-SLA: Parameter-Efficient CLIP Adaptation for Continuous Sign Language Recognition
by: Alyami, Sarah, et al.
Published: (2025)
by: Alyami, Sarah, et al.
Published: (2025)
IntraStyler: Intra-Domain Style Synthesis for Cross-Modality MRI Domain Adaptation
by: Liu, Han, et al.
Published: (2026)
by: Liu, Han, et al.
Published: (2026)
SRL-CLIP: Efficient CLIP Video Adaptation via Structured Semantic Role Labels
by: Singh, Darshan, et al.
Published: (2024)
by: Singh, Darshan, et al.
Published: (2024)
AF-CLIP: Zero-Shot Anomaly Detection via Anomaly-Focused CLIP Adaptation
by: Fang, Qingqing, et al.
Published: (2025)
by: Fang, Qingqing, et al.
Published: (2025)
Rethinking Domain Adaptation and Generalization in the Era of CLIP
by: Feng, Ruoyu, et al.
Published: (2024)
by: Feng, Ruoyu, et al.
Published: (2024)
Fusion2Print: Deep Flash-Non-Flash Fusion for Contactless Fingerprint Matching
by: Sahoo, Roja, et al.
Published: (2026)
by: Sahoo, Roja, et al.
Published: (2026)
Illumination-Aware Contactless Fingerprint Spoof Detection via Paired Flash-Non-Flash Imaging
by: Sahoo, Roja, et al.
Published: (2026)
by: Sahoo, Roja, et al.
Published: (2026)
CardiacCLIP: Video-based CLIP Adaptation for LVEF Prediction in a Few-shot Manner
by: Du, Yao, et al.
Published: (2025)
by: Du, Yao, et al.
Published: (2025)
CIBR: Cross-modal Information Bottleneck Regularization for Robust CLIP Generalization
by: Ji, Yingrui, et al.
Published: (2025)
by: Ji, Yingrui, et al.
Published: (2025)
WATT: Weight Average Test-Time Adaptation of CLIP
by: Osowiechi, David, et al.
Published: (2024)
by: Osowiechi, David, et al.
Published: (2024)
CLIP the Divergence: Language-guided Unsupervised Domain Adaptation
by: Zhu, Jinjing, et al.
Published: (2024)
by: Zhu, Jinjing, et al.
Published: (2024)
FLEX-CLIP: Feature-Level GEneration Network Enhanced CLIP for X-shot Cross-modal Retrieval
by: Xie, Jingyou, et al.
Published: (2024)
by: Xie, Jingyou, et al.
Published: (2024)
Detail-Enhanced Intra- and Inter-modal Interaction for Audio-Visual Emotion Recognition
by: Shi, Tong, et al.
Published: (2024)
by: Shi, Tong, et al.
Published: (2024)
Language Augmentation in CLIP for Improved Anatomy Detection on Multi-modal Medical Images
by: Kakkar, Mansi, et al.
Published: (2024)
by: Kakkar, Mansi, et al.
Published: (2024)
$\texttt{BATCLIP}$: Bimodal Online Test-Time Adaptation for CLIP
by: Maharana, Sarthak Kumar, et al.
Published: (2024)
by: Maharana, Sarthak Kumar, et al.
Published: (2024)
Rethinking Visual Content Refinement in Low-Shot CLIP Adaptation
by: Lu, Jinda, et al.
Published: (2024)
by: Lu, Jinda, et al.
Published: (2024)
Building a Multi-modal Spatiotemporal Expert for Zero-shot Action Recognition with CLIP
by: Yu, Yating, et al.
Published: (2024)
by: Yu, Yating, et al.
Published: (2024)
MFCLIP: Multi-modal Fine-grained CLIP for Generalizable Diffusion Face Forgery Detection
by: Zhang, Yaning, et al.
Published: (2024)
by: Zhang, Yaning, et al.
Published: (2024)
Similar Items
-
Rethinking Few Shot CLIP Benchmarks: A Critical Analysis in the Inductive Setting
by: Kravets, Alexey, et al.
Published: (2025) -
Zero-Shot Class Unlearning in CLIP with Synthetic Samples
by: Kravets, A., et al.
Published: (2024) -
Interpretability Transfer from Language to Vision via Sparse Autoencoders
by: Kravets, Alexey, et al.
Published: (2026) -
TalkLoRA: Low-Rank Adaptation for Speech-Driven Animation
by: Saunders, Jack, et al.
Published: (2024) -
MedFocusCLIP : Improving few shot classification in medical datasets using pixel wise attention
by: Arora, Aadya, et al.
Published: (2025)