Whitened CLIP as a Likelihood Surrogate of Images and Captions
Fuente:
arXiv
Saved in:
| Main Authors: | Betser, Roy, Levi, Meir Yossef, Gilboa, Guy |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Make it SING: Analyzing Semantic Invariants in Classifiers
by: Yadid, Harel, et al.
Published: (2026)
by: Yadid, Harel, et al.
Published: (2026)
The Universal Normal Embedding
by: Tasker, Chen, et al.
Published: (2026)
by: Tasker, Chen, et al.
Published: (2026)
The Double-Ellipsoid Geometry of CLIP
by: Levi, Meir Yossef, et al.
Published: (2024)
by: Levi, Meir Yossef, et al.
Published: (2024)
General and Domain-Specific Zero-shot Detection of Generated Images via Conditional Likelihood
by: Betser, Roy, et al.
Published: (2025)
by: Betser, Roy, et al.
Published: (2025)
Training-free Detection of Generated Videos via Spatial-Temporal Likelihoods
by: Hayun, Omer Ben, et al.
Published: (2026)
by: Hayun, Omer Ben, et al.
Published: (2026)
Interpretable Automatic Rosacea Detection with Whitened Cosine Similarity
by: Yang, Chengyu, et al.
Published: (2025)
by: Yang, Chengyu, et al.
Published: (2025)
Transfer CLIP for Generalizable Image Denoising
by: Cheng, Jun, et al.
Published: (2024)
by: Cheng, Jun, et al.
Published: (2024)
The Solution for the CVPR2023 NICE Image Captioning Challenge
by: Wu, Xiangyu, et al.
Published: (2023)
by: Wu, Xiangyu, et al.
Published: (2023)
Robustifying Point Cloud Networks by Refocusing
by: Levi, Meir Yossef, et al.
Published: (2023)
by: Levi, Meir Yossef, et al.
Published: (2023)
Fast and Simple Explainability for Point Cloud Networks
by: Levi, Meir Yossef, et al.
Published: (2024)
by: Levi, Meir Yossef, et al.
Published: (2024)
Robust COVID-19 Detection in CT Images with CLIP
by: Lin, Li, et al.
Published: (2024)
by: Lin, Li, et al.
Published: (2024)
Stacked Cross-modal Feature Consolidation Attention Networks for Image Captioning
by: Pourkeshavarz, Mozhgan, et al.
Published: (2023)
by: Pourkeshavarz, Mozhgan, et al.
Published: (2023)
Unsupervised Image Prior via Prompt Learning and CLIP Semantic Guidance for Low-Light Image Enhancement
by: Morawski, Igor, et al.
Published: (2024)
by: Morawski, Igor, et al.
Published: (2024)
RetinaLogos: Fine-Grained Synthesis of High-Resolution Retinal Images Through Captions
by: Ning, Junzhi, et al.
Published: (2025)
by: Ning, Junzhi, et al.
Published: (2025)
Omnidirectional Image Quality Captioning: A Large-scale Database and A New Model
by: Yan, Jiebin, et al.
Published: (2025)
by: Yan, Jiebin, et al.
Published: (2025)
CLIP Based Region-Aware Feature Fusion for Automated BBPS Scoring in Colonoscopy Images
by: Fu, Yujia, et al.
Published: (2025)
by: Fu, Yujia, et al.
Published: (2025)
Cardiac-CLIP: A Vision-Language Foundation Model for 3D Cardiac CT Images
by: Hu, Yutao, et al.
Published: (2025)
by: Hu, Yutao, et al.
Published: (2025)
VLSM-Ensemble: Ensembling CLIP-based Vision-Language Models for Enhanced Medical Image Segmentation
by: Dietlmeier, Julia, et al.
Published: (2025)
by: Dietlmeier, Julia, et al.
Published: (2025)
I2I-Galip: Unsupervised Medical Image Translation Using Generative Adversarial CLIP
by: Korkmaz, Yilmaz, et al.
Published: (2024)
by: Korkmaz, Yilmaz, et al.
Published: (2024)
Zero-Shot Solving of Imaging Inverse Problems via Noise-Refined Likelihood Guided Diffusion Models
by: Wang, Zhen, et al.
Published: (2025)
by: Wang, Zhen, et al.
Published: (2025)
CLIPVQA:Video Quality Assessment via CLIP
by: Xing, Fengchuang, et al.
Published: (2024)
by: Xing, Fengchuang, et al.
Published: (2024)
ML-CLIPSim: Multi-Layer CLIP Similarity for Machine-Oriented Image Quality
by: Ding, Feng, et al.
Published: (2026)
by: Ding, Feng, et al.
Published: (2026)
SCoCCA: Multi-modal Sparse Concept Decomposition via Canonical Correlation Analysis
by: Gordon, Ehud, et al.
Published: (2026)
by: Gordon, Ehud, et al.
Published: (2026)
FluoCLIP: Stain-Aware Focus Quality Assessment in Fluorescence Microscopy
by: Park, Hyejin, et al.
Published: (2026)
by: Park, Hyejin, et al.
Published: (2026)
Could We Generate Cytology Images from Histopathology Images? An Empirical Study
by: Dey, Soumyajyoti, et al.
Published: (2024)
by: Dey, Soumyajyoti, et al.
Published: (2024)
Generalized Gaussian Entropy Model for Point Cloud Attribute Compression with Dynamic Likelihood Intervals
by: Peng, Changhao, et al.
Published: (2025)
by: Peng, Changhao, et al.
Published: (2025)
Label-Efficient Chest X-ray Diagnosis via Partial CLIP Adaptation
by: Dalsania, Heet Nitinkumar
Published: (2025)
by: Dalsania, Heet Nitinkumar
Published: (2025)
SPECTRUM: Semantic Processing and Emotion-informed video-Captioning Through Retrieval and Understanding Modalities
by: Faghihi, Ehsan, et al.
Published: (2024)
by: Faghihi, Ehsan, et al.
Published: (2024)
MedBLIP: Fine-tuning BLIP for Medical Image Captioning
by: Limbu, Manshi, et al.
Published: (2025)
by: Limbu, Manshi, et al.
Published: (2025)
Mammo-CLIP: A Vision Language Foundation Model to Enhance Data Efficiency and Robustness in Mammography
by: Ghosh, Shantanu, et al.
Published: (2024)
by: Ghosh, Shantanu, et al.
Published: (2024)
Inserting Faces inside Captions: Image Captioning with Attention Guided Merging
by: Tevissen, Yannis, et al.
Published: (2024)
by: Tevissen, Yannis, et al.
Published: (2024)
MedFocusCLIP : Improving few shot classification in medical datasets using pixel wise attention
by: Arora, Aadya, et al.
Published: (2025)
by: Arora, Aadya, et al.
Published: (2025)
FACMIC: Federated Adaptative CLIP Model for Medical Image Classification
by: Wu, Yihang, et al.
Published: (2024)
by: Wu, Yihang, et al.
Published: (2024)
CanvOI, an Oncology Intelligence Foundation Model: Scaling FLOPS Differently
by: Zalach, Jonathan, et al.
Published: (2024)
by: Zalach, Jonathan, et al.
Published: (2024)
Mammo-CLIP: Leveraging Contrastive Language-Image Pre-training (CLIP) for Enhanced Breast Cancer Diagnosis with Multi-view Mammography
by: Chen, Xuxin, et al.
Published: (2024)
by: Chen, Xuxin, et al.
Published: (2024)
Transformers in Medicine: Improving Vision-Language Alignment for Medical Image Captioning
by: Suresh, Yogesh Thakku, et al.
Published: (2025)
by: Suresh, Yogesh Thakku, et al.
Published: (2025)
AD-Lite Net: A Lightweight and Concatenated CNN Model for Alzheimer's Detection from MRI Images
by: Roy, Santanu, et al.
Published: (2024)
by: Roy, Santanu, et al.
Published: (2024)
Surrogate-based cross-correlation for particle image velocimetry
by: Lee, Yong, et al.
Published: (2021)
by: Lee, Yong, et al.
Published: (2021)
Are Vision xLSTM Embedded UNet More Reliable in Medical 3D Image Segmentation?
by: Dutta, Pallabi, et al.
Published: (2024)
by: Dutta, Pallabi, et al.
Published: (2024)
Robust CLIP-Based Detector for Exposing Diffusion Model-Generated Images
by: Santosh, et al.
Published: (2024)
by: Santosh, et al.
Published: (2024)
Similar Items
-
Make it SING: Analyzing Semantic Invariants in Classifiers
by: Yadid, Harel, et al.
Published: (2026) -
The Universal Normal Embedding
by: Tasker, Chen, et al.
Published: (2026) -
The Double-Ellipsoid Geometry of CLIP
by: Levi, Meir Yossef, et al.
Published: (2024) -
General and Domain-Specific Zero-shot Detection of Generated Images via Conditional Likelihood
by: Betser, Roy, et al.
Published: (2025) -
Training-free Detection of Generated Videos via Spatial-Temporal Likelihoods
by: Hayun, Omer Ben, et al.
Published: (2026)