Concept Regions Matter: Benchmarking CLIP with a New Cluster-Importance Approach
Fuente:
arXiv
Saved in:
| Main Authors: | Agarwal, Aishwarya, Karanam, Srikrishna, Gandhi, Vineet |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LiteEmbed: Adapting CLIP to Rare Classes
by: Agarwal, Aishwarya, et al.
Published: (2026)
by: Agarwal, Aishwarya, et al.
Published: (2026)
TIDE: Training Locally Interpretable Domain Generalization Models Enables Test-time Correction
by: Agarwal, Aishwarya, et al.
Published: (2024)
by: Agarwal, Aishwarya, et al.
Published: (2024)
Training-free Color-Style Disentanglement for Constrained Text-to-Image Synthesis
by: Agarwal, Aishwarya, et al.
Published: (2024)
by: Agarwal, Aishwarya, et al.
Published: (2024)
AlignIT: Enhancing Prompt Alignment in Customization of Text-to-Image Models
by: Agarwal, Aishwarya, et al.
Published: (2024)
by: Agarwal, Aishwarya, et al.
Published: (2024)
CoCoNO: Attention Contrast-and-Complete for Initial Noise Optimization in Text-to-Image Synthesis
by: Sundaram, Aravindan, et al.
Published: (2024)
by: Sundaram, Aravindan, et al.
Published: (2024)
Composing Parts for Expressive Object Generation
by: Rangwani, Harsh, et al.
Published: (2024)
by: Rangwani, Harsh, et al.
Published: (2024)
Learning 3D Texture-Aware Representations for Parsing Diverse Human Clothing and Body Parts
by: Chhatre, Kiran, et al.
Published: (2025)
by: Chhatre, Kiran, et al.
Published: (2025)
Test-time Conditional Text-to-Image Synthesis Using Diffusion Models
by: Shukla, Tripti, et al.
Published: (2024)
by: Shukla, Tripti, et al.
Published: (2024)
SafaRi:Adaptive Sequence Transformer for Weakly Supervised Referring Expression Segmentation
by: Nag, Sayan, et al.
Published: (2024)
by: Nag, Sayan, et al.
Published: (2024)
Few Shot Class Incremental Learning using Vision-Language models
by: Kumar, Anurag, et al.
Published: (2024)
by: Kumar, Anurag, et al.
Published: (2024)
Enhancing Concept Localization in CLIP-based Concept Bottleneck Models
by: Kazmierczak, Rémi, et al.
Published: (2025)
by: Kazmierczak, Rémi, et al.
Published: (2025)
Simplifying Knowledge Transfer in Pretrained Models
by: Jain, Siddharth, et al.
Published: (2025)
by: Jain, Siddharth, et al.
Published: (2025)
Investigating Mechanisms for In-Context Vision Language Binding
by: Saravanan, Darshana, et al.
Published: (2025)
by: Saravanan, Darshana, et al.
Published: (2025)
Pseudo-labelling meets Label Smoothing for Noisy Partial Label Learning
by: Saravanan, Darshana, et al.
Published: (2024)
by: Saravanan, Darshana, et al.
Published: (2024)
Grid Jigsaw Representation with CLIP: A New Perspective on Image Clustering
by: Song, Zijie, et al.
Published: (2023)
by: Song, Zijie, et al.
Published: (2023)
CLIP-QDA: An Explainable Concept Bottleneck Model
by: Kazmierczak, Rémi, et al.
Published: (2023)
by: Kazmierczak, Rémi, et al.
Published: (2023)
CLIP-CID: Efficient CLIP Distillation via Cluster-Instance Discrimination
by: Yang, Kaicheng, et al.
Published: (2024)
by: Yang, Kaicheng, et al.
Published: (2024)
MedP-CLIP: Medical CLIP with Region-Aware Prompt Integration
by: Peng, Jiahui, et al.
Published: (2026)
by: Peng, Jiahui, et al.
Published: (2026)
Explaining CLIP Zero-shot Predictions Through Concepts
by: Ozdemir, Onat, et al.
Published: (2026)
by: Ozdemir, Onat, et al.
Published: (2026)
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues
by: Girmaji, Rohit, et al.
Published: (2025)
by: Girmaji, Rohit, et al.
Published: (2025)
Learning What Matters: Prioritized Concept Learning via Relative Error-driven Sample Selection
by: Chandhok, Shivam, et al.
Published: (2025)
by: Chandhok, Shivam, et al.
Published: (2025)
CLIP-Free, Label Free, Unsupervised Concept Bottleneck Models
by: Sammani, Fawaz, et al.
Published: (2025)
by: Sammani, Fawaz, et al.
Published: (2025)
Backdooring CLIP through Concept Confusion
by: Hu, Lijie, et al.
Published: (2025)
by: Hu, Lijie, et al.
Published: (2025)
Benchmarking PathCLIP for Pathology Image Analysis
by: Zheng, Sunyi, et al.
Published: (2024)
by: Zheng, Sunyi, et al.
Published: (2024)
Seeing What Matters: Empowering CLIP with Patch Generation-to-Selection
by: Pei, Gensheng, et al.
Published: (2025)
by: Pei, Gensheng, et al.
Published: (2025)
VELOCITI: Benchmarking Video-Language Compositional Reasoning with Strict Entailment
by: Saravanan, Darshana, et al.
Published: (2024)
by: Saravanan, Darshana, et al.
Published: (2024)
Mesh2SSM++: A Probabilistic Framework for Unsupervised Learning of Statistical Shape Model of Anatomies from Surface Meshes
by: Iyer, Krithika, et al.
Published: (2025)
by: Iyer, Krithika, et al.
Published: (2025)
MorphoFlow: Sparse-Supervised Generative Shape Modeling with Adaptive Latent Relevance
by: Karanam, Mokshagna Sai Teja, et al.
Published: (2026)
by: Karanam, Mokshagna Sai Teja, et al.
Published: (2026)
On the Provable Importance of Gradients for Language-Assisted Image Clustering
by: Peng, Bo, et al.
Published: (2025)
by: Peng, Bo, et al.
Published: (2025)
Mammo-CLIP Dissect: A Framework for Analysing Mammography Concepts in Vision-Language Models
by: Salahuddin, Suaiba Amina, et al.
Published: (2025)
by: Salahuddin, Suaiba Amina, et al.
Published: (2025)
From Weights to Concepts: Data-Free Interpretability of CLIP via Singular Vector Decomposition
by: Gentile, Francesco, et al.
Published: (2026)
by: Gentile, Francesco, et al.
Published: (2026)
StackCLIP: Clustering-Driven Stacked Prompt in Zero-Shot Industrial Anomaly Detection
by: Hou, Yanning, et al.
Published: (2025)
by: Hou, Yanning, et al.
Published: (2025)
Rethinking Few Shot CLIP Benchmarks: A Critical Analysis in the Inductive Setting
by: Kravets, Alexey, et al.
Published: (2025)
by: Kravets, Alexey, et al.
Published: (2025)
SuperCLIP: CLIP with Simple Classification Supervision
by: Zhao, Weiheng, et al.
Published: (2025)
by: Zhao, Weiheng, et al.
Published: (2025)
Zero Shot Context-Based Object Segmentation using SLIP (SAM+CLIP)
by: Gundavarapu, Saaketh Koundinya, et al.
Published: (2024)
by: Gundavarapu, Saaketh Koundinya, et al.
Published: (2024)
ACE: Action Concept Enhancement of Video-Language Models in Procedural Videos
by: Ghoddoosian, Reza, et al.
Published: (2024)
by: Ghoddoosian, Reza, et al.
Published: (2024)
Interpreting CLIP with Sparse Linear Concept Embeddings (SpLiCE)
by: Bhalla, Usha, et al.
Published: (2024)
by: Bhalla, Usha, et al.
Published: (2024)
TSCLIP: Robust CLIP Fine-Tuning for Worldwide Cross-Regional Traffic Sign Recognition
by: Zhao, Guoyang, et al.
Published: (2024)
by: Zhao, Guoyang, et al.
Published: (2024)
CLIP-IT: CLIP-based Pairing for Histology Images Classification
by: Karimian, Banafsheh, et al.
Published: (2025)
by: Karimian, Banafsheh, et al.
Published: (2025)
Long-CLIP: Unlocking the Long-Text Capability of CLIP
by: Zhang, Beichen, et al.
Published: (2024)
by: Zhang, Beichen, et al.
Published: (2024)
Similar Items
-
LiteEmbed: Adapting CLIP to Rare Classes
by: Agarwal, Aishwarya, et al.
Published: (2026) -
TIDE: Training Locally Interpretable Domain Generalization Models Enables Test-time Correction
by: Agarwal, Aishwarya, et al.
Published: (2024) -
Training-free Color-Style Disentanglement for Constrained Text-to-Image Synthesis
by: Agarwal, Aishwarya, et al.
Published: (2024) -
AlignIT: Enhancing Prompt Alignment in Customization of Text-to-Image Models
by: Agarwal, Aishwarya, et al.
Published: (2024) -
CoCoNO: Attention Contrast-and-Complete for Initial Noise Optimization in Text-to-Image Synthesis
by: Sundaram, Aravindan, et al.
Published: (2024)