ECOR: Explainable CLIP for Object Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Rasekh, Ali, Ranjbar, Sepehr Kazemi, Heidari, Milad, Nejdl, Wolfgang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Rationale Explainable Object Recognition via Contrastive Conditional Inference
by: Rasekh, Ali, et al.
Published: (2025)
by: Rasekh, Ali, et al.
Published: (2025)
ISLR101: an Iranian Word-Level Sign Language Recognition Dataset
by: Ranjbar, Hossein, et al.
Published: (2025)
by: Ranjbar, Hossein, et al.
Published: (2025)
ExIQA: Explainable Image Quality Assessment Using Distortion Attributes
by: Ranjbar, Sepehr Kazemi, et al.
Published: (2024)
by: Ranjbar, Sepehr Kazemi, et al.
Published: (2024)
Towards Precision Healthcare: Robust Fusion of Time Series and Image Data
by: Rasekh, Ali, et al.
Published: (2024)
by: Rasekh, Ali, et al.
Published: (2024)
DiffCLIP: Differential Attention Meets CLIP
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2025)
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2025)
CountCLIP -- [Re] Teaching CLIP to Count to Ten
by: Mestha, Harshvardhan, et al.
Published: (2024)
by: Mestha, Harshvardhan, et al.
Published: (2024)
Target-Oriented Single Domain Generalization
by: Heidari, Marzi, et al.
Published: (2025)
by: Heidari, Marzi, et al.
Published: (2025)
SynthCLIP: Are We Ready for a Fully Synthetic CLIP Training?
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2024)
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2024)
VeCLIP: Improving CLIP Training via Visual-enriched Captions
by: Lai, Zhengfeng, et al.
Published: (2023)
by: Lai, Zhengfeng, et al.
Published: (2023)
LeafLife: An Explainable Deep Learning Framework with Robustness for Grape Leaf Disease Recognition
by: Alam, B. M. Shahria, et al.
Published: (2026)
by: Alam, B. M. Shahria, et al.
Published: (2026)
CLIP Can Understand Depth
by: Kim, Sohee, et al.
Published: (2024)
by: Kim, Sohee, et al.
Published: (2024)
Provenance Networks: End-to-End Exemplar-Based Explainability
by: Kayyam, Ali, et al.
Published: (2025)
by: Kayyam, Ali, et al.
Published: (2025)
CLIP-MG: Guiding Semantic Attention with Skeletal Pose Features and RGB Data for Micro-Gesture Recognition on the iMiGUE Dataset
by: Patapati, Santosh, et al.
Published: (2025)
by: Patapati, Santosh, et al.
Published: (2025)
Prompt-Driven Feature Diffusion for Open-World Semi-Supervised Learning
by: Heidari, Marzi, et al.
Published: (2024)
by: Heidari, Marzi, et al.
Published: (2024)
Interpreting CLIP with Hierarchical Sparse Autoencoders
by: Zaigrajew, Vladimir, et al.
Published: (2025)
by: Zaigrajew, Vladimir, et al.
Published: (2025)
DivPrune: Diversity-based Visual Token Pruning for Large Multimodal Models
by: Alvar, Saeed Ranjbar, et al.
Published: (2025)
by: Alvar, Saeed Ranjbar, et al.
Published: (2025)
Incorporating Geo-Diverse Knowledge into Prompting for Increased Geographical Robustness in Object Recognition
by: Buettner, Kyle, et al.
Published: (2024)
by: Buettner, Kyle, et al.
Published: (2024)
Are Deep Learning Models Robust to Partial Object Occlusion in Visual Recognition Tasks?
by: Kassaw, Kaleb, et al.
Published: (2024)
by: Kassaw, Kaleb, et al.
Published: (2024)
Deep Models for Multi-View 3D Object Recognition: A Review
by: Alzahrani, Mona, et al.
Published: (2024)
by: Alzahrani, Mona, et al.
Published: (2024)
OBSER: Object-Based Sub-Environment Recognition for Zero-Shot Environmental Inference
by: Choi, Won-Seok, et al.
Published: (2025)
by: Choi, Won-Seok, et al.
Published: (2025)
Implicit Inversion turns CLIP into a Decoder
by: D'Orazio, Antonio, et al.
Published: (2025)
by: D'Orazio, Antonio, et al.
Published: (2025)
Steering CLIP's vision transformer with sparse autoencoders
by: Joseph, Sonia, et al.
Published: (2025)
by: Joseph, Sonia, et al.
Published: (2025)
IDEA: Image Description Enhanced CLIP-Adapter
by: Ye, Zhipeng, et al.
Published: (2025)
by: Ye, Zhipeng, et al.
Published: (2025)
ET tu, CLIP? Addressing Common Object Errors for Unseen Environments
by: Byun, Ye Won, et al.
Published: (2024)
by: Byun, Ye Won, et al.
Published: (2024)
Aligning Visual Contrastive learning models via Preference Optimization
by: Afzali, Amirabbas, et al.
Published: (2024)
by: Afzali, Amirabbas, et al.
Published: (2024)
RankCLIP: Ranking-Consistent Language-Image Pretraining
by: Zhang, Yiming, et al.
Published: (2024)
by: Zhang, Yiming, et al.
Published: (2024)
CLIPErase: Efficient Unlearning of Visual-Textual Associations in CLIP
by: Yang, Tianyu, et al.
Published: (2024)
by: Yang, Tianyu, et al.
Published: (2024)
WalkCLIP: Multimodal Learning for Urban Walkability Prediction
by: Xiang, Shilong, et al.
Published: (2025)
by: Xiang, Shilong, et al.
Published: (2025)
Learning Generalizable Prompt for CLIP with Class Similarity Knowledge
by: Jung, Sehun, et al.
Published: (2025)
by: Jung, Sehun, et al.
Published: (2025)
Memorization In Stable Diffusion Is Unexpectedly Driven by CLIP Embeddings
by: Kim, Bumjun, et al.
Published: (2026)
by: Kim, Bumjun, et al.
Published: (2026)
Captured by Captions: On Memorization and its Mitigation in CLIP Models
by: Wang, Wenhao, et al.
Published: (2025)
by: Wang, Wenhao, et al.
Published: (2025)
Investigating the Impact of Histopathological Foundation Models on Regressive Prediction of Homologous Recombination Deficiency
by: Blezinger, Alexander, et al.
Published: (2026)
by: Blezinger, Alexander, et al.
Published: (2026)
Bi-Level Optimization for Single Domain Generalization
by: Heidari, Marzi, et al.
Published: (2026)
by: Heidari, Marzi, et al.
Published: (2026)
Integration of Self-Supervised BYOL in Semi-Supervised Medical Image Recognition
by: Feng, Hao, et al.
Published: (2024)
by: Feng, Hao, et al.
Published: (2024)
FairerCLIP: Debiasing CLIP's Zero-Shot Predictions using Functions in RKHSs
by: Dehdashtian, Sepehr, et al.
Published: (2024)
by: Dehdashtian, Sepehr, et al.
Published: (2024)
Large-image Object Detection for Fine-grained Recognition of Punches Patterns in Medieval Panel Painting
by: Bruegger, Josh, et al.
Published: (2025)
by: Bruegger, Josh, et al.
Published: (2025)
Fine-tuning CLIP Text Encoders with Two-step Paraphrasing
by: Kim, Hyunjae, et al.
Published: (2024)
by: Kim, Hyunjae, et al.
Published: (2024)
On the Reproducibility of "FairCLIP: Harnessing Fairness in Vision-Language Learning''
by: Bakker, Hua Chang, et al.
Published: (2025)
by: Bakker, Hua Chang, et al.
Published: (2025)
Interpreting CLIP: Insights on the Robustness to ImageNet Distribution Shifts
by: Crabbé, Jonathan, et al.
Published: (2023)
by: Crabbé, Jonathan, et al.
Published: (2023)
A Hard-to-Beat Baseline for Training-free CLIP-based Adaptation
by: Wang, Zhengbo, et al.
Published: (2024)
by: Wang, Zhengbo, et al.
Published: (2024)
Similar Items
-
Multi-Rationale Explainable Object Recognition via Contrastive Conditional Inference
by: Rasekh, Ali, et al.
Published: (2025) -
ISLR101: an Iranian Word-Level Sign Language Recognition Dataset
by: Ranjbar, Hossein, et al.
Published: (2025) -
ExIQA: Explainable Image Quality Assessment Using Distortion Attributes
by: Ranjbar, Sepehr Kazemi, et al.
Published: (2024) -
Towards Precision Healthcare: Robust Fusion of Time Series and Image Data
by: Rasekh, Ali, et al.
Published: (2024) -
DiffCLIP: Differential Attention Meets CLIP
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2025)