Pruning the Paradox: How CLIP's Most Informative Heads Enhance Performance While Amplifying Bias
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Madasu, Avinash, Lal, Vasudev, Howard, Phillip |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Quantifying and Enabling the Interpretability of CLIP-like Models
von: Madasu, Avinash, et al.
Veröffentlicht: (2024)
von: Madasu, Avinash, et al.
Veröffentlicht: (2024)
Cultural Awareness in Vision-Language Models: A Cross-Country Exploration
von: Madasu, Avinash, et al.
Veröffentlicht: (2025)
von: Madasu, Avinash, et al.
Veröffentlicht: (2025)
ICSVR: Investigating Compositional and Syntactic Understanding in Video Retrieval Models
von: Madasu, Avinash, et al.
Veröffentlicht: (2023)
von: Madasu, Avinash, et al.
Veröffentlicht: (2023)
SocialCounterfactuals: Probing and Mitigating Intersectional Social Biases in Vision-Language Models with Counterfactual Examples
von: Howard, Phillip, et al.
Veröffentlicht: (2023)
von: Howard, Phillip, et al.
Veröffentlicht: (2023)
Training-Free Mitigation of Language Reasoning Degradation After Multimodal Instruction Tuning
von: Ratzlaff, Neale, et al.
Veröffentlicht: (2024)
von: Ratzlaff, Neale, et al.
Veröffentlicht: (2024)
Is Your Paper Being Reviewed by an LLM? Benchmarking AI Text Detection in Peer Review
von: Yu, Sungduk, et al.
Veröffentlicht: (2025)
von: Yu, Sungduk, et al.
Veröffentlicht: (2025)
Is Your Paper Being Reviewed by an LLM? Investigating AI Text Detectability in Peer Review
von: Yu, Sungduk, et al.
Veröffentlicht: (2024)
von: Yu, Sungduk, et al.
Veröffentlicht: (2024)
DPO Learning with LLMs-Judge Signal for Computer Use Agents
von: Luo, Man, et al.
Veröffentlicht: (2025)
von: Luo, Man, et al.
Veröffentlicht: (2025)
Locating Demographic Bias at the Attention-Head Level in CLIP's Vision Encoder
von: Yasser, Alaa, et al.
Veröffentlicht: (2026)
von: Yasser, Alaa, et al.
Veröffentlicht: (2026)
Uncovering Bias in Large Vision-Language Models with Counterfactuals
von: Howard, Phillip, et al.
Veröffentlicht: (2024)
von: Howard, Phillip, et al.
Veröffentlicht: (2024)
LVLM-Compress-Bench: Benchmarking the Broader Impact of Large Vision-Language Model Compression
von: Kundu, Souvik, et al.
Veröffentlicht: (2025)
von: Kundu, Souvik, et al.
Veröffentlicht: (2025)
Decoupling Template Bias in CLIP: Harnessing Empty Prompts for Enhanced Few-Shot Learning
von: Zhang, Zhenyu, et al.
Veröffentlicht: (2025)
von: Zhang, Zhenyu, et al.
Veröffentlicht: (2025)
Debiasing Classifiers by Amplifying Bias with Latent Diffusion and Large Language Models
von: Ko, Donggeun, et al.
Veröffentlicht: (2024)
von: Ko, Donggeun, et al.
Veröffentlicht: (2024)
Learning from Reasoning Failures via Synthetic Data Generation
von: Stan, Gabriela Ben Melech, et al.
Veröffentlicht: (2025)
von: Stan, Gabriela Ben Melech, et al.
Veröffentlicht: (2025)
Enhanced Structured Lasso Pruning with Class-wise Information
von: Liu, Xiang, et al.
Veröffentlicht: (2025)
von: Liu, Xiang, et al.
Veröffentlicht: (2025)
Geometry-Aware CLIP Retrieval via Local Cross-Modal Alignment and Steering
von: Prakash, Nirmalendu, et al.
Veröffentlicht: (2026)
von: Prakash, Nirmalendu, et al.
Veröffentlicht: (2026)
MoPE-CLIP: Structured Pruning for Efficient Vision-Language Models with Module-wise Pruning Error Metric
von: Lin, Haokun, et al.
Veröffentlicht: (2024)
von: Lin, Haokun, et al.
Veröffentlicht: (2024)
Automatic Channel Pruning for Multi-Head Attention
von: Lee, Eunho, et al.
Veröffentlicht: (2024)
von: Lee, Eunho, et al.
Veröffentlicht: (2024)
AA-CLIP: Enhancing Zero-shot Anomaly Detection via Anomaly-Aware CLIP
von: Ma, Wenxin, et al.
Veröffentlicht: (2025)
von: Ma, Wenxin, et al.
Veröffentlicht: (2025)
CLIP-MUSED: CLIP-Guided Multi-Subject Visual Neural Information Semantic Decoding
von: Zhou, Qiongyi, et al.
Veröffentlicht: (2024)
von: Zhou, Qiongyi, et al.
Veröffentlicht: (2024)
CLIP Tricks You: Training-free Token Pruning for Efficient Pixel Grounding in Large VIsion-Language Models
von: Lee, Sangin, et al.
Veröffentlicht: (2026)
von: Lee, Sangin, et al.
Veröffentlicht: (2026)
Omni-NegCLIP: Enhancing CLIP with Front-Layer Contrastive Fine-Tuning for Comprehensive Negation Understanding
von: Xu, Jingqi
Veröffentlicht: (2026)
von: Xu, Jingqi
Veröffentlicht: (2026)
Data Pruning by Information Maximization
von: Tan, Haoru, et al.
Veröffentlicht: (2025)
von: Tan, Haoru, et al.
Veröffentlicht: (2025)
SafeR-CLIP: Mitigating NSFW Content in Vision-Language Models While Preserving Pre-Trained Knowledge
von: Yousaf, Adeel, et al.
Veröffentlicht: (2025)
von: Yousaf, Adeel, et al.
Veröffentlicht: (2025)
FiVL: A Framework for Improved Vision-Language Alignment through the Lens of Training, Evaluation and Explainability
von: Aflalo, Estelle, et al.
Veröffentlicht: (2024)
von: Aflalo, Estelle, et al.
Veröffentlicht: (2024)
Grounding-Aware Token Pruning: Recovering from Drastic Performance Drops in Visual Grounding Caused by Pruning
von: Chien, Tzu-Chun, et al.
Veröffentlicht: (2025)
von: Chien, Tzu-Chun, et al.
Veröffentlicht: (2025)
PLPHP: Per-Layer Per-Head Vision Token Pruning for Efficient Large Vision-Language Models
von: Meng, Yu, et al.
Veröffentlicht: (2025)
von: Meng, Yu, et al.
Veröffentlicht: (2025)
One Head Eight Arms: Block Matrix based Low Rank Adaptation for CLIP-based Few-Shot Learning
von: Zhou, Chunpeng, et al.
Veröffentlicht: (2025)
von: Zhou, Chunpeng, et al.
Veröffentlicht: (2025)
Enhancing Multimodal Understanding with CLIP-Based Image-to-Text Transformation
von: Che, Chang, et al.
Veröffentlicht: (2024)
von: Che, Chang, et al.
Veröffentlicht: (2024)
Synergy and Diversity in CLIP: Enhancing Performance Through Adaptive Backbone Ensembling
von: Rodriguez-Opazo, Cristian, et al.
Veröffentlicht: (2024)
von: Rodriguez-Opazo, Cristian, et al.
Veröffentlicht: (2024)
How Bias Binds: Measuring Hidden Associations for Bias Control in Text-to-Image Compositions
von: Li, Jeng-Lin, et al.
Veröffentlicht: (2025)
von: Li, Jeng-Lin, et al.
Veröffentlicht: (2025)
Enhancing Compositional Reasoning in CLIP via Reconstruction and Alignment of Text Descriptions
von: Kwon, Jihoon, et al.
Veröffentlicht: (2025)
von: Kwon, Jihoon, et al.
Veröffentlicht: (2025)
InterCLIP-MEP: Interactive CLIP and Memory-Enhanced Predictor for Multi-modal Sarcasm Detection
von: Chen, Junjie, et al.
Veröffentlicht: (2024)
von: Chen, Junjie, et al.
Veröffentlicht: (2024)
Cross-Cultural Value Awareness in Large Vision-Language Models
von: Howard, Phillip, et al.
Veröffentlicht: (2026)
von: Howard, Phillip, et al.
Veröffentlicht: (2026)
AdaptCLIP: Adapting CLIP for Universal Visual Anomaly Detection
von: Gao, Bin-Bin, et al.
Veröffentlicht: (2025)
von: Gao, Bin-Bin, et al.
Veröffentlicht: (2025)
DesignCLIP: Multimodal Learning with CLIP for Design Patent Understanding
von: Wang, Zhu, et al.
Veröffentlicht: (2025)
von: Wang, Zhu, et al.
Veröffentlicht: (2025)
Semantic Relation-Enhanced CLIP Adapter for Domain Adaptive Zero-Shot Learning
von: Yu, Jiaao, et al.
Veröffentlicht: (2025)
von: Yu, Jiaao, et al.
Veröffentlicht: (2025)
Enhancing Layer Attention Efficiency through Pruning Redundant Retrievals
von: Li, Hanze, et al.
Veröffentlicht: (2025)
von: Li, Hanze, et al.
Veröffentlicht: (2025)
UNSEEN: Enhancing Dataset Pruning from a Generalization Perspective
von: Xu, Furui, et al.
Veröffentlicht: (2025)
von: Xu, Furui, et al.
Veröffentlicht: (2025)
The Reasoning Boundary Paradox: How Reinforcement Learning Constrains Language Models
von: Nguyen, Phuc Minh, et al.
Veröffentlicht: (2025)
von: Nguyen, Phuc Minh, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Quantifying and Enabling the Interpretability of CLIP-like Models
von: Madasu, Avinash, et al.
Veröffentlicht: (2024) -
Cultural Awareness in Vision-Language Models: A Cross-Country Exploration
von: Madasu, Avinash, et al.
Veröffentlicht: (2025) -
ICSVR: Investigating Compositional and Syntactic Understanding in Video Retrieval Models
von: Madasu, Avinash, et al.
Veröffentlicht: (2023) -
SocialCounterfactuals: Probing and Mitigating Intersectional Social Biases in Vision-Language Models with Counterfactual Examples
von: Howard, Phillip, et al.
Veröffentlicht: (2023) -
Training-Free Mitigation of Language Reasoning Degradation After Multimodal Instruction Tuning
von: Ratzlaff, Neale, et al.
Veröffentlicht: (2024)