A Closer Look at the Robustness of Contrastive Language-Image Pre-Training (CLIP)
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tu, Weijie, Deng, Weijian, Gedeon, Tom |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Toward a Holistic Evaluation of Robustness in CLIP Models
von: Tu, Weijie, et al.
Veröffentlicht: (2024)
von: Tu, Weijie, et al.
Veröffentlicht: (2024)
What Does Softmax Probability Tell Us about Classifiers Ranking Across Diverse Test Conditions?
von: Tu, Weijie, et al.
Veröffentlicht: (2024)
von: Tu, Weijie, et al.
Veröffentlicht: (2024)
An Empirical Study Into What Matters for Calibrating Vision-Language Models
von: Tu, Weijie, et al.
Veröffentlicht: (2024)
von: Tu, Weijie, et al.
Veröffentlicht: (2024)
A Closer Look at Multimodal Representation Collapse
von: Chaudhuri, Abhra, et al.
Veröffentlicht: (2025)
von: Chaudhuri, Abhra, et al.
Veröffentlicht: (2025)
A Closer Look at the Explainability of Contrastive Language-Image Pre-training
von: Li, Yi, et al.
Veröffentlicht: (2023)
von: Li, Yi, et al.
Veröffentlicht: (2023)
Contrastive Localized Language-Image Pre-Training
von: Chen, Hong-You, et al.
Veröffentlicht: (2024)
von: Chen, Hong-You, et al.
Veröffentlicht: (2024)
TopoFR: A Closer Look at Topology Alignment on Face Recognition
von: Dan, Jun, et al.
Veröffentlicht: (2024)
von: Dan, Jun, et al.
Veröffentlicht: (2024)
AdFair-CLIP: Adversarial Fair Contrastive Language-Image Pre-training for Chest X-rays
von: Yi, Chenlang, et al.
Veröffentlicht: (2025)
von: Yi, Chenlang, et al.
Veröffentlicht: (2025)
Embedding Geometries of Contrastive Language-Image Pre-Training
von: Chou, Jason Chuan-Chih, et al.
Veröffentlicht: (2024)
von: Chou, Jason Chuan-Chih, et al.
Veröffentlicht: (2024)
Adaptive Multi-head Contrastive Learning
von: Wang, Lei, et al.
Veröffentlicht: (2023)
von: Wang, Lei, et al.
Veröffentlicht: (2023)
Confidence and Dispersity as Signals: Unsupervised Model Evaluation and Ranking
von: Deng, Weijian, et al.
Veröffentlicht: (2025)
von: Deng, Weijian, et al.
Veröffentlicht: (2025)
Look Closer! An Adversarial Parametric Editing Framework for Hallucination Mitigation in VLMs
von: Hu, Jiayu, et al.
Veröffentlicht: (2025)
von: Hu, Jiayu, et al.
Veröffentlicht: (2025)
Probabilistic Language-Image Pre-Training
von: Chun, Sanghyuk, et al.
Veröffentlicht: (2024)
von: Chun, Sanghyuk, et al.
Veröffentlicht: (2024)
Ranked from Within: Ranking Large Multimodal Models Without Labels
von: Tu, Weijie, et al.
Veröffentlicht: (2024)
von: Tu, Weijie, et al.
Veröffentlicht: (2024)
Finding Closure: A Closer Look at the Gestalt Law of Closure in Convolutional Neural Networks
von: Zhang, Yuyan, et al.
Veröffentlicht: (2024)
von: Zhang, Yuyan, et al.
Veröffentlicht: (2024)
Denoising Fisher Training For Neural Implicit Samplers
von: Luo, Weijian, et al.
Veröffentlicht: (2024)
von: Luo, Weijian, et al.
Veröffentlicht: (2024)
TV100: A TV Series Dataset that Pre-Trained CLIP Has Not Seen
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2024)
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2024)
A Closer Look at In-Distribution vs. Out-of-Distribution Accuracy for Open-Set Test-time Adaptation
von: Li, Zefeng, et al.
Veröffentlicht: (2026)
von: Li, Zefeng, et al.
Veröffentlicht: (2026)
What Makes CLIP More Robust to Long-Tailed Pre-Training Data? A Controlled Study for Transferable Insights
von: Wen, Xin, et al.
Veröffentlicht: (2024)
von: Wen, Xin, et al.
Veröffentlicht: (2024)
How Does the Spatial Distribution of Pre-training Data Affect Geospatial Foundation Models?
von: Purohit, Mirali, et al.
Veröffentlicht: (2025)
von: Purohit, Mirali, et al.
Veröffentlicht: (2025)
Contrast-Aware Calibration for Fine-Tuned CLIP: Leveraging Image-Text Alignment
von: Lv, Song-Lin, et al.
Veröffentlicht: (2025)
von: Lv, Song-Lin, et al.
Veröffentlicht: (2025)
Taylor Videos for Action Recognition
von: Wang, Lei, et al.
Veröffentlicht: (2024)
von: Wang, Lei, et al.
Veröffentlicht: (2024)
Centered Masking for Language-Image Pre-Training
von: Liang, Mingliang, et al.
Veröffentlicht: (2024)
von: Liang, Mingliang, et al.
Veröffentlicht: (2024)
TrackNetV4: Enhancing Fast Sports Object Tracking with Motion Attention Maps
von: Raj, Arjun, et al.
Veröffentlicht: (2024)
von: Raj, Arjun, et al.
Veröffentlicht: (2024)
D4C: Data-Free Quantization for Contrastive Language-Image Pre-training Models
von: Zhang, Wenlun, et al.
Veröffentlicht: (2025)
von: Zhang, Wenlun, et al.
Veröffentlicht: (2025)
FastCLIP: A Suite of Optimization Techniques to Accelerate CLIP Training with Limited Resources
von: Wei, Xiyuan, et al.
Veröffentlicht: (2024)
von: Wei, Xiyuan, et al.
Veröffentlicht: (2024)
SpikeCLIP: A Contrastive Language-Image Pretrained Spiking Neural Network
von: Lv, Changze, et al.
Veröffentlicht: (2023)
von: Lv, Changze, et al.
Veröffentlicht: (2023)
NeuCLIP: Efficient Large-Scale CLIP Training with Neural Normalizer Optimization
von: Wei, Xiyuan, et al.
Veröffentlicht: (2025)
von: Wei, Xiyuan, et al.
Veröffentlicht: (2025)
CLIP-UP: A Simple and Efficient Mixture-of-Experts CLIP Training Recipe with Sparse Upcycling
von: Wang, Xinze, et al.
Veröffentlicht: (2025)
von: Wang, Xinze, et al.
Veröffentlicht: (2025)
AngularFuse: A Closer Look at Angle-based Perception for Spatial-Sensitive Multi-Modality Image Fusion
von: Liu, Xiaopeng, et al.
Veröffentlicht: (2025)
von: Liu, Xiaopeng, et al.
Veröffentlicht: (2025)
When Spatial meets Temporal in Action Recognition
von: Chen, Huilin, et al.
Veröffentlicht: (2024)
von: Chen, Huilin, et al.
Veröffentlicht: (2024)
Gems: Group Emotion Profiling Through Multimodal Situational Understanding
von: Kataria, Anubhav, et al.
Veröffentlicht: (2025)
von: Kataria, Anubhav, et al.
Veröffentlicht: (2025)
CSGaze: Context-aware Social Gaze Prediction
von: Madan, Surbhi, et al.
Veröffentlicht: (2025)
von: Madan, Surbhi, et al.
Veröffentlicht: (2025)
Fine-Grained Classification: Connecting Metadata via Cross-Contrastive Pre-Training
von: Mamtani, Sumit, et al.
Veröffentlicht: (2025)
von: Mamtani, Sumit, et al.
Veröffentlicht: (2025)
A Closer Look at Benchmarking Self-Supervised Pre-training with Image Classification
von: Marks, Markus, et al.
Veröffentlicht: (2024)
von: Marks, Markus, et al.
Veröffentlicht: (2024)
A Closer Look at Spatial-Slice Features Learning for COVID-19 Detection
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
von: Hsu, Chih-Chung, et al.
Veröffentlicht: (2024)
TiC-CLIP: Continual Training of CLIP Models
von: Garg, Saurabh, et al.
Veröffentlicht: (2023)
von: Garg, Saurabh, et al.
Veröffentlicht: (2023)
SafeR-CLIP: Mitigating NSFW Content in Vision-Language Models While Preserving Pre-Trained Knowledge
von: Yousaf, Adeel, et al.
Veröffentlicht: (2025)
von: Yousaf, Adeel, et al.
Veröffentlicht: (2025)
Prioritizing Image-Related Tokens Enhances Vision-Language Pre-Training
von: Chen, Yangyi, et al.
Veröffentlicht: (2025)
von: Chen, Yangyi, et al.
Veröffentlicht: (2025)
Diff-Instruct++: Training One-step Text-to-image Generator Model to Align with Human Preferences
von: Luo, Weijian
Veröffentlicht: (2024)
von: Luo, Weijian
Veröffentlicht: (2024)
Ähnliche Einträge
-
Toward a Holistic Evaluation of Robustness in CLIP Models
von: Tu, Weijie, et al.
Veröffentlicht: (2024) -
What Does Softmax Probability Tell Us about Classifiers Ranking Across Diverse Test Conditions?
von: Tu, Weijie, et al.
Veröffentlicht: (2024) -
An Empirical Study Into What Matters for Calibrating Vision-Language Models
von: Tu, Weijie, et al.
Veröffentlicht: (2024) -
A Closer Look at Multimodal Representation Collapse
von: Chaudhuri, Abhra, et al.
Veröffentlicht: (2025) -
A Closer Look at the Explainability of Contrastive Language-Image Pre-training
von: Li, Yi, et al.
Veröffentlicht: (2023)