How (Mis)calibrated is Your Federated CLIP and What To Do About It?
Fuente:
arXiv
Saved in:
| Main Authors: | Singha, Mainak, Aminbeidokhti, Masih, Casari, Paolo, Franchi, Gianni, Ricci, Elisa, Roy, Subhankar |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LT-Soups: Bridging Head and Tail Classes via Subsampled Model Soups
by: Aminbeidokhti, Masih, et al.
Published: (2025)
by: Aminbeidokhti, Masih, et al.
Published: (2025)
FedMVP: Federated Multimodal Visual Prompt Tuning for Vision-Language Models
by: Singha, Mainak, et al.
Published: (2025)
by: Singha, Mainak, et al.
Published: (2025)
CLIPoint3D: Language-Grounded Few-Shot Unsupervised 3D Point Cloud Domain Adaptation
by: Singha, Mainak, et al.
Published: (2026)
by: Singha, Mainak, et al.
Published: (2026)
Organizing Unstructured Image Collections using Natural Language
by: Liu, Mingxuan, et al.
Published: (2024)
by: Liu, Mingxuan, et al.
Published: (2024)
Ensembling Pruned Attention Heads For Uncertainty-Aware Efficient Transformers
by: Gabetni, Firas, et al.
Published: (2025)
by: Gabetni, Firas, et al.
Published: (2025)
AD-CLIP: Adapting Domains in Prompt Space Using CLIP
by: Singha, Mainak, et al.
Published: (2023)
by: Singha, Mainak, et al.
Published: (2023)
CLIP-QDA: An Explainable Concept Bottleneck Model
by: Kazmierczak, Rémi, et al.
Published: (2023)
by: Kazmierczak, Rémi, et al.
Published: (2023)
Enhancing Concept Localization in CLIP-based Concept Bottleneck Models
by: Kazmierczak, Rémi, et al.
Published: (2025)
by: Kazmierczak, Rémi, et al.
Published: (2025)
Frustratingly Easy Test-Time Adaptation of Vision-Language Models
by: Farina, Matteo, et al.
Published: (2024)
by: Farina, Matteo, et al.
Published: (2024)
MMLGNet: Cross-Modal Alignment of Remote Sensing Data using CLIP
by: Chaudhary, Aditya, et al.
Published: (2026)
by: Chaudhary, Aditya, et al.
Published: (2026)
Revisiting Mixout: An Overlooked Path to Robust Finetuning
by: Aminbeidokhti, Masih, et al.
Published: (2025)
by: Aminbeidokhti, Masih, et al.
Published: (2025)
Infrared Object Detection with Ultra Small ConvNets: Is ImageNet Pretraining Still Useful?
by: Muralidharan, Srikanth, et al.
Published: (2025)
by: Muralidharan, Srikanth, et al.
Published: (2025)
Unknown Prompt, the only Lacuna: Unveiling CLIP's Potential for Open Domain Generalization
by: Singha, Mainak, et al.
Published: (2024)
by: Singha, Mainak, et al.
Published: (2024)
COSMo: CLIP Talks on Open-Set Multi-Target Domain Adaptation
by: Monga, Munish, et al.
Published: (2024)
by: Monga, Munish, et al.
Published: (2024)
Source-Free Domain Adaptation for YOLO Object Detection
by: Varailhon, Simon, et al.
Published: (2024)
by: Varailhon, Simon, et al.
Published: (2024)
Unlearning Personal Data from a Single Image
by: De Min, Thomas, et al.
Published: (2024)
by: De Min, Thomas, et al.
Published: (2024)
ProactiveBench: Benchmarking Proactiveness in Multimodal Large Language Models
by: De Min, Thomas, et al.
Published: (2026)
by: De Min, Thomas, et al.
Published: (2026)
WiSE-OD: Benchmarking Robustness in Infrared Object Detection
by: Medeiros, Heitor R., et al.
Published: (2025)
by: Medeiros, Heitor R., et al.
Published: (2025)
High-Rate Mixout: Revisiting Mixout for Robust Domain Generalization
by: Aminbeidokhti, Masih, et al.
Published: (2025)
by: Aminbeidokhti, Masih, et al.
Published: (2025)
Democratizing Fine-grained Visual Recognition with Large Language Models
by: Liu, Mingxuan, et al.
Published: (2024)
by: Liu, Mingxuan, et al.
Published: (2024)
bi-modal textual prompt learning for vision-language models in remote sensing
by: Kashyap, Pankhi, et al.
Published: (2026)
by: Kashyap, Pankhi, et al.
Published: (2026)
Large-scale Pre-trained Models are Surprisingly Strong in Incremental Novel Class Discovery
by: Liu, Mingxuan, et al.
Published: (2023)
by: Liu, Mingxuan, et al.
Published: (2023)
Leveraging Visual Signals for Robust Token-Level Uncertainty in Vision-Language Generation
by: Hoche, Joseph, et al.
Published: (2026)
by: Hoche, Joseph, et al.
Published: (2026)
OSLoPrompt: Bridging Low-Supervision Challenges and Open-Set Domain Generalization in CLIP
by: C, Mohamad Hassan N, et al.
Published: (2025)
by: C, Mohamad Hassan N, et al.
Published: (2025)
Less is more: Summarizing Patch Tokens for efficient Multi-Label Class-Incremental Learning
by: De Min, Thomas, et al.
Published: (2024)
by: De Min, Thomas, et al.
Published: (2024)
COOkeD: Ensemble-based OOD detection in the era of zero-shot CLIP
by: Humblot-Renaux, Galadrielle, et al.
Published: (2025)
by: Humblot-Renaux, Galadrielle, et al.
Published: (2025)
WASH: Train your Ensemble with Communication-Efficient Weight Shuffling, then Average
by: Fournier, Louis, et al.
Published: (2024)
by: Fournier, Louis, et al.
Published: (2024)
Modality Translation for Object Detection Adaptation Without Forgetting Prior Knowledge
by: Medeiros, Heitor Rapela, et al.
Published: (2024)
by: Medeiros, Heitor Rapela, et al.
Published: (2024)
CTA: Cross-Task Alignment for Better Test Time Training
by: Barbeau, Samuel, et al.
Published: (2025)
by: Barbeau, Samuel, et al.
Published: (2025)
Hierarchical Light Transformer Ensembles for Multimodal Trajectory Forecasting
by: Lafage, Adrien, et al.
Published: (2024)
by: Lafage, Adrien, et al.
Published: (2024)
SS3D: End2End Self-Supervised 3D from Web Videos
by: Hariat, Marwane, et al.
Published: (2026)
by: Hariat, Marwane, et al.
Published: (2026)
Explainability for Vision Foundation Models: A Survey
by: Kazmierczak, Rémi, et al.
Published: (2025)
by: Kazmierczak, Rémi, et al.
Published: (2025)
Video, How Do Your Tokens Merge?
by: Pollard, Sam, et al.
Published: (2025)
by: Pollard, Sam, et al.
Published: (2025)
HalluciDet: Hallucinating RGB Modality for Person Detection Through Privileged Information
by: Medeiros, Heitor Rapela, et al.
Published: (2023)
by: Medeiros, Heitor Rapela, et al.
Published: (2023)
Domain Generalization by Rejecting Extreme Augmentations
by: Aminbeidokhti, Masih, et al.
Published: (2023)
by: Aminbeidokhti, Masih, et al.
Published: (2023)
From Weights to Concepts: Data-Free Interpretability of CLIP via Singular Vector Decomposition
by: Gentile, Francesco, et al.
Published: (2026)
by: Gentile, Francesco, et al.
Published: (2026)
EgoPrivacy: What Your First-Person Camera Says About You?
by: Li, Yijiang, et al.
Published: (2025)
by: Li, Yijiang, et al.
Published: (2025)
Elevating All Zero-Shot Sketch-Based Image Retrieval Through Multimodal Prompt Learning
by: Singha, Mainak, et al.
Published: (2024)
by: Singha, Mainak, et al.
Published: (2024)
Towards Understanding Why Label Smoothing Degrades Selective Classification and How to Fix It
by: Xia, Guoxuan, et al.
Published: (2024)
by: Xia, Guoxuan, et al.
Published: (2024)
How to Take a Memorable Picture? Empowering Users with Actionable Feedback
by: Laiti, Francesco, et al.
Published: (2026)
by: Laiti, Francesco, et al.
Published: (2026)
Similar Items
-
LT-Soups: Bridging Head and Tail Classes via Subsampled Model Soups
by: Aminbeidokhti, Masih, et al.
Published: (2025) -
FedMVP: Federated Multimodal Visual Prompt Tuning for Vision-Language Models
by: Singha, Mainak, et al.
Published: (2025) -
CLIPoint3D: Language-Grounded Few-Shot Unsupervised 3D Point Cloud Domain Adaptation
by: Singha, Mainak, et al.
Published: (2026) -
Organizing Unstructured Image Collections using Natural Language
by: Liu, Mingxuan, et al.
Published: (2024) -
Ensembling Pruned Attention Heads For Uncertainty-Aware Efficient Transformers
by: Gabetni, Firas, et al.
Published: (2025)