Calibration-Aware Prompt Learning for Medical Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Basu, Abhishek, Shamshad, Fahad, Sharifdeen, Ashshak, Nandakumar, Karthik, Khan, Muhammad Haris |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Calibrating Prompt Tuning of Vision-Language Models
von: Sharifdeen, Ashshak, et al.
Veröffentlicht: (2026)
von: Sharifdeen, Ashshak, et al.
Veröffentlicht: (2026)
O-TPT: Orthogonality Constraints for Calibrating Test-time Prompt Tuning in Vision-Language Models
von: Sharifdeen, Ashshak, et al.
Veröffentlicht: (2025)
von: Sharifdeen, Ashshak, et al.
Veröffentlicht: (2025)
PromptSmooth: Certifying Robustness of Medical Vision-Language Models via Prompt Learning
von: Hussein, Noor, et al.
Veröffentlicht: (2024)
von: Hussein, Noor, et al.
Veröffentlicht: (2024)
BAPLe: Backdoor Attacks on Medical Foundational Models using Prompt Learning
von: Hanif, Asif, et al.
Veröffentlicht: (2024)
von: Hanif, Asif, et al.
Veröffentlicht: (2024)
RAVEN: Erasing Invisible Watermarks via Novel View Synthesis
von: Shamshad, Fahad, et al.
Veröffentlicht: (2026)
von: Shamshad, Fahad, et al.
Veröffentlicht: (2026)
Robust-LLaVA: On the Effectiveness of Large-Scale Robust Image Encoders for Multi-modal Large Language Models
von: Malik, Hashmat Shadab, et al.
Veröffentlicht: (2025)
von: Malik, Hashmat Shadab, et al.
Veröffentlicht: (2025)
SafeDiffusion-R1: Online Reward Steering for Safe Diffusion Post-Training
von: Kumar, Komal, et al.
Veröffentlicht: (2026)
von: Kumar, Komal, et al.
Veröffentlicht: (2026)
Makeup-Guided Facial Privacy Protection via Untrained Neural Network Priors
von: Shamshad, Fahad, et al.
Veröffentlicht: (2024)
von: Shamshad, Fahad, et al.
Veröffentlicht: (2024)
Towards Evaluating the Robustness of Visual State Space Models
von: Malik, Hashmat Shadab, et al.
Veröffentlicht: (2024)
von: Malik, Hashmat Shadab, et al.
Veröffentlicht: (2024)
FaceAnonyMixer: Cancelable Faces via Identity Consistent Latent Space Mixing
von: Alam, Mohammed Talha, et al.
Veröffentlicht: (2025)
von: Alam, Mohammed Talha, et al.
Veröffentlicht: (2025)
STEREO: A Two-Stage Framework for Adversarially Robust Concept Erasing from Text-to-Image Diffusion Models
von: Srivatsan, Koushik, et al.
Veröffentlicht: (2024)
von: Srivatsan, Koushik, et al.
Veröffentlicht: (2024)
A-TPT: Angular Diversity Calibration Properties for Test-Time Prompt Tuning of Vision-Language Models
von: Ahamed, Shihab Aaqil, et al.
Veröffentlicht: (2025)
von: Ahamed, Shihab Aaqil, et al.
Veröffentlicht: (2025)
First-Place Solution to NeurIPS 2024 Invisible Watermark Removal Challenge
von: Shamshad, Fahad, et al.
Veröffentlicht: (2025)
von: Shamshad, Fahad, et al.
Veröffentlicht: (2025)
Test-Time Low Rank Adaptation via Confidence Maximization for Zero-Shot Generalization of Vision-Language Models
von: Imam, Raza, et al.
Veröffentlicht: (2024)
von: Imam, Raza, et al.
Veröffentlicht: (2024)
VFace: A Training-Free Approach for Diffusion-Based Video Face Swapping
von: Baliah, Sanoojan, et al.
Veröffentlicht: (2026)
von: Baliah, Sanoojan, et al.
Veröffentlicht: (2026)
DPA: Dual Prototypes Alignment for Unsupervised Adaptation of Vision-Language Models
von: Ali, Eman, et al.
Veröffentlicht: (2024)
von: Ali, Eman, et al.
Veröffentlicht: (2024)
Noise-Tolerant Few-Shot Unsupervised Adapter for Vision-Language Models
von: Ali, Eman, et al.
Veröffentlicht: (2023)
von: Ali, Eman, et al.
Veröffentlicht: (2023)
XrayGPT: Chest Radiographs Summarization using Medical Vision-Language Models
von: Thawakar, Omkar, et al.
Veröffentlicht: (2023)
von: Thawakar, Omkar, et al.
Veröffentlicht: (2023)
Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models
von: Maaz, Muhammad, et al.
Veröffentlicht: (2023)
von: Maaz, Muhammad, et al.
Veröffentlicht: (2023)
Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model
von: Chen, Shiming, et al.
Veröffentlicht: (2025)
von: Chen, Shiming, et al.
Veröffentlicht: (2025)
Chameleon: Images Are What You Need For Multimodal Learning Robust To Missing Modalities
von: Liaqat, Muhammad Irzam, et al.
Veröffentlicht: (2024)
von: Liaqat, Muhammad Irzam, et al.
Veröffentlicht: (2024)
Modality Invariant Multimodal Learning to Handle Missing Modalities: A Single-Branch Approach
von: Saeed, Muhammad Saad, et al.
Veröffentlicht: (2024)
von: Saeed, Muhammad Saad, et al.
Veröffentlicht: (2024)
SPQR: A Standardized Benchmark for Modern Safety Alignment Methods in Text-to-Image Diffusion Models
von: Alam, Mohammed Talha, et al.
Veröffentlicht: (2025)
von: Alam, Mohammed Talha, et al.
Veröffentlicht: (2025)
Not All Modalities Are Equal: Instruction-Aware Gating for Multimodal Videos
von: Ding, Bonan, et al.
Veröffentlicht: (2026)
von: Ding, Bonan, et al.
Veröffentlicht: (2026)
Learning to Prompt with Text Only Supervision for Vision-Language Models
von: Khattak, Muhammad Uzair, et al.
Veröffentlicht: (2024)
von: Khattak, Muhammad Uzair, et al.
Veröffentlicht: (2024)
Language Guided Domain Generalized Medical Image Segmentation
von: Kunhimon, Shahina, et al.
Veröffentlicht: (2024)
von: Kunhimon, Shahina, et al.
Veröffentlicht: (2024)
Prompt-Aware Adaptive Elastic Weight Consolidation for Continual Learning in Medical Vision-Language Models
von: Gao, Ziyuan, et al.
Veröffentlicht: (2025)
von: Gao, Ziyuan, et al.
Veröffentlicht: (2025)
GEOBench-VLM: Benchmarking Vision-Language Models for Geospatial Tasks
von: Danish, Muhammad Sohail, et al.
Veröffentlicht: (2024)
von: Danish, Muhammad Sohail, et al.
Veröffentlicht: (2024)
Diversity Covariance-Aware Prompt Learning for Vision-Language Models
von: Dong, Songlin, et al.
Veröffentlicht: (2025)
von: Dong, Songlin, et al.
Veröffentlicht: (2025)
Video-R2: Reinforcing Consistent and Grounded Reasoning in Multimodal Language Models
von: Maaz, Muhammad, et al.
Veröffentlicht: (2025)
von: Maaz, Muhammad, et al.
Veröffentlicht: (2025)
Improving Calibration in Test-Time Prompt Tuning for Vision-Language Models via Data-Free Flatness-Aware Prompt Pretraining
von: Jang, Hyeonseo, et al.
Veröffentlicht: (2026)
von: Jang, Hyeonseo, et al.
Veröffentlicht: (2026)
Robust and Label-Efficient Deep Waste Detection
von: Abid, Hassan, et al.
Veröffentlicht: (2025)
von: Abid, Hassan, et al.
Veröffentlicht: (2025)
IntCoOp: Interpretability-Aware Vision-Language Prompt Tuning
von: Ghosal, Soumya Suvra, et al.
Veröffentlicht: (2024)
von: Ghosal, Soumya Suvra, et al.
Veröffentlicht: (2024)
TerraFM: A Scalable Foundation Model for Unified Multisensor Earth Observation
von: Danish, Muhammad Sohail, et al.
Veröffentlicht: (2025)
von: Danish, Muhammad Sohail, et al.
Veröffentlicht: (2025)
DiffuseMix: Label-Preserving Data Augmentation with Diffusion Models
von: Islam, Khawar, et al.
Veröffentlicht: (2024)
von: Islam, Khawar, et al.
Veröffentlicht: (2024)
Multi-Attribute Vision Transformers are Efficient and Robust Learners
von: Gani, Hanan, et al.
Veröffentlicht: (2024)
von: Gani, Hanan, et al.
Veröffentlicht: (2024)
CountZES: Counting via Zero-Shot Exemplar Selection
von: Siddiqui, Muhammad Ibraheem, et al.
Veröffentlicht: (2025)
von: Siddiqui, Muhammad Ibraheem, et al.
Veröffentlicht: (2025)
Vocabulary-free Fine-grained Visual Recognition via Enriched Contextually Grounded Vision-Language Model
von: Demidov, Dmitry, et al.
Veröffentlicht: (2025)
von: Demidov, Dmitry, et al.
Veröffentlicht: (2025)
Agentic AI for Remote Sensing: Technical Challenges and Research Directions
von: Munir, Muhammad Akhtar, et al.
Veröffentlicht: (2026)
von: Munir, Muhammad Akhtar, et al.
Veröffentlicht: (2026)
Align Your Prompts: Test-Time Prompting with Distribution Alignment for Zero-Shot Generalization
von: Hassan, Jameel, et al.
Veröffentlicht: (2023)
von: Hassan, Jameel, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Towards Calibrating Prompt Tuning of Vision-Language Models
von: Sharifdeen, Ashshak, et al.
Veröffentlicht: (2026) -
O-TPT: Orthogonality Constraints for Calibrating Test-time Prompt Tuning in Vision-Language Models
von: Sharifdeen, Ashshak, et al.
Veröffentlicht: (2025) -
PromptSmooth: Certifying Robustness of Medical Vision-Language Models via Prompt Learning
von: Hussein, Noor, et al.
Veröffentlicht: (2024) -
BAPLe: Backdoor Attacks on Medical Foundational Models using Prompt Learning
von: Hanif, Asif, et al.
Veröffentlicht: (2024) -
RAVEN: Erasing Invisible Watermarks via Novel View Synthesis
von: Shamshad, Fahad, et al.
Veröffentlicht: (2026)