A Reality Check of Vision-Language Pre-training in Radiology: Have We Progressed Using Text?
Fuente:
arXiv
Saved in:
| Main Authors: | Silva-Rodríguez, Julio, Dolz, Jose, Ayed, Ismail Ben |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Robust Calibration of Large Vision-Language Adapters
by: Murugesan, Balamurali, et al.
Published: (2024)
by: Murugesan, Balamurali, et al.
Published: (2024)
A Closer Look at the Few-Shot Adaptation of Large Vision-Language Models
by: Silva-Rodríguez, Julio, et al.
Published: (2023)
by: Silva-Rodríguez, Julio, et al.
Published: (2023)
Towards Foundation Models and Few-Shot Parameter-Efficient Fine-Tuning for Volumetric Organ Segmentation
by: Silva-Rodríguez, Julio, et al.
Published: (2023)
by: Silva-Rodríguez, Julio, et al.
Published: (2023)
Trustworthy Few-Shot Transfer of Medical VLMs through Split Conformal Prediction
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
Conformal Prediction for Zero-Shot Models
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
A Foundation Language-Image Model of the Retina (FLAIR): Encoding Expert Knowledge in Text Supervision
by: Silva-Rodríguez, Julio, et al.
Published: (2023)
by: Silva-Rodríguez, Julio, et al.
Published: (2023)
Semantic Anchor Transport: Robust Test-Time Adaptation for Vision-Language Models
by: Mishra, Shambhavi, et al.
Published: (2024)
by: Mishra, Shambhavi, et al.
Published: (2024)
ORION: ORthonormal Text Encoding for Universal VLM AdaptatION
by: Chakraborty, Omprakash, et al.
Published: (2026)
by: Chakraborty, Omprakash, et al.
Published: (2026)
Few-shot Adaptation of Medical Vision-Language Models
by: Shakeri, Fereshteh, et al.
Published: (2024)
by: Shakeri, Fereshteh, et al.
Published: (2024)
Class and Region-Adaptive Constraints for Network Calibration
by: Murugesan, Balamurali, et al.
Published: (2024)
by: Murugesan, Balamurali, et al.
Published: (2024)
Full Conformal Adaptation of Medical Vision-Language Models
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
Pay Attention to Your Neighbours: Training-Free Open-Vocabulary Semantic Segmentation
by: Hajimiri, Sina, et al.
Published: (2024)
by: Hajimiri, Sina, et al.
Published: (2024)
Few-Shot, Now for Real: Medical VLMs Adaptation without Balanced Sets or Validation
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
Regularized Low-Rank Adaptation for Few-Shot Organ Segmentation
by: Baklouti, Ghassen, et al.
Published: (2025)
by: Baklouti, Ghassen, et al.
Published: (2025)
ViLU: Learning Vision-Language Uncertainties for Failure Prediction
by: Lafon, Marc, et al.
Published: (2025)
by: Lafon, Marc, et al.
Published: (2025)
Low-Rank Few-Shot Adaptation of Vision-Language Models
by: Zanella, Maxime, et al.
Published: (2024)
by: Zanella, Maxime, et al.
Published: (2024)
Locality-Attending Vision Transformer
by: Hajimiri, Sina, et al.
Published: (2026)
by: Hajimiri, Sina, et al.
Published: (2026)
Boosting Vision-Language Models with Transduction
by: Zanella, Maxime, et al.
Published: (2024)
by: Zanella, Maxime, et al.
Published: (2024)
Exploring the Transferability of a Foundation Model for Fundus Images: Application to Hypertensive Retinopathy
by: Silva-Rodriguez, Julio, et al.
Published: (2024)
by: Silva-Rodriguez, Julio, et al.
Published: (2024)
Prompting classes: Exploring the Power of Prompt Class Learning in Weakly Supervised Semantic Segmentation
by: Murugesan, Balamurali, et al.
Published: (2023)
by: Murugesan, Balamurali, et al.
Published: (2023)
Calibrating Segmentation Networks with Margin-based Label Smoothing
by: Murugesan, Balamurali, et al.
Published: (2022)
by: Murugesan, Balamurali, et al.
Published: (2022)
Are foundation models for computer vision good conformal predictors?
by: Fillioux, Leo, et al.
Published: (2024)
by: Fillioux, Leo, et al.
Published: (2024)
Class Adaptive Conformal Training
by: Marani, Badr-Eddine, et al.
Published: (2026)
by: Marani, Badr-Eddine, et al.
Published: (2026)
LP++: A Surprisingly Strong Linear Probe for Few-Shot CLIP
by: Huang, Yunshi, et al.
Published: (2024)
by: Huang, Yunshi, et al.
Published: (2024)
Realistic Test-Time Adaptation of Vision-Language Models
by: Zanella, Maxime, et al.
Published: (2025)
by: Zanella, Maxime, et al.
Published: (2025)
Do not trust what you trust: Miscalibration in Semi-supervised Learning
by: Mishra, Shambhavi, et al.
Published: (2024)
by: Mishra, Shambhavi, et al.
Published: (2024)
Vocabulary-free few-shot learning for Vision-Language Models
by: Zanella, Maxime, et al.
Published: (2025)
by: Zanella, Maxime, et al.
Published: (2025)
Trust your neighbours: Penalty-based constraints for model calibration
by: Murugesan, Balamurali, et al.
Published: (2023)
by: Murugesan, Balamurali, et al.
Published: (2023)
Neighbor-Aware Calibration of Segmentation Networks with Penalty-Based Constraints
by: Murugesan, Balamurali, et al.
Published: (2024)
by: Murugesan, Balamurali, et al.
Published: (2024)
Boosting Vision-Language Models for Histopathology Classification: Predict all at once
by: Zanella, Maxime, et al.
Published: (2024)
by: Zanella, Maxime, et al.
Published: (2024)
Language-Aware Information Maximization for Transductive Few-Shot CLIP
by: Baklouti, Ghassen, et al.
Published: (2025)
by: Baklouti, Ghassen, et al.
Published: (2025)
On the test-time zero-shot generalization of vision-language models: Do we really need prompt learning?
by: Zanella, Maxime, et al.
Published: (2024)
by: Zanella, Maxime, et al.
Published: (2024)
SoC: Semantic Orthogonal Calibration for Test-Time Prompt Tuning
by: Fillioux, Leo, et al.
Published: (2026)
by: Fillioux, Leo, et al.
Published: (2026)
REFLECT: Rectified Flows for Efficient Brain Anomaly Correction Transport
by: Beizaee, Farzad, et al.
Published: (2025)
by: Beizaee, Farzad, et al.
Published: (2025)
MedCutMix: A Data-Centric Approach to Improve Radiology Vision-Language Pre-training with Disease Awareness
by: Wang, Sinuo, et al.
Published: (2025)
by: Wang, Sinuo, et al.
Published: (2025)
Can LLM-Generated Text Empower Surgical Vision-Language Pre-training?
by: Che, Chengan, et al.
Published: (2026)
by: Che, Chengan, et al.
Published: (2026)
TAP-VL: Text Layout-Aware Pre-training for Enriched Vision-Language Models
by: Fhima, Jonathan, et al.
Published: (2024)
by: Fhima, Jonathan, et al.
Published: (2024)
Multilingual Vision-Language Pre-training for the Remote Sensing Domain
by: Silva, João Daniel, et al.
Published: (2024)
by: Silva, João Daniel, et al.
Published: (2024)
Information Maximization for Long-Tailed Semi-Supervised Domain Generalization
by: Fillioux, Leo, et al.
Published: (2026)
by: Fillioux, Leo, et al.
Published: (2026)
Deep Active Learning: A Reality Check
by: Gashi, Edrina, et al.
Published: (2024)
by: Gashi, Edrina, et al.
Published: (2024)
Similar Items
-
Robust Calibration of Large Vision-Language Adapters
by: Murugesan, Balamurali, et al.
Published: (2024) -
A Closer Look at the Few-Shot Adaptation of Large Vision-Language Models
by: Silva-Rodríguez, Julio, et al.
Published: (2023) -
Towards Foundation Models and Few-Shot Parameter-Efficient Fine-Tuning for Volumetric Organ Segmentation
by: Silva-Rodríguez, Julio, et al.
Published: (2023) -
Trustworthy Few-Shot Transfer of Medical VLMs through Split Conformal Prediction
by: Silva-Rodríguez, Julio, et al.
Published: (2025) -
Conformal Prediction for Zero-Shot Models
by: Silva-Rodríguez, Julio, et al.
Published: (2025)