ViLU: Learning Vision-Language Uncertainties for Failure Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Lafon, Marc, Karmim, Yannis, Silva-Rodríguez, Julio, Couairon, Paul, Rambour, Clément, Fournier-Sniehotta, Raphaël, Ayed, Ismail Ben, Dolz, Jose, Thome, Nicolas |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Supra-Laplacian Encoding for Transformer on Dynamic Graphs
by: Karmim, Yannis, et al.
Published: (2024)
by: Karmim, Yannis, et al.
Published: (2024)
ITEM: Improving Training and Evaluation of Message-Passing based GNNs for top-k recommendation
by: Karmim, Yannis, et al.
Published: (2024)
by: Karmim, Yannis, et al.
Published: (2024)
Energy Correction Model in the Feature Space for Out-of-Distribution Detection
by: Lafon, Marc, et al.
Published: (2024)
by: Lafon, Marc, et al.
Published: (2024)
GalLoP: Learning Global and Local Prompts for Vision-Language Models
by: Lafon, Marc, et al.
Published: (2024)
by: Lafon, Marc, et al.
Published: (2024)
VidEdit: Zero-Shot and Spatially Aware Text-Driven Video Editing
by: Couairon, Paul, et al.
Published: (2023)
by: Couairon, Paul, et al.
Published: (2023)
CLIPTTA: Robust Contrastive Vision-Language Test-Time Adaptation
by: Lafon, Marc, et al.
Published: (2025)
by: Lafon, Marc, et al.
Published: (2025)
Temporal receptive field in dynamic graph learning: A comprehensive analysis
by: Karmim, Yannis, et al.
Published: (2024)
by: Karmim, Yannis, et al.
Published: (2024)
Conformal Prediction for Zero-Shot Models
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
A Reality Check of Vision-Language Pre-training in Radiology: Have We Progressed Using Text?
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
Trustworthy Few-Shot Transfer of Medical VLMs through Split Conformal Prediction
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
Robust Calibration of Large Vision-Language Adapters
by: Murugesan, Balamurali, et al.
Published: (2024)
by: Murugesan, Balamurali, et al.
Published: (2024)
A Closer Look at the Few-Shot Adaptation of Large Vision-Language Models
by: Silva-Rodríguez, Julio, et al.
Published: (2023)
by: Silva-Rodríguez, Julio, et al.
Published: (2023)
Towards Foundation Models and Few-Shot Parameter-Efficient Fine-Tuning for Volumetric Organ Segmentation
by: Silva-Rodríguez, Julio, et al.
Published: (2023)
by: Silva-Rodríguez, Julio, et al.
Published: (2023)
Semantic Anchor Transport: Robust Test-Time Adaptation for Vision-Language Models
by: Mishra, Shambhavi, et al.
Published: (2024)
by: Mishra, Shambhavi, et al.
Published: (2024)
ORION: ORthonormal Text Encoding for Universal VLM AdaptatION
by: Chakraborty, Omprakash, et al.
Published: (2026)
by: Chakraborty, Omprakash, et al.
Published: (2026)
Pay Attention to Your Neighbours: Training-Free Open-Vocabulary Semantic Segmentation
by: Hajimiri, Sina, et al.
Published: (2024)
by: Hajimiri, Sina, et al.
Published: (2024)
Class and Region-Adaptive Constraints for Network Calibration
by: Murugesan, Balamurali, et al.
Published: (2024)
by: Murugesan, Balamurali, et al.
Published: (2024)
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms
by: Asri, Zakariae El, et al.
Published: (2025)
by: Asri, Zakariae El, et al.
Published: (2025)
D-MODD: A Diffusion Model of Opinion Dynamics Derived from Online Data
by: Achitouv, Ixandra, et al.
Published: (2026)
by: Achitouv, Ixandra, et al.
Published: (2026)
Full Conformal Adaptation of Medical Vision-Language Models
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
Few-shot Adaptation of Medical Vision-Language Models
by: Shakeri, Fereshteh, et al.
Published: (2024)
by: Shakeri, Fereshteh, et al.
Published: (2024)
Locality-Attending Vision Transformer
by: Hajimiri, Sina, et al.
Published: (2026)
by: Hajimiri, Sina, et al.
Published: (2026)
Few-Shot, Now for Real: Medical VLMs Adaptation without Balanced Sets or Validation
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
A Foundation Language-Image Model of the Retina (FLAIR): Encoding Expert Knowledge in Text Supervision
by: Silva-Rodríguez, Julio, et al.
Published: (2023)
by: Silva-Rodríguez, Julio, et al.
Published: (2023)
Regularized Low-Rank Adaptation for Few-Shot Organ Segmentation
by: Baklouti, Ghassen, et al.
Published: (2025)
by: Baklouti, Ghassen, et al.
Published: (2025)
Low-Rank Few-Shot Adaptation of Vision-Language Models
by: Zanella, Maxime, et al.
Published: (2024)
by: Zanella, Maxime, et al.
Published: (2024)
Realistic Test-Time Adaptation of Vision-Language Models
by: Zanella, Maxime, et al.
Published: (2025)
by: Zanella, Maxime, et al.
Published: (2025)
Vocabulary-free few-shot learning for Vision-Language Models
by: Zanella, Maxime, et al.
Published: (2025)
by: Zanella, Maxime, et al.
Published: (2025)
DiffCut: Catalyzing Zero-Shot Semantic Segmentation with Diffusion Features and Recursive Normalized Cut
by: Couairon, Paul, et al.
Published: (2024)
by: Couairon, Paul, et al.
Published: (2024)
DAFTED: Decoupled Asymmetric Fusion of Tabular and Echocardiographic Data for Cardiac Hypertension Diagnosis
by: Stym-Popper, Jérémie, et al.
Published: (2025)
by: Stym-Popper, Jérémie, et al.
Published: (2025)
Prompting classes: Exploring the Power of Prompt Class Learning in Weakly Supervised Semantic Segmentation
by: Murugesan, Balamurali, et al.
Published: (2023)
by: Murugesan, Balamurali, et al.
Published: (2023)
Do not trust what you trust: Miscalibration in Semi-supervised Learning
by: Mishra, Shambhavi, et al.
Published: (2024)
by: Mishra, Shambhavi, et al.
Published: (2024)
Calibrating Segmentation Networks with Margin-based Label Smoothing
by: Murugesan, Balamurali, et al.
Published: (2022)
by: Murugesan, Balamurali, et al.
Published: (2022)
Boosting Vision-Language Models with Transduction
by: Zanella, Maxime, et al.
Published: (2024)
by: Zanella, Maxime, et al.
Published: (2024)
NAF: Zero-Shot Feature Upsampling via Neighborhood Attention Filtering
by: Chambon, Loick, et al.
Published: (2025)
by: Chambon, Loick, et al.
Published: (2025)
Are foundation models for computer vision good conformal predictors?
by: Fillioux, Leo, et al.
Published: (2024)
by: Fillioux, Leo, et al.
Published: (2024)
Class Adaptive Conformal Training
by: Marani, Badr-Eddine, et al.
Published: (2026)
by: Marani, Badr-Eddine, et al.
Published: (2026)
Boosting Vision-Language Models for Histopathology Classification: Predict all at once
by: Zanella, Maxime, et al.
Published: (2024)
by: Zanella, Maxime, et al.
Published: (2024)
JAFAR: Jack up Any Feature at Any Resolution
by: Couairon, Paul, et al.
Published: (2025)
by: Couairon, Paul, et al.
Published: (2025)
Asmptotic of the eigenvalues of Toeplitz matrices with even symbol
by: Rambour, Philippe
Published: (2021)
by: Rambour, Philippe
Published: (2021)
Similar Items
-
Supra-Laplacian Encoding for Transformer on Dynamic Graphs
by: Karmim, Yannis, et al.
Published: (2024) -
ITEM: Improving Training and Evaluation of Message-Passing based GNNs for top-k recommendation
by: Karmim, Yannis, et al.
Published: (2024) -
Energy Correction Model in the Feature Space for Out-of-Distribution Detection
by: Lafon, Marc, et al.
Published: (2024) -
GalLoP: Learning Global and Local Prompts for Vision-Language Models
by: Lafon, Marc, et al.
Published: (2024) -
VidEdit: Zero-Shot and Spatially Aware Text-Driven Video Editing
by: Couairon, Paul, et al.
Published: (2023)