Vocabulary-free few-shot learning for Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zanella, Maxime, Fuchs, Clément, Ayed, Ismail Ben, De Vleeschouwer, Christophe |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Realistic Test-Time Adaptation of Vision-Language Models
von: Zanella, Maxime, et al.
Veröffentlicht: (2025)
von: Zanella, Maxime, et al.
Veröffentlicht: (2025)
Online Gaussian Test-Time Adaptation of Vision-Language Models
von: Fuchs, Clément, et al.
Veröffentlicht: (2025)
von: Fuchs, Clément, et al.
Veröffentlicht: (2025)
On the test-time zero-shot generalization of vision-language models: Do we really need prompt learning?
von: Zanella, Maxime, et al.
Veröffentlicht: (2024)
von: Zanella, Maxime, et al.
Veröffentlicht: (2024)
Low-Rank Few-Shot Adaptation of Vision-Language Models
von: Zanella, Maxime, et al.
Veröffentlicht: (2024)
von: Zanella, Maxime, et al.
Veröffentlicht: (2024)
Boosting Vision-Language Models with Transduction
von: Zanella, Maxime, et al.
Veröffentlicht: (2024)
von: Zanella, Maxime, et al.
Veröffentlicht: (2024)
Enhancing Remote Sensing Vision-Language Models for Zero-Shot Scene Classification
von: Khoury, Karim El, et al.
Veröffentlicht: (2024)
von: Khoury, Karim El, et al.
Veröffentlicht: (2024)
Language-Aware Information Maximization for Transductive Few-Shot CLIP
von: Baklouti, Ghassen, et al.
Veröffentlicht: (2025)
von: Baklouti, Ghassen, et al.
Veröffentlicht: (2025)
Boosting Vision-Language Models for Histopathology Classification: Predict all at once
von: Zanella, Maxime, et al.
Veröffentlicht: (2024)
von: Zanella, Maxime, et al.
Veröffentlicht: (2024)
Few-Shot Adaptation Benchmark for Remote Sensing Vision-Language Models
von: Khoury, Karim El, et al.
Veröffentlicht: (2025)
von: Khoury, Karim El, et al.
Veröffentlicht: (2025)
Few-shot Adaptation of Medical Vision-Language Models
von: Shakeri, Fereshteh, et al.
Veröffentlicht: (2024)
von: Shakeri, Fereshteh, et al.
Veröffentlicht: (2024)
Pay Attention to Your Neighbours: Training-Free Open-Vocabulary Semantic Segmentation
von: Hajimiri, Sina, et al.
Veröffentlicht: (2024)
von: Hajimiri, Sina, et al.
Veröffentlicht: (2024)
Test-Time Adaptation of Vision-Language Models for Open-Vocabulary Semantic Segmentation
von: Noori, Mehrdad, et al.
Veröffentlicht: (2025)
von: Noori, Mehrdad, et al.
Veröffentlicht: (2025)
A Closer Look at the Few-Shot Adaptation of Large Vision-Language Models
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2023)
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2023)
A Reality Check of Vision-Language Pre-training in Radiology: Have We Progressed Using Text?
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2025)
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2025)
Robust Calibration of Large Vision-Language Adapters
von: Murugesan, Balamurali, et al.
Veröffentlicht: (2024)
von: Murugesan, Balamurali, et al.
Veröffentlicht: (2024)
Semantic Anchor Transport: Robust Test-Time Adaptation for Vision-Language Models
von: Mishra, Shambhavi, et al.
Veröffentlicht: (2024)
von: Mishra, Shambhavi, et al.
Veröffentlicht: (2024)
Full Conformal Adaptation of Medical Vision-Language Models
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2025)
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2025)
ViLU: Learning Vision-Language Uncertainties for Failure Prediction
von: Lafon, Marc, et al.
Veröffentlicht: (2025)
von: Lafon, Marc, et al.
Veröffentlicht: (2025)
Conformal Prediction for Zero-Shot Models
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2025)
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2025)
Conditional Random Fields for Interactive Refinement of Histopathological Predictions
von: Godelaine, Tiffanie, et al.
Veröffentlicht: (2026)
von: Godelaine, Tiffanie, et al.
Veröffentlicht: (2026)
Towards Foundation Models and Few-Shot Parameter-Efficient Fine-Tuning for Volumetric Organ Segmentation
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2023)
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2023)
ORION: ORthonormal Text Encoding for Universal VLM AdaptatION
von: Chakraborty, Omprakash, et al.
Veröffentlicht: (2026)
von: Chakraborty, Omprakash, et al.
Veröffentlicht: (2026)
Locality-Attending Vision Transformer
von: Hajimiri, Sina, et al.
Veröffentlicht: (2026)
von: Hajimiri, Sina, et al.
Veröffentlicht: (2026)
Conformal Predictions for Human Action Recognition with Vision-Language Models
von: Tim, Bary, et al.
Veröffentlicht: (2025)
von: Tim, Bary, et al.
Veröffentlicht: (2025)
Spatial frequency information fusion network for few-shot learning
von: Zhao, Wenqing, et al.
Veröffentlicht: (2025)
von: Zhao, Wenqing, et al.
Veröffentlicht: (2025)
A Foundation Language-Image Model of the Retina (FLAIR): Encoding Expert Knowledge in Text Supervision
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2023)
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2023)
Confidence-calibrated covariate shift correction for few-shot classification in Vision-Language Models
von: Khan, Behraj, et al.
Veröffentlicht: (2025)
von: Khan, Behraj, et al.
Veröffentlicht: (2025)
Trustworthy Few-Shot Transfer of Medical VLMs through Split Conformal Prediction
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2025)
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2025)
LPN: Language-guided Prototypical Network for few-shot classification
von: Cheng, Kaihui, et al.
Veröffentlicht: (2023)
von: Cheng, Kaihui, et al.
Veröffentlicht: (2023)
UNEM: UNrolled Generalized EM for Transductive Few-Shot Learning
von: Zhou, Long, et al.
Veröffentlicht: (2024)
von: Zhou, Long, et al.
Veröffentlicht: (2024)
Transductive Zero-Shot and Few-Shot CLIP
von: Martin, Ségolène, et al.
Veröffentlicht: (2024)
von: Martin, Ségolène, et al.
Veröffentlicht: (2024)
On the Geometric Accuracy of Implicit and Primitive-based Representations Derived from View Rendering Constraints
von: De Smijter, Elias, et al.
Veröffentlicht: (2025)
von: De Smijter, Elias, et al.
Veröffentlicht: (2025)
Histopath-C: Towards Realistic Domain Shifts for Histopathology Vision-Language Adaptation
von: Noori, Mehrdad, et al.
Veröffentlicht: (2026)
von: Noori, Mehrdad, et al.
Veröffentlicht: (2026)
NeRF-based Visualization of 3D Cues Supporting Data-Driven Spacecraft Pose Estimation
von: Legrand, Antoine, et al.
Veröffentlicht: (2025)
von: Legrand, Antoine, et al.
Veröffentlicht: (2025)
NeRF-based Spacecraft Reconstruction from Monocular Imagery Under Illumination Variability and Pose Uncertainty
von: Legrand, Antoine, et al.
Veröffentlicht: (2026)
von: Legrand, Antoine, et al.
Veröffentlicht: (2026)
Body Part-Based Representation Learning for Occluded Person Re-Identification
von: Somers, Vladimir, et al.
Veröffentlicht: (2022)
von: Somers, Vladimir, et al.
Veröffentlicht: (2022)
CAD-Free Learning of Spacecraft Pose Estimators via NeRF-Based Augmentations
von: Legrand, Antoine, et al.
Veröffentlicht: (2026)
von: Legrand, Antoine, et al.
Veröffentlicht: (2026)
Keypoint Promptable Re-Identification
von: Somers, Vladimir, et al.
Veröffentlicht: (2024)
von: Somers, Vladimir, et al.
Veröffentlicht: (2024)
Vocabulary-free Fine-grained Visual Recognition via Enriched Contextually Grounded Vision-Language Model
von: Demidov, Dmitry, et al.
Veröffentlicht: (2025)
von: Demidov, Dmitry, et al.
Veröffentlicht: (2025)
Class and Region-Adaptive Constraints for Network Calibration
von: Murugesan, Balamurali, et al.
Veröffentlicht: (2024)
von: Murugesan, Balamurali, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Realistic Test-Time Adaptation of Vision-Language Models
von: Zanella, Maxime, et al.
Veröffentlicht: (2025) -
Online Gaussian Test-Time Adaptation of Vision-Language Models
von: Fuchs, Clément, et al.
Veröffentlicht: (2025) -
On the test-time zero-shot generalization of vision-language models: Do we really need prompt learning?
von: Zanella, Maxime, et al.
Veröffentlicht: (2024) -
Low-Rank Few-Shot Adaptation of Vision-Language Models
von: Zanella, Maxime, et al.
Veröffentlicht: (2024) -
Boosting Vision-Language Models with Transduction
von: Zanella, Maxime, et al.
Veröffentlicht: (2024)