Label Propagation for Zero-shot Classification with Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Stojnić, Vladan, Kalantidis, Yannis, Tolias, Giorgos |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LPOSS: Label Propagation Over Patches and Pixels for Open-vocabulary Semantic Segmentation
by: Stojnić, Vladan, et al.
Published: (2025)
by: Stojnić, Vladan, et al.
Published: (2025)
ELViS: Efficient Visual Similarity from Local Descriptors that Generalizes Across Domains
by: Suma, Pavel, et al.
Published: (2026)
by: Suma, Pavel, et al.
Published: (2026)
Category-level Text-to-Image Retrieval Improved: Bridging the Domain Gap with Diffusion Models and Vision Encoders
by: Khan, Faizan Farooq, et al.
Published: (2025)
by: Khan, Faizan Farooq, et al.
Published: (2025)
Retrieve and Segment: Are a Few Examples Enough to Bridge the Supervision Gap in Open-Vocabulary Segmentation?
by: Aravanis, Tilemachos, et al.
Published: (2026)
by: Aravanis, Tilemachos, et al.
Published: (2026)
Processing and acquisition traces in visual encoders: What does CLIP know about your camera?
by: Ramos, Ryan, et al.
Published: (2025)
by: Ramos, Ryan, et al.
Published: (2025)
Zero-shot image privacy classification with Vision-Language Models
by: Baia, Alina Elena, et al.
Published: (2025)
by: Baia, Alina Elena, et al.
Published: (2025)
Zero-shot Classification using Hyperdimensional Computing
by: Ruffino, Samuele, et al.
Published: (2024)
by: Ruffino, Samuele, et al.
Published: (2024)
DUNE: Distilling a Universal Encoder from Heterogeneous 2D and 3D Teachers
by: Sariyildiz, Mert Bulent, et al.
Published: (2025)
by: Sariyildiz, Mert Bulent, et al.
Published: (2025)
Text2Model: Text-based Model Induction for Zero-shot Image Classification
by: Amosy, Ohad, et al.
Published: (2022)
by: Amosy, Ohad, et al.
Published: (2022)
Hierarchically Robust Zero-shot Vision-language Models
by: Dong, Junhao, et al.
Published: (2026)
by: Dong, Junhao, et al.
Published: (2026)
ILIAS: Instance-Level Image retrieval At Scale
by: Kordopatis-Zilos, Giorgos, et al.
Published: (2025)
by: Kordopatis-Zilos, Giorgos, et al.
Published: (2025)
Explaining Vision GNNs: A Semantic and Visual Analysis of Graph-based Image Classification
by: Chaidos, Nikolaos, et al.
Published: (2025)
by: Chaidos, Nikolaos, et al.
Published: (2025)
Bridging Mini-Batch and Asymptotic Analysis in Contrastive Learning: From InfoNCE to Kernel-Based Losses
by: Koromilas, Panagiotis, et al.
Published: (2024)
by: Koromilas, Panagiotis, et al.
Published: (2024)
Vision-Language Models are Strong Noisy Label Detectors
by: Wei, Tong, et al.
Published: (2024)
by: Wei, Tong, et al.
Published: (2024)
UNIC: Universal Classification Models via Multi-teacher Distillation
by: Sariyildiz, Mert Bulent, et al.
Published: (2024)
by: Sariyildiz, Mert Bulent, et al.
Published: (2024)
NLPrompt: Noise-Label Prompt Learning for Vision-Language Models
by: Pan, Bikang, et al.
Published: (2024)
by: Pan, Bikang, et al.
Published: (2024)
MM-Zero: Self-Evolving Multi-Model Vision Language Models From Zero Data
by: Li, Zongxia, et al.
Published: (2026)
by: Li, Zongxia, et al.
Published: (2026)
Efficient and Versatile Robust Fine-Tuning of Zero-shot Models
by: Kim, Sungyeon, et al.
Published: (2024)
by: Kim, Sungyeon, et al.
Published: (2024)
Confidence-calibrated covariate shift correction for few-shot classification in Vision-Language Models
by: Khan, Behraj, et al.
Published: (2025)
by: Khan, Behraj, et al.
Published: (2025)
Negative Label Guided OOD Detection with Pretrained Vision-Language Models
by: Jiang, Xue, et al.
Published: (2024)
by: Jiang, Xue, et al.
Published: (2024)
World Action Models are Zero-shot Policies
by: Ye, Seonghyeon, et al.
Published: (2026)
by: Ye, Seonghyeon, et al.
Published: (2026)
A Principled Framework for Multi-View Contrastive Learning
by: Koromilas, Panagiotis, et al.
Published: (2025)
by: Koromilas, Panagiotis, et al.
Published: (2025)
Zero-shot Concept Bottleneck Models
by: Yamaguchi, Shin'ya, et al.
Published: (2025)
by: Yamaguchi, Shin'ya, et al.
Published: (2025)
Unleashing the Power of Vision-Language Models for Long-Tailed Multi-Label Visual Recognition
by: Tang, Wei, et al.
Published: (2025)
by: Tang, Wei, et al.
Published: (2025)
Sharpness-Aware Data Generation for Zero-shot Quantization
by: Hoang-Anh, Dung, et al.
Published: (2025)
by: Hoang-Anh, Dung, et al.
Published: (2025)
Diorama: Unleashing Zero-shot Single-view 3D Indoor Scene Modeling
by: Wu, Qirui, et al.
Published: (2024)
by: Wu, Qirui, et al.
Published: (2024)
Active Zero: Self-Evolving Vision-Language Models through Active Environment Exploration
by: He, Jinghan, et al.
Published: (2026)
by: He, Jinghan, et al.
Published: (2026)
Unlabeled Data Improves Fine-Grained Image Zero-shot Classification with Multimodal LLMs
by: Hong, Yunqi, et al.
Published: (2025)
by: Hong, Yunqi, et al.
Published: (2025)
No Labels Needed: Zero-Shot Image Classification with Collaborative Self-Learning
by: Todescato, Matheus Vinícius, et al.
Published: (2025)
by: Todescato, Matheus Vinícius, et al.
Published: (2025)
Flatness Improves Backbone Generalisation in Few-shot Classification
by: Li, Rui, et al.
Published: (2024)
by: Li, Rui, et al.
Published: (2024)
Zero-shot domain adaptation based on dual-level mix and contrast
by: Zhe, Yu, et al.
Published: (2024)
by: Zhe, Yu, et al.
Published: (2024)
Enhancing Zero-Shot Image Recognition in Vision-Language Models through Human-like Concept Guidance
by: Liu, Hui, et al.
Published: (2025)
by: Liu, Hui, et al.
Published: (2025)
Benchmarking Counterfactual Image Generation
by: Melistas, Thomas, et al.
Published: (2024)
by: Melistas, Thomas, et al.
Published: (2024)
Prompting without Panic: Attribute-aware, Zero-shot, Test-Time Calibration
by: Hebbalaguppe, Ramya, et al.
Published: (2025)
by: Hebbalaguppe, Ramya, et al.
Published: (2025)
MGPATH: Vision-Language Model with Multi-Granular Prompt Learning for Few-Shot WSI Classification
by: Nguyen, Anh-Tien, et al.
Published: (2025)
by: Nguyen, Anh-Tien, et al.
Published: (2025)
Weakly Supervised Point Cloud Segmentation via Conservative Propagation of Scene-level Labels
by: Xia, Shaobo, et al.
Published: (2023)
by: Xia, Shaobo, et al.
Published: (2023)
Cooperative Pseudo Labeling for Unsupervised Federated Classification
by: Guo, Kuangpu, et al.
Published: (2025)
by: Guo, Kuangpu, et al.
Published: (2025)
Online Zero-Shot Classification with CLIP
by: Qian, Qi, et al.
Published: (2024)
by: Qian, Qi, et al.
Published: (2024)
Adapting Vision-Language Models Without Labels: A Comprehensive Survey
by: Dong, Hao, et al.
Published: (2025)
by: Dong, Hao, et al.
Published: (2025)
Multi-Label Plant Species Classification with Self-Supervised Vision Transformers
by: Gustineli, Murilo, et al.
Published: (2024)
by: Gustineli, Murilo, et al.
Published: (2024)
Similar Items
-
LPOSS: Label Propagation Over Patches and Pixels for Open-vocabulary Semantic Segmentation
by: Stojnić, Vladan, et al.
Published: (2025) -
ELViS: Efficient Visual Similarity from Local Descriptors that Generalizes Across Domains
by: Suma, Pavel, et al.
Published: (2026) -
Category-level Text-to-Image Retrieval Improved: Bridging the Domain Gap with Diffusion Models and Vision Encoders
by: Khan, Faizan Farooq, et al.
Published: (2025) -
Retrieve and Segment: Are a Few Examples Enough to Bridge the Supervision Gap in Open-Vocabulary Segmentation?
by: Aravanis, Tilemachos, et al.
Published: (2026) -
Processing and acquisition traces in visual encoders: What does CLIP know about your camera?
by: Ramos, Ryan, et al.
Published: (2025)