Zero-Shot Product Attribute Labeling with Vision-Language Models: A Three-Tier Evaluation Framework
Fuente:
arXiv
Saved in:
| Main Authors: | Shukla, Shubham, Sonalkar, Kunal |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can GPT-4o mini and Gemini 2.0 Flash Predict Fine-Grained Fashion Product Attributes? A Zero-Shot Analysis
by: Shukla, Shubham, et al.
Published: (2025)
by: Shukla, Shubham, et al.
Published: (2025)
PVLM: Parsing-Aware Vision Language Model with Dynamic Contrastive Learning for Zero-Shot Deepfake Attribution
by: Zhang, Yaning, et al.
Published: (2025)
by: Zhang, Yaning, et al.
Published: (2025)
WisWheat: A Three-Tiered Vision-Language Dataset for Wheat Management
by: Yuan, Bowen, et al.
Published: (2025)
by: Yuan, Bowen, et al.
Published: (2025)
Model Synthesis for Zero-Shot Model Attribution
by: Yang, Tianyun, et al.
Published: (2023)
by: Yang, Tianyun, et al.
Published: (2023)
SPARC: Score Prompting and Adaptive Fusion for Zero-Shot Multi-Label Recognition in Vision-Language Models
by: Miller, Kevin, et al.
Published: (2025)
by: Miller, Kevin, et al.
Published: (2025)
Efficient and Context-Aware Label Propagation for Zero-/Few-Shot Training-Free Adaptation of Vision-Language Model
by: Li, Yushu, et al.
Published: (2024)
by: Li, Yushu, et al.
Published: (2024)
Insert In Style: A Zero-Shot Generative Framework for Harmonious Cross-Domain Object Composition
by: Chittersu, Raghu Vamsi, et al.
Published: (2025)
by: Chittersu, Raghu Vamsi, et al.
Published: (2025)
Noise is an Efficient Learner for Zero-Shot Vision-Language Models
by: Imam, Raza, et al.
Published: (2025)
by: Imam, Raza, et al.
Published: (2025)
Enhancing Zero-Shot Vision Models by Label-Free Prompt Distribution Learning and Bias Correcting
by: Zhu, Xingyu, et al.
Published: (2024)
by: Zhu, Xingyu, et al.
Published: (2024)
SleepWalk: A Three-Tier Benchmark for Stress-Testing Instruction-Guided Vision-Language Navigation
by: Rawal, Niyati, et al.
Published: (2026)
by: Rawal, Niyati, et al.
Published: (2026)
Evaluating Attribute Comprehension in Large Vision-Language Models
by: Zhang, Haiwen, et al.
Published: (2024)
by: Zhang, Haiwen, et al.
Published: (2024)
Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model
by: Chen, Shiming, et al.
Published: (2025)
by: Chen, Shiming, et al.
Published: (2025)
Three-Step Nav: A Hierarchical Global-Local Planner for Zero-Shot Vision-and-Language Navigation
by: Zheng, Wanrong, et al.
Published: (2026)
by: Zheng, Wanrong, et al.
Published: (2026)
Evaluating Cascaded Methods of Vision-Language Models for Zero-Shot Detection and Association of Hardhats for Increased Construction Safety
by: Choi, Lucas, et al.
Published: (2024)
by: Choi, Lucas, et al.
Published: (2024)
Exploring Vision-Language Models for Open-Vocabulary Zero-Shot Action Segmentation
by: Unmesh, Asim, et al.
Published: (2026)
by: Unmesh, Asim, et al.
Published: (2026)
Constrained Prompt Enhancement for Improving Zero-Shot Generalization of Vision-Language Models
by: Yin, Xiaojie, et al.
Published: (2025)
by: Yin, Xiaojie, et al.
Published: (2025)
Zero-Shot Robustness of Vision Language Models Via Confidence-Aware Weighting
by: Naghavian, Nikoo, et al.
Published: (2025)
by: Naghavian, Nikoo, et al.
Published: (2025)
Exploring the Zero-Shot Capabilities of Vision-Language Models for Improving Gaze Following
by: Gupta, Anshul, et al.
Published: (2024)
by: Gupta, Anshul, et al.
Published: (2024)
Enhancing Remote Sensing Vision-Language Models for Zero-Shot Scene Classification
by: Khoury, Karim El, et al.
Published: (2024)
by: Khoury, Karim El, et al.
Published: (2024)
Training-Free Zero-Shot Temporal Action Detection with Vision-Language Models
by: Han, Chaolei, et al.
Published: (2025)
by: Han, Chaolei, et al.
Published: (2025)
Visual Zero-Shot E-Commerce Product Attribute Value Extraction
by: Gong, Jiaying, et al.
Published: (2025)
by: Gong, Jiaying, et al.
Published: (2025)
Towards Zero-Shot Annotation of the Built Environment with Vision-Language Models (Vision Paper)
by: Han, Bin, et al.
Published: (2024)
by: Han, Bin, et al.
Published: (2024)
ContextVLM: Zero-Shot and Few-Shot Context Understanding for Autonomous Driving using Vision Language Models
by: Sural, Shounak, et al.
Published: (2024)
by: Sural, Shounak, et al.
Published: (2024)
Exploring Vision-Language Models for Online Signature Verification: A Zero-Shot Capability Study
by: Robledo-Moreno, Marta, et al.
Published: (2026)
by: Robledo-Moreno, Marta, et al.
Published: (2026)
MAC: A Benchmark for Multiple Attributes Compositional Zero-Shot Learning
by: Xu, Shuo, et al.
Published: (2024)
by: Xu, Shuo, et al.
Published: (2024)
Investigating Zero-Shot Diagnostic Pathology in Vision-Language Models with Efficient Prompt Design
by: Sharma, Vasudev, et al.
Published: (2025)
by: Sharma, Vasudev, et al.
Published: (2025)
Zero-Shot Fine-Grained Image Classification Using Large Vision-Language Models
by: Atabuzzaman, Md., et al.
Published: (2025)
by: Atabuzzaman, Md., et al.
Published: (2025)
$S^3$: Synonymous Semantic Space for Improving Zero-Shot Generalization of Vision-Language Models
by: Yin, Xiaojie, et al.
Published: (2024)
by: Yin, Xiaojie, et al.
Published: (2024)
Fine-Grained Zero-Shot Learning with Attribute-Centric Representations
by: Chen, Zhi, et al.
Published: (2025)
by: Chen, Zhi, et al.
Published: (2025)
Super-class guided Transformer for Zero-Shot Attribute Classification
by: Kim, Sehyung, et al.
Published: (2025)
by: Kim, Sehyung, et al.
Published: (2025)
Label Propagation for Zero-shot Classification with Vision-Language Models
by: Stojnić, Vladan, et al.
Published: (2024)
by: Stojnić, Vladan, et al.
Published: (2024)
Evaluating Vision-Language Models for Zero-Shot Detection, Classification, and Association of Motorcycles, Passengers, and Helmets
by: Choi, Lucas, et al.
Published: (2024)
by: Choi, Lucas, et al.
Published: (2024)
TINA: Think, Interaction, and Action Framework for Zero-Shot Vision Language Navigation
by: Li, Dingbang, et al.
Published: (2024)
by: Li, Dingbang, et al.
Published: (2024)
Leveraging Vision-Language Embeddings for Zero-Shot Learning in Histopathology Images
by: Rahaman, Md Mamunur, et al.
Published: (2025)
by: Rahaman, Md Mamunur, et al.
Published: (2025)
Zero-Shot 3D Visual Grounding from Vision-Language Models
by: Li, Rong, et al.
Published: (2025)
by: Li, Rong, et al.
Published: (2025)
Benchmarking Foundation Models for Zero-Shot Biometric Tasks
by: Sony, Redwan, et al.
Published: (2025)
by: Sony, Redwan, et al.
Published: (2025)
BaFTA: Backprop-Free Test-Time Adaptation For Zero-Shot Vision-Language Models
by: Hu, Xuefeng, et al.
Published: (2024)
by: Hu, Xuefeng, et al.
Published: (2024)
Zero-Shot Scene Understanding for Automatic Target Recognition Using Large Vision-Language Models
by: Ranasinghe, Yasiru, et al.
Published: (2025)
by: Ranasinghe, Yasiru, et al.
Published: (2025)
High-Discriminative Attribute Feature Learning for Generalized Zero-Shot Learning
by: Lei, Yu, et al.
Published: (2024)
by: Lei, Yu, et al.
Published: (2024)
MADS: Multi-Attribute Document Supervision for Zero-Shot Image Classification
by: Qu, Xiangyan, et al.
Published: (2025)
by: Qu, Xiangyan, et al.
Published: (2025)
Similar Items
-
Can GPT-4o mini and Gemini 2.0 Flash Predict Fine-Grained Fashion Product Attributes? A Zero-Shot Analysis
by: Shukla, Shubham, et al.
Published: (2025) -
PVLM: Parsing-Aware Vision Language Model with Dynamic Contrastive Learning for Zero-Shot Deepfake Attribution
by: Zhang, Yaning, et al.
Published: (2025) -
WisWheat: A Three-Tiered Vision-Language Dataset for Wheat Management
by: Yuan, Bowen, et al.
Published: (2025) -
Model Synthesis for Zero-Shot Model Attribution
by: Yang, Tianyun, et al.
Published: (2023) -
SPARC: Score Prompting and Adaptive Fusion for Zero-Shot Multi-Label Recognition in Vision-Language Models
by: Miller, Kevin, et al.
Published: (2025)