Large Language Models are Good Prompt Learners for Low-Shot Image Classification
Fuente:
arXiv
Salvato in:
| Autori principali: | Zheng, Zhaoheng, Wei, Jingmin, Hu, Xuefeng, Zhu, Haidong, Nevatia, Ram |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GaitSTR: Gait Recognition with Sequential Two-stream Refinement
di: Zheng, Wanrong, et al.
Pubblicazione: (2024)
di: Zheng, Wanrong, et al.
Pubblicazione: (2024)
BaFTA: Backprop-Free Test-Time Adaptation For Zero-Shot Vision-Language Models
di: Hu, Xuefeng, et al.
Pubblicazione: (2024)
di: Hu, Xuefeng, et al.
Pubblicazione: (2024)
DiVE-k: Differential Visual Reasoning for Fine-grained Image Recognition
di: Kumar, Raja, et al.
Pubblicazione: (2025)
di: Kumar, Raja, et al.
Pubblicazione: (2025)
CaesarNeRF: Calibrated Semantic Representation for Few-shot Generalizable Neural Rendering
di: Zhu, Haidong, et al.
Pubblicazione: (2023)
di: Zhu, Haidong, et al.
Pubblicazione: (2023)
Zero-Shot Prompting and Few-Shot Fine-Tuning: Revisiting Document Image Classification Using Large Language Models
di: Scius-Bertrand, Anna, et al.
Pubblicazione: (2024)
di: Scius-Bertrand, Anna, et al.
Pubblicazione: (2024)
Making Large Vision Language Models to be Good Few-shot Learners
di: Liu, Fan, et al.
Pubblicazione: (2024)
di: Liu, Fan, et al.
Pubblicazione: (2024)
Open-Source Image Editing Models Are Zero-Shot Vision Learners
di: Liu, Wei, et al.
Pubblicazione: (2026)
di: Liu, Wei, et al.
Pubblicazione: (2026)
Are Image-to-Video Models Good Zero-Shot Image Editors?
di: Zhang, Zechuan, et al.
Pubblicazione: (2025)
di: Zhang, Zechuan, et al.
Pubblicazione: (2025)
Zero-Shot Fine-Grained Image Classification Using Large Vision-Language Models
di: Atabuzzaman, Md., et al.
Pubblicazione: (2025)
di: Atabuzzaman, Md., et al.
Pubblicazione: (2025)
Noise is an Efficient Learner for Zero-Shot Vision-Language Models
di: Imam, Raza, et al.
Pubblicazione: (2025)
di: Imam, Raza, et al.
Pubblicazione: (2025)
Are Multimodal Large Language Models Good Annotators for Image Tagging?
di: Xie, Ming-Kun, et al.
Pubblicazione: (2026)
di: Xie, Ming-Kun, et al.
Pubblicazione: (2026)
Exploring Low-Resource Medical Image Classification with Weakly Supervised Prompt Learning
di: Zheng, Fudan, et al.
Pubblicazione: (2024)
di: Zheng, Fudan, et al.
Pubblicazione: (2024)
Are Video Models Emerging as Zero-Shot Learners and Reasoners in Medical Imaging?
di: Lai, Yuxiang, et al.
Pubblicazione: (2025)
di: Lai, Yuxiang, et al.
Pubblicazione: (2025)
What Do You See? Enhancing Zero-Shot Image Classification with Multimodal Large Language Models
di: Abdelhamed, Abdelrahman, et al.
Pubblicazione: (2024)
di: Abdelhamed, Abdelrahman, et al.
Pubblicazione: (2024)
Constrained Prompt Enhancement for Improving Zero-Shot Generalization of Vision-Language Models
di: Yin, Xiaojie, et al.
Pubblicazione: (2025)
di: Yin, Xiaojie, et al.
Pubblicazione: (2025)
Foundation Models as Class-Incremental Learners for Dermatological Image Classification
di: Elkhayat, Mohamed, et al.
Pubblicazione: (2025)
di: Elkhayat, Mohamed, et al.
Pubblicazione: (2025)
Prompts to Summaries: Zero-Shot Language-Guided Video Summarization with Large Language and Video Models
di: Barbara, Mario, et al.
Pubblicazione: (2025)
di: Barbara, Mario, et al.
Pubblicazione: (2025)
Complementary Subspace Low-Rank Adaptation of Vision-Language Models for Few-Shot Classification
di: Wang, Zhongqi, et al.
Pubblicazione: (2025)
di: Wang, Zhongqi, et al.
Pubblicazione: (2025)
TransMed: Large Language Models Enhance Vision Transformer for Biomedical Image Classification
di: Zheng, Kaipeng, et al.
Pubblicazione: (2023)
di: Zheng, Kaipeng, et al.
Pubblicazione: (2023)
User-Aware Prefix-Tuning is a Good Learner for Personalized Image Captioning
di: Wang, Xuan, et al.
Pubblicazione: (2023)
di: Wang, Xuan, et al.
Pubblicazione: (2023)
Prompt-Guided Adaptive Model Transformation for Whole Slide Image Classification
di: Lin, Yi, et al.
Pubblicazione: (2024)
di: Lin, Yi, et al.
Pubblicazione: (2024)
Automatic Pruning Discovery for Large Language Models
di: Kang, Haidong, et al.
Pubblicazione: (2025)
di: Kang, Haidong, et al.
Pubblicazione: (2025)
MLP Can Be A Good Transformer Learner
di: Lin, Sihao, et al.
Pubblicazione: (2024)
di: Lin, Sihao, et al.
Pubblicazione: (2024)
Few-Shot Learner Generalizes Across AI-Generated Image Detection
di: Wu, Shiyu, et al.
Pubblicazione: (2025)
di: Wu, Shiyu, et al.
Pubblicazione: (2025)
Inter- and Intra-image Refinement for Few Shot Segmentation
di: Fu, Ourui, et al.
Pubblicazione: (2025)
di: Fu, Ourui, et al.
Pubblicazione: (2025)
Local-Prompt: Extensible Local Prompts for Few-Shot Out-of-Distribution Detection
di: Zeng, Fanhu, et al.
Pubblicazione: (2024)
di: Zeng, Fanhu, et al.
Pubblicazione: (2024)
Making Better Mistakes in CLIP-Based Zero-Shot Classification with Hierarchy-Aware Language Prompts
di: Liang, Tong, et al.
Pubblicazione: (2025)
di: Liang, Tong, et al.
Pubblicazione: (2025)
Task-Related Token Compression in Multimodal Large Language Models from an Explainability Perspective
di: Lei, Lei, et al.
Pubblicazione: (2025)
di: Lei, Lei, et al.
Pubblicazione: (2025)
CoS: Chain-of-Shot Prompting for Long Video Understanding
di: Hu, Jian, et al.
Pubblicazione: (2025)
di: Hu, Jian, et al.
Pubblicazione: (2025)
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners
di: Yang, Jingyi, et al.
Pubblicazione: (2025)
di: Yang, Jingyi, et al.
Pubblicazione: (2025)
AntifakePrompt: Prompt-Tuned Vision-Language Models are Fake Image Detectors
di: Chang, You-Ming, et al.
Pubblicazione: (2023)
di: Chang, You-Ming, et al.
Pubblicazione: (2023)
Efficient Prompt Tuning of Large Vision-Language Model for Fine-Grained Ship Classification
di: Lan, Long, et al.
Pubblicazione: (2024)
di: Lan, Long, et al.
Pubblicazione: (2024)
Pre-trained Vision and Language Transformers Are Few-Shot Incremental Learners
di: Park, Keon-Hee, et al.
Pubblicazione: (2024)
di: Park, Keon-Hee, et al.
Pubblicazione: (2024)
Few-Shot Image Classification and Segmentation as Visual Question Answering Using Vision-Language Models
di: Meng, Tian, et al.
Pubblicazione: (2024)
di: Meng, Tian, et al.
Pubblicazione: (2024)
Learning to Obstruct Few-Shot Image Classification over Restricted Classes
di: Zheng, Amber Yijia, et al.
Pubblicazione: (2024)
di: Zheng, Amber Yijia, et al.
Pubblicazione: (2024)
Multimodal Large Language Models for Medical Report Generation via Customized Prompt Tuning
di: Li, Chunlei, et al.
Pubblicazione: (2025)
di: Li, Chunlei, et al.
Pubblicazione: (2025)
Boosting Few-Shot Detection with Large Language Models and Layout-to-Image Synthesis
di: Abdullah, Ahmed, et al.
Pubblicazione: (2024)
di: Abdullah, Ahmed, et al.
Pubblicazione: (2024)
Prompt-Induced Score Variance in Zero-Shot Binary Vision-Language Safety Classification
di: Weng, Charles, et al.
Pubblicazione: (2026)
di: Weng, Charles, et al.
Pubblicazione: (2026)
ST-LLM: Large Language Models Are Effective Temporal Learners
di: Liu, Ruyang, et al.
Pubblicazione: (2024)
di: Liu, Ruyang, et al.
Pubblicazione: (2024)
Hierarchical Mask-Enhanced Dual Reconstruction Network for Few-Shot Fine-Grained Image Classification
di: Luo, Ning, et al.
Pubblicazione: (2025)
di: Luo, Ning, et al.
Pubblicazione: (2025)
Documenti analoghi
-
GaitSTR: Gait Recognition with Sequential Two-stream Refinement
di: Zheng, Wanrong, et al.
Pubblicazione: (2024) -
BaFTA: Backprop-Free Test-Time Adaptation For Zero-Shot Vision-Language Models
di: Hu, Xuefeng, et al.
Pubblicazione: (2024) -
DiVE-k: Differential Visual Reasoning for Fine-grained Image Recognition
di: Kumar, Raja, et al.
Pubblicazione: (2025) -
CaesarNeRF: Calibrated Semantic Representation for Few-shot Generalizable Neural Rendering
di: Zhu, Haidong, et al.
Pubblicazione: (2023) -
Zero-Shot Prompting and Few-Shot Fine-Tuning: Revisiting Document Image Classification Using Large Language Models
di: Scius-Bertrand, Anna, et al.
Pubblicazione: (2024)