Concept-Guided Prompt Learning for Generalization in Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Yi, Zhang, Ce, Yu, Ke, Tang, Yushun, He, Zhihai |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Conceptual Codebook Learning for Vision-Language Models
von: Zhang, Yi, et al.
Veröffentlicht: (2024)
von: Zhang, Yi, et al.
Veröffentlicht: (2024)
Domain-Conditioned Transformer for Fully Test-time Adaptation
von: Tang, Yushun, et al.
Veröffentlicht: (2024)
von: Tang, Yushun, et al.
Veröffentlicht: (2024)
Progressive Conditioned Scale-Shift Recalibration of Self-Attention for Online Test-time Adaptation
von: Tang, Yushun, et al.
Veröffentlicht: (2025)
von: Tang, Yushun, et al.
Veröffentlicht: (2025)
Learning Visual Conditioning Tokens to Correct Domain Shift for Fully Test-time Adaptation
von: Tang, Yushun, et al.
Veröffentlicht: (2024)
von: Tang, Yushun, et al.
Veröffentlicht: (2024)
Open-World Test-Time Adaptation with Hierarchical Feature Aggregation and Attention Affine
von: Liu, Ziqiong, et al.
Veröffentlicht: (2025)
von: Liu, Ziqiong, et al.
Veröffentlicht: (2025)
Window-based Channel Attention for Wavelet-enhanced Learned Image Compression
von: Xu, Heng, et al.
Veröffentlicht: (2024)
von: Xu, Heng, et al.
Veröffentlicht: (2024)
Dual-Path Adversarial Lifting for Domain Shift Correction in Online Test-time Adaptation
von: Tang, Yushun, et al.
Veröffentlicht: (2024)
von: Tang, Yushun, et al.
Veröffentlicht: (2024)
NODE-Adapter: Neural Ordinary Differential Equations for Better Vision-Language Reasoning
von: Zhang, Yi, et al.
Veröffentlicht: (2024)
von: Zhang, Yi, et al.
Veröffentlicht: (2024)
LatentEdit: Adaptive Latent Control for Consistent Semantic Editing
von: Liu, Siyi, et al.
Veröffentlicht: (2025)
von: Liu, Siyi, et al.
Veröffentlicht: (2025)
Training-Free Dual Hyperbolic Adapters for Better Cross-Modal Reasoning
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
ArGue: Attribute-Guided Prompt Tuning for Vision-Language Models
von: Tian, Xinyu, et al.
Veröffentlicht: (2023)
von: Tian, Xinyu, et al.
Veröffentlicht: (2023)
Mixture of Prompt Learning for Vision Language Models
von: Du, Yu, et al.
Veröffentlicht: (2024)
von: Du, Yu, et al.
Veröffentlicht: (2024)
Training-Free Test-Time Adaptation with Brownian Distance Covariance in Vision-Language Models
von: Zhang, Yi, et al.
Veröffentlicht: (2026)
von: Zhang, Yi, et al.
Veröffentlicht: (2026)
Optimization of Prompt Learning via Multi-Knowledge Representation for Vision-Language Models
von: Zhang, Enming, et al.
Veröffentlicht: (2024)
von: Zhang, Enming, et al.
Veröffentlicht: (2024)
Guiding Medical Vision-Language Models with Explicit Visual Prompts: Framework Design and Comprehensive Exploration of Prompt Variations
von: Zhu, Kangyu, et al.
Veröffentlicht: (2025)
von: Zhu, Kangyu, et al.
Veröffentlicht: (2025)
GeneralVLA: Generalizable Vision-Language-Action Models with Knowledge-Guided Trajectory Planning
von: Ma, Guoqing, et al.
Veröffentlicht: (2026)
von: Ma, Guoqing, et al.
Veröffentlicht: (2026)
MePT: Multi-Representation Guided Prompt Tuning for Vision-Language Model
von: Wang, Xinyang, et al.
Veröffentlicht: (2024)
von: Wang, Xinyang, et al.
Veröffentlicht: (2024)
Test-time Distribution Learning Adapter for Cross-modal Visual Reasoning
von: Zhang, Yi, et al.
Veröffentlicht: (2024)
von: Zhang, Yi, et al.
Veröffentlicht: (2024)
Generalizable Prompt Tuning for Vision-Language Models
von: Zhang, Qian
Veröffentlicht: (2024)
von: Zhang, Qian
Veröffentlicht: (2024)
Style-Pro: Style-Guided Prompt Learning for Generalizable Vision-Language Models
von: Talemi, Niloufar Alipour, et al.
Veröffentlicht: (2024)
von: Talemi, Niloufar Alipour, et al.
Veröffentlicht: (2024)
Light Up the Shadows: Enhance Long-Tailed Entity Grounding with Concept-Guided Vision-Language Models
von: Zhang, Yikai, et al.
Veröffentlicht: (2024)
von: Zhang, Yikai, et al.
Veröffentlicht: (2024)
Cascade Prompt Learning for Vision-Language Model Adaptation
von: Wu, Ge, et al.
Veröffentlicht: (2024)
von: Wu, Ge, et al.
Veröffentlicht: (2024)
Modular Prompt Learning Improves Vision-Language Models
von: Huang, Zhenhan, et al.
Veröffentlicht: (2025)
von: Huang, Zhenhan, et al.
Veröffentlicht: (2025)
RESTORE: Towards Feature Shift for Vision-Language Prompt Learning
von: Yang, Yuncheng, et al.
Veröffentlicht: (2024)
von: Yang, Yuncheng, et al.
Veröffentlicht: (2024)
Dual Prototype Evolving for Test-Time Generalization of Vision-Language Models
von: Zhang, Ce, et al.
Veröffentlicht: (2024)
von: Zhang, Ce, et al.
Veröffentlicht: (2024)
Language Guided Concept Bottleneck Models for Interpretable Continual Learning
von: Yu, Lu, et al.
Veröffentlicht: (2025)
von: Yu, Lu, et al.
Veröffentlicht: (2025)
Hierarchical Cross-modal Prompt Learning for Vision-Language Models
von: Zheng, Hao, et al.
Veröffentlicht: (2025)
von: Zheng, Hao, et al.
Veröffentlicht: (2025)
Vision-Language Model IP Protection via Prompt-based Learning
von: Wang, Lianyu, et al.
Veröffentlicht: (2025)
von: Wang, Lianyu, et al.
Veröffentlicht: (2025)
Enhancing Vision-Language Few-Shot Adaptation with Negative Learning
von: Zhang, Ce, et al.
Veröffentlicht: (2024)
von: Zhang, Ce, et al.
Veröffentlicht: (2024)
Modeling Variants of Prompts for Vision-Language Models
von: Li, Ao, et al.
Veröffentlicht: (2025)
von: Li, Ao, et al.
Veröffentlicht: (2025)
In the Era of Prompt Learning with Vision-Language Models
von: Jha, Ankit
Veröffentlicht: (2024)
von: Jha, Ankit
Veröffentlicht: (2024)
PromptKD: Unsupervised Prompt Distillation for Vision-Language Models
von: Li, Zheng, et al.
Veröffentlicht: (2024)
von: Li, Zheng, et al.
Veröffentlicht: (2024)
Active Prompt Learning in Vision Language Models
von: Bang, Jihwan, et al.
Veröffentlicht: (2023)
von: Bang, Jihwan, et al.
Veröffentlicht: (2023)
Interpretable Few-Shot Retinal Disease Diagnosis with Concept-Guided Prompting of Vision-Language Models
von: Mehta, Deval, et al.
Veröffentlicht: (2025)
von: Mehta, Deval, et al.
Veröffentlicht: (2025)
VScan: Rethinking Visual Token Reduction for Efficient Large Vision-Language Models
von: Zhang, Ce, et al.
Veröffentlicht: (2025)
von: Zhang, Ce, et al.
Veröffentlicht: (2025)
Language-Guided Token Compression with Reinforcement Learning in Large Vision-Language Models
von: Cao, Sihan, et al.
Veröffentlicht: (2026)
von: Cao, Sihan, et al.
Veröffentlicht: (2026)
Dual Prompt Learning for Adapting Vision-Language Models to Downstream Image-Text Retrieval
von: Wang, Yifan, et al.
Veröffentlicht: (2025)
von: Wang, Yifan, et al.
Veröffentlicht: (2025)
Quantized Prompt for Efficient Generalization of Vision-Language Models
von: Hao, Tianxiang, et al.
Veröffentlicht: (2024)
von: Hao, Tianxiang, et al.
Veröffentlicht: (2024)
Generalizing Vision-Language Models with Dedicated Prompt Guidance
von: Li, Xinyao, et al.
Veröffentlicht: (2025)
von: Li, Xinyao, et al.
Veröffentlicht: (2025)
InPK: Infusing Prior Knowledge into Prompt for Vision-Language Models
von: Zhou, Shuchang, et al.
Veröffentlicht: (2025)
von: Zhou, Shuchang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Conceptual Codebook Learning for Vision-Language Models
von: Zhang, Yi, et al.
Veröffentlicht: (2024) -
Domain-Conditioned Transformer for Fully Test-time Adaptation
von: Tang, Yushun, et al.
Veröffentlicht: (2024) -
Progressive Conditioned Scale-Shift Recalibration of Self-Attention for Online Test-time Adaptation
von: Tang, Yushun, et al.
Veröffentlicht: (2025) -
Learning Visual Conditioning Tokens to Correct Domain Shift for Fully Test-time Adaptation
von: Tang, Yushun, et al.
Veröffentlicht: (2024) -
Open-World Test-Time Adaptation with Hierarchical Feature Aggregation and Attention Affine
von: Liu, Ziqiong, et al.
Veröffentlicht: (2025)