Learning to Prompt Your Domain for Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wei, Guoyizhe, Wang, Feng, Shah, Anshul, Chellappa, Rama |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ViT-Linearizer: Distilling Quadratic Knowledge into Linear-Time Vision Models
von: Wei, Guoyizhe, et al.
Veröffentlicht: (2025)
von: Wei, Guoyizhe, et al.
Veröffentlicht: (2025)
Latent Domain Prompt Learning for Vision-Language Models
von: Li, Zhixing, et al.
Veröffentlicht: (2025)
von: Li, Zhixing, et al.
Veröffentlicht: (2025)
Unsupervised Video Domain Adaptation with Masked Pre-Training and Collaborative Self-Training
von: Reddy, Arun, et al.
Veröffentlicht: (2023)
von: Reddy, Arun, et al.
Veröffentlicht: (2023)
Transitive Vision-Language Prompt Learning for Domain Generalization
von: Wang, Liyuan, et al.
Veröffentlicht: (2024)
von: Wang, Liyuan, et al.
Veröffentlicht: (2024)
FLAME: Adaptive Mixture-of-Experts for Continual Multimodal Multi-Task Learning
von: Han, Xing, et al.
Veröffentlicht: (2026)
von: Han, Xing, et al.
Veröffentlicht: (2026)
Weighted Risk Invariance: Domain Generalization under Invariant Feature Shift
von: Wong, Gina, et al.
Veröffentlicht: (2024)
von: Wong, Gina, et al.
Veröffentlicht: (2024)
NLPrompt: Noise-Label Prompt Learning for Vision-Language Models
von: Pan, Bikang, et al.
Veröffentlicht: (2024)
von: Pan, Bikang, et al.
Veröffentlicht: (2024)
Understanding and Mitigating Miscalibration in Prompt Tuning for Vision-Language Models
von: Wang, Shuoyuan, et al.
Veröffentlicht: (2024)
von: Wang, Shuoyuan, et al.
Veröffentlicht: (2024)
An Empirical Study of Federated Prompt Learning for Vision Language Model
von: Wang, Zhihao, et al.
Veröffentlicht: (2025)
von: Wang, Zhihao, et al.
Veröffentlicht: (2025)
Certified Robustness via Dynamic Margin Maximization and Improved Lipschitz Regularization
von: Fazlyab, Mahyar, et al.
Veröffentlicht: (2023)
von: Fazlyab, Mahyar, et al.
Veröffentlicht: (2023)
Candidate Pseudolabel Learning: Enhancing Vision-Language Models by Prompt Tuning with Unlabeled Data
von: Zhang, Jiahan, et al.
Veröffentlicht: (2024)
von: Zhang, Jiahan, et al.
Veröffentlicht: (2024)
Shared LoRA Subspaces for almost Strict Continual Learning
von: Kaushik, Prakhar, et al.
Veröffentlicht: (2026)
von: Kaushik, Prakhar, et al.
Veröffentlicht: (2026)
Improving Coverage in Combined Prediction Sets with Weighted p-values
von: Wong, Gina, et al.
Veröffentlicht: (2025)
von: Wong, Gina, et al.
Veröffentlicht: (2025)
Mesh-Gait: A Unified Framework for Gait Recognition Through Multi-Modal Representation Learning from 2D Silhouettes
von: Wang, Zhao-Yang, et al.
Veröffentlicht: (2025)
von: Wang, Zhao-Yang, et al.
Veröffentlicht: (2025)
Neutral-Reference Prompting for Vision-Language Models
von: Tian, Senmao, et al.
Veröffentlicht: (2026)
von: Tian, Senmao, et al.
Veröffentlicht: (2026)
Fool Your (Vision and) Language Model With Embarrassingly Simple Permutations
von: Zong, Yongshuo, et al.
Veröffentlicht: (2023)
von: Zong, Yongshuo, et al.
Veröffentlicht: (2023)
Make Prompts Adaptable: Bayesian Modeling for Vision-Language Prompt Learning with Data-Dependent Prior
von: Cho, Youngjae, et al.
Veröffentlicht: (2024)
von: Cho, Youngjae, et al.
Veröffentlicht: (2024)
Tree of Attributes Prompt Learning for Vision-Language Models
von: Ding, Tong, et al.
Veröffentlicht: (2024)
von: Ding, Tong, et al.
Veröffentlicht: (2024)
Differentiable Prompt Learning for Vision Language Models
von: Huang, Zhenhan, et al.
Veröffentlicht: (2024)
von: Huang, Zhenhan, et al.
Veröffentlicht: (2024)
Scaling State-Space Models on Multiple GPUs with Tensor Parallelism
von: Dutt, Anurag, et al.
Veröffentlicht: (2026)
von: Dutt, Anurag, et al.
Veröffentlicht: (2026)
The Universal Weight Subspace Hypothesis
von: Kaushik, Prakhar, et al.
Veröffentlicht: (2025)
von: Kaushik, Prakhar, et al.
Veröffentlicht: (2025)
When Adaptation Fails: A Gradient-Based Diagnosis of Collapsed Gating in Vision-Language Prompt Learning
von: Fang, Yunxuan, et al.
Veröffentlicht: (2026)
von: Fang, Yunxuan, et al.
Veröffentlicht: (2026)
Privacy-Preserving Personalized Federated Prompt Learning for Multimodal Large Language Models
von: Tran, Linh, et al.
Veröffentlicht: (2025)
von: Tran, Linh, et al.
Veröffentlicht: (2025)
Commute Your Domains: Trajectory Optimality Criterion for Multi-Domain Learning
von: Rukhovich, Alexey, et al.
Veröffentlicht: (2025)
von: Rukhovich, Alexey, et al.
Veröffentlicht: (2025)
Pix2Key: Controllable Open-Vocabulary Retrieval with Semantic Decomposition and Self-Supervised Visual Dictionary Learning
von: Wei, Guoyizhe, et al.
Veröffentlicht: (2026)
von: Wei, Guoyizhe, et al.
Veröffentlicht: (2026)
Decoupling Augmentation Bias in Prompt Learning for Vision-Language Models
von: Kim, Gahyeon, et al.
Veröffentlicht: (2025)
von: Kim, Gahyeon, et al.
Veröffentlicht: (2025)
AAPL: Adding Attributes to Prompt Learning for Vision-Language Models
von: Kim, Gahyeon, et al.
Veröffentlicht: (2024)
von: Kim, Gahyeon, et al.
Veröffentlicht: (2024)
DriVLM: Domain Adaptation of Vision-Language Models in Autonomous Driving
von: Zheng, Xuran, et al.
Veröffentlicht: (2025)
von: Zheng, Xuran, et al.
Veröffentlicht: (2025)
Patch-Prompt Aligned Bayesian Prompt Tuning for Vision-Language Models
von: Liu, Xinyang, et al.
Veröffentlicht: (2023)
von: Liu, Xinyang, et al.
Veröffentlicht: (2023)
PLeaS -- Merging Models with Permutations and Least Squares
von: Nasery, Anshul, et al.
Veröffentlicht: (2024)
von: Nasery, Anshul, et al.
Veröffentlicht: (2024)
To Trust Or Not To Trust Your Vision-Language Model's Prediction
von: Dong, Hao, et al.
Veröffentlicht: (2025)
von: Dong, Hao, et al.
Veröffentlicht: (2025)
CAPT: Class-Aware Prompt Tuning for Federated Long-Tailed Learning with Vision-Language Model
von: Hou, Shihao, et al.
Veröffentlicht: (2025)
von: Hou, Shihao, et al.
Veröffentlicht: (2025)
ChordPrompt: Orchestrating Cross-Modal Prompt Synergy for Multi-Domain Incremental Learning in CLIP
von: Wang, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Wang, Zhiyuan, et al.
Veröffentlicht: (2025)
CP-Prompt: Composition-Based Cross-modal Prompting for Domain-Incremental Continual Learning
von: Feng, Yu, et al.
Veröffentlicht: (2024)
von: Feng, Yu, et al.
Veröffentlicht: (2024)
Few-Shot Adversarial Prompt Learning on Vision-Language Models
von: Zhou, Yiwei, et al.
Veröffentlicht: (2024)
von: Zhou, Yiwei, et al.
Veröffentlicht: (2024)
Vision and Language Integration for Domain Generalization
von: Wang, Yanmei, et al.
Veröffentlicht: (2025)
von: Wang, Yanmei, et al.
Veröffentlicht: (2025)
Context-Adaptive Multi-Prompt Embedding with Large Language Models for Vision-Language Alignment
von: Kim, Dahun, et al.
Veröffentlicht: (2025)
von: Kim, Dahun, et al.
Veröffentlicht: (2025)
OmniVLA: An Omni-Modal Vision-Language-Action Model for Robot Navigation
von: Hirose, Noriaki, et al.
Veröffentlicht: (2025)
von: Hirose, Noriaki, et al.
Veröffentlicht: (2025)
Effects of Prompt Length on Domain-specific Tasks for Large Language Models
von: Liu, Qibang, et al.
Veröffentlicht: (2025)
von: Liu, Qibang, et al.
Veröffentlicht: (2025)
Graph Your Own Prompt
von: Ding, Xi, et al.
Veröffentlicht: (2025)
von: Ding, Xi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ViT-Linearizer: Distilling Quadratic Knowledge into Linear-Time Vision Models
von: Wei, Guoyizhe, et al.
Veröffentlicht: (2025) -
Latent Domain Prompt Learning for Vision-Language Models
von: Li, Zhixing, et al.
Veröffentlicht: (2025) -
Unsupervised Video Domain Adaptation with Masked Pre-Training and Collaborative Self-Training
von: Reddy, Arun, et al.
Veröffentlicht: (2023) -
Transitive Vision-Language Prompt Learning for Domain Generalization
von: Wang, Liyuan, et al.
Veröffentlicht: (2024) -
FLAME: Adaptive Mixture-of-Experts for Continual Multimodal Multi-Task Learning
von: Han, Xing, et al.
Veröffentlicht: (2026)