Understanding Model Reprogramming for CLIP via Decoupling Visual Prompts
Fuente:
arXiv
Saved in:
| Main Authors: | Cai, Chengyi, Ye, Zesheng, Feng, Lei, Qi, Jianzhong, Liu, Feng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sample-specific Masks for Visual Reprogramming-based Prompting
by: Cai, Chengyi, et al.
Published: (2024)
by: Cai, Chengyi, et al.
Published: (2024)
Attribute-based Visual Reprogramming for Vision-Language Models
by: Cai, Chengyi, et al.
Published: (2025)
by: Cai, Chengyi, et al.
Published: (2025)
Bayesian-guided Label Mapping for Visual Reprogramming
by: Cai, Chengyi, et al.
Published: (2024)
by: Cai, Chengyi, et al.
Published: (2024)
Visual-Guided Key-Token Regularization for Multimodal Large Language Model Unlearning
by: Cai, Chengyi, et al.
Published: (2026)
by: Cai, Chengyi, et al.
Published: (2026)
Prime Once, then Reprogram Locally: An Efficient Alternative to Black-Box Service Model Adaptation
by: Zhang, Yunbei, et al.
Published: (2026)
by: Zhang, Yunbei, et al.
Published: (2026)
Neural Network Reprogrammability: A Unified Theme on Model Reprogramming, Prompt Tuning, and Prompt Instruction
by: Ye, Zesheng, et al.
Published: (2025)
by: Ye, Zesheng, et al.
Published: (2025)
Adapting in the Dark: Efficient and Stable Test-Time Adaptation for Black-Box Models
by: Zhang, Yunbei, et al.
Published: (2026)
by: Zhang, Yunbei, et al.
Published: (2026)
Sample-Specific Noise Injection For Diffusion-Based Adversarial Purification
by: Sun, Yuhao, et al.
Published: (2025)
by: Sun, Yuhao, et al.
Published: (2025)
Candidate Pseudolabel Learning: Enhancing Vision-Language Models by Prompt Tuning with Unlabeled Data
by: Zhang, Jiahan, et al.
Published: (2024)
by: Zhang, Jiahan, et al.
Published: (2024)
Test-Time Multimodal Backdoor Detection by Contrastive Prompting
by: Niu, Yuwei, et al.
Published: (2024)
by: Niu, Yuwei, et al.
Published: (2024)
X2CT-CLIP: Enable Multi-Abnormality Detection in Computed Tomography from Chest Radiography via Tri-Modal Contrastive Learning
by: You, Jianzhong, et al.
Published: (2025)
by: You, Jianzhong, et al.
Published: (2025)
MoP-CLIP: A Mixture of Prompt-Tuned CLIP Models for Domain Incremental Learning
by: Nicolas, Julien, et al.
Published: (2023)
by: Nicolas, Julien, et al.
Published: (2023)
Enhancing CLIP with CLIP: Exploring Pseudolabeling for Limited-Label Prompt Tuning
by: Menghini, Cristina, et al.
Published: (2023)
by: Menghini, Cristina, et al.
Published: (2023)
GRIT-LP: Graph Transformer with Long-Range Skip Connection and Partitioned Spatial Graphs for Accurate Ice Layer Thickness Prediction
by: Liu, Zesheng, et al.
Published: (2025)
by: Liu, Zesheng, et al.
Published: (2025)
ST-GRIT: Spatio-Temporal Graph Transformer For Internal Ice Layer Thickness Prediction
by: Liu, Zesheng, et al.
Published: (2025)
by: Liu, Zesheng, et al.
Published: (2025)
Multi-branch Spatio-Temporal Graph Neural Network For Efficient Ice Layer Thickness Prediction
by: Liu, Zesheng, et al.
Published: (2024)
by: Liu, Zesheng, et al.
Published: (2024)
K-STEMIT: Knowledge-Informed Spatio-Temporal Efficient Multi-Branch Graph Neural Network for Subsurface Stratigraphy Thickness Estimation from Radar Data
by: Liu, Zesheng, et al.
Published: (2026)
by: Liu, Zesheng, et al.
Published: (2026)
Learning Spatio-Temporal Patterns of Polar Ice Layers With Physics-Informed Graph Neural Network
by: Liu, Zesheng, et al.
Published: (2024)
by: Liu, Zesheng, et al.
Published: (2024)
Let's Roll a BiFTA: Bi-refinement for Fine-grained Text-visual Alignment in Vision-Language Models
by: Sun, Yuhao, et al.
Published: (2026)
by: Sun, Yuhao, et al.
Published: (2026)
Visually Prompted Benchmarks Are Surprisingly Fragile
by: Feng, Haiwen, et al.
Published: (2025)
by: Feng, Haiwen, et al.
Published: (2025)
VeCLIP: Improving CLIP Training via Visual-enriched Captions
by: Lai, Zhengfeng, et al.
Published: (2023)
by: Lai, Zhengfeng, et al.
Published: (2023)
IDEA: Image Description Enhanced CLIP-Adapter
by: Ye, Zhipeng, et al.
Published: (2025)
by: Ye, Zhipeng, et al.
Published: (2025)
NLPrompt: Noise-Label Prompt Learning for Vision-Language Models
by: Pan, Bikang, et al.
Published: (2024)
by: Pan, Bikang, et al.
Published: (2024)
Aggregate-and-Adapt Natural Language Prompts for Downstream Generalization of CLIP
by: Huang, Chen, et al.
Published: (2024)
by: Huang, Chen, et al.
Published: (2024)
Towards Visuospatial Cognition via Hierarchical Fusion of Visual Experts
by: Feng, Qi
Published: (2025)
by: Feng, Qi
Published: (2025)
SAM-CLIP: Merging Vision Foundation Models towards Semantic and Spatial Understanding
by: Wang, Haoxiang, et al.
Published: (2023)
by: Wang, Haoxiang, et al.
Published: (2023)
K Nearest Neighbor-Guided Trajectory Similarity Learning
by: Chang, Yanchuan, et al.
Published: (2025)
by: Chang, Yanchuan, et al.
Published: (2025)
Robust Adaptation of Foundation Models with Black-Box Visual Prompting
by: Oh, Changdae, et al.
Published: (2024)
by: Oh, Changdae, et al.
Published: (2024)
Online Zero-Shot Classification with CLIP
by: Qian, Qi, et al.
Published: (2024)
by: Qian, Qi, et al.
Published: (2024)
AutoVP: An Automated Visual Prompting Framework and Benchmark
by: Tsao, Hsi-Ai, et al.
Published: (2023)
by: Tsao, Hsi-Ai, et al.
Published: (2023)
PromptHub: Enhancing Multi-Prompt Visual In-Context Learning with Locality-Aware Fusion, Concentration and Alignment
by: Luo, Tianci, et al.
Published: (2026)
by: Luo, Tianci, et al.
Published: (2026)
ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts
by: Cai, Mu, et al.
Published: (2023)
by: Cai, Mu, et al.
Published: (2023)
PRO-VPT: Distribution-Adaptive Visual Prompt Tuning via Prompt Relocation
by: Shang, Chikai, et al.
Published: (2025)
by: Shang, Chikai, et al.
Published: (2025)
FastCLIP: A Suite of Optimization Techniques to Accelerate CLIP Training with Limited Resources
by: Wei, Xiyuan, et al.
Published: (2024)
by: Wei, Xiyuan, et al.
Published: (2024)
TV100: A TV Series Dataset that Pre-Trained CLIP Has Not Seen
by: Zhou, Da-Wei, et al.
Published: (2024)
by: Zhou, Da-Wei, et al.
Published: (2024)
CLIP Can Understand Depth
by: Kim, Sohee, et al.
Published: (2024)
by: Kim, Sohee, et al.
Published: (2024)
C^2Prompt: Class-aware Client Knowledge Interaction for Federated Continual Learning
by: Xu, Kunlun, et al.
Published: (2025)
by: Xu, Kunlun, et al.
Published: (2025)
Any-Order GPT as Masked Diffusion Model: Decoupling Formulation and Architecture
by: Xue, Shuchen, et al.
Published: (2025)
by: Xue, Shuchen, et al.
Published: (2025)
From Semantics to Pixels: Coarse-to-Fine Masked Autoencoders for Hierarchical Visual Understanding
by: Xiang, Wenzhao, et al.
Published: (2026)
by: Xiang, Wenzhao, et al.
Published: (2026)
CLIP-Inspector: Model-Level Backdoor Detection for Prompt-Tuned CLIP via OOD Trigger Inversion
by: Jindal, Akshit, et al.
Published: (2026)
by: Jindal, Akshit, et al.
Published: (2026)
Similar Items
-
Sample-specific Masks for Visual Reprogramming-based Prompting
by: Cai, Chengyi, et al.
Published: (2024) -
Attribute-based Visual Reprogramming for Vision-Language Models
by: Cai, Chengyi, et al.
Published: (2025) -
Bayesian-guided Label Mapping for Visual Reprogramming
by: Cai, Chengyi, et al.
Published: (2024) -
Visual-Guided Key-Token Regularization for Multimodal Large Language Model Unlearning
by: Cai, Chengyi, et al.
Published: (2026) -
Prime Once, then Reprogram Locally: An Efficient Alternative to Black-Box Service Model Adaptation
by: Zhang, Yunbei, et al.
Published: (2026)