Exploring Interpretability for Visual Prompt Tuning with Cross-layer Concepts
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Yubin, Jiang, Xinyang, Cheng, De, Zhao, Xiangqian, Wang, Zilong, Li, Dongsheng, Zhao, Cairong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ActPrompt: In-Domain Feature Adaptation via Action Cues for Video Temporal Grounding
by: Wang, Yubin, et al.
Published: (2024)
by: Wang, Yubin, et al.
Published: (2024)
HPT++: Hierarchically Prompting Vision-Language Models with Multi-Granularity Knowledge Generation and Improved Structure Modeling
by: Wang, Yubin, et al.
Published: (2024)
by: Wang, Yubin, et al.
Published: (2024)
DiffPhysBA: Diffusion-based Physical Backdoor Attack against Person Re-Identification in Real-World
by: Sun, Wenli, et al.
Published: (2024)
by: Sun, Wenli, et al.
Published: (2024)
One Object, Multiple Lies: A Benchmark for Cross-task Adversarial Attack on Unified Vision-Language Models
by: Zhao, Jiale, et al.
Published: (2025)
by: Zhao, Jiale, et al.
Published: (2025)
Online Video Quality Enhancement with Spatial-Temporal Look-up Tables
by: Qu, Zefan, et al.
Published: (2023)
by: Qu, Zefan, et al.
Published: (2023)
Uni$^2$Det: Unified and Universal Framework for Prompt-Guided Multi-dataset 3D Detection
by: Wang, Yubin, et al.
Published: (2024)
by: Wang, Yubin, et al.
Published: (2024)
Prompt Disentanglement via Language Guidance and Representation Alignment for Domain Generalization
by: Cheng, De, et al.
Published: (2025)
by: Cheng, De, et al.
Published: (2025)
Similarity Distribution based Membership Inference Attack on Person Re-identification
by: Gao, Junyao, et al.
Published: (2022)
by: Gao, Junyao, et al.
Published: (2022)
Patch-Prompt Aligned Bayesian Prompt Tuning for Vision-Language Models
by: Liu, Xinyang, et al.
Published: (2023)
by: Liu, Xinyang, et al.
Published: (2023)
CVPT: Cross Visual Prompt Tuning
by: Huang, Lingyun, et al.
Published: (2024)
by: Huang, Lingyun, et al.
Published: (2024)
LSPT: Long-term Spatial Prompt Tuning for Visual Representation Learning
by: Mo, Shentong, et al.
Published: (2024)
by: Mo, Shentong, et al.
Published: (2024)
Visual Variational Autoencoder Prompt Tuning
by: Xiao, Xi, et al.
Published: (2025)
by: Xiao, Xi, et al.
Published: (2025)
Adaptive Discriminative Regularization for Visual Classification
by: Zhao, Qingsong, et al.
Published: (2022)
by: Zhao, Qingsong, et al.
Published: (2022)
Embedded Visual Prompt Tuning
by: Zu, Wenqiang, et al.
Published: (2024)
by: Zu, Wenqiang, et al.
Published: (2024)
DreamDistribution: Learning Prompt Distribution for Diverse In-distribution Generation
by: Zhao, Brian Nlong, et al.
Published: (2023)
by: Zhao, Brian Nlong, et al.
Published: (2023)
Visual Prompt Tuning in Null Space for Continual Learning
by: Lu, Yue, et al.
Published: (2024)
by: Lu, Yue, et al.
Published: (2024)
iVPT: Improving Task-relevant Information Sharing in Visual Prompt Tuning by Cross-layer Dynamic Connection
by: Zhou, Nan, et al.
Published: (2024)
by: Zhou, Nan, et al.
Published: (2024)
Revisiting the Power of Prompt for Visual Tuning
by: Wang, Yuzhu, et al.
Published: (2024)
by: Wang, Yuzhu, et al.
Published: (2024)
Instruction Tuning-free Visual Token Complement for Multimodal LLMs
by: Wang, Dongsheng, et al.
Published: (2024)
by: Wang, Dongsheng, et al.
Published: (2024)
Attention to the Burstiness in Visual Prompt Tuning!
by: Wang, Yuzhu, et al.
Published: (2025)
by: Wang, Yuzhu, et al.
Published: (2025)
Facing the Elephant in the Room: Visual Prompt Tuning or Full Finetuning?
by: Han, Cheng, et al.
Published: (2024)
by: Han, Cheng, et al.
Published: (2024)
Cross-Resolution Land Cover Classification Using Outdated Products and Transformers
by: Ni, Huan, et al.
Published: (2024)
by: Ni, Huan, et al.
Published: (2024)
Correlative and Discriminative Label Grouping for Multi-Label Visual Prompt Tuning
by: Ma, LeiLei, et al.
Published: (2025)
by: Ma, LeiLei, et al.
Published: (2025)
MoviePuzzle: Visual Narrative Reasoning through Multimodal Order Learning
by: Wang, Jianghui, et al.
Published: (2023)
by: Wang, Jianghui, et al.
Published: (2023)
TRAIL: Transferable Robust Adversarial Images via Latent diffusion
by: Xue, Yuhao, et al.
Published: (2025)
by: Xue, Yuhao, et al.
Published: (2025)
Exploring Visual Prompting: Robustness Inheritance and Beyond
by: Li, Qi, et al.
Published: (2025)
by: Li, Qi, et al.
Published: (2025)
Visual Fourier Prompt Tuning
by: Zeng, Runjia, et al.
Published: (2024)
by: Zeng, Runjia, et al.
Published: (2024)
Exploring Typographic Visual Prompts Injection Threats in Cross-Modality Generation Models
by: Cheng, Hao, et al.
Published: (2025)
by: Cheng, Hao, et al.
Published: (2025)
MePT: Multi-Representation Guided Prompt Tuning for Vision-Language Model
by: Wang, Xinyang, et al.
Published: (2024)
by: Wang, Xinyang, et al.
Published: (2024)
Joint Semantic Token Selection and Prompt Optimization for Interpretable Prompt Learning
by: Wang, Yating, et al.
Published: (2026)
by: Wang, Yating, et al.
Published: (2026)
Visual Instance-aware Prompt Tuning
by: Xiao, Xi, et al.
Published: (2025)
by: Xiao, Xi, et al.
Published: (2025)
DA-VPT: Semantic-Guided Visual Prompt Tuning for Vision Transformers
by: Ren, Li, et al.
Published: (2025)
by: Ren, Li, et al.
Published: (2025)
DEVICE: Depth and Visual Concepts Aware Transformer for OCR-based Image Captioning
by: Xu, Dongsheng, et al.
Published: (2023)
by: Xu, Dongsheng, et al.
Published: (2023)
SDVPT: Semantic-Driven Visual Prompt Tuning for Open-World Object Counting
by: Zhao, Yiming, et al.
Published: (2025)
by: Zhao, Yiming, et al.
Published: (2025)
Molecular Identifier Visual Prompt and Verifiable Reinforcement Learning for Chemical Reaction Diagram Parsing
by: Song, Jiahe, et al.
Published: (2026)
by: Song, Jiahe, et al.
Published: (2026)
Exploring Task-Solving Paradigm for Generalized Cross-Domain Face Anti-Spoofing via Reinforcement Fine-Tuning
by: Jiang, Fangling, et al.
Published: (2025)
by: Jiang, Fangling, et al.
Published: (2025)
Self-Supervised Visual Prompting for Cross-Domain Road Damage Detection
by: Xiao, Xi, et al.
Published: (2025)
by: Xiao, Xi, et al.
Published: (2025)
Defending Multimodal Backdoored Models by Repulsive Visual Prompt Tuning
by: Zhang, Zhifang, et al.
Published: (2024)
by: Zhang, Zhifang, et al.
Published: (2024)
CamPVG: Camera-Controlled Panoramic Video Generation with Epipolar-Aware Diffusion
by: Ji, Chenhao, et al.
Published: (2025)
by: Ji, Chenhao, et al.
Published: (2025)
Visual Spatial Tuning
by: Yang, Rui, et al.
Published: (2025)
by: Yang, Rui, et al.
Published: (2025)
Similar Items
-
ActPrompt: In-Domain Feature Adaptation via Action Cues for Video Temporal Grounding
by: Wang, Yubin, et al.
Published: (2024) -
HPT++: Hierarchically Prompting Vision-Language Models with Multi-Granularity Knowledge Generation and Improved Structure Modeling
by: Wang, Yubin, et al.
Published: (2024) -
DiffPhysBA: Diffusion-based Physical Backdoor Attack against Person Re-Identification in Real-World
by: Sun, Wenli, et al.
Published: (2024) -
One Object, Multiple Lies: A Benchmark for Cross-task Adversarial Attack on Unified Vision-Language Models
by: Zhao, Jiale, et al.
Published: (2025) -
Online Video Quality Enhancement with Spatial-Temporal Look-up Tables
by: Qu, Zefan, et al.
Published: (2023)