ADAPT to Robustify Prompt Tuning Vision Transformers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Eskandar, Masih, Imtiaz, Tooba, Wang, Zifeng, Dy, Jennifer |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
STAR: Stability-Inducing Weight Perturbation for Continual Learning
von: Eskandar, Masih, et al.
Veröffentlicht: (2025)
von: Eskandar, Masih, et al.
Veröffentlicht: (2025)
SAIF: Sparse Adversarial and Imperceptible Attack Framework
von: Imtiaz, Tooba, et al.
Veröffentlicht: (2022)
von: Imtiaz, Tooba, et al.
Veröffentlicht: (2022)
LVT: Large-Scale Scene Reconstruction via Local View Transformers
von: Imtiaz, Tooba, et al.
Veröffentlicht: (2025)
von: Imtiaz, Tooba, et al.
Veröffentlicht: (2025)
Adaptive Prompt Tuning: Vision Guided Prompt Tuning with Cross-Attention for Fine-Grained Few-Shot Learning
von: Brouwer, Eric, et al.
Veröffentlicht: (2024)
von: Brouwer, Eric, et al.
Veröffentlicht: (2024)
Grounding Multimodal Large Language Models with Quantitative Skin Attributes: A Retrieval Study
von: Torop, Max, et al.
Veröffentlicht: (2025)
von: Torop, Max, et al.
Veröffentlicht: (2025)
Doubly Debiased Test-Time Prompt Tuning for Vision-Language Models
von: Song, Fei, et al.
Veröffentlicht: (2025)
von: Song, Fei, et al.
Veröffentlicht: (2025)
Candidate Pseudolabel Learning: Enhancing Vision-Language Models by Prompt Tuning with Unlabeled Data
von: Zhang, Jiahan, et al.
Veröffentlicht: (2024)
von: Zhang, Jiahan, et al.
Veröffentlicht: (2024)
LAPT: Label-driven Automated Prompt Tuning for OOD Detection with Vision-Language Models
von: Zhang, Yabin, et al.
Veröffentlicht: (2024)
von: Zhang, Yabin, et al.
Veröffentlicht: (2024)
AdapterTune: Zero-Initialized Low-Rank Adapters for Frozen Vision Transformers
von: Khazem, Salim
Veröffentlicht: (2026)
von: Khazem, Salim
Veröffentlicht: (2026)
PanoWorld: Geometry-Consistent Panoramic Video World Modeling
von: Jiang, Le, et al.
Veröffentlicht: (2026)
von: Jiang, Le, et al.
Veröffentlicht: (2026)
MM-Prompt: Cross-Modal Prompt Tuning for Continual Visual Question Answering
von: Li, Xu, et al.
Veröffentlicht: (2025)
von: Li, Xu, et al.
Veröffentlicht: (2025)
LSPT: Long-term Spatial Prompt Tuning for Visual Representation Learning
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
Source-Free Domain Adaptation for YOLO Object Detection
von: Varailhon, Simon, et al.
Veröffentlicht: (2024)
von: Varailhon, Simon, et al.
Veröffentlicht: (2024)
Prompt Diffusion Robustifies Any-Modality Prompt Learning
von: Du, Yingjun, et al.
Veröffentlicht: (2024)
von: Du, Yingjun, et al.
Veröffentlicht: (2024)
Improving Interpretation Faithfulness for Vision Transformers
von: Hu, Lijie, et al.
Veröffentlicht: (2023)
von: Hu, Lijie, et al.
Veröffentlicht: (2023)
BiPrompt: Bilateral Prompt Optimization for Visual and Textual Debiasing in Vision-Language Models
von: Gupta, Sunny, et al.
Veröffentlicht: (2026)
von: Gupta, Sunny, et al.
Veröffentlicht: (2026)
The devil is in discretization discrepancy. Robustifying Differentiable NAS with Single-Stage Searching Protocol
von: Subbotko, Konstanty, et al.
Veröffentlicht: (2024)
von: Subbotko, Konstanty, et al.
Veröffentlicht: (2024)
Dynamic Cross-Modal Prompt Generation for Multimodal Continual Instruction Tuning
von: Hu, Tao, et al.
Veröffentlicht: (2026)
von: Hu, Tao, et al.
Veröffentlicht: (2026)
Tree of Attributes Prompt Learning for Vision-Language Models
von: Ding, Tong, et al.
Veröffentlicht: (2024)
von: Ding, Tong, et al.
Veröffentlicht: (2024)
C-TPT: Calibrated Test-Time Prompt Tuning for Vision-Language Models via Text Feature Dispersion
von: Yoon, Hee Suk, et al.
Veröffentlicht: (2024)
von: Yoon, Hee Suk, et al.
Veröffentlicht: (2024)
Discovering Influential Neuron Path in Vision Transformers
von: Wang, Yifan, et al.
Veröffentlicht: (2025)
von: Wang, Yifan, et al.
Veröffentlicht: (2025)
Visual Modality Prompt for Adapting Vision-Language Object Detectors
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2024)
von: Medeiros, Heitor R., et al.
Veröffentlicht: (2024)
AAPL: Adding Attributes to Prompt Learning for Vision-Language Models
von: Kim, Gahyeon, et al.
Veröffentlicht: (2024)
von: Kim, Gahyeon, et al.
Veröffentlicht: (2024)
Decoupling Augmentation Bias in Prompt Learning for Vision-Language Models
von: Kim, Gahyeon, et al.
Veröffentlicht: (2025)
von: Kim, Gahyeon, et al.
Veröffentlicht: (2025)
Deep Networks Always Grok and Here is Why
von: Humayun, Ahmed Imtiaz, et al.
Veröffentlicht: (2024)
von: Humayun, Ahmed Imtiaz, et al.
Veröffentlicht: (2024)
On the Geometry of Deep Learning
von: Balestriero, Randall, et al.
Veröffentlicht: (2024)
von: Balestriero, Randall, et al.
Veröffentlicht: (2024)
DiPrompT: Disentangled Prompt Tuning for Multiple Latent Domain Generalization in Federated Learning
von: Bai, Sikai, et al.
Veröffentlicht: (2024)
von: Bai, Sikai, et al.
Veröffentlicht: (2024)
A Vision-Enabled Prosthetic Hand for Children with Upper Limb Disabilities
von: Sarker, Md Abdul Baset, et al.
Veröffentlicht: (2025)
von: Sarker, Md Abdul Baset, et al.
Veröffentlicht: (2025)
Block-Recurrent Dynamics in Vision Transformers
von: Jacobs, Mozes, et al.
Veröffentlicht: (2025)
von: Jacobs, Mozes, et al.
Veröffentlicht: (2025)
Robust CLIP: Unsupervised Adversarial Fine-Tuning of Vision Embeddings for Robust Large Vision-Language Models
von: Schlarmann, Christian, et al.
Veröffentlicht: (2024)
von: Schlarmann, Christian, et al.
Veröffentlicht: (2024)
HGTDP-DTA: Hybrid Graph-Transformer with Dynamic Prompt for Drug-Target Binding Affinity Prediction
von: Xiao, Xi, et al.
Veröffentlicht: (2024)
von: Xiao, Xi, et al.
Veröffentlicht: (2024)
When Prompts Override Vision: Prompt-Induced Hallucinations in LVLMs
von: Khayatan, Pegah, et al.
Veröffentlicht: (2026)
von: Khayatan, Pegah, et al.
Veröffentlicht: (2026)
Self-Calibrated Tuning of Vision-Language Models for Out-of-Distribution Detection
von: Yu, Geng, et al.
Veröffentlicht: (2024)
von: Yu, Geng, et al.
Veröffentlicht: (2024)
CLEFT: Language-Image Contrastive Learning with Efficient Large Language Model and Prompt Fine-Tuning
von: Du, Yuexi, et al.
Veröffentlicht: (2024)
von: Du, Yuexi, et al.
Veröffentlicht: (2024)
Continual Adaptation of Vision Transformers for Federated Learning
von: Halbe, Shaunak, et al.
Veröffentlicht: (2023)
von: Halbe, Shaunak, et al.
Veröffentlicht: (2023)
Mechanisms of Non-Monotonic Scaling in Vision Transformers
von: Kumar, Anantha Padmanaban Krishna
Veröffentlicht: (2025)
von: Kumar, Anantha Padmanaban Krishna
Veröffentlicht: (2025)
DiffiT: Diffusion Vision Transformers for Image Generation
von: Hatamizadeh, Ali, et al.
Veröffentlicht: (2023)
von: Hatamizadeh, Ali, et al.
Veröffentlicht: (2023)
Accelerating Vision Transformers with Adaptive Patch Sizes
von: Choudhury, Rohan, et al.
Veröffentlicht: (2025)
von: Choudhury, Rohan, et al.
Veröffentlicht: (2025)
Class-Discriminative Attention Maps for Vision Transformers
von: Brocki, Lennart, et al.
Veröffentlicht: (2023)
von: Brocki, Lennart, et al.
Veröffentlicht: (2023)
Proactive Adversarial Defense: Harnessing Prompt Tuning in Vision-Language Models to Detect Unseen Backdoored Images
von: Stein, Kyle, et al.
Veröffentlicht: (2024)
von: Stein, Kyle, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
STAR: Stability-Inducing Weight Perturbation for Continual Learning
von: Eskandar, Masih, et al.
Veröffentlicht: (2025) -
SAIF: Sparse Adversarial and Imperceptible Attack Framework
von: Imtiaz, Tooba, et al.
Veröffentlicht: (2022) -
LVT: Large-Scale Scene Reconstruction via Local View Transformers
von: Imtiaz, Tooba, et al.
Veröffentlicht: (2025) -
Adaptive Prompt Tuning: Vision Guided Prompt Tuning with Cross-Attention for Fine-Grained Few-Shot Learning
von: Brouwer, Eric, et al.
Veröffentlicht: (2024) -
Grounding Multimodal Large Language Models with Quantitative Skin Attributes: A Retrieval Study
von: Torop, Max, et al.
Veröffentlicht: (2025)