X-Transfer Attacks: Towards Super Transferable Adversarial Attacks on CLIP
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Hanxun, Erfani, Sarah, Li, Yige, Ma, Xingjun, Bailey, James |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Toward Universal and Transferable Jailbreak Attacks on Vision-Language Models
by: Cui, Kaiyuan, et al.
Published: (2026)
by: Cui, Kaiyuan, et al.
Published: (2026)
Detecting Backdoor Samples in Contrastive Language Image Pretraining
by: Huang, Hanxun, et al.
Published: (2025)
by: Huang, Hanxun, et al.
Published: (2025)
Towards Million-Scale Adversarial Robustness Evaluation With Stronger Individual Attacks
by: Xie, Yong, et al.
Published: (2024)
by: Xie, Yong, et al.
Published: (2024)
Downstream Transfer Attack: Adversarial Attacks on Downstream Models with Pre-trained Vision Transformers
by: Zheng, Weijie, et al.
Published: (2024)
by: Zheng, Weijie, et al.
Published: (2024)
LDReg: Local Dimensionality Regularized Self-Supervised Learning
by: Huang, Hanxun, et al.
Published: (2024)
by: Huang, Hanxun, et al.
Published: (2024)
Prompt-Driven Contrastive Learning for Transferable Adversarial Attacks
by: Yang, Hunmin, et al.
Published: (2024)
by: Yang, Hunmin, et al.
Published: (2024)
FACL-Attack: Frequency-Aware Contrastive Learning for Transferable Adversarial Attacks
by: Yang, Hunmin, et al.
Published: (2024)
by: Yang, Hunmin, et al.
Published: (2024)
AIM: Additional Image Guided Generation of Transferable Adversarial Attacks
by: Li, Teng, et al.
Published: (2025)
by: Li, Teng, et al.
Published: (2025)
BlueSuffix: Reinforced Blue Teaming for Vision-Language Models Against Jailbreak Attacks
by: Zhao, Yunhan, et al.
Published: (2024)
by: Zhao, Yunhan, et al.
Published: (2024)
Improving the Transferability of Adversarial Attacks by an Input Transpose
by: Wan, Qing, et al.
Published: (2025)
by: Wan, Qing, et al.
Published: (2025)
Sparse Autoencoders as Plug-and-Play Firewalls for Adversarial Attack Detection in VLMs
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
AttackVLA: Benchmarking Adversarial and Backdoor Attacks on Vision-Language-Action Models
by: Li, Jiayu, et al.
Published: (2025)
by: Li, Jiayu, et al.
Published: (2025)
Benchmarking Transferable Adversarial Attacks
by: Jin, Zhibo, et al.
Published: (2024)
by: Jin, Zhibo, et al.
Published: (2024)
Sim-CLIP: Unsupervised Siamese Adversarial Fine-Tuning for Robust and Semantically-Rich Vision-Language Models
by: Hossain, Md Zarif, et al.
Published: (2024)
by: Hossain, Md Zarif, et al.
Published: (2024)
Attacking the Spike: On the Transferability and Security of Spiking Neural Networks to Adversarial Examples
by: Xu, Nuo, et al.
Published: (2022)
by: Xu, Nuo, et al.
Published: (2022)
MoDE: CLIP Data Experts via Clustering
by: Ma, Jiawei, et al.
Published: (2024)
by: Ma, Jiawei, et al.
Published: (2024)
Density-Aware Translation of Spurious Correlations in Zero-Shot VLMs
by: Hasanebrahimi, Afsaneh, et al.
Published: (2026)
by: Hasanebrahimi, Afsaneh, et al.
Published: (2026)
WebInject: Prompt Injection Attack to Web Agents
by: Wang, Xilong, et al.
Published: (2025)
by: Wang, Xilong, et al.
Published: (2025)
DarkLLM: Learning Language-Driven Adversarial Attacks with Large Language Models
by: Sun, Ye, et al.
Published: (2026)
by: Sun, Ye, et al.
Published: (2026)
Towards Predicting the Success of Transfer-based Attacks by Quantifying Shared Feature Representations
by: Dale, Ashley S., et al.
Published: (2024)
by: Dale, Ashley S., et al.
Published: (2024)
PerceptionCLIP: Visual Classification by Inferring and Conditioning on Contexts
by: An, Bang, et al.
Published: (2023)
by: An, Bang, et al.
Published: (2023)
Steering the Verifiability of Multimodal AI Hallucinations
by: Pang, Jianhong, et al.
Published: (2026)
by: Pang, Jianhong, et al.
Published: (2026)
Robust Prompt Optimization for Defending Language Models Against Jailbreaking Attacks
by: Zhou, Andy, et al.
Published: (2024)
by: Zhou, Andy, et al.
Published: (2024)
ADBA:Approximation Decision Boundary Approach for Black-Box Adversarial Attacks
by: Wang, Feiyang, et al.
Published: (2024)
by: Wang, Feiyang, et al.
Published: (2024)
PubDef: Defending Against Transfer Attacks From Public Models
by: Sitawarin, Chawin, et al.
Published: (2023)
by: Sitawarin, Chawin, et al.
Published: (2023)
Zero-Shot Defense Against Toxic Images via Inherent Multimodal Alignment in LVLMs
by: Zhao, Wei, et al.
Published: (2025)
by: Zhao, Wei, et al.
Published: (2025)
Failures to Find Transferable Image Jailbreaks Between Vision-Language Models
by: Schaeffer, Rylan, et al.
Published: (2024)
by: Schaeffer, Rylan, et al.
Published: (2024)
Pushing the Frontier of Black-Box LVLM Attacks via Fine-Grained Detail Targeting
by: Zhao, Xiaohan, et al.
Published: (2026)
by: Zhao, Xiaohan, et al.
Published: (2026)
eXIAA: eXplainable Injections for Adversarial Attack
by: Pesce, Leonardo, et al.
Published: (2025)
by: Pesce, Leonardo, et al.
Published: (2025)
Impact of Adversarial Attacks on Deep Learning Model Explainability
by: Nur, Gazi Nazia, et al.
Published: (2024)
by: Nur, Gazi Nazia, et al.
Published: (2024)
The Anatomy of Adversarial Attacks: Concept-based XAI Dissection
by: Mikriukov, Georgii, et al.
Published: (2024)
by: Mikriukov, Georgii, et al.
Published: (2024)
Adversarial Attacks Leverage Interference Between Features in Superposition
by: Stevinson, Edward, et al.
Published: (2025)
by: Stevinson, Edward, et al.
Published: (2025)
Can LLMs Deceive CLIP? Benchmarking Adversarial Compositionality of Pre-trained Multimodal Representation via Text Updates
by: Ahn, Jaewoo, et al.
Published: (2025)
by: Ahn, Jaewoo, et al.
Published: (2025)
Transferring Textual Preferences to Vision-Language Understanding through Model Merging
by: Li, Chen-An, et al.
Published: (2025)
by: Li, Chen-An, et al.
Published: (2025)
Transferable Adversarial Face Attack with Text Controlled Attribute
by: Li, Wenyun, et al.
Published: (2024)
by: Li, Wenyun, et al.
Published: (2024)
Exploring Transfer Learning in Medical Image Segmentation using Vision-Language Models
by: Poudel, Kanchan, et al.
Published: (2023)
by: Poudel, Kanchan, et al.
Published: (2023)
Enhancing CLIP Conceptual Embedding through Knowledge Distillation
by: Kao, Kuei-Chun
Published: (2024)
by: Kao, Kuei-Chun
Published: (2024)
Adversarial Semantic and Label Perturbation Attack for Pedestrian Attribute Recognition
by: Kong, Weizhe, et al.
Published: (2025)
by: Kong, Weizhe, et al.
Published: (2025)
SORA: Free Second-Order Attacks in Fast Adversarial Training
by: Teymourian, Mazdak, et al.
Published: (2026)
by: Teymourian, Mazdak, et al.
Published: (2026)
MobileCLIP2: Improving Multi-Modal Reinforced Training
by: Faghri, Fartash, et al.
Published: (2025)
by: Faghri, Fartash, et al.
Published: (2025)
Similar Items
-
Toward Universal and Transferable Jailbreak Attacks on Vision-Language Models
by: Cui, Kaiyuan, et al.
Published: (2026) -
Detecting Backdoor Samples in Contrastive Language Image Pretraining
by: Huang, Hanxun, et al.
Published: (2025) -
Towards Million-Scale Adversarial Robustness Evaluation With Stronger Individual Attacks
by: Xie, Yong, et al.
Published: (2024) -
Downstream Transfer Attack: Adversarial Attacks on Downstream Models with Pre-trained Vision Transformers
by: Zheng, Weijie, et al.
Published: (2024) -
LDReg: Local Dimensionality Regularized Self-Supervised Learning
by: Huang, Hanxun, et al.
Published: (2024)