Why Does Little Robustness Help? A Further Step Towards Understanding Adversarial Transferability
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yechao, Hu, Shengshan, Zhang, Leo Yu, Shi, Junyu, Li, Minghui, Liu, Xiaogeng, Wan, Wei, Jin, Hai |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Why does weak-OOD help? A Further Step Towards Understanding Jailbreaking VLMs
by: Zhou, Yuxuan, et al.
Published: (2025)
by: Zhou, Yuxuan, et al.
Published: (2025)
Secure Transfer Learning: Training Clean Models Against Backdoor in (Both) Pre-trained Encoders and Downstream Datasets
by: Zhang, Yechao, et al.
Published: (2025)
by: Zhang, Yechao, et al.
Published: (2025)
Spa-VLM: Stealthy Poisoning Attacks on RAG-based VLM
by: Yu, Lei, et al.
Published: (2025)
by: Yu, Lei, et al.
Published: (2025)
ECLIPSE: Expunging Clean-label Indiscriminate Poisons via Sparse Diffusion Purification
by: Wang, Xianlong, et al.
Published: (2024)
by: Wang, Xianlong, et al.
Published: (2024)
"MCP Does Not Stand for Misuse Cryptography Protocol": Uncovering Cryptographic Misuse in Model Context Protocol at Scale
by: Yan, Biwei, et al.
Published: (2025)
by: Yan, Biwei, et al.
Published: (2025)
MARS: A Malignity-Aware Backdoor Defense in Federated Learning
by: Wan, Wei, et al.
Published: (2025)
by: Wan, Wei, et al.
Published: (2025)
Masked Language Model Based Textual Adversarial Example Detection
by: Zhang, Xiaomei, et al.
Published: (2023)
by: Zhang, Xiaomei, et al.
Published: (2023)
Towards Model Extraction Attacks in GAN-Based Image Translation via Domain Shift Mitigation
by: Mi, Di, et al.
Published: (2024)
by: Mi, Di, et al.
Published: (2024)
DarkFed: A Data-Free Backdoor Attack in Federated Learning
by: Li, Minghui, et al.
Published: (2024)
by: Li, Minghui, et al.
Published: (2024)
Low Rank Comes with Low Security: Gradient Assembly Poisoning Attacks against Distributed LoRA-based LLM Systems
by: Dong, Yueyan, et al.
Published: (2026)
by: Dong, Yueyan, et al.
Published: (2026)
Give Them an Inch and They Will Take a Mile:Understanding and Measuring Caller Identity Confusion in MCP-Based AI Systems
by: Huang, Yuhang, et al.
Published: (2026)
by: Huang, Yuhang, et al.
Published: (2026)
ReasoningBomb: A Stealthy Denial-of-Service Attack by Inducing Pathologically Long Reasoning in Large Reasoning Models
by: Liu, Xiaogeng, et al.
Published: (2026)
by: Liu, Xiaogeng, et al.
Published: (2026)
Mind Your HEARTBEAT! Claw Background Execution Inherently Enables Silent Memory Pollution
by: Zhang, Yechao, et al.
Published: (2026)
by: Zhang, Yechao, et al.
Published: (2026)
Take a Step Further: Understanding Page Spray in Linux Kernel Exploitation
by: Guo, Ziyi, et al.
Published: (2024)
by: Guo, Ziyi, et al.
Published: (2024)
Adversarial Example Based Fingerprinting for Robust Copyright Protection in Split Learning
by: Lin, Zhangting, et al.
Published: (2025)
by: Lin, Zhangting, et al.
Published: (2025)
Don't Listen To Me: Understanding and Exploring Jailbreak Prompts of Large Language Models
by: Yu, Zhiyuan, et al.
Published: (2024)
by: Yu, Zhiyuan, et al.
Published: (2024)
Revisiting Gradient Pruning: A Dual Realization for Defending against Gradient Attacks
by: Xue, Lulu, et al.
Published: (2024)
by: Xue, Lulu, et al.
Published: (2024)
Attention Masks Help Adversarial Attacks to Bypass Safety Detectors
by: Shi, Yunfan
Published: (2024)
by: Shi, Yunfan
Published: (2024)
AutoDAN-Reasoning: Enhancing Strategies Exploration based Jailbreak Attacks with Test-Time Scaling
by: Liu, Xiaogeng, et al.
Published: (2025)
by: Liu, Xiaogeng, et al.
Published: (2025)
Large Language Model Watermark Stealing With Mixed Integer Programming
by: Zhang, Zhaoxi, et al.
Published: (2024)
by: Zhang, Zhaoxi, et al.
Published: (2024)
UnlearnShield: Shielding Forgotten Privacy against Unlearning Inversion
by: Xue, Lulu, et al.
Published: (2026)
by: Xue, Lulu, et al.
Published: (2026)
Robustness Over Time: Understanding Adversarial Examples' Effectiveness on Longitudinal Versions of Large Language Models
by: Liu, Yugeng, et al.
Published: (2023)
by: Liu, Yugeng, et al.
Published: (2023)
Transferable Backdoor Attacks for Code Models via Sharpness-Aware Adversarial Perturbation
by: Chang, Shuyu, et al.
Published: (2026)
by: Chang, Shuyu, et al.
Published: (2026)
DMS: Addressing Information Loss with More Steps for Pragmatic Adversarial Attacks
by: Zhu, Zhiyu, et al.
Published: (2024)
by: Zhu, Zhiyu, et al.
Published: (2024)
Denial-of-Service or Fine-Grained Control: Towards Flexible Model Poisoning Attacks on Federated Learning
by: Zhang, Hangtao, et al.
Published: (2023)
by: Zhang, Hangtao, et al.
Published: (2023)
Exploring the Robustness and Transferability of Patch-Based Adversarial Attacks in Quantized Neural Networks
by: Guesmi, Amira, et al.
Published: (2024)
by: Guesmi, Amira, et al.
Published: (2024)
Prediction Inconsistency Helps Achieve Generalizable Detection of Adversarial Examples
by: Han, Sicong, et al.
Published: (2025)
by: Han, Sicong, et al.
Published: (2025)
Enhancing Adversarial Transferability with Adversarial Weight Tuning
by: Chen, Jiahao, et al.
Published: (2024)
by: Chen, Jiahao, et al.
Published: (2024)
Toward Realistic Adversarial Attacks in IDS: A Novel Feasibility Metric for Transferability
by: Ennaji, Sabrine, et al.
Published: (2025)
by: Ennaji, Sabrine, et al.
Published: (2025)
Understanding Help Seeking for Digital Privacy, Safety, and Security
by: Thomas, Kurt, et al.
Published: (2026)
by: Thomas, Kurt, et al.
Published: (2026)
Adaptive Randomized Smoothing: Certified Adversarial Robustness for Multi-Step Defences
by: Lyu, Saiyue, et al.
Published: (2024)
by: Lyu, Saiyue, et al.
Published: (2024)
OET: Optimization-based prompt injection Evaluation Toolkit
by: Pan, Jinsheng, et al.
Published: (2025)
by: Pan, Jinsheng, et al.
Published: (2025)
RePD: Defending Jailbreak Attack through a Retrieval-based Prompt Decomposition Process
by: Wang, Peiran, et al.
Published: (2024)
by: Wang, Peiran, et al.
Published: (2024)
InjecGuard: Benchmarking and Mitigating Over-defense in Prompt Injection Guardrail Models
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
RerouteGuard: Understanding and Mitigating Adversarial Risks for LLM Routing
by: Zhang, Wenhui, et al.
Published: (2026)
by: Zhang, Wenhui, et al.
Published: (2026)
Explainable and Transferable Adversarial Attack for ML-Based Network Intrusion Detectors
by: Zhang, Hangsheng, et al.
Published: (2024)
by: Zhang, Hangsheng, et al.
Published: (2024)
Adversarially Robust Assembly Language Model for Packed Executables Detection
by: Li, Shijia, et al.
Published: (2025)
by: Li, Shijia, et al.
Published: (2025)
Understanding Sensitivity of Differential Attention through the Lens of Adversarial Robustness
by: Takahashi, Tsubasa, et al.
Published: (2025)
by: Takahashi, Tsubasa, et al.
Published: (2025)
Revisiting the Robust Alignment of Circuit Breakers
by: Schwinn, Leo, et al.
Published: (2024)
by: Schwinn, Leo, et al.
Published: (2024)
Understanding and Characterizing Obfuscated Funds Transfers in Ethereum Smart Contracts
by: Sheng, Zhang, et al.
Published: (2025)
by: Sheng, Zhang, et al.
Published: (2025)
Similar Items
-
Why does weak-OOD help? A Further Step Towards Understanding Jailbreaking VLMs
by: Zhou, Yuxuan, et al.
Published: (2025) -
Secure Transfer Learning: Training Clean Models Against Backdoor in (Both) Pre-trained Encoders and Downstream Datasets
by: Zhang, Yechao, et al.
Published: (2025) -
Spa-VLM: Stealthy Poisoning Attacks on RAG-based VLM
by: Yu, Lei, et al.
Published: (2025) -
ECLIPSE: Expunging Clean-label Indiscriminate Poisons via Sparse Diffusion Purification
by: Wang, Xianlong, et al.
Published: (2024) -
"MCP Does Not Stand for Misuse Cryptography Protocol": Uncovering Cryptographic Misuse in Model Context Protocol at Scale
by: Yan, Biwei, et al.
Published: (2025)