Enhancing Adversarial Transferability with Adversarial Weight Tuning
Fuente:
arXiv
Guardado en:
| Autores principales: | Chen, Jiahao, Feng, Zhou, Zeng, Rui, Pu, Yuwen, Zhou, Chunyi, Jiang, Yi, Gan, Yuyou, Li, Jinbao, Ji, Shouling |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CAMH: Advancing Model Hijacking Attack in Machine Learning
por: He, Xing, et al.
Publicado: (2024)
por: He, Xing, et al.
Publicado: (2024)
Mellivora Capensis: A Backdoor-Free Training Framework on the Poisoned Dataset without Auxiliary Data
por: Pu, Yuwen, et al.
Publicado: (2024)
por: Pu, Yuwen, et al.
Publicado: (2024)
Unveiling the Security Risks of Federated Learning in the Wild: From Research to Practice
por: Chen, Jiahao, et al.
Publicado: (2026)
por: Chen, Jiahao, et al.
Publicado: (2026)
FreeTalk:A plug-and-play and black-box defense against speech synthesis attacks
por: Pu, Yuwen, et al.
Publicado: (2025)
por: Pu, Yuwen, et al.
Publicado: (2025)
Dullahan: Stealthy Backdoor Attack against Without-Label-Sharing Split Learning
por: Pu, Yuwen, et al.
Publicado: (2024)
por: Pu, Yuwen, et al.
Publicado: (2024)
Rethinking the Vulnerabilities of Face Recognition Systems:From a Practical Perspective
por: Chen, Jiahao, et al.
Publicado: (2024)
por: Chen, Jiahao, et al.
Publicado: (2024)
Poison in the Well: Feature Embedding Disruption in Backdoor Attacks
por: Feng, Zhou, et al.
Publicado: (2025)
por: Feng, Zhou, et al.
Publicado: (2025)
Profiling for Pennies: Unveiling the Privacy Iceberg of LLM Agents
por: Chen, Jiahao, et al.
Publicado: (2026)
por: Chen, Jiahao, et al.
Publicado: (2026)
UNIDOOR: A Universal Framework for Action-Level Backdoor Attacks in Deep Reinforcement Learning
por: Ma, Oubo, et al.
Publicado: (2025)
por: Ma, Oubo, et al.
Publicado: (2025)
Compiling Activation Steering into Weights via Null-Space Constraints for Stealthy Backdoors
por: Yin, Rui, et al.
Publicado: (2026)
por: Yin, Rui, et al.
Publicado: (2026)
Auditing M-LLMs for Privacy Risks: A Synthetic Benchmark and Evaluation Framework
por: Li, Junhao, et al.
Publicado: (2025)
por: Li, Junhao, et al.
Publicado: (2025)
CLIBE: Detecting Dynamic Backdoors in Transformer-based NLP Models
por: Zeng, Rui, et al.
Publicado: (2024)
por: Zeng, Rui, et al.
Publicado: (2024)
SUB-PLAY: Adversarial Policies against Partially Observed Multi-Agent Reinforcement Learning Systems
por: Ma, Oubo, et al.
Publicado: (2024)
por: Ma, Oubo, et al.
Publicado: (2024)
LoRAShield: Data-Free Editing Alignment for Secure Personalized LoRA Sharing
por: Chen, Jiahao, et al.
Publicado: (2025)
por: Chen, Jiahao, et al.
Publicado: (2025)
Enhancing Adversarial Attacks via Parameter Adaptive Adversarial Attack
por: Jin, Zhibo, et al.
Publicado: (2024)
por: Jin, Zhibo, et al.
Publicado: (2024)
AdvSQLi: Generating Adversarial SQL Injections against Real-world WAF-as-a-service
por: Qu, Zhenqing, et al.
Publicado: (2024)
por: Qu, Zhenqing, et al.
Publicado: (2024)
Boosting Adversarial Transferability with Spatial Adversarial Alignment
por: Chen, Zhaoyu, et al.
Publicado: (2025)
por: Chen, Zhaoyu, et al.
Publicado: (2025)
Quantization Aware Attack: Enhancing Transferable Adversarial Attacks by Model Quantization
por: Yang, Yulong, et al.
Publicado: (2023)
por: Yang, Yulong, et al.
Publicado: (2023)
TrapSuffix: Proactive Defense Against Adversarial Suffixes in Jailbreaking
por: Du, Mengyao, et al.
Publicado: (2026)
por: Du, Mengyao, et al.
Publicado: (2026)
Angel or Demon: Investigating the Plasticity Interventions' Impact on Backdoor Threats in Deep Reinforcement Learning
por: Ma, Oubo, et al.
Publicado: (2026)
por: Ma, Oubo, et al.
Publicado: (2026)
ArmSSL: Adversarial Robust Black-Box Watermarking for Self-Supervised Learning Pre-trained Encoders
por: Jiang, Yongqi, et al.
Publicado: (2026)
por: Jiang, Yongqi, et al.
Publicado: (2026)
NeuroBreak: Unveil Internal Jailbreak Mechanisms in Large Language Models
por: Zhang, Chuhan, et al.
Publicado: (2025)
por: Zhang, Chuhan, et al.
Publicado: (2025)
AEIOU: A Unified Defense Framework against NSFW Prompts in Text-to-Image Models
por: Wang, Yiming, et al.
Publicado: (2024)
por: Wang, Yiming, et al.
Publicado: (2024)
Defending against Adversarial Malware Attacks on ML-based Android Malware Detection Systems
por: He, Ping, et al.
Publicado: (2025)
por: He, Ping, et al.
Publicado: (2025)
Adversary-Aware DPO: Enhancing Safety Alignment in Vision Language Models via Adversarial Training
por: Weng, Fenghua, et al.
Publicado: (2025)
por: Weng, Fenghua, et al.
Publicado: (2025)
A Wolf in Sheep's Clothing: Practical Black-box Adversarial Attacks for Evading Learning-based Windows Malware Detection in the Wild
por: Ling, Xiang, et al.
Publicado: (2024)
por: Ling, Xiang, et al.
Publicado: (2024)
ReCIT: Reconstructing Full Private Data from Gradient in Parameter-Efficient Fine-Tuning of Large Language Models
por: Xie, Jin, et al.
Publicado: (2025)
por: Xie, Jin, et al.
Publicado: (2025)
IPIGuard: A Novel Tool Dependency Graph-Based Defense Against Indirect Prompt Injection in LLM Agents
por: An, Hengyu, et al.
Publicado: (2025)
por: An, Hengyu, et al.
Publicado: (2025)
Investigating Deep Watermark Security: An Adversarial Transferability Perspective
por: Qi, Biqing, et al.
Publicado: (2024)
por: Qi, Biqing, et al.
Publicado: (2024)
Foe for Fraud: Transferable Adversarial Attacks in Credit Card Fraud Detection
por: Fok, Jan Lum, et al.
Publicado: (2025)
por: Fok, Jan Lum, et al.
Publicado: (2025)
"No Matter What You Do": Purifying GNN Models via Backdoor Unlearning
por: Zhang, Jiale, et al.
Publicado: (2024)
por: Zhang, Jiale, et al.
Publicado: (2024)
SwitchPatch: Physical Adversarial Attack Strategy with Switchable Adversarial Objectives
por: Jiang, Hanrui, et al.
Publicado: (2025)
por: Jiang, Hanrui, et al.
Publicado: (2025)
Transferable Adversarial Examples with Bayes Approach
por: Fan, Mingyuan, et al.
Publicado: (2022)
por: Fan, Mingyuan, et al.
Publicado: (2022)
Hijack Vertical Federated Learning Models As One Party
por: Qiu, Pengyu, et al.
Publicado: (2022)
por: Qiu, Pengyu, et al.
Publicado: (2022)
Transferability Bound Theory: Exploring Relationship between Adversarial Transferability and Flatness
por: Fan, Mingyuan, et al.
Publicado: (2023)
por: Fan, Mingyuan, et al.
Publicado: (2023)
Enhancing TinyML Security: Study of Adversarial Attack Transferability
por: Shah, Parin, et al.
Publicado: (2024)
por: Shah, Parin, et al.
Publicado: (2024)
LRS: Enhancing Adversarial Transferability through Lipschitz Regularized Surrogate
por: Wu, Tao, et al.
Publicado: (2023)
por: Wu, Tao, et al.
Publicado: (2023)
One Prompt to Verify Your Models: Black-Box Text-to-Image Models Verification via Non-Transferable Adversarial Attacks
por: Guo, Ji, et al.
Publicado: (2024)
por: Guo, Ji, et al.
Publicado: (2024)
Transferability Ranking of Adversarial Examples
por: Levy, Mosh, et al.
Publicado: (2022)
por: Levy, Mosh, et al.
Publicado: (2022)
Enhancing Adversarial Attacks: The Similar Target Method
por: Zhang, Shuo, et al.
Publicado: (2023)
por: Zhang, Shuo, et al.
Publicado: (2023)
Ejemplares similares
-
CAMH: Advancing Model Hijacking Attack in Machine Learning
por: He, Xing, et al.
Publicado: (2024) -
Mellivora Capensis: A Backdoor-Free Training Framework on the Poisoned Dataset without Auxiliary Data
por: Pu, Yuwen, et al.
Publicado: (2024) -
Unveiling the Security Risks of Federated Learning in the Wild: From Research to Practice
por: Chen, Jiahao, et al.
Publicado: (2026) -
FreeTalk:A plug-and-play and black-box defense against speech synthesis attacks
por: Pu, Yuwen, et al.
Publicado: (2025) -
Dullahan: Stealthy Backdoor Attack against Without-Label-Sharing Split Learning
por: Pu, Yuwen, et al.
Publicado: (2024)