How to Enhance Downstream Adversarial Robustness (almost) without Touching the Pre-Trained Foundation Model?
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Meiqi, Huang, Zhuoqun, Xing, Yue |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adversarial Attacks on Downstream Weather Forecasting Models: Application to Tropical Cyclone Trajectory Prediction
by: Deng, Yue, et al.
Published: (2025)
by: Deng, Yue, et al.
Published: (2025)
Attacking Attention of Foundation Models Disrupts Downstream Tasks
by: Silva, Hondamunige Prasanna, et al.
Published: (2025)
by: Silva, Hondamunige Prasanna, et al.
Published: (2025)
Secure Transfer Learning: Training Clean Models Against Backdoor in (Both) Pre-trained Encoders and Downstream Datasets
by: Zhang, Yechao, et al.
Published: (2025)
by: Zhang, Yechao, et al.
Published: (2025)
Vulnerability-Aware Robust Multimodal Adversarial Training
by: Zhang, Junrui, et al.
Published: (2025)
by: Zhang, Junrui, et al.
Published: (2025)
Robustness-Congruent Adversarial Training for Secure Machine Learning Model Updates
by: Angioni, Daniele, et al.
Published: (2024)
by: Angioni, Daniele, et al.
Published: (2024)
Pre-trained Encoder Inference: Revealing Upstream Encoders In Downstream Machine Learning Services
by: Fu, Shaopeng, et al.
Published: (2024)
by: Fu, Shaopeng, et al.
Published: (2024)
Stealing the Invisible: Unveiling Pre-Trained CNN Models through Adversarial Examples and Timing Side-Channels
by: Shukla, Shubhi, et al.
Published: (2024)
by: Shukla, Shubhi, et al.
Published: (2024)
Effect of Ambient-Intrinsic Dimension Gap on Adversarial Vulnerability
by: Haldar, Rajdeep, et al.
Published: (2024)
by: Haldar, Rajdeep, et al.
Published: (2024)
Embedding Hidden Adversarial Capabilities in Pre-Trained Diffusion Models
by: Beerens, Lucas, et al.
Published: (2025)
by: Beerens, Lucas, et al.
Published: (2025)
RS-Del: Edit Distance Robustness Certificates for Sequence Classifiers via Randomized Deletion
by: Huang, Zhuoqun, et al.
Published: (2023)
by: Huang, Zhuoqun, et al.
Published: (2023)
Mitigating Downstream Model Risks via Model Provenance
by: Wang, Keyu, et al.
Published: (2024)
by: Wang, Keyu, et al.
Published: (2024)
Adaptive Meta-learning-based Adversarial Training for Robust Automatic Modulation Classification
by: Bamdad, Amirmohammad, et al.
Published: (2025)
by: Bamdad, Amirmohammad, et al.
Published: (2025)
CERT-ED: Certifiably Robust Text Classification for Edit Distance
by: Huang, Zhuoqun, et al.
Published: (2024)
by: Huang, Zhuoqun, et al.
Published: (2024)
Transferable Adversarial Attacks on SAM and Its Downstream Models
by: Xia, Song, et al.
Published: (2024)
by: Xia, Song, et al.
Published: (2024)
AdaptDel: Adaptable Deletion Rate Randomized Smoothing for Certified Robustness
by: Huang, Zhuoqun, et al.
Published: (2025)
by: Huang, Zhuoqun, et al.
Published: (2025)
PromptRobust: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts
by: Zhu, Kaijie, et al.
Published: (2023)
by: Zhu, Kaijie, et al.
Published: (2023)
Adversarial Robustness of Link Sign Prediction in Signed Graphs
by: Zhou, Jialong, et al.
Published: (2024)
by: Zhou, Jialong, et al.
Published: (2024)
Enhancing Adversarial Attacks via Parameter Adaptive Adversarial Attack
by: Jin, Zhibo, et al.
Published: (2024)
by: Jin, Zhibo, et al.
Published: (2024)
Low-Cost Hard-Label Adversarial Attack with Theoretical Foundations
by: Liu, Jun, et al.
Published: (2026)
by: Liu, Jun, et al.
Published: (2026)
DeMem: Privacy-Enhanced Robust Adversarial Learning via De-Memorization
by: Luo, Xiaoyu, et al.
Published: (2024)
by: Luo, Xiaoyu, et al.
Published: (2024)
Elevating Defenses: Bridging Adversarial Training and Watermarking for Model Resilience
by: Thakkar, Janvi, et al.
Published: (2023)
by: Thakkar, Janvi, et al.
Published: (2023)
When and How to Fool Explainable Models (and Humans) with Adversarial Examples
by: Vadillo, Jon, et al.
Published: (2021)
by: Vadillo, Jon, et al.
Published: (2021)
Enabling Adversarial Robustness in AI Models through Kubeflow MLOps
by: Bouras, Stavros, et al.
Published: (2026)
by: Bouras, Stavros, et al.
Published: (2026)
Introducing Adaptive Continuous Adversarial Training (ACAT) to Enhance ML Robustness
by: elShehaby, Mohamed, et al.
Published: (2024)
by: elShehaby, Mohamed, et al.
Published: (2024)
Generalization Properties of Adversarial Training for $\ell_0$-Bounded Adversarial Attacks
by: Delgosha, Payam, et al.
Published: (2024)
by: Delgosha, Payam, et al.
Published: (2024)
Privacy without Noisy Gradients: Slicing Mechanism for Generative Model Training
by: Greenewald, Kristjan, et al.
Published: (2024)
by: Greenewald, Kristjan, et al.
Published: (2024)
On the Effectiveness of Adversarial Training on Malware Classifiers
by: Bostani, Hamid, et al.
Published: (2024)
by: Bostani, Hamid, et al.
Published: (2024)
A Systematic Study of Model Extraction Attacks on Graph Foundation Models
by: Xu, Haoyan, et al.
Published: (2025)
by: Xu, Haoyan, et al.
Published: (2025)
Robustness Against Adversarial Attacks via Learning Confined Adversarial Polytopes
by: Hamidi, Shayan Mohajer, et al.
Published: (2024)
by: Hamidi, Shayan Mohajer, et al.
Published: (2024)
Enhancing Adversarial Robustness in Network Intrusion Detection: A Layer-wise Adaptive Regularization Approach
by: Nasir, Hira, et al.
Published: (2026)
by: Nasir, Hira, et al.
Published: (2026)
On the Robustness of Malware Detectors to Adversarial Samples
by: Salman, Muhammad, et al.
Published: (2024)
by: Salman, Muhammad, et al.
Published: (2024)
Steering Beyond the Support: Adversarial Training on Unsupervised Jailbroken Activation Simulation
by: Chen, Luoyu, et al.
Published: (2026)
by: Chen, Luoyu, et al.
Published: (2026)
Mitigating Error Amplification in Fast Adversarial Training
by: Zhao, Mengnan, et al.
Published: (2026)
by: Zhao, Mengnan, et al.
Published: (2026)
Efficient Adversarial Training in LLMs with Continuous Attacks
by: Xhonneux, Sophie, et al.
Published: (2024)
by: Xhonneux, Sophie, et al.
Published: (2024)
Winter Soldier: Backdooring Language Models at Pre-Training with Indirect Data Poisoning
by: Bouaziz, Wassim, et al.
Published: (2025)
by: Bouaziz, Wassim, et al.
Published: (2025)
ProDiF: Protecting Domain-Invariant Features to Secure Pre-Trained Models Against Extraction
by: Zhou, Tong, et al.
Published: (2025)
by: Zhou, Tong, et al.
Published: (2025)
CARE: Ensemble Adversarial Robustness Evaluation Against Adaptive Attackers for Security Applications
by: Zhang, Hangsheng, et al.
Published: (2024)
by: Zhang, Hangsheng, et al.
Published: (2024)
Trading Inference-Time Compute for Adversarial Robustness
by: Zaremba, Wojciech, et al.
Published: (2025)
by: Zaremba, Wojciech, et al.
Published: (2025)
The Surprising Harmfulness of Benign Overfitting for Adversarial Robustness
by: Hao, Yifan, et al.
Published: (2024)
by: Hao, Yifan, et al.
Published: (2024)
IDEA: Invariant Defense for Graph Adversarial Robustness
by: Tao, Shuchang, et al.
Published: (2023)
by: Tao, Shuchang, et al.
Published: (2023)
Similar Items
-
Adversarial Attacks on Downstream Weather Forecasting Models: Application to Tropical Cyclone Trajectory Prediction
by: Deng, Yue, et al.
Published: (2025) -
Attacking Attention of Foundation Models Disrupts Downstream Tasks
by: Silva, Hondamunige Prasanna, et al.
Published: (2025) -
Secure Transfer Learning: Training Clean Models Against Backdoor in (Both) Pre-trained Encoders and Downstream Datasets
by: Zhang, Yechao, et al.
Published: (2025) -
Vulnerability-Aware Robust Multimodal Adversarial Training
by: Zhang, Junrui, et al.
Published: (2025) -
Robustness-Congruent Adversarial Training for Secure Machine Learning Model Updates
by: Angioni, Daniele, et al.
Published: (2024)