Zero-Sacrifice Persistent-Robustness Adversarial Defense for Pre-Trained Encoders
Fuente:
arXiv
Saved in:
| Main Authors: | Lei, Zhuxin, Yang, Ziyuan, Zhang, Yi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PrismAgent: Illuminating Harm in Memes via a Zero-Shot Interpretable Multi-Agent Framework
by: Ding, Zihan, et al.
Published: (2026)
by: Ding, Zihan, et al.
Published: (2026)
Toward a Graph Foundation Model: Pre-Training Transformers With Random Walks
by: Tang, Ziyuan, et al.
Published: (2025)
by: Tang, Ziyuan, et al.
Published: (2025)
From Static Analysis to Audience Dissemination: A Training-Free Multimodal Controversy Detection Multi-Agent Framework
by: Ding, Zihan, et al.
Published: (2026)
by: Ding, Zihan, et al.
Published: (2026)
Adversarial Training for Defense Against Label Poisoning Attacks
by: Bal, Melis Ilayda, et al.
Published: (2025)
by: Bal, Melis Ilayda, et al.
Published: (2025)
Latent Adversarial Training Improves Robustness to Persistent Harmful Behaviors in LLMs
by: Sheshadri, Abhay, et al.
Published: (2024)
by: Sheshadri, Abhay, et al.
Published: (2024)
Trained Persistent Memory for Frozen Encoder--Decoder LLMs: Six Architectural Methods
by: Jeong, Hong
Published: (2026)
by: Jeong, Hong
Published: (2026)
Targeted Downstream-Agnostic Attack
by: Lei, Zhuxin, et al.
Published: (2026)
by: Lei, Zhuxin, et al.
Published: (2026)
Ignition Phase : Standard Training for Fast Adversarial Robustness
by: Yu-Hang, Wang, et al.
Published: (2025)
by: Yu-Hang, Wang, et al.
Published: (2025)
Zero-Shot Reinforcement Learning via Function Encoders
by: Ingebrand, Tyler, et al.
Published: (2024)
by: Ingebrand, Tyler, et al.
Published: (2024)
Zero-shot Meta-learning for Tabular Prediction Tasks with Adversarially Pre-trained Transformer
by: Wu, Yulun, et al.
Published: (2025)
by: Wu, Yulun, et al.
Published: (2025)
Towards Robust Policy: Enhancing Offline Reinforcement Learning with Adversarial Attacks and Defenses
by: Nguyen, Thanh, et al.
Published: (2024)
by: Nguyen, Thanh, et al.
Published: (2024)
SeqFusion: Sequential Fusion of Pre-Trained Models for Zero-Shot Time-Series Forecasting
by: Huang, Ting-Ji, et al.
Published: (2025)
by: Huang, Ting-Ji, et al.
Published: (2025)
Defense without Forgetting: Continual Adversarial Defense with Anisotropic & Isotropic Pseudo Replay
by: Zhou, Yuhang, et al.
Published: (2024)
by: Zhou, Yuhang, et al.
Published: (2024)
Explanation-Guided Adversarial Training for Robust and Interpretable Models
by: Chen, Chao, et al.
Published: (2026)
by: Chen, Chao, et al.
Published: (2026)
Regret-Based Defense in Adversarial Reinforcement Learning
by: Belaire, Roman, et al.
Published: (2023)
by: Belaire, Roman, et al.
Published: (2023)
Towards Interpretability Without Sacrifice: Faithful Dense Layer Decomposition with Mixture of Decoders
by: Oldfield, James, et al.
Published: (2025)
by: Oldfield, James, et al.
Published: (2025)
NPAT Null-Space Projected Adversarial Training Towards Zero Deterioration
by: Hu, Hanyi, et al.
Published: (2024)
by: Hu, Hanyi, et al.
Published: (2024)
Adversarial Instance Generation and Robust Training for Neural Combinatorial Optimization with Multiple Objectives
by: Liu, Wei, et al.
Published: (2026)
by: Liu, Wei, et al.
Published: (2026)
Adversarial Training for Robust Coverage Network under Worst-case Facility Losses
by: Miao, Changhao, et al.
Published: (2026)
by: Miao, Changhao, et al.
Published: (2026)
Adversarial Reinforcement Learning for Offensive and Defensive Agents in a Simulated Zero-Sum Network Environment
by: Shahid, Abrar, et al.
Published: (2025)
by: Shahid, Abrar, et al.
Published: (2025)
Adversarial Robustness in Financial Machine Learning: Defenses, Economic Impact, and Governance Evidence
by: Baviskar, Samruddhi
Published: (2025)
by: Baviskar, Samruddhi
Published: (2025)
Ensemble of Pre-Trained Models for Long-Tailed Trajectory Prediction
by: Thuremella, Divya, et al.
Published: (2025)
by: Thuremella, Divya, et al.
Published: (2025)
Mitigating the Structural Bias in Graph Adversarial Defenses
by: Fang, Junyuan, et al.
Published: (2025)
by: Fang, Junyuan, et al.
Published: (2025)
Adversarial Latent-State Training for Robust Policies in Partially Observable Domains
by: Ahuja, Angad Singh
Published: (2026)
by: Ahuja, Angad Singh
Published: (2026)
Bridging Symmetry and Robustness: On the Role of Equivariance in Enhancing Adversarial Robustness
by: Wang, Longwei, et al.
Published: (2025)
by: Wang, Longwei, et al.
Published: (2025)
Towards Reliable Evaluation of Adversarial Robustness for Spiking Neural Networks
by: Wang, Jihang, et al.
Published: (2025)
by: Wang, Jihang, et al.
Published: (2025)
Provably Robust Pre-Trained Ensembles for Biomarker-Based Cancer Classification
by: Lee, Chongmin, et al.
Published: (2024)
by: Lee, Chongmin, et al.
Published: (2024)
Embedding Hidden Adversarial Capabilities in Pre-Trained Diffusion Models
by: Beerens, Lucas, et al.
Published: (2025)
by: Beerens, Lucas, et al.
Published: (2025)
Language-Driven Anchors for Zero-Shot Adversarial Robustness
by: Li, Xiao, et al.
Published: (2023)
by: Li, Xiao, et al.
Published: (2023)
Adversarial Training for Process Reward Models
by: Juneja, Gurusha, et al.
Published: (2025)
by: Juneja, Gurusha, et al.
Published: (2025)
SMILE: Zero-Shot Sparse Mixture of Low-Rank Experts Construction From Pre-Trained Foundation Models
by: Tang, Anke, et al.
Published: (2024)
by: Tang, Anke, et al.
Published: (2024)
The Spectral Lifecycle of Transformer Training: Transient Compression Waves, Persistent Spectral Gradients, and the Q/K--V Asymmetry
by: Liu, Yi
Published: (2026)
by: Liu, Yi
Published: (2026)
Adversarial Training: A Survey
by: Zhao, Mengnan, et al.
Published: (2024)
by: Zhao, Mengnan, et al.
Published: (2024)
Rethinking Invariance Regularization in Adversarial Training to Improve Robustness-Accuracy Trade-off
by: Waseda, Futa, et al.
Published: (2024)
by: Waseda, Futa, et al.
Published: (2024)
Robust Deep Reinforcement Learning Through Adversarial Attacks and Training : A Survey
by: Schott, Lucas, et al.
Published: (2024)
by: Schott, Lucas, et al.
Published: (2024)
FedProphet: Memory-Efficient Federated Adversarial Training via Robust and Consistent Cascade Learning
by: Tang, Minxue, et al.
Published: (2024)
by: Tang, Minxue, et al.
Published: (2024)
Bridging Models to Defend: A Population-Based Strategy for Robust Adversarial Defense
by: Wang, Ren, et al.
Published: (2023)
by: Wang, Ren, et al.
Published: (2023)
MaskPure: Improving Defense Against Text Adversaries with Stochastic Purification
by: Gietz, Harrison, et al.
Published: (2024)
by: Gietz, Harrison, et al.
Published: (2024)
Benign Overfitting in Adversarial Training for Vision Transformers
by: Zhang, Jiaming, et al.
Published: (2026)
by: Zhang, Jiaming, et al.
Published: (2026)
Can You Break RLVER? Probing Adversarial Robustness of RL-Trained Empathetic Agents
by: K, Deeraj S, et al.
Published: (2026)
by: K, Deeraj S, et al.
Published: (2026)
Similar Items
-
PrismAgent: Illuminating Harm in Memes via a Zero-Shot Interpretable Multi-Agent Framework
by: Ding, Zihan, et al.
Published: (2026) -
Toward a Graph Foundation Model: Pre-Training Transformers With Random Walks
by: Tang, Ziyuan, et al.
Published: (2025) -
From Static Analysis to Audience Dissemination: A Training-Free Multimodal Controversy Detection Multi-Agent Framework
by: Ding, Zihan, et al.
Published: (2026) -
Adversarial Training for Defense Against Label Poisoning Attacks
by: Bal, Melis Ilayda, et al.
Published: (2025) -
Latent Adversarial Training Improves Robustness to Persistent Harmful Behaviors in LLMs
by: Sheshadri, Abhay, et al.
Published: (2024)