Adapting to Evolving Adversaries with Regularized Continual Robust Training
Fuente:
arXiv
Saved in:
| Main Authors: | Dai, Sihui, Cianfarani, Christian, Bhagoji, Arjun, Sehwag, Vikash, Mittal, Prateek |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Position: Towards Resilience Against Adversarial Examples
by: Dai, Sihui, et al.
Published: (2024)
by: Dai, Sihui, et al.
Published: (2024)
A New Linear Scaling Rule for Private Adaptive Hyperparameter Optimization
by: Panda, Ashwinee, et al.
Published: (2022)
by: Panda, Ashwinee, et al.
Published: (2022)
Towards Scalable and Robust Model Versioning
by: Ding, Wenxin, et al.
Published: (2024)
by: Ding, Wenxin, et al.
Published: (2024)
PatchDEMUX: A Certifiably Robust Framework for Multi-label Classifiers Against Adversarial Patches
by: Jacob, Dennis, et al.
Published: (2025)
by: Jacob, Dennis, et al.
Published: (2025)
Adaptive and Stratified Subsampling for High-Dimensional Robust Estimation
by: Mittal, Prateek, et al.
Published: (2024)
by: Mittal, Prateek, et al.
Published: (2024)
Stretching Each Dollar: Diffusion Training from Scratch on a Micro-Budget
by: Sehwag, Vikash, et al.
Published: (2024)
by: Sehwag, Vikash, et al.
Published: (2024)
Data Shapley in One Training Run
by: Wang, Jiachen T., et al.
Published: (2024)
by: Wang, Jiachen T., et al.
Published: (2024)
Capturing the Temporal Dependence of Training Data Influence
by: Wang, Jiachen T., et al.
Published: (2024)
by: Wang, Jiachen T., et al.
Published: (2024)
PropensityBench: Evaluating Latent Safety Risks in Large Language Models via an Agentic Approach
by: Sehwag, Udari Madhushani, et al.
Published: (2025)
by: Sehwag, Udari Madhushani, et al.
Published: (2025)
MYCROFT: Towards Effective and Efficient External Data Augmentation
by: Sarwar, Zain, et al.
Published: (2024)
by: Sarwar, Zain, et al.
Published: (2024)
Regularization for Adversarial Robust Learning
by: Wang, Jie, et al.
Published: (2024)
by: Wang, Jie, et al.
Published: (2024)
Rethinking Invariance Regularization in Adversarial Training to Improve Robustness-Accuracy Trade-off
by: Waseda, Futa, et al.
Published: (2024)
by: Waseda, Futa, et al.
Published: (2024)
Does More Inference-Time Compute Really Help Robustness?
by: Wu, Tong, et al.
Published: (2025)
by: Wu, Tong, et al.
Published: (2025)
Bhargava Cube--Inspired Quadratic Regularization for Structured Neural Embeddings
by: Sairam, S, et al.
Published: (2025)
by: Sairam, S, et al.
Published: (2025)
Adversarial Robustness of VAEs across Intersectional Subgroups
by: Ramanaik, Chethan Krishnamurthy, et al.
Published: (2024)
by: Ramanaik, Chethan Krishnamurthy, et al.
Published: (2024)
PatchCURE: Improving Certifiable Robustness, Model Utility, and Computation Efficiency of Adversarial Patch Defenses
by: Xiang, Chong, et al.
Published: (2023)
by: Xiang, Chong, et al.
Published: (2023)
AdvBDGen: Adversarially Fortified Prompt-Specific Fuzzy Backdoor Generator Against LLM Alignment
by: Pathmanathan, Pankayaraj, et al.
Published: (2024)
by: Pathmanathan, Pankayaraj, et al.
Published: (2024)
Optimal Transport Regularized Divergences: Application to Adversarial Robustness
by: Birrell, Jeremiah, et al.
Published: (2023)
by: Birrell, Jeremiah, et al.
Published: (2023)
Efficient Data Shapley for Weighted Nearest Neighbor Algorithms
by: Wang, Jiachen T., et al.
Published: (2024)
by: Wang, Jiachen T., et al.
Published: (2024)
Silencing Empowerment, Allowing Bigotry: Auditing the Moderation of Hate Speech on Twitch
by: Shukla, Prarabdh, et al.
Published: (2025)
by: Shukla, Prarabdh, et al.
Published: (2025)
Self-Comparison for Dataset-Level Membership Inference in Large (Vision-)Language Models
by: Ren, Jie, et al.
Published: (2024)
by: Ren, Jie, et al.
Published: (2024)
Certifiably Robust RAG against Retrieval Corruption
by: Xiang, Chong, et al.
Published: (2024)
by: Xiang, Chong, et al.
Published: (2024)
On the Robustness of Adversarial Training Against Uncertainty Attacks
by: Ledda, Emanuele, et al.
Published: (2024)
by: Ledda, Emanuele, et al.
Published: (2024)
How to Trace Latent Generative Model Generated Images without Artificial Watermark?
by: Wang, Zhenting, et al.
Published: (2024)
by: Wang, Zhenting, et al.
Published: (2024)
Adapting to Fragmented and Evolving Data: A Fisher Information Perspective
by: Khan, Behraj, et al.
Published: (2025)
by: Khan, Behraj, et al.
Published: (2025)
Holistic Adversarially Robust Pruning
by: Zhao, Qi, et al.
Published: (2024)
by: Zhao, Qi, et al.
Published: (2024)
JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models
by: Chao, Patrick, et al.
Published: (2024)
by: Chao, Patrick, et al.
Published: (2024)
Transfer Orthology Networks
by: Singh, Vikash
Published: (2025)
by: Singh, Vikash
Published: (2025)
Contrastive Adversarial Training for Unsupervised Domain Adaptation
by: Chen, Jiahong, et al.
Published: (2024)
by: Chen, Jiahong, et al.
Published: (2024)
Introducing Adaptive Continuous Adversarial Training (ACAT) to Enhance ML Robustness
by: elShehaby, Mohamed, et al.
Published: (2024)
by: elShehaby, Mohamed, et al.
Published: (2024)
Maintaining Adversarial Robustness in Continuous Learning
by: Ru, Xiaolei, et al.
Published: (2024)
by: Ru, Xiaolei, et al.
Published: (2024)
Can Small Training Runs Reliably Guide Data Curation? Rethinking Proxy-Model Practice
by: Wang, Jiachen T., et al.
Published: (2025)
by: Wang, Jiachen T., et al.
Published: (2025)
Truly Adapting to Adversarial Constraints in Constrained MABs
by: Stradi, Francesco Emanuele, et al.
Published: (2026)
by: Stradi, Francesco Emanuele, et al.
Published: (2026)
Exploration by Optimization with Hybrid Regularizers: Logarithmic Regret with Adversarial Robustness in Partial Monitoring
by: Tsuchiya, Taira, et al.
Published: (2024)
by: Tsuchiya, Taira, et al.
Published: (2024)
GCAL: Adapting Graph Models to Evolving Domain Shifts
by: Qiao, Ziyue, et al.
Published: (2025)
by: Qiao, Ziyue, et al.
Published: (2025)
GReAT: A Graph Regularized Adversarial Training Method
by: Bayram, Samet, et al.
Published: (2023)
by: Bayram, Samet, et al.
Published: (2023)
Adversarial Déjà Vu: Jailbreak Dictionary Learning for Stronger Generalization to Unseen Attacks
by: Dabas, Mahavir, et al.
Published: (2025)
by: Dabas, Mahavir, et al.
Published: (2025)
Hybrid Quantum-Classical GANs for the Generation of Adversarial Network Flows
by: Paudel, Prateek, et al.
Published: (2026)
by: Paudel, Prateek, et al.
Published: (2026)
Towards Robust Continual Learning with Bayesian Adaptive Moment Regularization
by: Foster, Jack, et al.
Published: (2023)
by: Foster, Jack, et al.
Published: (2023)
Efficient Adversarial Training in LLMs with Continuous Attacks
by: Xhonneux, Sophie, et al.
Published: (2024)
by: Xhonneux, Sophie, et al.
Published: (2024)
Similar Items
-
Position: Towards Resilience Against Adversarial Examples
by: Dai, Sihui, et al.
Published: (2024) -
A New Linear Scaling Rule for Private Adaptive Hyperparameter Optimization
by: Panda, Ashwinee, et al.
Published: (2022) -
Towards Scalable and Robust Model Versioning
by: Ding, Wenxin, et al.
Published: (2024) -
PatchDEMUX: A Certifiably Robust Framework for Multi-label Classifiers Against Adversarial Patches
by: Jacob, Dennis, et al.
Published: (2025) -
Adaptive and Stratified Subsampling for High-Dimensional Robust Estimation
by: Mittal, Prateek, et al.
Published: (2024)