On the Duality Between Sharpness-Aware Minimization and Adversarial Training
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Yihao, He, Hangzhou, Zhu, Jingyu, Chen, Huanran, Wang, Yifei, Wei, Zeming |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Boosting Jailbreak Attack with Momentum
di: Zhang, Yihao, et al.
Pubblicazione: (2024)
di: Zhang, Yihao, et al.
Pubblicazione: (2024)
Adversarial Representation Engineering: A General Model Editing Framework for Large Language Models
di: Zhang, Yihao, et al.
Pubblicazione: (2024)
di: Zhang, Yihao, et al.
Pubblicazione: (2024)
Secure LLM Fine-Tuning via Safety-Aware Probing
di: Wu, Chengcan, et al.
Pubblicazione: (2025)
di: Wu, Chengcan, et al.
Pubblicazione: (2025)
Identifying and Understanding Cross-Class Features in Adversarial Training
di: Wei, Zeming, et al.
Pubblicazione: (2025)
di: Wei, Zeming, et al.
Pubblicazione: (2025)
RAPO: Risk-Aware Preference Optimization for Generalizable Safe Reasoning
di: Wei, Zeming, et al.
Pubblicazione: (2026)
di: Wei, Zeming, et al.
Pubblicazione: (2026)
Exploring the Robustness of In-Context Learning with Noisy Labels
di: Cheng, Chen, et al.
Pubblicazione: (2024)
di: Cheng, Chen, et al.
Pubblicazione: (2024)
Dynamic Orthogonal Continual Fine-tuning for Mitigating Catastrophic Forgettings
di: Zhang, Zhixin, et al.
Pubblicazione: (2025)
di: Zhang, Zhixin, et al.
Pubblicazione: (2025)
Calibrated Adversarial Sampling: Multi-Armed Bandit-Guided Generalization Against Unforeseen Attacks
di: Wang, Rui, et al.
Pubblicazione: (2025)
di: Wang, Rui, et al.
Pubblicazione: (2025)
Securing Multi-Agent Systems Against Corruptions via Node Contribution Backpropagation
di: Wu, Chengcan, et al.
Pubblicazione: (2025)
di: Wu, Chengcan, et al.
Pubblicazione: (2025)
Characterizing the Training Dynamics of Private Fine-tuning with Langevin diffusion
di: Ke, Shuqi, et al.
Pubblicazione: (2024)
di: Ke, Shuqi, et al.
Pubblicazione: (2024)
Clip-and-Verify: Linear Constraint-Driven Domain Clipping for Accelerating Neural Network Verification
di: Zhou, Duo, et al.
Pubblicazione: (2025)
di: Zhou, Duo, et al.
Pubblicazione: (2025)
Efficient Optimization Algorithms for Linear Adversarial Training
di: RIbeiro, Antônio H., et al.
Pubblicazione: (2024)
di: RIbeiro, Antônio H., et al.
Pubblicazione: (2024)
Approximate and Weighted Data Reconstruction Attack in Federated Learning
di: Song, Yongcun, et al.
Pubblicazione: (2023)
di: Song, Yongcun, et al.
Pubblicazione: (2023)
Correlated Noise Provably Beats Independent Noise for Differentially Private Learning
di: Choquette-Choo, Christopher A., et al.
Pubblicazione: (2023)
di: Choquette-Choo, Christopher A., et al.
Pubblicazione: (2023)
SMI: Statistical Membership Inference for Reliable Unlearned Model Auditing
di: Sun, Jialong, et al.
Pubblicazione: (2026)
di: Sun, Jialong, et al.
Pubblicazione: (2026)
A Queueing-Theoretic Framework for Dynamic Attack Surfaces: Data-Integrated Risk Analysis and Adaptive Defense
di: Yun, Jihyeon, et al.
Pubblicazione: (2026)
di: Yun, Jihyeon, et al.
Pubblicazione: (2026)
On Model Protection in Federated Learning against Eavesdropping Attacks
di: Maity, Dipankar, et al.
Pubblicazione: (2025)
di: Maity, Dipankar, et al.
Pubblicazione: (2025)
Kernel Learning with Adversarial Features: Numerical Efficiency and Adaptive Regularization
di: Ribeiro, Antônio H., et al.
Pubblicazione: (2025)
di: Ribeiro, Antônio H., et al.
Pubblicazione: (2025)
ViSTR-GP: Online Cyberattack Detection via Vision-to-State Tensor Regression and Gaussian Processes in Automated Robotic Operations
di: Aftabi, Navid, et al.
Pubblicazione: (2025)
di: Aftabi, Navid, et al.
Pubblicazione: (2025)
Optimal Differentially Private Model Training with Public Data
di: Lowy, Andrew, et al.
Pubblicazione: (2023)
di: Lowy, Andrew, et al.
Pubblicazione: (2023)
MILE: A Mutation Testing Framework of In-Context Learning Systems
di: Wei, Zeming, et al.
Pubblicazione: (2024)
di: Wei, Zeming, et al.
Pubblicazione: (2024)
DPZero: Private Fine-Tuning of Language Models without Backpropagation
di: Zhang, Liang, et al.
Pubblicazione: (2023)
di: Zhang, Liang, et al.
Pubblicazione: (2023)
RACC: Representation-Aware Coverage Criteria for LLM Safety Testing
di: Wei, Zeming, et al.
Pubblicazione: (2026)
di: Wei, Zeming, et al.
Pubblicazione: (2026)
ClawWorm: Self-Propagating Attacks Across LLM Agent Ecosystems
di: Zhang, Yihao, et al.
Pubblicazione: (2026)
di: Zhang, Yihao, et al.
Pubblicazione: (2026)
Private Zeroth-Order Nonsmooth Nonconvex Optimization
di: Zhang, Qinzi, et al.
Pubblicazione: (2024)
di: Zhang, Qinzi, et al.
Pubblicazione: (2024)
Scalable Neural Network Verification with Branch-and-bound Inferred Cutting Planes
di: Zhou, Duo, et al.
Pubblicazione: (2024)
di: Zhou, Duo, et al.
Pubblicazione: (2024)
Fight Back Against Jailbreaking via Prompt Adversarial Tuning
di: Mo, Yichuan, et al.
Pubblicazione: (2024)
di: Mo, Yichuan, et al.
Pubblicazione: (2024)
Automata-Based Steering of Large Language Models for Diverse Structured Generation
di: Luan, Xiaokun, et al.
Pubblicazione: (2025)
di: Luan, Xiaokun, et al.
Pubblicazione: (2025)
An Unsupervised Adversarial Autoencoder for Cyber Attack Detection in Power Distribution Grids
di: Zideh, Mehdi Jabbari, et al.
Pubblicazione: (2024)
di: Zideh, Mehdi Jabbari, et al.
Pubblicazione: (2024)
RL-Based Method for Benchmarking the Adversarial Resilience and Robustness of Deep Reinforcement Learning Policies
di: Behzadan, Vahid, et al.
Pubblicazione: (2019)
di: Behzadan, Vahid, et al.
Pubblicazione: (2019)
TAO: Tolerance-Aware Optimistic Verification for Floating-Point Neural Networks
di: Yao, Jianzhu, et al.
Pubblicazione: (2025)
di: Yao, Jianzhu, et al.
Pubblicazione: (2025)
Mirror Descent Algorithms with Nearly Dimension-Independent Rates for Differentially-Private Stochastic Saddle-Point Problems
di: González, Tomás, et al.
Pubblicazione: (2024)
di: González, Tomás, et al.
Pubblicazione: (2024)
Byzantine-Robust and Differentially Private Federated Optimization under Weaker Assumptions
di: Islamov, Rustem, et al.
Pubblicazione: (2026)
di: Islamov, Rustem, et al.
Pubblicazione: (2026)
The Power of Sampling: Dimension-free Risk Bounds in Private ERM
di: Lee, Yin Tat, et al.
Pubblicazione: (2021)
di: Lee, Yin Tat, et al.
Pubblicazione: (2021)
Attention-Enhanced Graph Filtering for False Data Injection Attack Detection and Localization
di: Abdulin, Ruslan, et al.
Pubblicazione: (2026)
di: Abdulin, Ruslan, et al.
Pubblicazione: (2026)
The Relative Gaussian Mechanism and its Application to Private Gradient Descent
di: Hendrikx, Hadrien, et al.
Pubblicazione: (2023)
di: Hendrikx, Hadrien, et al.
Pubblicazione: (2023)
Differentially Private Non-Convex Optimization under the KL Condition with Optimal Rates
di: Menart, Michael, et al.
Pubblicazione: (2023)
di: Menart, Michael, et al.
Pubblicazione: (2023)
Differentially Private Bilevel Optimization
di: Kornowski, Guy
Pubblicazione: (2024)
di: Kornowski, Guy
Pubblicazione: (2024)
Clipped SGD Algorithms for Performative Prediction: Tight Bounds for Clipping Bias and Remedies
di: Li, Qiang, et al.
Pubblicazione: (2024)
di: Li, Qiang, et al.
Pubblicazione: (2024)
Differential Privacy via Distributionally Robust Optimization
di: Selvi, Aras, et al.
Pubblicazione: (2023)
di: Selvi, Aras, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Boosting Jailbreak Attack with Momentum
di: Zhang, Yihao, et al.
Pubblicazione: (2024) -
Adversarial Representation Engineering: A General Model Editing Framework for Large Language Models
di: Zhang, Yihao, et al.
Pubblicazione: (2024) -
Secure LLM Fine-Tuning via Safety-Aware Probing
di: Wu, Chengcan, et al.
Pubblicazione: (2025) -
Identifying and Understanding Cross-Class Features in Adversarial Training
di: Wei, Zeming, et al.
Pubblicazione: (2025) -
RAPO: Risk-Aware Preference Optimization for Generalizable Safe Reasoning
di: Wei, Zeming, et al.
Pubblicazione: (2026)