Identifying and Understanding Cross-Class Features in Adversarial Training
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Zeming, Guo, Yiwen, Wang, Yisen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Calibrated Adversarial Sampling: Multi-Armed Bandit-Guided Generalization Against Unforeseen Attacks
by: Wang, Rui, et al.
Published: (2025)
by: Wang, Rui, et al.
Published: (2025)
SMI: Statistical Membership Inference for Reliable Unlearned Model Auditing
by: Sun, Jialong, et al.
Published: (2026)
by: Sun, Jialong, et al.
Published: (2026)
On the Duality Between Sharpness-Aware Minimization and Adversarial Training
by: Zhang, Yihao, et al.
Published: (2024)
by: Zhang, Yihao, et al.
Published: (2024)
Adversarial Representation Engineering: A General Model Editing Framework for Large Language Models
by: Zhang, Yihao, et al.
Published: (2024)
by: Zhang, Yihao, et al.
Published: (2024)
GSE: Group-wise Sparse and Explainable Adversarial Attacks
by: Sadiku, Shpresim, et al.
Published: (2023)
by: Sadiku, Shpresim, et al.
Published: (2023)
Boosting Jailbreak Attack with Momentum
by: Zhang, Yihao, et al.
Published: (2024)
by: Zhang, Yihao, et al.
Published: (2024)
Towards Understanding Dual BN In Hybrid Adversarial Training
by: Zhang, Chenshuang, et al.
Published: (2024)
by: Zhang, Chenshuang, et al.
Published: (2024)
Dynamic Orthogonal Continual Fine-tuning for Mitigating Catastrophic Forgettings
by: Zhang, Zhixin, et al.
Published: (2025)
by: Zhang, Zhixin, et al.
Published: (2025)
Understanding Key Point Cloud Features for Development Three-dimensional Adversarial Attacks
by: Naderi, Hanieh, et al.
Published: (2022)
by: Naderi, Hanieh, et al.
Published: (2022)
RAPO: Risk-Aware Preference Optimization for Generalizable Safe Reasoning
by: Wei, Zeming, et al.
Published: (2026)
by: Wei, Zeming, et al.
Published: (2026)
Secure LLM Fine-Tuning via Safety-Aware Probing
by: Wu, Chengcan, et al.
Published: (2025)
by: Wu, Chengcan, et al.
Published: (2025)
The Impact of Scaling Training Data on Adversarial Robustness
by: Zimmerli, Marco, et al.
Published: (2025)
by: Zimmerli, Marco, et al.
Published: (2025)
Exploring the Robustness of In-Context Learning with Noisy Labels
by: Cheng, Chen, et al.
Published: (2024)
by: Cheng, Chen, et al.
Published: (2024)
Fight Back Against Jailbreaking via Prompt Adversarial Tuning
by: Mo, Yichuan, et al.
Published: (2024)
by: Mo, Yichuan, et al.
Published: (2024)
Improving Adversarial Training using Vulnerability-Aware Perturbation Budget
by: Fakorede, Olukorede, et al.
Published: (2024)
by: Fakorede, Olukorede, et al.
Published: (2024)
Exploring the Adversarial Frontier: Quantifying Robustness via Adversarial Hypervolume
by: Guo, Ping, et al.
Published: (2024)
by: Guo, Ping, et al.
Published: (2024)
Fingerprinting Image-to-Image Generative Adversarial Networks
by: Li, Guanlin, et al.
Published: (2021)
by: Li, Guanlin, et al.
Published: (2021)
Efficient Semi-Supervised Adversarial Training via Latent Clustering-Based Data Reduction
by: Ghosh, Somrita, et al.
Published: (2025)
by: Ghosh, Somrita, et al.
Published: (2025)
MOS-Attack: A Scalable Multi-objective Adversarial Attack Framework
by: Guo, Ping, et al.
Published: (2025)
by: Guo, Ping, et al.
Published: (2025)
DIFFender: Diffusion-Based Adversarial Defense against Patch Attacks
by: Kang, Caixin, et al.
Published: (2023)
by: Kang, Caixin, et al.
Published: (2023)
Real-world Adversarial Defense against Patch Attacks based on Diffusion Model
by: Wei, Xingxing, et al.
Published: (2024)
by: Wei, Xingxing, et al.
Published: (2024)
Improving the Transferability of Adversarial Attacks by an Input Transpose
by: Wan, Qing, et al.
Published: (2025)
by: Wan, Qing, et al.
Published: (2025)
Securing Multi-Agent Systems Against Corruptions via Node Contribution Backpropagation
by: Wu, Chengcan, et al.
Published: (2025)
by: Wu, Chengcan, et al.
Published: (2025)
L-AutoDA: Leveraging Large Language Models for Automated Decision-based Adversarial Attacks
by: Guo, Ping, et al.
Published: (2024)
by: Guo, Ping, et al.
Published: (2024)
PuriDefense: Randomized Local Implicit Adversarial Purification for Defending Black-box Query-based Attacks
by: Guo, Ping, et al.
Published: (2024)
by: Guo, Ping, et al.
Published: (2024)
Adversarial Nibbler: An Open Red-Teaming Method for Identifying Diverse Harms in Text-to-Image Generation
by: Quaye, Jessica, et al.
Published: (2024)
by: Quaye, Jessica, et al.
Published: (2024)
A White-Box False Positive Adversarial Attack Method on Contrastive Loss Based Offline Handwritten Signature Verification Models
by: Guo, Zhongliang, et al.
Published: (2023)
by: Guo, Zhongliang, et al.
Published: (2023)
A Survey on Physical Adversarial Attacks against Face Recognition Systems
by: Wang, Mingsi, et al.
Published: (2024)
by: Wang, Mingsi, et al.
Published: (2024)
Adversarial Detection by Approximation of Ensemble Boundary
by: Windeatt, T.
Published: (2022)
by: Windeatt, T.
Published: (2022)
Crafting Adversarial Inputs for Large Vision-Language Models Using Black-Box Optimization
by: Guan, Jiwei, et al.
Published: (2026)
by: Guan, Jiwei, et al.
Published: (2026)
DarkLLM: Learning Language-Driven Adversarial Attacks with Large Language Models
by: Sun, Ye, et al.
Published: (2026)
by: Sun, Ye, et al.
Published: (2026)
Characterizing the Training Dynamics of Private Fine-tuning with Langevin diffusion
by: Ke, Shuqi, et al.
Published: (2024)
by: Ke, Shuqi, et al.
Published: (2024)
On the Importance of Backbone to the Adversarial Robustness of Object Detectors
by: Li, Xiao, et al.
Published: (2023)
by: Li, Xiao, et al.
Published: (2023)
Evaluating the Evaluators: Trust in Adversarial Robustness Tests
by: Cinà, Antonio Emanuele, et al.
Published: (2025)
by: Cinà, Antonio Emanuele, et al.
Published: (2025)
GLEAN: Generative Learning for Eliminating Adversarial Noise
by: Kim, Justin Lyu, et al.
Published: (2024)
by: Kim, Justin Lyu, et al.
Published: (2024)
Scalable Secure Biometric Authentication without Auxiliary Identifiers
by: Bienstock, Alexander, et al.
Published: (2026)
by: Bienstock, Alexander, et al.
Published: (2026)
AR-GAN: Generative Adversarial Network-Based Defense Method Against Adversarial Attacks on the Traffic Sign Classification System of Autonomous Vehicles
by: Salek, M Sabbir, et al.
Published: (2023)
by: Salek, M Sabbir, et al.
Published: (2023)
Sy-FAR: Symmetry-based Fair Adversarial Robustness
by: Najjar, Haneen, et al.
Published: (2025)
by: Najjar, Haneen, et al.
Published: (2025)
Standard-Deviation-Inspired Regularization for Improving Adversarial Robustness
by: Fakorede, Olukorede, et al.
Published: (2024)
by: Fakorede, Olukorede, et al.
Published: (2024)
Impact of Architectural Modifications on Deep Learning Adversarial Robustness
by: Juraev, Firuz, et al.
Published: (2024)
by: Juraev, Firuz, et al.
Published: (2024)
Similar Items
-
Calibrated Adversarial Sampling: Multi-Armed Bandit-Guided Generalization Against Unforeseen Attacks
by: Wang, Rui, et al.
Published: (2025) -
SMI: Statistical Membership Inference for Reliable Unlearned Model Auditing
by: Sun, Jialong, et al.
Published: (2026) -
On the Duality Between Sharpness-Aware Minimization and Adversarial Training
by: Zhang, Yihao, et al.
Published: (2024) -
Adversarial Representation Engineering: A General Model Editing Framework for Large Language Models
by: Zhang, Yihao, et al.
Published: (2024) -
GSE: Group-wise Sparse and Explainable Adversarial Attacks
by: Sadiku, Shpresim, et al.
Published: (2023)