MaskPure: Improving Defense Against Text Adversaries with Stochastic Purification
Fuente:
arXiv
Saved in:
| Main Authors: | Gietz, Harrison, Kalita, Jugal |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Abstractive Text Summarization: State of the Art, Challenges, and Improvements
by: Shakil, Hassan, et al.
Published: (2024)
by: Shakil, Hassan, et al.
Published: (2024)
CNN-LSTM and Transfer Learning Models for Malware Classification based on Opcodes and API Calls
by: Bensaoud, Ahmed, et al.
Published: (2024)
by: Bensaoud, Ahmed, et al.
Published: (2024)
A Novel Active Learning Approach to Label One Million Unknown Malware Variants
by: Bensaoud, Ahmed, et al.
Published: (2025)
by: Bensaoud, Ahmed, et al.
Published: (2025)
FlowPure: Continuous Normalizing Flows for Adversarial Purification
by: Collaert, Elias, et al.
Published: (2025)
by: Collaert, Elias, et al.
Published: (2025)
Adversarial Text Purification: A Large Language Model Approach for Defense
by: Moraffah, Raha, et al.
Published: (2024)
by: Moraffah, Raha, et al.
Published: (2024)
Adversarial Training for Defense Against Label Poisoning Attacks
by: Bal, Melis Ilayda, et al.
Published: (2025)
by: Bal, Melis Ilayda, et al.
Published: (2025)
Adapting Feature Attenuation to NLP
by: Yang, Tianshuo, et al.
Published: (2026)
by: Yang, Tianshuo, et al.
Published: (2026)
PureGen: Universal Data Purification for Train-Time Poison Defense via Generative Model Dynamics
by: Bhat, Sunay, et al.
Published: (2024)
by: Bhat, Sunay, et al.
Published: (2024)
Explainable AI: Context-Aware Layer-Wise Integrated Gradients for Explaining Transformer Models
by: Mersha, Melkamu Abay, et al.
Published: (2026)
by: Mersha, Melkamu Abay, et al.
Published: (2026)
Alert-ME: An Explainability-Driven Defense Against Adversarial Examples in Transformer-Based Text Classification
by: Sabir, Bushra, et al.
Published: (2023)
by: Sabir, Bushra, et al.
Published: (2023)
Deep Adversarial Defense Against Multilevel-Lp Attacks
by: Wang, Ren, et al.
Published: (2024)
by: Wang, Ren, et al.
Published: (2024)
Test-Time Defense Against Adversarial Attacks via Stochastic Resonance of Latent Ensembles
by: Lao, Dong, et al.
Published: (2025)
by: Lao, Dong, et al.
Published: (2025)
Robust Graph Learning Against Adversarial Evasion Attacks via Prior-Free Diffusion-Based Structure Purification
by: Luo, Jiayi, et al.
Published: (2025)
by: Luo, Jiayi, et al.
Published: (2025)
Evaluating the Effectiveness of XAI Techniques for Encoder-Based Language Models
by: Mersha, Melkamu Abay, et al.
Published: (2025)
by: Mersha, Melkamu Abay, et al.
Published: (2025)
PureEBM: Universal Poison Purification via Mid-Run Dynamics of Energy-Based Models
by: Pooladzandi, Omead, et al.
Published: (2024)
by: Pooladzandi, Omead, et al.
Published: (2024)
Adversarial Attacks on Transformers-Based Malware Detectors
by: Jakhotiya, Yash, et al.
Published: (2022)
by: Jakhotiya, Yash, et al.
Published: (2022)
Defense without Forgetting: Continual Adversarial Defense with Anisotropic & Isotropic Pseudo Replay
by: Zhou, Yuhang, et al.
Published: (2024)
by: Zhou, Yuhang, et al.
Published: (2024)
Regret-Based Defense in Adversarial Reinforcement Learning
by: Belaire, Roman, et al.
Published: (2023)
by: Belaire, Roman, et al.
Published: (2023)
One Step to the Side: Why Defenses Against Malicious Finetuning Fail Under Adaptive Adversaries
by: Zloczower, Itay, et al.
Published: (2026)
by: Zloczower, Itay, et al.
Published: (2026)
Diffusion-based Adversarial Purification for Intrusion Detection
by: Merzouk, Mohamed Amine, et al.
Published: (2024)
by: Merzouk, Mohamed Amine, et al.
Published: (2024)
Instant Adversarial Purification with Adversarial Consistency Distillation
by: Lei, Chun Tong, et al.
Published: (2024)
by: Lei, Chun Tong, et al.
Published: (2024)
TopoReformer: Mitigating Adversarial Attacks Using Topological Purification in OCR Models
by: Kumar, Bhagyesh, et al.
Published: (2025)
by: Kumar, Bhagyesh, et al.
Published: (2025)
Adversarial Vulnerability Transcends Computational Paradigms: Feature Engineering Provides No Defense Against Neural Adversarial Transfer
by: Hsain, Achraf, et al.
Published: (2026)
by: Hsain, Achraf, et al.
Published: (2026)
Explainable Artificial Intelligence: A Survey of Needs, Techniques, Applications, and Future Direction
by: Mersha, Melkamu, et al.
Published: (2024)
by: Mersha, Melkamu, et al.
Published: (2024)
Attacks and Defenses Against LLM Fingerprinting
by: Kurian, Kevin, et al.
Published: (2025)
by: Kurian, Kevin, et al.
Published: (2025)
PuriDefense: Randomized Local Implicit Adversarial Purification for Defending Black-box Query-based Attacks
by: Guo, Ping, et al.
Published: (2024)
by: Guo, Ping, et al.
Published: (2024)
LoRID: Low-Rank Iterative Diffusion for Adversarial Purification
by: Zollicoffer, Geigh, et al.
Published: (2024)
by: Zollicoffer, Geigh, et al.
Published: (2024)
Zero-Sacrifice Persistent-Robustness Adversarial Defense for Pre-Trained Encoders
by: Lei, Zhuxin, et al.
Published: (2026)
by: Lei, Zhuxin, et al.
Published: (2026)
Optimal Defenses Against Gradient Reconstruction Attacks
by: Chen, Yuxiao, et al.
Published: (2024)
by: Chen, Yuxiao, et al.
Published: (2024)
Task-Informed Anti-Curriculum by Masking Improves Downstream Performance on Text
by: Jarca, Andrei, et al.
Published: (2025)
by: Jarca, Andrei, et al.
Published: (2025)
Multi-level Certified Defense Against Poisoning Attacks in Offline Reinforcement Learning
by: Liu, Shijie, et al.
Published: (2025)
by: Liu, Shijie, et al.
Published: (2025)
It's all about PR -- Smart Benchmarking AI Accelerators using Performance Representatives
by: Jung, Alexander Louis-Ferdinand, et al.
Published: (2024)
by: Jung, Alexander Louis-Ferdinand, et al.
Published: (2024)
Rethinking Adversarial Policies: A Generalized Attack Formulation and Provable Defense in RL
by: Liu, Xiangyu, et al.
Published: (2023)
by: Liu, Xiangyu, et al.
Published: (2023)
Mitigating the Structural Bias in Graph Adversarial Defenses
by: Fang, Junyuan, et al.
Published: (2025)
by: Fang, Junyuan, et al.
Published: (2025)
Random Sampling for Diffusion-based Adversarial Purification
by: Zhang, Jiancheng, et al.
Published: (2024)
by: Zhang, Jiancheng, et al.
Published: (2024)
A Unified Framework with Novel Metrics for Evaluating the Effectiveness of XAI Techniques in LLMs
by: Mersha, Melkamu Abay, et al.
Published: (2025)
by: Mersha, Melkamu Abay, et al.
Published: (2025)
FLAegis: A Two-Layer Defense Framework for Federated Learning Against Poisoning Attacks
by: Campos, Enrique Mármol, et al.
Published: (2025)
by: Campos, Enrique Mármol, et al.
Published: (2025)
An Embarrassingly Simple Defense Against LLM Abliteration Attacks
by: Shairah, Harethah Abu, et al.
Published: (2025)
by: Shairah, Harethah Abu, et al.
Published: (2025)
Certifying Language Model Robustness with Fuzzed Randomized Smoothing: An Efficient Defense Against Backdoor Attacks
by: He, Bowei, et al.
Published: (2025)
by: He, Bowei, et al.
Published: (2025)
Contrastive ECOC: Learning Output Codes for Adversarial Defense
by: Chou, Che-Yu, et al.
Published: (2025)
by: Chou, Che-Yu, et al.
Published: (2025)
Similar Items
-
Abstractive Text Summarization: State of the Art, Challenges, and Improvements
by: Shakil, Hassan, et al.
Published: (2024) -
CNN-LSTM and Transfer Learning Models for Malware Classification based on Opcodes and API Calls
by: Bensaoud, Ahmed, et al.
Published: (2024) -
A Novel Active Learning Approach to Label One Million Unknown Malware Variants
by: Bensaoud, Ahmed, et al.
Published: (2025) -
FlowPure: Continuous Normalizing Flows for Adversarial Purification
by: Collaert, Elias, et al.
Published: (2025) -
Adversarial Text Purification: A Large Language Model Approach for Defense
by: Moraffah, Raha, et al.
Published: (2024)