Adversarial Training for Defense Against Label Poisoning Attacks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bal, Melis Ilayda, Cevher, Volkan, Muehlebach, Michael |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining
von: Bal, Melis Ilayda, et al.
Veröffentlicht: (2025)
von: Bal, Melis Ilayda, et al.
Veröffentlicht: (2025)
Revisiting Character-level Adversarial Attacks for Language Models
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2024)
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2024)
Multi-level Certified Defense Against Poisoning Attacks in Offline Reinforcement Learning
von: Liu, Shijie, et al.
Veröffentlicht: (2025)
von: Liu, Shijie, et al.
Veröffentlicht: (2025)
FLAegis: A Two-Layer Defense Framework for Federated Learning Against Poisoning Attacks
von: Campos, Enrique Mármol, et al.
Veröffentlicht: (2025)
von: Campos, Enrique Mármol, et al.
Veröffentlicht: (2025)
Deep Adversarial Defense Against Multilevel-Lp Attacks
von: Wang, Ren, et al.
Veröffentlicht: (2024)
von: Wang, Ren, et al.
Veröffentlicht: (2024)
Addressing Label Shift in Distributed Learning via Entropy Regularization
von: Wu, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Wu, Zhiyuan, et al.
Veröffentlicht: (2025)
Efficient Large Language Model Inference with Neural Block Linearization
von: Erdogan, Mete, et al.
Veröffentlicht: (2025)
von: Erdogan, Mete, et al.
Veröffentlicht: (2025)
Leveraging the Context through Multi-Round Interactions for Jailbreaking Attacks
von: Cheng, Yixin, et al.
Veröffentlicht: (2024)
von: Cheng, Yixin, et al.
Veröffentlicht: (2024)
Optimistic Games for Combinatorial Bayesian Optimization with Application to Protein Design
von: Bal, Melis Ilayda, et al.
Veröffentlicht: (2024)
von: Bal, Melis Ilayda, et al.
Veröffentlicht: (2024)
Stealthy Poisoning Attacks Bypass Defenses in Regression Settings
von: Carnerero-Cano, Javier, et al.
Veröffentlicht: (2026)
von: Carnerero-Cano, Javier, et al.
Veröffentlicht: (2026)
MT-NAM: An Efficient and Adaptive Model for Epileptic Seizure Detection
von: Afzal, Arshia, et al.
Veröffentlicht: (2025)
von: Afzal, Arshia, et al.
Veröffentlicht: (2025)
Cost-Minimized Label-Flipping Poisoning Attack to LLM Alignment
von: Kusaka, Shigeki, et al.
Veröffentlicht: (2025)
von: Kusaka, Shigeki, et al.
Veröffentlicht: (2025)
Learning to Attack: A Bandit Approach to Adversarial Context Poisoning
von: Telikani, Ray, et al.
Veröffentlicht: (2026)
von: Telikani, Ray, et al.
Veröffentlicht: (2026)
SoK: Benchmarking Poisoning Attacks and Defenses in Federated Learning
von: Zhang, Heyi, et al.
Veröffentlicht: (2025)
von: Zhang, Heyi, et al.
Veröffentlicht: (2025)
Attacks and Defenses Against LLM Fingerprinting
von: Kurian, Kevin, et al.
Veröffentlicht: (2025)
von: Kurian, Kevin, et al.
Veröffentlicht: (2025)
Ascent Fails to Forget
von: Mavrothalassitis, Ioannis, et al.
Veröffentlicht: (2025)
von: Mavrothalassitis, Ioannis, et al.
Veröffentlicht: (2025)
Semantic Chameleon: Corpus-Dependent Poisoning Attacks and Defenses in RAG Systems
von: Thornton, Scott
Veröffentlicht: (2026)
von: Thornton, Scott
Veröffentlicht: (2026)
MaskPure: Improving Defense Against Text Adversaries with Stochastic Purification
von: Gietz, Harrison, et al.
Veröffentlicht: (2024)
von: Gietz, Harrison, et al.
Veröffentlicht: (2024)
Optimal Defenses Against Gradient Reconstruction Attacks
von: Chen, Yuxiao, et al.
Veröffentlicht: (2024)
von: Chen, Yuxiao, et al.
Veröffentlicht: (2024)
A Novel Defense Against Poisoning Attacks on Federated Learning: LayerCAM Augmented with Autoencoder
von: Zheng, Jingjing, et al.
Veröffentlicht: (2024)
von: Zheng, Jingjing, et al.
Veröffentlicht: (2024)
Poisoning the Inner Prediction Logic of Graph Neural Networks for Clean-Label Backdoor Attacks
von: Zhang, Yuxiang, et al.
Veröffentlicht: (2026)
von: Zhang, Yuxiang, et al.
Veröffentlicht: (2026)
Certified Robustness Under Bounded Levenshtein Distance
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2025)
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2025)
REST: Efficient and Accelerated EEG Seizure Analysis through Residual State Updates
von: Afzal, Arshia, et al.
Veröffentlicht: (2024)
von: Afzal, Arshia, et al.
Veröffentlicht: (2024)
Test-Time Defense Against Adversarial Attacks via Stochastic Resonance of Latent Ensembles
von: Lao, Dong, et al.
Veröffentlicht: (2025)
von: Lao, Dong, et al.
Veröffentlicht: (2025)
An Embarrassingly Simple Defense Against LLM Abliteration Attacks
von: Shairah, Harethah Abu, et al.
Veröffentlicht: (2025)
von: Shairah, Harethah Abu, et al.
Veröffentlicht: (2025)
SALAAD: Sparse And Low-Rank Adaptation via ADMM for Large Language Model Inference
von: Ma, Hao, et al.
Veröffentlicht: (2026)
von: Ma, Hao, et al.
Veröffentlicht: (2026)
Topology-Independent Robustness of the Weighted Mean under Label Poisoning Attacks in Heterogeneous Decentralized Learning
von: Peng, Jie, et al.
Veröffentlicht: (2026)
von: Peng, Jie, et al.
Veröffentlicht: (2026)
Shadowcast: Stealthy Data Poisoning Attacks Against Vision-Language Models
von: Xu, Yuancheng, et al.
Veröffentlicht: (2024)
von: Xu, Yuancheng, et al.
Veröffentlicht: (2024)
Rethinking Adversarial Policies: A Generalized Attack Formulation and Provable Defense in RL
von: Liu, Xiangyu, et al.
Veröffentlicht: (2023)
von: Liu, Xiangyu, et al.
Veröffentlicht: (2023)
Defending Against Poisoning Attacks in Federated Learning with Blockchain
von: Dong, Nanqing, et al.
Veröffentlicht: (2023)
von: Dong, Nanqing, et al.
Veröffentlicht: (2023)
ARMOR: Adaptive Resilience Against Model Poisoning Attacks in Continual Federated Learning for Mobile Indoor Localization
von: Gufran, Danish, et al.
Veröffentlicht: (2026)
von: Gufran, Danish, et al.
Veröffentlicht: (2026)
Zero-Sacrifice Persistent-Robustness Adversarial Defense for Pre-Trained Encoders
von: Lei, Zhuxin, et al.
Veröffentlicht: (2026)
von: Lei, Zhuxin, et al.
Veröffentlicht: (2026)
Rate optimal learning of equilibria from data
von: Freihaut, Till, et al.
Veröffentlicht: (2025)
von: Freihaut, Till, et al.
Veröffentlicht: (2025)
Certifying Language Model Robustness with Fuzzed Randomized Smoothing: An Efficient Defense Against Backdoor Attacks
von: He, Bowei, et al.
Veröffentlicht: (2025)
von: He, Bowei, et al.
Veröffentlicht: (2025)
Towards Robust Policy: Enhancing Offline Reinforcement Learning with Adversarial Attacks and Defenses
von: Nguyen, Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen, Thanh, et al.
Veröffentlicht: (2024)
Robust NAS under adversarial training: benchmark, theory, and beyond
von: Wu, Yongtao, et al.
Veröffentlicht: (2024)
von: Wu, Yongtao, et al.
Veröffentlicht: (2024)
Conformal Generative Modeling with Improved Sample Efficiency through Sequential Greedy Filtering
von: Kladny, Klaus-Rudolf, et al.
Veröffentlicht: (2024)
von: Kladny, Klaus-Rudolf, et al.
Veröffentlicht: (2024)
Be Kind, Rewrite: Benign Projections via Rewriting Defend Against LLM Data Poisoning Attacks
von: Halloran, John T., et al.
Veröffentlicht: (2026)
von: Halloran, John T., et al.
Veröffentlicht: (2026)
Peak-Controlled Logits Poisoning Attack in Federated Distillation
von: Tang, Yuhan, et al.
Veröffentlicht: (2024)
von: Tang, Yuhan, et al.
Veröffentlicht: (2024)
PureGen: Universal Data Purification for Train-Time Poison Defense via Generative Model Dynamics
von: Bhat, Sunay, et al.
Veröffentlicht: (2024)
von: Bhat, Sunay, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining
von: Bal, Melis Ilayda, et al.
Veröffentlicht: (2025) -
Revisiting Character-level Adversarial Attacks for Language Models
von: Rocamora, Elias Abad, et al.
Veröffentlicht: (2024) -
Multi-level Certified Defense Against Poisoning Attacks in Offline Reinforcement Learning
von: Liu, Shijie, et al.
Veröffentlicht: (2025) -
FLAegis: A Two-Layer Defense Framework for Federated Learning Against Poisoning Attacks
von: Campos, Enrique Mármol, et al.
Veröffentlicht: (2025) -
Deep Adversarial Defense Against Multilevel-Lp Attacks
von: Wang, Ren, et al.
Veröffentlicht: (2024)