Cutting through buggy adversarial example defenses: fixing 1 line of code breaks Sabre
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Carlini, Nicholas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AutoAdvExBench: Benchmarking autonomous exploitation of adversarial example defenses
von: Carlini, Nicholas, et al.
Veröffentlicht: (2025)
von: Carlini, Nicholas, et al.
Veröffentlicht: (2025)
Robustness bounds on the successful adversarial examples in probabilistic models: Implications from Gaussian processes
von: Maeshima, Hiroaki, et al.
Veröffentlicht: (2024)
von: Maeshima, Hiroaki, et al.
Veröffentlicht: (2024)
Remote Timing Attacks on Efficient Language Model Inference
von: Carlini, Nicholas, et al.
Veröffentlicht: (2024)
von: Carlini, Nicholas, et al.
Veröffentlicht: (2024)
Evading Black-box Classifiers Without Breaking Eggs
von: Debenedetti, Edoardo, et al.
Veröffentlicht: (2023)
von: Debenedetti, Edoardo, et al.
Veröffentlicht: (2023)
Position: Considerations for Differentially Private Learning with Large-Scale Public Pretraining
von: Tramèr, Florian, et al.
Veröffentlicht: (2022)
von: Tramèr, Florian, et al.
Veröffentlicht: (2022)
Robust width: A lightweight and certifiable adversarial defense
von: Peck, Jonathan, et al.
Veröffentlicht: (2024)
von: Peck, Jonathan, et al.
Veröffentlicht: (2024)
Privacy Backdoors: Enhancing Membership Inference through Poisoning Pre-trained Models
von: Wen, Yuxin, et al.
Veröffentlicht: (2024)
von: Wen, Yuxin, et al.
Veröffentlicht: (2024)
IF-GUIDE: Influence Function-Guided Detoxification of LLMs
von: Coalson, Zachary, et al.
Veröffentlicht: (2025)
von: Coalson, Zachary, et al.
Veröffentlicht: (2025)
Adversarial ML Problems Are Getting Harder to Solve and to Evaluate
von: Rando, Javier, et al.
Veröffentlicht: (2025)
von: Rando, Javier, et al.
Veröffentlicht: (2025)
RAB$^2$-DEF: Dynamic and explainable defense against adversarial attacks in Federated Learning to fair poor clients
von: Rodríguez-Barroso, Nuria, et al.
Veröffentlicht: (2024)
von: Rodríguez-Barroso, Nuria, et al.
Veröffentlicht: (2024)
Power side-channel leakage localization through adversarial training of deep neural networks
von: Gammell, Jimmy, et al.
Veröffentlicht: (2024)
von: Gammell, Jimmy, et al.
Veröffentlicht: (2024)
Are aligned neural networks adversarially aligned?
von: Carlini, Nicholas, et al.
Veröffentlicht: (2023)
von: Carlini, Nicholas, et al.
Veröffentlicht: (2023)
SHLIME: Foiling adversarial attacks fooling SHAP and LIME
von: Chauhan, Sam, et al.
Veröffentlicht: (2025)
von: Chauhan, Sam, et al.
Veröffentlicht: (2025)
Evaluating the Robustness of the "Ensemble Everything Everywhere" Defense
von: Zhang, Jie, et al.
Veröffentlicht: (2024)
von: Zhang, Jie, et al.
Veröffentlicht: (2024)
Performance Evaluation of Knowledge Graph Embedding Approaches under Non-adversarial Attacks
von: Kapoor, Sourabh, et al.
Veröffentlicht: (2024)
von: Kapoor, Sourabh, et al.
Veröffentlicht: (2024)
Evaluating the robustness of adversarial defenses in malware detection systems
von: Jafari, Mostafa, et al.
Veröffentlicht: (2025)
von: Jafari, Mostafa, et al.
Veröffentlicht: (2025)
Elevating Intrusion Detection and Security Fortification in Intelligent Networks through Cutting-Edge Machine Learning Paradigms
von: Munna, Md Minhazul Islam, et al.
Veröffentlicht: (2025)
von: Munna, Md Minhazul Islam, et al.
Veröffentlicht: (2025)
Privacy Side Channels in Machine Learning Systems
von: Debenedetti, Edoardo, et al.
Veröffentlicht: (2023)
von: Debenedetti, Edoardo, et al.
Veröffentlicht: (2023)
Understanding Deep Learning defenses Against Adversarial Examples Through Visualizations for Dynamic Risk Assessment
von: Echeberria-Barrio, Xabier, et al.
Veröffentlicht: (2024)
von: Echeberria-Barrio, Xabier, et al.
Veröffentlicht: (2024)
Certified Robustness to Clean-Label Poisoning Using Diffusion Denoising
von: Hong, Sanghyun, et al.
Veröffentlicht: (2024)
von: Hong, Sanghyun, et al.
Veröffentlicht: (2024)
Poisoning Web-Scale Training Datasets is Practical
von: Carlini, Nicholas, et al.
Veröffentlicht: (2023)
von: Carlini, Nicholas, et al.
Veröffentlicht: (2023)
Backdoor defense, learnability and obfuscation
von: Christiano, Paul, et al.
Veröffentlicht: (2024)
von: Christiano, Paul, et al.
Veröffentlicht: (2024)
New intelligent defense systems to reduce the risks of Selfish Mining and Double-Spending attacks using Learning Automata
von: Ghoreishi, Seyed Ardalan, et al.
Veröffentlicht: (2023)
von: Ghoreishi, Seyed Ardalan, et al.
Veröffentlicht: (2023)
Evaluating the Robustness of a Production Malware Detection System to Transferable Adversarial Attacks
von: Nasr, Milad, et al.
Veröffentlicht: (2025)
von: Nasr, Milad, et al.
Veröffentlicht: (2025)
Enhancing the Robustness of QMIX against State-adversarial Attacks
von: Guo, Weiran, et al.
Veröffentlicht: (2023)
von: Guo, Weiran, et al.
Veröffentlicht: (2023)
Large-scale online deanonymization with LLMs
von: Lermen, Simon, et al.
Veröffentlicht: (2026)
von: Lermen, Simon, et al.
Veröffentlicht: (2026)
Robust LLM safeguarding via refusal feature adversarial training
von: Yu, Lei, et al.
Veröffentlicht: (2024)
von: Yu, Lei, et al.
Veröffentlicht: (2024)
Stealing User Prompts from Mixture of Experts
von: Yona, Itay, et al.
Veröffentlicht: (2024)
von: Yona, Itay, et al.
Veröffentlicht: (2024)
Nonideality-aware training makes memristive networks more robust to adversarial attacks
von: Joksas, Dovydas, et al.
Veröffentlicht: (2024)
von: Joksas, Dovydas, et al.
Veröffentlicht: (2024)
The Attacker Moves Second: Stronger Adaptive Attacks Bypass Defenses Against Llm Jailbreaks and Prompt Injections
von: Nasr, Milad, et al.
Veröffentlicht: (2025)
von: Nasr, Milad, et al.
Veröffentlicht: (2025)
Exploring and Mitigating Adversarial Manipulation of Voting-Based Leaderboards
von: Huang, Yangsibo, et al.
Veröffentlicht: (2025)
von: Huang, Yangsibo, et al.
Veröffentlicht: (2025)
Command-line Obfuscation Detection using Small Language Models
von: Outrata, Vojtech, et al.
Veröffentlicht: (2024)
von: Outrata, Vojtech, et al.
Veröffentlicht: (2024)
Query-Based Adversarial Prompt Generation
von: Hayase, Jonathan, et al.
Veröffentlicht: (2024)
von: Hayase, Jonathan, et al.
Veröffentlicht: (2024)
Malacopula: adversarial automatic speaker verification attacks using a neural-based generalised Hammerstein model
von: Todisco, Massimiliano, et al.
Veröffentlicht: (2024)
von: Todisco, Massimiliano, et al.
Veröffentlicht: (2024)
Accuracy of TextFooler black box adversarial attacks on 01 loss sign activation neural network ensemble
von: Xue, Yunzhe, et al.
Veröffentlicht: (2024)
von: Xue, Yunzhe, et al.
Veröffentlicht: (2024)
Precision-Varying Prediction (PVP): Robustifying ASR systems against adversarial attacks
von: Pizarro, Matías, et al.
Veröffentlicht: (2026)
von: Pizarro, Matías, et al.
Veröffentlicht: (2026)
Designing an attack-defense game: how to increase robustness of financial transaction models via a competition
von: Zaytsev, Alexey, et al.
Veröffentlicht: (2023)
von: Zaytsev, Alexey, et al.
Veröffentlicht: (2023)
Trainwreck: A damaging adversarial attack on image classifiers
von: Zahálka, Jan
Veröffentlicht: (2023)
von: Zahálka, Jan
Veröffentlicht: (2023)
AutoAdv: Automated Adversarial Prompting for Multi-Turn Jailbreaking of Large Language Models
von: Reddy, Aashray, et al.
Veröffentlicht: (2025)
von: Reddy, Aashray, et al.
Veröffentlicht: (2025)
Tracing the Chain: Deep Learning for Stepping-Stone Intrusion Detection
von: Mathews, Nate, et al.
Veröffentlicht: (2026)
von: Mathews, Nate, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
AutoAdvExBench: Benchmarking autonomous exploitation of adversarial example defenses
von: Carlini, Nicholas, et al.
Veröffentlicht: (2025) -
Robustness bounds on the successful adversarial examples in probabilistic models: Implications from Gaussian processes
von: Maeshima, Hiroaki, et al.
Veröffentlicht: (2024) -
Remote Timing Attacks on Efficient Language Model Inference
von: Carlini, Nicholas, et al.
Veröffentlicht: (2024) -
Evading Black-box Classifiers Without Breaking Eggs
von: Debenedetti, Edoardo, et al.
Veröffentlicht: (2023) -
Position: Considerations for Differentially Private Learning with Large-Scale Public Pretraining
von: Tramèr, Florian, et al.
Veröffentlicht: (2022)