Counter-Samples: A Stateless Strategy to Neutralize Black Box Adversarial Attacks
Fuente:
arXiv
Salvato in:
| Autori principali: | Bokobza, Roey, Mirsky, Yisroel |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PEAS: A Strategy for Crafting Transferable Adversarial Examples
di: Avraham, Bar, et al.
Pubblicazione: (2024)
di: Avraham, Bar, et al.
Pubblicazione: (2024)
One Step to the Side: Why Defenses Against Malicious Finetuning Fail Under Adaptive Adversaries
di: Zloczower, Itay, et al.
Pubblicazione: (2026)
di: Zloczower, Itay, et al.
Pubblicazione: (2026)
The Best Defense is a Good Offense: Countering LLM-Powered Cyberattacks
di: Ayzenshteyn, Daniel, et al.
Pubblicazione: (2024)
di: Ayzenshteyn, Daniel, et al.
Pubblicazione: (2024)
Efficient Model Extraction via Boundary Sampling
di: Dor, Maor Biton, et al.
Pubblicazione: (2024)
di: Dor, Maor Biton, et al.
Pubblicazione: (2024)
Transpose Attack: Stealing Datasets with Bidirectional Training
di: Amit, Guy, et al.
Pubblicazione: (2023)
di: Amit, Guy, et al.
Pubblicazione: (2023)
GAVEL: Towards Rule-Based Safety Through Activation Monitoring
di: Rozenfeld, Shir, et al.
Pubblicazione: (2026)
di: Rozenfeld, Shir, et al.
Pubblicazione: (2026)
Transferability Ranking of Adversarial Examples
di: Levy, Mosh, et al.
Pubblicazione: (2022)
di: Levy, Mosh, et al.
Pubblicazione: (2022)
Runtime Detection of Adversarial Attacks in AI Accelerators Using Performance Counters
di: Rahaman, Habibur, et al.
Pubblicazione: (2025)
di: Rahaman, Habibur, et al.
Pubblicazione: (2025)
A White-Box Adversarial Attack Against a Digital Twin
di: Patterson, Wilson, et al.
Pubblicazione: (2022)
di: Patterson, Wilson, et al.
Pubblicazione: (2022)
What Was Your Prompt? A Remote Keylogging Attack on AI Assistants
di: Weiss, Roy, et al.
Pubblicazione: (2024)
di: Weiss, Roy, et al.
Pubblicazione: (2024)
SurvAttack: Black-Box Attack On Survival Models through Ontology-Informed EHR Perturbation
di: Kerdabadi, Mohsen Nayebi, et al.
Pubblicazione: (2024)
di: Kerdabadi, Mohsen Nayebi, et al.
Pubblicazione: (2024)
Universal Black-Box Reward Poisoning Attack against Offline Reinforcement Learning
di: Xu, Yinglun, et al.
Pubblicazione: (2024)
di: Xu, Yinglun, et al.
Pubblicazione: (2024)
Prompt-Induced Over-Generation as Denial-of-Service: A Black-Box Attack-Side Benchmark
di: Manu, et al.
Pubblicazione: (2025)
di: Manu, et al.
Pubblicazione: (2025)
A General Black-box Adversarial Attack on Graph-based Fake News Detectors
di: Zhu, Peican, et al.
Pubblicazione: (2024)
di: Zhu, Peican, et al.
Pubblicazione: (2024)
Turning Black Box into White Box: Dataset Distillation Leaks
di: Chen, Huajie, et al.
Pubblicazione: (2026)
di: Chen, Huajie, et al.
Pubblicazione: (2026)
Tree of Attacks: Jailbreaking Black-Box LLMs Automatically
di: Mehrotra, Anay, et al.
Pubblicazione: (2023)
di: Mehrotra, Anay, et al.
Pubblicazione: (2023)
Stateless Yet Not Forgetful: Implicit Memory as a Hidden Channel in LLMs
di: Salem, Ahmed, et al.
Pubblicazione: (2026)
di: Salem, Ahmed, et al.
Pubblicazione: (2026)
Black-box Adversarial Attacks on Network-wide Multi-step Traffic State Prediction Models
di: Poudel, Bibek, et al.
Pubblicazione: (2021)
di: Poudel, Bibek, et al.
Pubblicazione: (2021)
TRAP: Targeted Random Adversarial Prompt Honeypot for Black-Box Identification
di: Gubri, Martin, et al.
Pubblicazione: (2024)
di: Gubri, Martin, et al.
Pubblicazione: (2024)
PAL: Proxy-Guided Black-Box Attack on Large Language Models
di: Sitawarin, Chawin, et al.
Pubblicazione: (2024)
di: Sitawarin, Chawin, et al.
Pubblicazione: (2024)
Untargeted Adversarial Attack on Knowledge Graph Embeddings
di: Zhao, Tianzhe, et al.
Pubblicazione: (2024)
di: Zhao, Tianzhe, et al.
Pubblicazione: (2024)
Relationship between Uncertainty in DNNs and Adversarial Attacks
di: Ogonna, Mabel, et al.
Pubblicazione: (2024)
di: Ogonna, Mabel, et al.
Pubblicazione: (2024)
On the Robustness of Bayesian Neural Networks to Adversarial Attacks
di: Bortolussi, Luca, et al.
Pubblicazione: (2022)
di: Bortolussi, Luca, et al.
Pubblicazione: (2022)
Adversarial Attacks on Transformers-Based Malware Detectors
di: Jakhotiya, Yash, et al.
Pubblicazione: (2022)
di: Jakhotiya, Yash, et al.
Pubblicazione: (2022)
Adversarial Agents: Black-Box Evasion Attacks with Reinforcement Learning
di: Domico, Kyle, et al.
Pubblicazione: (2025)
di: Domico, Kyle, et al.
Pubblicazione: (2025)
Adversarial Attacks on Reinforcement Learning-based Medical Questionnaire Systems: Input-level Perturbation Strategies and Medical Constraint Validation
di: Liu, Peizhuo
Pubblicazione: (2025)
di: Liu, Peizhuo
Pubblicazione: (2025)
Vision Transformer with Adversarial Indicator Token against Adversarial Attacks in Radio Signal Classifications
di: Zhang, Lu, et al.
Pubblicazione: (2025)
di: Zhang, Lu, et al.
Pubblicazione: (2025)
Trust Me, I Know This Function: Hijacking LLM Static Analysis using Bias
di: Bernstein, Shir, et al.
Pubblicazione: (2025)
di: Bernstein, Shir, et al.
Pubblicazione: (2025)
POT: Inducing Overthinking in LLMs via Black-Box Iterative Optimization
di: Li, Xinyu, et al.
Pubblicazione: (2025)
di: Li, Xinyu, et al.
Pubblicazione: (2025)
Quantifying the Noise of Structural Perturbations on Graph Adversarial Attacks
di: Fang, Junyuan, et al.
Pubblicazione: (2025)
di: Fang, Junyuan, et al.
Pubblicazione: (2025)
Vulnerability Disclosure through Adaptive Black-Box Adversarial Attacks on NIDS
di: Ennaji, Sabrine, et al.
Pubblicazione: (2025)
di: Ennaji, Sabrine, et al.
Pubblicazione: (2025)
Disttack: Graph Adversarial Attacks Toward Distributed GNN Training
di: Zhang, Yuxiang, et al.
Pubblicazione: (2024)
di: Zhang, Yuxiang, et al.
Pubblicazione: (2024)
Investigating Imperceptibility of Adversarial Attacks on Tabular Data: An Empirical Analysis
di: He, Zhipeng, et al.
Pubblicazione: (2024)
di: He, Zhipeng, et al.
Pubblicazione: (2024)
A Generative Approach to Surrogate-based Black-box Attacks
di: Moraffah, Raha, et al.
Pubblicazione: (2024)
di: Moraffah, Raha, et al.
Pubblicazione: (2024)
SoK: Pitfalls in Evaluating Black-Box Attacks
di: Suya, Fnu, et al.
Pubblicazione: (2023)
di: Suya, Fnu, et al.
Pubblicazione: (2023)
Variational Randomized Smoothing for Sample-Wise Adversarial Robustness
di: Hase, Ryo, et al.
Pubblicazione: (2024)
di: Hase, Ryo, et al.
Pubblicazione: (2024)
MF-CLIP: Leveraging CLIP as Surrogate Models for No-box Adversarial Attacks
di: Zhang, Jiaming, et al.
Pubblicazione: (2023)
di: Zhang, Jiaming, et al.
Pubblicazione: (2023)
Claudini: Autoresearch Discovers State-of-the-Art Adversarial Attack Algorithms for LLMs
di: Panfilov, Alexander, et al.
Pubblicazione: (2026)
di: Panfilov, Alexander, et al.
Pubblicazione: (2026)
Practical Adversarial Attacks on Stochastic Bandits via Fake Data Injection
di: Zeng, Qirun, et al.
Pubblicazione: (2025)
di: Zeng, Qirun, et al.
Pubblicazione: (2025)
Remote Rowhammer Attack using Adversarial Observations on Federated Learning Clients
di: Yuan, Jinsheng, et al.
Pubblicazione: (2025)
di: Yuan, Jinsheng, et al.
Pubblicazione: (2025)
Documenti analoghi
-
PEAS: A Strategy for Crafting Transferable Adversarial Examples
di: Avraham, Bar, et al.
Pubblicazione: (2024) -
One Step to the Side: Why Defenses Against Malicious Finetuning Fail Under Adaptive Adversaries
di: Zloczower, Itay, et al.
Pubblicazione: (2026) -
The Best Defense is a Good Offense: Countering LLM-Powered Cyberattacks
di: Ayzenshteyn, Daniel, et al.
Pubblicazione: (2024) -
Efficient Model Extraction via Boundary Sampling
di: Dor, Maor Biton, et al.
Pubblicazione: (2024) -
Transpose Attack: Stealing Datasets with Bidirectional Training
di: Amit, Guy, et al.
Pubblicazione: (2023)