How adversarial attacks can disrupt seemingly stable accurate classifiers
Fuente:
arXiv
Saved in:
| Main Authors: | Sutton, Oliver J., Zhou, Qinghua, Tyukin, Ivan Y., Gorban, Alexander N., Bastounis, Alexander, Higham, Desmond J. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Stealth edits to large language models
by: Sutton, Oliver J., et al.
Published: (2024)
by: Sutton, Oliver J., et al.
Published: (2024)
The Boundaries of Verifiable Accuracy, Robustness, and Generalisation in Deep Learning
by: Bastounis, Alexander, et al.
Published: (2023)
by: Bastounis, Alexander, et al.
Published: (2023)
Staining and locking computer vision models without retraining
by: Sutton, Oliver J., et al.
Published: (2025)
by: Sutton, Oliver J., et al.
Published: (2025)
Weakly Supervised Learners for Correction of AI Errors with Provable Performance Guarantees
by: Tyukin, Ivan Y., et al.
Published: (2024)
by: Tyukin, Ivan Y., et al.
Published: (2024)
Harnessing non-adversarial robustness in large language models
by: Zhou, Qinghua, et al.
Published: (2026)
by: Zhou, Qinghua, et al.
Published: (2026)
StepProof: Step-by-step verification of natural language mathematical proofs
by: Hu, Xiaolin, et al.
Published: (2025)
by: Hu, Xiaolin, et al.
Published: (2025)
The mathematics of adversarial attacks in AI -- Why deep learning is unstable despite the existence of stable neural networks
by: Bastounis, Alexander, et al.
Published: (2021)
by: Bastounis, Alexander, et al.
Published: (2021)
Embedding Hidden Adversarial Capabilities in Pre-Trained Diffusion Models
by: Beerens, Lucas, et al.
Published: (2025)
by: Beerens, Lucas, et al.
Published: (2025)
Deceptive Diffusion: Generating Synthetic Adversarial Examples
by: Beerens, Lucas, et al.
Published: (2024)
by: Beerens, Lucas, et al.
Published: (2024)
Deterministic versus stochastic dynamical classifiers: opposing random adversarial attacks with noise
by: Chicchi, Lorenzo, et al.
Published: (2024)
by: Chicchi, Lorenzo, et al.
Published: (2024)
Are We Measuring Oversmoothing in Graph Neural Networks Correctly?
by: Zhang, Kaicheng, et al.
Published: (2025)
by: Zhang, Kaicheng, et al.
Published: (2025)
Enhancing Inference Efficiency of Large Language Models: Investigating Optimization Strategies and Architectural Innovations
by: Tyukin, Georgy
Published: (2024)
by: Tyukin, Georgy
Published: (2024)
Bias in the Mirror: Are LLMs opinions robust to their own adversarial attacks ?
by: Rennard, Virgile, et al.
Published: (2024)
by: Rennard, Virgile, et al.
Published: (2024)
Fixed-point graph convolutional networks against adversarial attacks
by: Khan, Shakib, et al.
Published: (2025)
by: Khan, Shakib, et al.
Published: (2025)
When fractional quasi p-norms concentrate
by: Tyukin, Ivan Y., et al.
Published: (2025)
by: Tyukin, Ivan Y., et al.
Published: (2025)
Discrete optimal transport is a strong audio adversarial attack
by: Selitskiy, Anton, et al.
Published: (2025)
by: Selitskiy, Anton, et al.
Published: (2025)
Effective faking of verbal deception detection with target-aligned adversarial attacks
by: Kleinberg, Bennett, et al.
Published: (2025)
by: Kleinberg, Bennett, et al.
Published: (2025)
Deep generative models as an adversarial attack strategy for tabular machine learning
by: Dyrmishi, Salijona, et al.
Published: (2024)
by: Dyrmishi, Salijona, et al.
Published: (2024)
Properties that allow or prohibit transferability of adversarial attacks among quantized networks
by: Shrestha, Abhishek, et al.
Published: (2024)
by: Shrestha, Abhishek, et al.
Published: (2024)
Empirical evaluation of the Frank-Wolfe methods for constructing white-box adversarial attacks
by: Korotkova, Kristina, et al.
Published: (2025)
by: Korotkova, Kristina, et al.
Published: (2025)
Can adversarial attacks by large language models be attributed?
by: Cebrian, Manuel, et al.
Published: (2024)
by: Cebrian, Manuel, et al.
Published: (2024)
Multi-style conversion for semantic segmentation of lesions in fundus images by adversarial attacks
by: Playout, Clément, et al.
Published: (2024)
by: Playout, Clément, et al.
Published: (2024)
Krum Federated Chain (KFC): Using blockchain to defend against adversarial attacks in Federated Learning
by: García-Márquez, Mario, et al.
Published: (2025)
by: García-Márquez, Mario, et al.
Published: (2025)
GPT-ology, Computational Models, Silicon Sampling: How should we think about LLMs in Cognitive Science?
by: Ong, Desmond C.
Published: (2024)
by: Ong, Desmond C.
Published: (2024)
FRAUD-RLA: A new reinforcement learning adversarial attack against credit card fraud detection
by: Lunghi, Daniele, et al.
Published: (2025)
by: Lunghi, Daniele, et al.
Published: (2025)
Problem space structural adversarial attacks for Network Intrusion Detection Systems based on Graph Neural Networks
by: Venturi, Andrea, et al.
Published: (2024)
by: Venturi, Andrea, et al.
Published: (2024)
On the consistent reasoning paradox of intelligence and optimal trust in AI: The power of 'I don't know'
by: Bastounis, Alexander, et al.
Published: (2024)
by: Bastounis, Alexander, et al.
Published: (2024)
Cross-site scripting adversarial attacks based on deep reinforcement learning: Evaluation and extension study
by: Pasini, Samuele, et al.
Published: (2025)
by: Pasini, Samuele, et al.
Published: (2025)
Unveiling the optimization process of Physics Informed Neural Networks: How accurate and competitive can PINNs be?
by: Urbán, Jorge F., et al.
Published: (2024)
by: Urbán, Jorge F., et al.
Published: (2024)
Can an AI-Powered Presentation Platform Based On The Game "Just a Minute" Be Used To Improve Students' Public Speaking Skills?
by: Higham, Frederic, et al.
Published: (2025)
by: Higham, Frederic, et al.
Published: (2025)
Analysis of the vulnerability of machine learning regression models to adversarial attacks using data from 5G wireless networks
by: Legashev, Leonid, et al.
Published: (2025)
by: Legashev, Leonid, et al.
Published: (2025)
The Neuromorphic Supremacy
by: Tsybina, Yuliya, et al.
Published: (2026)
by: Tsybina, Yuliya, et al.
Published: (2026)
PBCAT: Patch-based composite adversarial training against physically realizable attacks on object detection
by: Li, Xiao, et al.
Published: (2025)
by: Li, Xiao, et al.
Published: (2025)
Accuracy of TextFooler black box adversarial attacks on 01 loss sign activation neural network ensemble
by: Xue, Yunzhe, et al.
Published: (2024)
by: Xue, Yunzhe, et al.
Published: (2024)
RAB$^2$-DEF: Dynamic and explainable defense against adversarial attacks in Federated Learning to fair poor clients
by: Rodríguez-Barroso, Nuria, et al.
Published: (2024)
by: Rodríguez-Barroso, Nuria, et al.
Published: (2024)
To what extent can ASV systems naturally defend against spoofing attacks?
by: Jung, Jee-weon, et al.
Published: (2024)
by: Jung, Jee-weon, et al.
Published: (2024)
A robust three-way classifier with shadowed granular-balls based on justifiable granularity
by: Yang, Jie, et al.
Published: (2024)
by: Yang, Jie, et al.
Published: (2024)
Vision Transformers: the threat of realistic adversarial patches
by: Cools, Kasper, et al.
Published: (2025)
by: Cools, Kasper, et al.
Published: (2025)
Trainwreck: A damaging adversarial attack on image classifiers
by: Zahálka, Jan
Published: (2023)
by: Zahálka, Jan
Published: (2023)
What is Hiding in Medicine's Dark Matter? Learning with Missing Data in Medical Practices
by: Suzen, Neslihan, et al.
Published: (2024)
by: Suzen, Neslihan, et al.
Published: (2024)
Similar Items
-
Stealth edits to large language models
by: Sutton, Oliver J., et al.
Published: (2024) -
The Boundaries of Verifiable Accuracy, Robustness, and Generalisation in Deep Learning
by: Bastounis, Alexander, et al.
Published: (2023) -
Staining and locking computer vision models without retraining
by: Sutton, Oliver J., et al.
Published: (2025) -
Weakly Supervised Learners for Correction of AI Errors with Provable Performance Guarantees
by: Tyukin, Ivan Y., et al.
Published: (2024) -
Harnessing non-adversarial robustness in large language models
by: Zhou, Qinghua, et al.
Published: (2026)