Seeing Is Not Always Believing: Invisible Collision Attack and Defence on Pre-Trained Models
Fuente:
arXiv
Saved in:
| Main Authors: | Deng, Minghang, Zhang, Zhong, Shao, Junming |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Large Language Models are Autonomous Cyber Defenders
by: Castro, Sebastián R., et al.
Published: (2025)
by: Castro, Sebastián R., et al.
Published: (2025)
TrojanTime: Backdoor Attacks on Time Series Classification
by: Dong, Chang, et al.
Published: (2025)
by: Dong, Chang, et al.
Published: (2025)
PatchBlock: A Lightweight Defense Against Adversarial Patches for Embedded EdgeAI Devices
by: Chattopadhyay, Nandish, et al.
Published: (2026)
by: Chattopadhyay, Nandish, et al.
Published: (2026)
David and Goliath: An Empirical Evaluation of Attacks and Defenses for QNNs at the Deep Edge
by: Costa, Miguel, et al.
Published: (2024)
by: Costa, Miguel, et al.
Published: (2024)
Attack Selection Reduces Safety in Concentrated AI Control Settings against Trusted Monitoring
by: Schaeffer, Joachim, et al.
Published: (2026)
by: Schaeffer, Joachim, et al.
Published: (2026)
Combating Phone Scams with LLM-based Detection: Where Do We Stand?
by: Shen, Zitong, et al.
Published: (2024)
by: Shen, Zitong, et al.
Published: (2024)
Machine Learning-Based Detection of MCP Attacks
by: Mattsson, Tobias, et al.
Published: (2026)
by: Mattsson, Tobias, et al.
Published: (2026)
TSFool: Crafting Highly-Imperceptible Adversarial Time Series through Multi-Objective Attack
by: Wang, Yanyun, et al.
Published: (2022)
by: Wang, Yanyun, et al.
Published: (2022)
Scalable APT Malware Classification via Parallel Feature Extraction and GPU-Accelerated Learning
by: Subedar, Noah, et al.
Published: (2025)
by: Subedar, Noah, et al.
Published: (2025)
Correlation Analysis of Adversarial Attack in Time Series Classification
by: Li, Zhengyang, et al.
Published: (2024)
by: Li, Zhengyang, et al.
Published: (2024)
An Efficient Gradient-Based Inference Attack for Federated Learning
by: Montaña-Fernández, Pablo, et al.
Published: (2025)
by: Montaña-Fernández, Pablo, et al.
Published: (2025)
Multiclass Classification Procedure for Detecting Attacks on MQTT-IoT Protocol
by: Alaiz-Moreton, Hector, et al.
Published: (2024)
by: Alaiz-Moreton, Hector, et al.
Published: (2024)
SMSI: System Model Security Inference: Automated Threat Modeling for Cyber-Physical Systems
by: Radaideh, RoÝah, et al.
Published: (2026)
by: Radaideh, RoÝah, et al.
Published: (2026)
Kill-Chain Canaries: Stage-Level Tracking of Prompt Injection Across Attack Surfaces and Model Safety Tiers
by: Wang, Haochuan Kevin, et al.
Published: (2026)
by: Wang, Haochuan Kevin, et al.
Published: (2026)
Defense Against the Dark Prompts: Mitigating Best-of-N Jailbreaking with Prompt Evaluation
by: Armstrong, Stuart, et al.
Published: (2025)
by: Armstrong, Stuart, et al.
Published: (2025)
Safety, Security, and Cognitive Risks in State-Space Models: A Systematic Threat Analysis with Spectral, Stateful, and Capacity Attacks
by: Parmar, Manoj
Published: (2026)
by: Parmar, Manoj
Published: (2026)
Forecasting Anonymized Electricity Load Profiles
by: Fernandez, Joaquin Delgado, et al.
Published: (2025)
by: Fernandez, Joaquin Delgado, et al.
Published: (2025)
A Protocol-Language Model for Network Intrusion (Without Deep Packet Inspection)
by: Sharma, Vivek Kumar
Published: (2026)
by: Sharma, Vivek Kumar
Published: (2026)
Deep Learning-Based Intrusion Detection for Automotive Ethernet: Evaluating & Optimizing Fast Inference Techniques for Deployment on Low-Cost Platform
by: Carmo, Pedro R. X., et al.
Published: (2025)
by: Carmo, Pedro R. X., et al.
Published: (2025)
A Taxonomy of Data Risks in AI and Quantum Computing (QAI) - A Systematic Review
by: Billiris, Grace, et al.
Published: (2025)
by: Billiris, Grace, et al.
Published: (2025)
Ollabench: Evaluating LLMs' Reasoning for Human-centric Interdependent Cybersecurity
by: Nguyen, Tam n.
Published: (2024)
by: Nguyen, Tam n.
Published: (2024)
Impact of Phonetics on Speaker Identity in Adversarial Voice Attack
by: Dar, Daniyal Kabir, et al.
Published: (2025)
by: Dar, Daniyal Kabir, et al.
Published: (2025)
Facebook Report on Privacy of fNIRS data
by: Hossen, Md Imran, et al.
Published: (2024)
by: Hossen, Md Imran, et al.
Published: (2024)
Jailbreak Mimicry: Automated Discovery of Narrative-Based Jailbreaks for Large Language Models
by: Ntais, Pavlos
Published: (2025)
by: Ntais, Pavlos
Published: (2025)
Variables are a Curse in Software Vulnerability Prediction
by: Groppe, Jinghua, et al.
Published: (2024)
by: Groppe, Jinghua, et al.
Published: (2024)
Lightweight LLMs for Network Attack Detection in IoT Networks
by: Sudasinghe, Piyumi Bhagya, et al.
Published: (2026)
by: Sudasinghe, Piyumi Bhagya, et al.
Published: (2026)
Robust DDoS-Attack Classification with 3D CNNs Against Adversarial Methods
by: Bragg, Landon, et al.
Published: (2025)
by: Bragg, Landon, et al.
Published: (2025)
Preliminary study on artificial intelligence methods for cybersecurity threat detection in computer networks based on raw data packets
by: Ogonowski, Aleksander, et al.
Published: (2024)
by: Ogonowski, Aleksander, et al.
Published: (2024)
Influence of Autoencoder Latent Space on Classifying IoT CoAP Attacks
by: García-Ordás, María Teresa, et al.
Published: (2026)
by: García-Ordás, María Teresa, et al.
Published: (2026)
AI-Driven Security Alert Screening and Alert Fatigue Mitigation in Security Operations Centers: A Survey
by: Ndichu, Samuel, et al.
Published: (2026)
by: Ndichu, Samuel, et al.
Published: (2026)
Peak + Accumulation: A Proxy-Level Scoring Formula for Multi-Turn LLM Attack Detection
by: Corll, J Alex
Published: (2026)
by: Corll, J Alex
Published: (2026)
Whisper Leak: a side-channel attack on Large Language Models
by: McDonald, Geoff, et al.
Published: (2025)
by: McDonald, Geoff, et al.
Published: (2025)
Temporal Attack Pattern Detection in Multi-Agent AI Workflows: An Open Framework for Training Trace-Based Security Models
by: Del Rosario, Ron F.
Published: (2025)
by: Del Rosario, Ron F.
Published: (2025)
How Worrying Are Privacy Attacks Against Machine Learning?
by: Domingo-Ferrer, Josep
Published: (2025)
by: Domingo-Ferrer, Josep
Published: (2025)
MaPPing Your Model: Assessing the Impact of Adversarial Attacks on LLM-based Programming Assistants
by: Heibel, John, et al.
Published: (2024)
by: Heibel, John, et al.
Published: (2024)
PARD-SSM: Probabilistic Cyber-Attack Regime Detection via Variational Switching State-Space Models
by: Hiremath, Prakul Sunil, et al.
Published: (2026)
by: Hiremath, Prakul Sunil, et al.
Published: (2026)
Privacy in the Age of AI: A Taxonomy of Data Risks
by: Billiris, Grace, et al.
Published: (2025)
by: Billiris, Grace, et al.
Published: (2025)
Sparse vs Contiguous Adversarial Pixel Perturbations in Multimodal Models: An Empirical Analysis
by: Botocan, Cristian-Alexandru, et al.
Published: (2024)
by: Botocan, Cristian-Alexandru, et al.
Published: (2024)
A Survey on the Security of Long-Term Memory in LLM Agents: Toward Mnemonic Sovereignty
by: Lin, Zehao, et al.
Published: (2026)
by: Lin, Zehao, et al.
Published: (2026)
Organizational Adaptation to Generative AI in Cybersecurity
by: Nott, Christopher
Published: (2025)
by: Nott, Christopher
Published: (2025)
Similar Items
-
Large Language Models are Autonomous Cyber Defenders
by: Castro, Sebastián R., et al.
Published: (2025) -
TrojanTime: Backdoor Attacks on Time Series Classification
by: Dong, Chang, et al.
Published: (2025) -
PatchBlock: A Lightweight Defense Against Adversarial Patches for Embedded EdgeAI Devices
by: Chattopadhyay, Nandish, et al.
Published: (2026) -
David and Goliath: An Empirical Evaluation of Attacks and Defenses for QNNs at the Deep Edge
by: Costa, Miguel, et al.
Published: (2024) -
Attack Selection Reduces Safety in Concentrated AI Control Settings against Trusted Monitoring
by: Schaeffer, Joachim, et al.
Published: (2026)