A general approach to enhance the survivability of backdoor attacks by decision path coupling
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhao, Yufei, Wang, Dingji, Chen, Bihuan, Chen, Ziqian, Peng, Xin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Effective backdoor attack on graph neural networks in link prediction tasks
di: Dai, Jiazhu, et al.
Pubblicazione: (2024)
di: Dai, Jiazhu, et al.
Pubblicazione: (2024)
The last Dance : Robust backdoor attack via diffusion models and bayesian approach
di: Mengara, Orson
Pubblicazione: (2024)
di: Mengara, Orson
Pubblicazione: (2024)
SDD: Self-Degraded Defense against Malicious Fine-tuning
di: Chen, Zixuan, et al.
Pubblicazione: (2025)
di: Chen, Zixuan, et al.
Pubblicazione: (2025)
FL-CLEANER: byzantine and backdoor defense by CLustering Errors of Activation maps in Non-iid fedErated leaRning
di: Ghali, Mehdi Ben, et al.
Pubblicazione: (2025)
di: Ghali, Mehdi Ben, et al.
Pubblicazione: (2025)
MEASER: Malware embedding attacks on open-source LLMs
di: Tan, Ming, et al.
Pubblicazione: (2025)
di: Tan, Ming, et al.
Pubblicazione: (2025)
DLP: towards active defense against backdoor attacks with decoupled learning process
di: Ying, Zonghao, et al.
Pubblicazione: (2024)
di: Ying, Zonghao, et al.
Pubblicazione: (2024)
PrivacyRestore: Privacy-Preserving Inference in Large Language Models via Privacy Removal and Restoration
di: Zeng, Ziqian, et al.
Pubblicazione: (2024)
di: Zeng, Ziqian, et al.
Pubblicazione: (2024)
Adversarial attacks against Modern Vision-Language Models
di: La Torre, Alejandro Paredes
Pubblicazione: (2026)
di: La Torre, Alejandro Paredes
Pubblicazione: (2026)
AutoAttacker: A Large Language Model Guided System to Implement Automatic Cyber-attacks
di: Xu, Jiacen, et al.
Pubblicazione: (2024)
di: Xu, Jiacen, et al.
Pubblicazione: (2024)
Context manipulation attacks : Web agents are susceptible to corrupted memory
di: Patlan, Atharv Singh, et al.
Pubblicazione: (2025)
di: Patlan, Atharv Singh, et al.
Pubblicazione: (2025)
RewardDS: Privacy-Preserving Fine-Tuning for Large Language Models via Reward Driven Data Synthesis
di: Wang, Jianwei, et al.
Pubblicazione: (2025)
di: Wang, Jianwei, et al.
Pubblicazione: (2025)
Detection of ransomware attacks using federated learning based on the CNN model
di: Nguyen, Hong-Nhung, et al.
Pubblicazione: (2024)
di: Nguyen, Hong-Nhung, et al.
Pubblicazione: (2024)
Unleash the Power of Ellipsis: Accuracy-enhanced Sparse Vector Technique with Exponential Noise
di: Liu, Yuhan, et al.
Pubblicazione: (2024)
di: Liu, Yuhan, et al.
Pubblicazione: (2024)
Capturing the security expert knowledge in feature selection for web application attack detection
di: Riverol, Amanda, et al.
Pubblicazione: (2024)
di: Riverol, Amanda, et al.
Pubblicazione: (2024)
Optimized detection of cyber-attacks on IoT networks via hybrid deep learning models
di: Bensaoud, Ahmed, et al.
Pubblicazione: (2025)
di: Bensaoud, Ahmed, et al.
Pubblicazione: (2025)
LLMs unlock new paths to monetizing exploits
di: Carlini, Nicholas, et al.
Pubblicazione: (2025)
di: Carlini, Nicholas, et al.
Pubblicazione: (2025)
DAIRE: A lightweight AI model for real-time detection of Controller Area Network attacks in the Internet of Vehicles
di: Alam, Shahid, et al.
Pubblicazione: (2026)
di: Alam, Shahid, et al.
Pubblicazione: (2026)
A clean-label graph backdoor attack method in node classification task
di: Xing, Xiaogang, et al.
Pubblicazione: (2023)
di: Xing, Xiaogang, et al.
Pubblicazione: (2023)
Retrieval-Confused Generation is a Good Defender for Privacy Violation Attack of Large Language Models
di: Peng, Wanli, et al.
Pubblicazione: (2025)
di: Peng, Wanli, et al.
Pubblicazione: (2025)
An incremental hybrid adaptive network-based IDS in Software Defined Networks to detect stealth attacks
di: Alqahtani, Abdullah H
Pubblicazione: (2024)
di: Alqahtani, Abdullah H
Pubblicazione: (2024)
BF-Meta: Secure Blockchain-enhanced Privacy-preserving Federated Learning for Metaverse
di: Liu, Wenbo, et al.
Pubblicazione: (2024)
di: Liu, Wenbo, et al.
Pubblicazione: (2024)
Problem space structural adversarial attacks for Network Intrusion Detection Systems based on Graph Neural Networks
di: Venturi, Andrea, et al.
Pubblicazione: (2024)
di: Venturi, Andrea, et al.
Pubblicazione: (2024)
Swallowing the Poison Pills: Insights from Vulnerability Disparity Among LLMs
di: Yifeng, Peng, et al.
Pubblicazione: (2025)
di: Yifeng, Peng, et al.
Pubblicazione: (2025)
Synthetic is all you need: removing the auxiliary data assumption for membership inference attacks against synthetic data
di: Guépin, Florent, et al.
Pubblicazione: (2023)
di: Guépin, Florent, et al.
Pubblicazione: (2023)
The dark deep side of DeepSeek: Fine-tuning attacks against the safety alignment of CoT-enabled models
di: Xu, Zhiyuan, et al.
Pubblicazione: (2025)
di: Xu, Zhiyuan, et al.
Pubblicazione: (2025)
Analysis of the vulnerability of machine learning regression models to adversarial attacks using data from 5G wireless networks
di: Legashev, Leonid, et al.
Pubblicazione: (2025)
di: Legashev, Leonid, et al.
Pubblicazione: (2025)
VTarbel: Targeted Label Attack with Minimal Knowledge on Detector-enhanced Vertical Federated Learning
di: Tan, Juntao, et al.
Pubblicazione: (2025)
di: Tan, Juntao, et al.
Pubblicazione: (2025)
AIAuditTrack: A Framework for AI Security system
di: Luo, Zixun, et al.
Pubblicazione: (2025)
di: Luo, Zixun, et al.
Pubblicazione: (2025)
Activation-Guided Local Editing for Jailbreaking Attacks
di: Wang, Jiecong, et al.
Pubblicazione: (2025)
di: Wang, Jiecong, et al.
Pubblicazione: (2025)
Don't Click That: Teaching Web Agents to Resist Deceptive Interfaces
di: Zhang, Yilin, et al.
Pubblicazione: (2026)
di: Zhang, Yilin, et al.
Pubblicazione: (2026)
Persistent Backdoor Attacks under Continual Fine-Tuning of LLMs
di: Cui, Jing, et al.
Pubblicazione: (2025)
di: Cui, Jing, et al.
Pubblicazione: (2025)
Analysis and prevention of AI-based phishing email attacks
di: Eze, Chibuike Samuel, et al.
Pubblicazione: (2024)
di: Eze, Chibuike Samuel, et al.
Pubblicazione: (2024)
On the use of neurosymbolic AI for defending against cyber attacks
di: Grov, Gudmund, et al.
Pubblicazione: (2024)
di: Grov, Gudmund, et al.
Pubblicazione: (2024)
On the critical path to implant backdoors and the effectiveness of potential mitigation techniques: Early learnings from XZ
di: Lins, Mario, et al.
Pubblicazione: (2024)
di: Lins, Mario, et al.
Pubblicazione: (2024)
Sandwich attack: Multi-language Mixture Adaptive Attack on LLMs
di: Upadhayay, Bibek, et al.
Pubblicazione: (2024)
di: Upadhayay, Bibek, et al.
Pubblicazione: (2024)
AIRGuard: Guarding Agent Actions with Runtime Authority Control
di: Qin, Suliu, et al.
Pubblicazione: (2026)
di: Qin, Suliu, et al.
Pubblicazione: (2026)
Trading Devil: Robust backdoor attack via Stochastic investment models and Bayesian approach
di: Mengara, Orson
Pubblicazione: (2024)
di: Mengara, Orson
Pubblicazione: (2024)
How Does Naming Affect LLMs on Code Analysis Tasks?
di: Wang, Zhilong, et al.
Pubblicazione: (2023)
di: Wang, Zhilong, et al.
Pubblicazione: (2023)
LLM App Squatting and Cloning
di: Xie, Yinglin, et al.
Pubblicazione: (2024)
di: Xie, Yinglin, et al.
Pubblicazione: (2024)
Large Language Models for Cyber Security: A Systematic Literature Review
di: Xu, Hanxiang, et al.
Pubblicazione: (2024)
di: Xu, Hanxiang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Effective backdoor attack on graph neural networks in link prediction tasks
di: Dai, Jiazhu, et al.
Pubblicazione: (2024) -
The last Dance : Robust backdoor attack via diffusion models and bayesian approach
di: Mengara, Orson
Pubblicazione: (2024) -
SDD: Self-Degraded Defense against Malicious Fine-tuning
di: Chen, Zixuan, et al.
Pubblicazione: (2025) -
FL-CLEANER: byzantine and backdoor defense by CLustering Errors of Activation maps in Non-iid fedErated leaRning
di: Ghali, Mehdi Ben, et al.
Pubblicazione: (2025) -
MEASER: Malware embedding attacks on open-source LLMs
di: Tan, Ming, et al.
Pubblicazione: (2025)