Breaking Down the Defenses: A Comparative Survey of Attacks on Large Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Chowdhury, Arijit Ghosh, Islam, Md Mofijul, Kumar, Vaibhav, Shezan, Faysal Hossain, Jain, Vinija, Chadha, Aman |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Breaking the illusion: Automated Reasoning of GDPR Consent Violations
por: Li, Ying, et al.
Publicado: (2025)
por: Li, Ying, et al.
Publicado: (2025)
A Novel XAI-Enhanced Quantum Adversarial Networks for Velocity Dispersion Modeling in MaNGA Galaxies
por: Narkedimilli, Sathwik, et al.
Publicado: (2025)
por: Narkedimilli, Sathwik, et al.
Publicado: (2025)
AntiFLipper: A Secure and Efficient Defense Against Label-Flipping Attacks in Federated Learning
por: Rahman, Aashnan, et al.
Publicado: (2025)
por: Rahman, Aashnan, et al.
Publicado: (2025)
Quantum-Resistant Authentication Scheme for RFID Systems Using Lattice-Based Cryptography
por: Kumar, Vaibhav, et al.
Publicado: (2025)
por: Kumar, Vaibhav, et al.
Publicado: (2025)
DSTAN-Med: Dual-Channel Spatiotemporal Attention with Physiological Plausibility Filtering for False Data Injection Attack Detection in IoT-Based Medical Devices
por: Hasan, Md Mehedi, et al.
Publicado: (2026)
por: Hasan, Md Mehedi, et al.
Publicado: (2026)
Latent Geometry as a Structural Monitor: Eigenspace Alignment for Anomaly Detection in Anonymity Networks
por: Chhabra, Vaibhav
Publicado: (2026)
por: Chhabra, Vaibhav
Publicado: (2026)
Options, Not Clicks: Lattice Refinement for Consent-Driven MCP Authorization
por: Li, Ying, et al.
Publicado: (2026)
por: Li, Ying, et al.
Publicado: (2026)
LLM-Driven SAST-Genius: A Hybrid Static Analysis Framework for Comprehensive and Actionable Security
por: Agrawal, Vaibhav, et al.
Publicado: (2025)
por: Agrawal, Vaibhav, et al.
Publicado: (2025)
How Resilient is QUIC to Security and Privacy Attacks?
por: Sengupta, Jayasree, et al.
Publicado: (2024)
por: Sengupta, Jayasree, et al.
Publicado: (2024)
PRvL: Quantifying the Capabilities and Risks of Large Language Models for PII Redaction
por: Garza, Leon, et al.
Publicado: (2025)
por: Garza, Leon, et al.
Publicado: (2025)
System Prompt Extraction Attacks and Defenses in Large Language Models
por: Das, Badhan Chandra, et al.
Publicado: (2025)
por: Das, Badhan Chandra, et al.
Publicado: (2025)
Adversarial Threats in Quantum Machine Learning: A Survey of Attacks and Defenses
por: Ghosh, Archisman, et al.
Publicado: (2025)
por: Ghosh, Archisman, et al.
Publicado: (2025)
A Survey on Backdoor Threats in Large Language Models (LLMs): Attacks, Defenses, and Evaluations
por: Zhou, Yihe, et al.
Publicado: (2025)
por: Zhou, Yihe, et al.
Publicado: (2025)
"I Strongly Suspect This Website Is a Scam": Benchmarking PII Leakage and Detection without Defense in Autonomous Web Agents
por: Roy, Soham, et al.
Publicado: (2026)
por: Roy, Soham, et al.
Publicado: (2026)
Adversarial Attacks on Multimodal Large Language Models: A Comprehensive Survey
por: Jain, Bhavuk, et al.
Publicado: (2026)
por: Jain, Bhavuk, et al.
Publicado: (2026)
A Comparative Analysis of Ensemble-Based Machine Learning Approaches with Explainable AI for Multi-Class Intrusion Detection in Drone Networks
por: Hossain, Md. Alamgir, et al.
Publicado: (2025)
por: Hossain, Md. Alamgir, et al.
Publicado: (2025)
A Learning-Based Attack Framework to Break SOTA Poisoning Defenses in Federated Learning
por: Yang, Yuxin, et al.
Publicado: (2024)
por: Yang, Yuxin, et al.
Publicado: (2024)
Semantic Stealth: Adversarial Text Attacks on NLP Using Several Methods
por: Dey, Roopkatha, et al.
Publicado: (2024)
por: Dey, Roopkatha, et al.
Publicado: (2024)
Membership Inference Attacks and Defenses in Federated Learning: A Survey
por: Bai, Li, et al.
Publicado: (2024)
por: Bai, Li, et al.
Publicado: (2024)
Backdoor Attacks and Defenses in Computer Vision Domain: A Survey
por: Abbasi, Bilal Hussain, et al.
Publicado: (2025)
por: Abbasi, Bilal Hussain, et al.
Publicado: (2025)
A Survey of Recent Backdoor Attacks and Defenses in Large Language Models
por: Zhao, Shuai, et al.
Publicado: (2024)
por: Zhao, Shuai, et al.
Publicado: (2024)
A Survey on Model Extraction Attacks and Defenses for Large Language Models
por: Zhao, Kaixiang, et al.
Publicado: (2025)
por: Zhao, Kaixiang, et al.
Publicado: (2025)
A Multi-Agent LLM Defense Pipeline Against Prompt Injection Attacks
por: Hossain, S M Asif, et al.
Publicado: (2025)
por: Hossain, S M Asif, et al.
Publicado: (2025)
Deep Learning Model Inversion Attacks and Defenses: A Comprehensive Survey
por: Yang, Wencheng, et al.
Publicado: (2025)
por: Yang, Wencheng, et al.
Publicado: (2025)
Adaptive Attacks Break Defenses Against Indirect Prompt Injection Attacks on LLM Agents
por: Zhan, Qiusi, et al.
Publicado: (2025)
por: Zhan, Qiusi, et al.
Publicado: (2025)
A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations
por: Ye, Mang, et al.
Publicado: (2025)
por: Ye, Mang, et al.
Publicado: (2025)
Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey
por: Huang, Tiansheng, et al.
Publicado: (2024)
por: Huang, Tiansheng, et al.
Publicado: (2024)
A Comparative Study of Watering Hole Attack Detection Using Supervised Neural Network
por: Aktar, Mst. Nishita, et al.
Publicado: (2023)
por: Aktar, Mst. Nishita, et al.
Publicado: (2023)
Modelling Attacks in Blockchain Systems using Petri Nets
por: Shahriar, Md. Atik, et al.
Publicado: (2020)
por: Shahriar, Md. Atik, et al.
Publicado: (2020)
Quantum Cyber-Attack on Blockchain-based VANET
por: Shakib, Kazi Hassan, et al.
Publicado: (2023)
por: Shakib, Kazi Hassan, et al.
Publicado: (2023)
A Novel Framework for Transmitter Privacy in Integrated Sensing and Communication
por: Kumar, Vaibhav, et al.
Publicado: (2026)
por: Kumar, Vaibhav, et al.
Publicado: (2026)
Secure Audio Embedding in Images using Nature-Inspired Optimization
por: Kumar, Aman, et al.
Publicado: (2025)
por: Kumar, Aman, et al.
Publicado: (2025)
Privacy in Large Language Models: Attacks, Defenses and Future Directions
por: Li, Haoran, et al.
Publicado: (2023)
por: Li, Haoran, et al.
Publicado: (2023)
Recent Advances in Attack and Defense Approaches of Large Language Models
por: Cui, Jing, et al.
Publicado: (2024)
por: Cui, Jing, et al.
Publicado: (2024)
Rotated Robustness: A Training-Free Defense against Bit-Flip Attacks on Large Language Models
por: Liu, Deng, et al.
Publicado: (2026)
por: Liu, Deng, et al.
Publicado: (2026)
$\textit{MMJ-Bench}$: A Comprehensive Study on Jailbreak Attacks and Defenses for Multimodal Large Language Models
por: Weng, Fenghua, et al.
Publicado: (2024)
por: Weng, Fenghua, et al.
Publicado: (2024)
Tit-for-Tat: Safeguarding Large Vision-Language Models Against Jailbreak Attacks via Adversarial Defense
por: Hao, Shuyang, et al.
Publicado: (2025)
por: Hao, Shuyang, et al.
Publicado: (2025)
PhishGuard: A Multi-Layered Ensemble Model for Optimal Phishing Website Detection
por: Ovi, Md Sultanul Islam, et al.
Publicado: (2024)
por: Ovi, Md Sultanul Islam, et al.
Publicado: (2024)
Towards Robust IoT Defense: Comparative Statistics of Attack Detection in Resource-Constrained Scenarios
por: Alwaisi, Zainab, et al.
Publicado: (2024)
por: Alwaisi, Zainab, et al.
Publicado: (2024)
A Comprehensive Survey of Website Fingerprinting Attacks and Defenses in Tor: Advances and Open Challenges
por: Cui, Yuwen, et al.
Publicado: (2025)
por: Cui, Yuwen, et al.
Publicado: (2025)
Ejemplares similares
-
Breaking the illusion: Automated Reasoning of GDPR Consent Violations
por: Li, Ying, et al.
Publicado: (2025) -
A Novel XAI-Enhanced Quantum Adversarial Networks for Velocity Dispersion Modeling in MaNGA Galaxies
por: Narkedimilli, Sathwik, et al.
Publicado: (2025) -
AntiFLipper: A Secure and Efficient Defense Against Label-Flipping Attacks in Federated Learning
por: Rahman, Aashnan, et al.
Publicado: (2025) -
Quantum-Resistant Authentication Scheme for RFID Systems Using Lattice-Based Cryptography
por: Kumar, Vaibhav, et al.
Publicado: (2025) -
DSTAN-Med: Dual-Channel Spatiotemporal Attention with Physiological Plausibility Filtering for False Data Injection Attack Detection in IoT-Based Medical Devices
por: Hasan, Md Mehedi, et al.
Publicado: (2026)