MaPPing Your Model: Assessing the Impact of Adversarial Attacks on LLM-based Programming Assistants
Fuente:
arXiv
Guardado en:
| Autores principales: | Heibel, John, Lowd, Daniel |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Train to Defend: First Defense Against Cryptanalytic Neural Network Parameter Extraction Attacks
por: Kurian, Ashley, et al.
Publicado: (2025)
por: Kurian, Ashley, et al.
Publicado: (2025)
Evaluating the efficacy of LLM Safety Solutions : The Palit Benchmark Dataset
por: Palit, Sayon, et al.
Publicado: (2025)
por: Palit, Sayon, et al.
Publicado: (2025)
Biometrics Employing Neural Network
por: Bhuiyan, Sajjad
Publicado: (2024)
por: Bhuiyan, Sajjad
Publicado: (2024)
Guardians of the Web: The Evolution and Future of Website Information Security
por: Islam, Md Saiful, et al.
Publicado: (2025)
por: Islam, Md Saiful, et al.
Publicado: (2025)
Mitigating the Impact of Malware Evolution on API Sequence-based Windows Malware Detector
por: Wei, Xingyuan, et al.
Publicado: (2024)
por: Wei, Xingyuan, et al.
Publicado: (2024)
PromptSAM+: Malware Detection based on Prompt Segment Anything Model
por: Wei, Xingyuan, et al.
Publicado: (2024)
por: Wei, Xingyuan, et al.
Publicado: (2024)
Seeing Is Not Always Believing: Invisible Collision Attack and Defence on Pre-Trained Models
por: Deng, Minghang, et al.
Publicado: (2023)
por: Deng, Minghang, et al.
Publicado: (2023)
Eliminating Backdoors in Neural Code Models for Secure Code Understanding
por: Sun, Weisong, et al.
Publicado: (2024)
por: Sun, Weisong, et al.
Publicado: (2024)
Toward Intelligent and Secure Cloud: Large Language Model Empowered Proactive Defense
por: Zhou, Yuyang, et al.
Publicado: (2024)
por: Zhou, Yuyang, et al.
Publicado: (2024)
Temporal Attack Pattern Detection in Multi-Agent AI Workflows: An Open Framework for Training Trace-Based Security Models
por: Del Rosario, Ron F.
Publicado: (2025)
por: Del Rosario, Ron F.
Publicado: (2025)
PatchBlock: A Lightweight Defense Against Adversarial Patches for Embedded EdgeAI Devices
por: Chattopadhyay, Nandish, et al.
Publicado: (2026)
por: Chattopadhyay, Nandish, et al.
Publicado: (2026)
Real AI Agents with Fake Memories: Fatal Context Manipulation Attacks on Web3 Agents
por: Patlan, Atharv Singh, et al.
Publicado: (2025)
por: Patlan, Atharv Singh, et al.
Publicado: (2025)
A Self-Improving Architecture for Dynamic Safety in Large Language Models
por: Slater, Tyler
Publicado: (2025)
por: Slater, Tyler
Publicado: (2025)
Mitigating Trojanized Prompt Chains in Educational LLM Use Cases: Experimental Findings and Detection Tool Design
por: Charles, Richard M., et al.
Publicado: (2025)
por: Charles, Richard M., et al.
Publicado: (2025)
Hacking, The Lazy Way: LLM Augmented Pentesting
por: Goyal, Dhruva, et al.
Publicado: (2024)
por: Goyal, Dhruva, et al.
Publicado: (2024)
The Quantum State Continuity Problem and Temporal Enforcement Against Fork Attacks
por: Ünsal, Samet
Publicado: (2025)
por: Ünsal, Samet
Publicado: (2025)
Adversarial Attacks on Large Language Models Using Regularized Relaxation
por: Chacko, Samuel Jacob, et al.
Publicado: (2024)
por: Chacko, Samuel Jacob, et al.
Publicado: (2024)
Efficient LLM Safety Evaluation through Multi-Agent Debate
por: Lin, Dachuan, et al.
Publicado: (2025)
por: Lin, Dachuan, et al.
Publicado: (2025)
VFLAIR-LLM: A Comprehensive Framework and Benchmark for Split Learning of LLMs
por: Gu, Zixuan, et al.
Publicado: (2025)
por: Gu, Zixuan, et al.
Publicado: (2025)
Impact of Phonetics on Speaker Identity in Adversarial Voice Attack
por: Dar, Daniyal Kabir, et al.
Publicado: (2025)
por: Dar, Daniyal Kabir, et al.
Publicado: (2025)
BadVLA: Towards Backdoor Attacks on Vision-Language-Action Models via Objective-Decoupled Optimization
por: Zhou, Xueyang, et al.
Publicado: (2025)
por: Zhou, Xueyang, et al.
Publicado: (2025)
NatGVD: Natural Adversarial Example Attack towards Graph-based Vulnerability Detection
por: Rath, Avilash, et al.
Publicado: (2025)
por: Rath, Avilash, et al.
Publicado: (2025)
MalPurifier: Enhancing Android Malware Detection with Adversarial Purification against Evasion Attacks
por: Zhou, Yuyang, et al.
Publicado: (2023)
por: Zhou, Yuyang, et al.
Publicado: (2023)
Jailbreaking Attacks vs. Content Safety Filters: How Far Are We in the LLM Safety Arms Race?
por: Xin, Yuan, et al.
Publicado: (2025)
por: Xin, Yuan, et al.
Publicado: (2025)
Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems
por: Hackett, William, et al.
Publicado: (2025)
por: Hackett, William, et al.
Publicado: (2025)
Privacy-preserving Universal Adversarial Defense for Black-box Models
por: Li, Qiao, et al.
Publicado: (2024)
por: Li, Qiao, et al.
Publicado: (2024)
Power-Softmax: Towards Secure LLM Inference over Encrypted Data
por: Zimerman, Itamar, et al.
Publicado: (2024)
por: Zimerman, Itamar, et al.
Publicado: (2024)
Blind Spots in the Guard: How Domain-Camouflaged Injection Attacks Evade Detection in Multi-Agent LLM Systems
por: Pai, Aaditya
Publicado: (2026)
por: Pai, Aaditya
Publicado: (2026)
Differential Robustness in Transformer Language Models: Empirical Evaluation Under Adversarial Text Attacks
por: Gidatkar, Taniya, et al.
Publicado: (2025)
por: Gidatkar, Taniya, et al.
Publicado: (2025)
Orion: Fuzzing Workflow Automation
por: Bazalii, Max, et al.
Publicado: (2025)
por: Bazalii, Max, et al.
Publicado: (2025)
Adversarial Feeds Steer LLM Agent Decisions Against Their Defaults
por: Usman, Rana Muhammad
Publicado: (2026)
por: Usman, Rana Muhammad
Publicado: (2026)
CVE-Bench: A Benchmark for AI Agents' Ability to Exploit Real-World Web Application Vulnerabilities
por: Zhu, Yuxuan, et al.
Publicado: (2025)
por: Zhu, Yuxuan, et al.
Publicado: (2025)
Large Language Models are Autonomous Cyber Defenders
por: Castro, Sebastián R., et al.
Publicado: (2025)
por: Castro, Sebastián R., et al.
Publicado: (2025)
An Intelligent Native Network Slicing Security Architecture Empowered by Federated Learning
por: Moreira, Rodrigo, et al.
Publicado: (2024)
por: Moreira, Rodrigo, et al.
Publicado: (2024)
Safeguarding Efficacy in Large Language Models: Evaluating Resistance to Human-Written and Algorithmic Adversarial Prompts
por: Downey-Webb, Tiarnaigh, et al.
Publicado: (2025)
por: Downey-Webb, Tiarnaigh, et al.
Publicado: (2025)
ConfusionPrompt: Practical Private Inference for Online Large Language Models
por: Mai, Peihua, et al.
Publicado: (2023)
por: Mai, Peihua, et al.
Publicado: (2023)
Super Suffixes: Bypassing Text Generation Alignment and Guard Models Simultaneously
por: Adiletta, Andrew, et al.
Publicado: (2025)
por: Adiletta, Andrew, et al.
Publicado: (2025)
Is My Data in Your Retrieval Database? Membership Inference Attacks Against Retrieval Augmented Generation
por: Anderson, Maya, et al.
Publicado: (2024)
por: Anderson, Maya, et al.
Publicado: (2024)
Enabling Transparent Cyber Threat Intelligence Combining Large Language Models and Domain Ontologies
por: Cotti, Luca, et al.
Publicado: (2025)
por: Cotti, Luca, et al.
Publicado: (2025)
Exploiting Latent Space Discontinuities for Building Universal LLM Jailbreaks and Data Extraction Attacks
por: Paim, Kayua Oleques, et al.
Publicado: (2025)
por: Paim, Kayua Oleques, et al.
Publicado: (2025)
Ejemplares similares
-
Train to Defend: First Defense Against Cryptanalytic Neural Network Parameter Extraction Attacks
por: Kurian, Ashley, et al.
Publicado: (2025) -
Evaluating the efficacy of LLM Safety Solutions : The Palit Benchmark Dataset
por: Palit, Sayon, et al.
Publicado: (2025) -
Biometrics Employing Neural Network
por: Bhuiyan, Sajjad
Publicado: (2024) -
Guardians of the Web: The Evolution and Future of Website Information Security
por: Islam, Md Saiful, et al.
Publicado: (2025) -
Mitigating the Impact of Malware Evolution on API Sequence-based Windows Malware Detector
por: Wei, Xingyuan, et al.
Publicado: (2024)