ShadowCode: Towards (Automatic) External Prompt Injection Attack against Code LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Yang, Yuchen, Li, Yiming, Yao, Hongwei, Yang, Bingrun, He, Yiling, Zhang, Tianwei, Tao, Dacheng, Qin, Zhan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Red-Teaming Agent Execution Contexts: Open-World Security Evaluation on OpenClaw
di: Yao, Hongwei, et al.
Pubblicazione: (2026)
di: Yao, Hongwei, et al.
Pubblicazione: (2026)
External Data Extraction Attacks against Retrieval-Augmented Large Language Models
di: He, Yu, et al.
Pubblicazione: (2025)
di: He, Yu, et al.
Pubblicazione: (2025)
AttriGuard: Defeating Indirect Prompt Injection in LLM Agents via Causal Attribution of Tool Invocations
di: He, Yu, et al.
Pubblicazione: (2026)
di: He, Yu, et al.
Pubblicazione: (2026)
Towards Effective Prompt Stealing Attack against Text-to-Image Diffusion Models
di: Zhao, Shiqian, et al.
Pubblicazione: (2025)
di: Zhao, Shiqian, et al.
Pubblicazione: (2025)
On Benchmarking Code LLMs for Android Malware Analysis
di: He, Yiling, et al.
Pubblicazione: (2025)
di: He, Yiling, et al.
Pubblicazione: (2025)
PINA: Prompt Injection Attack against Navigation Agents
di: Liu, Jiani, et al.
Pubblicazione: (2026)
di: Liu, Jiani, et al.
Pubblicazione: (2026)
VortexPIA: Indirect Prompt Injection Attack against LLMs for Efficient Extraction of User Privacy
di: Cui, Yu, et al.
Pubblicazione: (2025)
di: Cui, Yu, et al.
Pubblicazione: (2025)
Reading Between the Lines: Towards Reliable Black-box LLM Fingerprinting via Zeroth-order Gradient Estimation
di: Shao, Shuo, et al.
Pubblicazione: (2025)
di: Shao, Shuo, et al.
Pubblicazione: (2025)
Explanation as a Watermark: Towards Harmless and Multi-bit Model Ownership Verification via Watermarking Feature Attribution
di: Shao, Shuo, et al.
Pubblicazione: (2024)
di: Shao, Shuo, et al.
Pubblicazione: (2024)
FIT-Print: Towards False-claim-resistant Model Ownership Verification via Targeted Fingerprint
di: Shao, Shuo, et al.
Pubblicazione: (2025)
di: Shao, Shuo, et al.
Pubblicazione: (2025)
Demonstration Attack against In-Context Learning for Code Intelligence
di: Ge, Yifei, et al.
Pubblicazione: (2024)
di: Ge, Yifei, et al.
Pubblicazione: (2024)
SoK: Large Language Model Copyright Auditing via Fingerprinting
di: Shao, Shuo, et al.
Pubblicazione: (2025)
di: Shao, Shuo, et al.
Pubblicazione: (2025)
AgentTypo: Adaptive Typographic Prompt Injection Attacks against Black-box Multimodal Agents
di: Li, Yanjie, et al.
Pubblicazione: (2025)
di: Li, Yanjie, et al.
Pubblicazione: (2025)
BitHydra: Towards Bit-flip Inference Cost Attack against Large Language Models
di: Yan, Xiaobei, et al.
Pubblicazione: (2025)
di: Yan, Xiaobei, et al.
Pubblicazione: (2025)
Prompt Injection Attacks on Agentic Coding Assistants: A Systematic Analysis of Vulnerabilities in Skills, Tools, and Protocol Ecosystems
di: Maloyan, Narek, et al.
Pubblicazione: (2026)
di: Maloyan, Narek, et al.
Pubblicazione: (2026)
Explainer-guided Targeted Adversarial Attacks against Binary Code Similarity Detection Models
di: Chen, Mingjie, et al.
Pubblicazione: (2025)
di: Chen, Mingjie, et al.
Pubblicazione: (2025)
AdapTools: Adaptive Tool-based Indirect Prompt Injection Attacks on Agentic LLMs
di: Wang, Che, et al.
Pubblicazione: (2026)
di: Wang, Che, et al.
Pubblicazione: (2026)
Robustness via Referencing: Defending against Prompt Injection Attacks by Referencing the Executed Instruction
di: Chen, Yulin, et al.
Pubblicazione: (2025)
di: Chen, Yulin, et al.
Pubblicazione: (2025)
A Critical Evaluation of Defenses against Prompt Injection Attacks
di: Jia, Yuqi, et al.
Pubblicazione: (2025)
di: Jia, Yuqi, et al.
Pubblicazione: (2025)
Analysis of LLMs Against Prompt Injection and Jailbreak Attacks
di: Jaiswal, Piyush, et al.
Pubblicazione: (2026)
di: Jaiswal, Piyush, et al.
Pubblicazione: (2026)
"Your AI, My Shell": Demystifying Prompt Injection Attacks on Agentic AI Coding Editors
di: Liu, Yue, et al.
Pubblicazione: (2025)
di: Liu, Yue, et al.
Pubblicazione: (2025)
PromptShield: Deployable Detection for Prompt Injection Attacks
di: Jacob, Dennis, et al.
Pubblicazione: (2025)
di: Jacob, Dennis, et al.
Pubblicazione: (2025)
System-Level Defense against Indirect Prompt Injection Attacks: An Information Flow Control Perspective
di: Wu, Fangzhou, et al.
Pubblicazione: (2024)
di: Wu, Fangzhou, et al.
Pubblicazione: (2024)
Attention is All You Need to Defend Against Indirect Prompt Injection Attacks in LLMs
di: Zhong, Yinan, et al.
Pubblicazione: (2025)
di: Zhong, Yinan, et al.
Pubblicazione: (2025)
Multimodal Prompt Injection Attacks: Risks and Defenses for Modern LLMs
di: Yeo, Andrew, et al.
Pubblicazione: (2025)
di: Yeo, Andrew, et al.
Pubblicazione: (2025)
Defense Against Prompt Injection Attack by Leveraging Attack Techniques
di: Chen, Yulin, et al.
Pubblicazione: (2024)
di: Chen, Yulin, et al.
Pubblicazione: (2024)
Prompt Inversion Attack against Collaborative Inference of Large Language Models
di: Qu, Wenjie, et al.
Pubblicazione: (2025)
di: Qu, Wenjie, et al.
Pubblicazione: (2025)
ReasAlign: Reasoning Enhanced Safety Alignment against Prompt Injection Attack
di: Li, Hao, et al.
Pubblicazione: (2026)
di: Li, Hao, et al.
Pubblicazione: (2026)
The Vulnerability of LLM Rankers to Prompt Injection Attacks
di: Yin, Yu, et al.
Pubblicazione: (2026)
di: Yin, Yu, et al.
Pubblicazione: (2026)
Adaptive Attacks Break Defenses Against Indirect Prompt Injection Attacks on LLM Agents
di: Zhan, Qiusi, et al.
Pubblicazione: (2025)
di: Zhan, Qiusi, et al.
Pubblicazione: (2025)
Systematic Categorization, Construction and Evaluation of New Attacks against Multi-modal Mobile GUI Agents
di: Yang, Yulong, et al.
Pubblicazione: (2024)
di: Yang, Yulong, et al.
Pubblicazione: (2024)
AegisAgent: An Autonomous Defense Agent Against Prompt Injection Attacks in LLM-HARs
di: Wang, Yihan, et al.
Pubblicazione: (2025)
di: Wang, Yihan, et al.
Pubblicazione: (2025)
Protecting Cryptographic Libraries against Side-Channel and Code-Reuse Attacks
di: Tsoupidi, Rodothea Myrsini, et al.
Pubblicazione: (2024)
di: Tsoupidi, Rodothea Myrsini, et al.
Pubblicazione: (2024)
Benchmarking LLMs and LLM-based Agents in Practical Vulnerability Detection for Code Repositories
di: Yildiz, Alperen, et al.
Pubblicazione: (2025)
di: Yildiz, Alperen, et al.
Pubblicazione: (2025)
RTL-Breaker: Assessing the Security of LLMs against Backdoor Attacks on HDL Code Generation
di: Mankali, Lakshmi Likhitha, et al.
Pubblicazione: (2024)
di: Mankali, Lakshmi Likhitha, et al.
Pubblicazione: (2024)
Is Your Prompt Safe? Investigating Prompt Injection Attacks Against Open-Source LLMs
di: Wang, Jiawen, et al.
Pubblicazione: (2025)
di: Wang, Jiawen, et al.
Pubblicazione: (2025)
TopicAttack: An Indirect Prompt Injection Attack via Topic Transition
di: Chen, Yulin, et al.
Pubblicazione: (2025)
di: Chen, Yulin, et al.
Pubblicazione: (2025)
Prompt Injection Attack to Tool Selection in LLM Agents
di: Shi, Jiawen, et al.
Pubblicazione: (2025)
di: Shi, Jiawen, et al.
Pubblicazione: (2025)
Can Indirect Prompt Injection Attacks Be Detected and Removed?
di: Chen, Yulin, et al.
Pubblicazione: (2025)
di: Chen, Yulin, et al.
Pubblicazione: (2025)
SWAP: Towards Copyright Auditing of Soft Prompts via Sequential Watermarking
di: Yang, Wenyuan, et al.
Pubblicazione: (2025)
di: Yang, Wenyuan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Red-Teaming Agent Execution Contexts: Open-World Security Evaluation on OpenClaw
di: Yao, Hongwei, et al.
Pubblicazione: (2026) -
External Data Extraction Attacks against Retrieval-Augmented Large Language Models
di: He, Yu, et al.
Pubblicazione: (2025) -
AttriGuard: Defeating Indirect Prompt Injection in LLM Agents via Causal Attribution of Tool Invocations
di: He, Yu, et al.
Pubblicazione: (2026) -
Towards Effective Prompt Stealing Attack against Text-to-Image Diffusion Models
di: Zhao, Shiqian, et al.
Pubblicazione: (2025) -
On Benchmarking Code LLMs for Android Malware Analysis
di: He, Yiling, et al.
Pubblicazione: (2025)