Smoke and Mirrors: Jailbreaking LLM-based Code Generation via Implicit Malicious Prompts
Fuente:
arXiv
Saved in:
| Main Authors: | Ouyang, Sheng, Qin, Yihao, Lin, Bo, Chen, Liqian, Mao, Xiaoguang, Wang, Shangwen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Give LLMs a Security Course: Securing Retrieval-Augmented Code Generation via Knowledge Injection
by: Lin, Bo, et al.
Published: (2025)
by: Lin, Bo, et al.
Published: (2025)
Keep It Simple: Towards Accurate Vulnerability Detection for Large Code Graphs
by: Peng, Xin, et al.
Published: (2024)
by: Peng, Xin, et al.
Published: (2024)
Exploring the Security Threats of Knowledge Base Poisoning in Retrieval-Augmented Code Generation
by: Lin, Bo, et al.
Published: (2025)
by: Lin, Bo, et al.
Published: (2025)
Large Language Models-Aided Program Debloating
by: Lin, Bo, et al.
Published: (2025)
by: Lin, Bo, et al.
Published: (2025)
There are More Fish in the Sea: Automated Vulnerability Repair via Binary Templates
by: Lin, Bo, et al.
Published: (2024)
by: Lin, Bo, et al.
Published: (2024)
Fault Localization from the Semantic Code Search Perspective
by: Qin, Yihao, et al.
Published: (2024)
by: Qin, Yihao, et al.
Published: (2024)
AgentFL: Scaling LLM-based Fault Localization to Project-Level Context
by: Qin, Yihao, et al.
Published: (2024)
by: Qin, Yihao, et al.
Published: (2024)
Atomizer: An LLM-based Collaborative Multi-Agent Framework for Intent-Driven Commit Untangling
by: Zhu, Kangchen, et al.
Published: (2026)
by: Zhu, Kangchen, et al.
Published: (2026)
Exploring the Security Threats of Retriever Backdoors in Retrieval-Augmented Code Generation
by: Li, Tian, et al.
Published: (2025)
by: Li, Tian, et al.
Published: (2025)
Three Heads Are Better Than One: A Multi-perspective Reasoning Framework for Enhanced Vulnerability Detection
by: Peng, Xin, et al.
Published: (2026)
by: Peng, Xin, et al.
Published: (2026)
Large Language Models are Qualified Benchmark Builders: Rebuilding Pre-Training Datasets for Advancing Code Intelligence Tasks
by: Yang, Kang, et al.
Published: (2025)
by: Yang, Kang, et al.
Published: (2025)
Debug Like a Human: Scaling LLM-based Fault Localization to Processor Design via Block-Level Instruction-Oriented Slicing
by: Liu, Zizhen, et al.
Published: (2026)
by: Liu, Zizhen, et al.
Published: (2026)
Prompt Optimization for LLM Code Generation via Reinforcement Learning
by: Esfahani, Ali Mohammadi, et al.
Published: (2026)
by: Esfahani, Ali Mohammadi, et al.
Published: (2026)
MCGMark: An Encodable and Robust Online Watermark for Tracing LLM-Generated Malicious Code
by: Ning, Kaiwen, et al.
Published: (2024)
by: Ning, Kaiwen, et al.
Published: (2024)
Shapley-Guided Neural Repair Approach via Derivative-Free Optimization
by: Sun, Xinyu, et al.
Published: (2026)
by: Sun, Xinyu, et al.
Published: (2026)
Assessing the Impact of Requirement Ambiguity on LLM-based Function-Level Code Generation
by: Yang, Di, et al.
Published: (2026)
by: Yang, Di, et al.
Published: (2026)
Prompt-based Code Completion via Multi-Retrieval Augmented Generation
by: Tan, Hanzhuo, et al.
Published: (2024)
by: Tan, Hanzhuo, et al.
Published: (2024)
SolSearch: An LLM-Driven Framework for Efficient SAT-Solving Code Generation
by: Sheng, Junjie, et al.
Published: (2025)
by: Sheng, Junjie, et al.
Published: (2025)
GeoJSON Agents:A Multi-Agent LLM Architecture for Geospatial Analysis-Function Calling vs Code Generation
by: Luo, Qianqian, et al.
Published: (2025)
by: Luo, Qianqian, et al.
Published: (2025)
Tuning LLM-based Code Optimization via Meta-Prompting: An Industrial Perspective
by: Gong, Jingzhi, et al.
Published: (2025)
by: Gong, Jingzhi, et al.
Published: (2025)
CODEPROMPTZIP: Code-specific Prompt Compression for Retrieval-Augmented Generation in Coding Tasks with LMs
by: He, Pengfei, et al.
Published: (2025)
by: He, Pengfei, et al.
Published: (2025)
Instruct or Interact? Exploring and Eliciting LLMs' Capability in Code Snippet Adaptation Through Prompt Engineering
by: Zhang, Tanghaoran, et al.
Published: (2024)
by: Zhang, Tanghaoran, et al.
Published: (2024)
When Prompt Under-Specification Improves Code Correctness: An Exploratory Study of Prompt Wording and Structure Effects on LLM-Based Code Generation
by: AKLI, Amal, et al.
Published: (2026)
by: AKLI, Amal, et al.
Published: (2026)
Are They All Good? Evaluating the Quality of CoTs in LLM-based Code Generation
by: Zhang, Binquan, et al.
Published: (2025)
by: Zhang, Binquan, et al.
Published: (2025)
Prompt Alchemy: Automatic Prompt Refinement for Enhancing Code Generation
by: Ye, Sixiang, et al.
Published: (2025)
by: Ye, Sixiang, et al.
Published: (2025)
BitsAI-CR: Automated Code Review via LLM in Practice
by: Sun, Tao, et al.
Published: (2025)
by: Sun, Tao, et al.
Published: (2025)
On the Effectiveness of Training Data Optimization for LLM-based Code Generation: An Empirical Study
by: Kuang, Shiqi, et al.
Published: (2025)
by: Kuang, Shiqi, et al.
Published: (2025)
Probing Privacy Leaks in LLM-based Code Generation via Test Generation
by: Ge, Yifei, et al.
Published: (2026)
by: Ge, Yifei, et al.
Published: (2026)
R2Code: A Self-Reflective LLM Framework for Requirements-to-Code Traceability
by: Wang, Yifei, et al.
Published: (2026)
by: Wang, Yifei, et al.
Published: (2026)
Code Roulette: How Prompt Variability Affects LLM Code Generation
by: Paleyes, Andrei, et al.
Published: (2025)
by: Paleyes, Andrei, et al.
Published: (2025)
Bias Unveiled: Investigating Social Bias in LLM-Generated Code
by: Ling, Lin, et al.
Published: (2024)
by: Ling, Lin, et al.
Published: (2024)
Models Are Codes: Towards Measuring Malicious Code Poisoning Attacks on Pre-trained Model Hubs
by: Zhao, Jian, et al.
Published: (2024)
by: Zhao, Jian, et al.
Published: (2024)
Understanding the Effectiveness of Coverage Criteria for Large Language Models: A Special Angle from Jailbreak Attacks
by: Zhou, Shide, et al.
Published: (2024)
by: Zhou, Shide, et al.
Published: (2024)
Chart2Code-MoLA: Efficient Multi-Modal Code Generation via Adaptive Expert Routing
by: Wang, Yifei, et al.
Published: (2025)
by: Wang, Yifei, et al.
Published: (2025)
Automated Prompt Generation for Code Intelligence: An Empirical study and Experience in WeChat
by: Ji, Kexing, et al.
Published: (2025)
by: Ji, Kexing, et al.
Published: (2025)
HintPilot: LLM-based Compiler Hint Synthesis for Code Optimization
by: Jiang, Hanyun, et al.
Published: (2026)
by: Jiang, Hanyun, et al.
Published: (2026)
When Neural Code Completion Models Size up the Situation: Attaining Cheaper and Faster Completion through Dynamic Model Inference
by: Sun, Zhensu, et al.
Published: (2024)
by: Sun, Zhensu, et al.
Published: (2024)
Don't Complete It! Preventing Unhelpful Code Completion for Productive and Sustainable Neural Code Completion Systems
by: Sun, Zhensu, et al.
Published: (2022)
by: Sun, Zhensu, et al.
Published: (2022)
An Empirical Study of the Non-determinism of ChatGPT in Code Generation
by: Ouyang, Shuyin, et al.
Published: (2023)
by: Ouyang, Shuyin, et al.
Published: (2023)
Reducing Hallucinations in LLM-Generated Code via Semantic Triangulation
by: Dai, Yihan, et al.
Published: (2025)
by: Dai, Yihan, et al.
Published: (2025)
Similar Items
-
Give LLMs a Security Course: Securing Retrieval-Augmented Code Generation via Knowledge Injection
by: Lin, Bo, et al.
Published: (2025) -
Keep It Simple: Towards Accurate Vulnerability Detection for Large Code Graphs
by: Peng, Xin, et al.
Published: (2024) -
Exploring the Security Threats of Knowledge Base Poisoning in Retrieval-Augmented Code Generation
by: Lin, Bo, et al.
Published: (2025) -
Large Language Models-Aided Program Debloating
by: Lin, Bo, et al.
Published: (2025) -
There are More Fish in the Sea: Automated Vulnerability Repair via Binary Templates
by: Lin, Bo, et al.
Published: (2024)