MOCHA: Are Code Language Models Robust Against Multi-Turn Malicious Coding Prompts?
Fuente:
arXiv
Saved in:
| Main Authors: | Wahed, Muntasir, Zhou, Xiaona, Nguyen, Kiet A., Yu, Tianjiao, Diwan, Nirav, Wang, Gang, Hakkani-Tür, Dilek, Lourentzou, Ismini |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PurpCode: Reasoning for Safer Code Generation
by: Liu, Jiawei, et al.
Published: (2025)
by: Liu, Jiawei, et al.
Published: (2025)
CALICO: Part-Focused Semantic Co-Segmentation with Large Vision-Language Models
by: Nguyen, Kiet A., et al.
Published: (2024)
by: Nguyen, Kiet A., et al.
Published: (2024)
Tiny but Trusted: Efficient Vision-Language Reasoning for Time-Series Anomaly Detection
by: Zhou, Xiaona, et al.
Published: (2026)
by: Zhou, Xiaona, et al.
Published: (2026)
Beyond BeautifulSoup: Benchmarking LLM-Powered Web Scraping for Everyday Users
by: Bhardwaj, Arth, et al.
Published: (2026)
by: Bhardwaj, Arth, et al.
Published: (2026)
Part$^{2}$GS: Part-aware Modeling of Articulated Objects using 3D Gaussian Splatting
by: Yu, Tianjiao, et al.
Published: (2025)
by: Yu, Tianjiao, et al.
Published: (2025)
Taint-Based Code Slicing for LLMs-based Malicious NPM Package Detection
by: Nguyen, Dang-Khoa, et al.
Published: (2025)
by: Nguyen, Dang-Khoa, et al.
Published: (2025)
Multi-Agent Systems Execute Arbitrary Malicious Code
by: Triedman, Harold, et al.
Published: (2025)
by: Triedman, Harold, et al.
Published: (2025)
PRIMA: Multi-Image Vision-Language Models for Reasoning Segmentation
by: Wahed, Muntasir, et al.
Published: (2024)
by: Wahed, Muntasir, et al.
Published: (2024)
Malicious Code Detection in Smart Contracts via Opcode Vectorization
by: Zou, Huanhuan, et al.
Published: (2025)
by: Zou, Huanhuan, et al.
Published: (2025)
One Turn Too Late: Response-Aware Defense Against Hidden Malicious Intent in Multi-Turn Dialogue
by: Shen, Xinjie, et al.
Published: (2026)
by: Shen, Xinjie, et al.
Published: (2026)
MCGMark: An Encodable and Robust Online Watermark for Tracing LLM-Generated Malicious Code
by: Ning, Kaiwen, et al.
Published: (2024)
by: Ning, Kaiwen, et al.
Published: (2024)
Assessing LLMs in Malicious Code Deobfuscation of Real-world Malware Campaigns
by: Patsakis, Constantinos, et al.
Published: (2024)
by: Patsakis, Constantinos, et al.
Published: (2024)
Explainable Android Malware Detection and Malicious Code Localization Using Graph Attention
by: Ipek, Merve Cigdem, et al.
Published: (2025)
by: Ipek, Merve Cigdem, et al.
Published: (2025)
Hallucinating AI Hijacking Attack: Large Language Models and Malicious Code Recommenders
by: Noever, David, et al.
Published: (2024)
by: Noever, David, et al.
Published: (2024)
Towards Quantum Machine Learning for Malicious Code Analysis
by: Lopez, Jesus, et al.
Published: (2025)
by: Lopez, Jesus, et al.
Published: (2025)
Malicious Agent Detection for Robust Multi-Agent Collaborative Perception
by: Zhao, Yangheng, et al.
Published: (2023)
by: Zhao, Yangheng, et al.
Published: (2023)
Uncertainty in Action: Confidence Elicitation in Embodied Agents
by: Yu, Tianjiao, et al.
Published: (2025)
by: Yu, Tianjiao, et al.
Published: (2025)
Models Are Codes: Towards Measuring Malicious Code Poisoning Attacks on Pre-trained Model Hubs
by: Zhao, Jian, et al.
Published: (2024)
by: Zhao, Jian, et al.
Published: (2024)
Malicious and Unintentional Disclosure Risks in Large Language Models for Code Generation
by: Rabin, Rafiqul, et al.
Published: (2025)
by: Rabin, Rafiqul, et al.
Published: (2025)
Localizing Malicious Outputs from CodeLLM
by: Borana, Mayukh, et al.
Published: (2025)
by: Borana, Mayukh, et al.
Published: (2025)
FreeMOCA: Memory-Free Continual Learning for Malicious Code Analysis
by: Asadi, Zahra, et al.
Published: (2026)
by: Asadi, Zahra, et al.
Published: (2026)
WAFBOOSTER: Automatic Boosting of WAF Security Against Mutated Malicious Payloads
by: Wu, Cong, et al.
Published: (2025)
by: Wu, Cong, et al.
Published: (2025)
JavaSith: A Client-Side Framework for Analyzing Potentially Malicious Extensions in Browsers, VS Code, and NPM Packages
by: Cohen, Avihay
Published: (2025)
by: Cohen, Avihay
Published: (2025)
CoT-Guard: Small Models for Strong Monitoring
by: Diwan, Nirav, et al.
Published: (2026)
by: Diwan, Nirav, et al.
Published: (2026)
Towards Classifying Benign And Malicious Packages Using Machine Learning
by: Nguyen, Thanh-Cong, et al.
Published: (2025)
by: Nguyen, Thanh-Cong, et al.
Published: (2025)
Turn-Based Structural Triggers: Prompt-Free Backdoors in Multi-Turn LLMs
by: Lu, Yiyang, et al.
Published: (2026)
by: Lu, Yiyang, et al.
Published: (2026)
WARD: Adversarially Robust Defense of Web Agents Against Prompt Injections
by: Cao, Tri, et al.
Published: (2026)
by: Cao, Tri, et al.
Published: (2026)
A Novel Approach to Malicious Code Detection Using CNN-BiLSTM and Feature Fusion
by: Zhang, Lixia, et al.
Published: (2024)
by: Zhang, Lixia, et al.
Published: (2024)
From Past to Present: A Survey of Malicious URL Detection Techniques, Datasets and Code Repositories
by: Tian, Ye, et al.
Published: (2025)
by: Tian, Ye, et al.
Published: (2025)
AutoAdv: Automated Adversarial Prompting for Multi-Turn Jailbreaking of Large Language Models
by: Reddy, Aashray, et al.
Published: (2025)
by: Reddy, Aashray, et al.
Published: (2025)
CoTDeceptor:Adversarial Code Obfuscation Against CoT-Enhanced LLM Code Agents
by: Li, Haoyang, et al.
Published: (2025)
by: Li, Haoyang, et al.
Published: (2025)
DeepSeek Robustness Against Semantic-Character Dual-Space Mutated Prompt Injection
by: Ren, Junyu, et al.
Published: (2026)
by: Ren, Junyu, et al.
Published: (2026)
CircuitGuard: Mitigating LLM Memorization in RTL Code Generation Against IP Leakage
by: Mashnoor, Nowfel, et al.
Published: (2025)
by: Mashnoor, Nowfel, et al.
Published: (2025)
Robust Federated Learning for Malicious Clients using Loss Trend Deviation Detection
by: Bhaskar, Deepthy K, et al.
Published: (2026)
by: Bhaskar, Deepthy K, et al.
Published: (2026)
Chain-of-Code Collapse: Reasoning Failures in LLMs via Adversarial Prompting in Code Generation
by: Roh, Jaechul, et al.
Published: (2025)
by: Roh, Jaechul, et al.
Published: (2025)
ShadowCode: Towards (Automatic) External Prompt Injection Attack against Code LLMs
by: Yang, Yuchen, et al.
Published: (2024)
by: Yang, Yuchen, et al.
Published: (2024)
On the Adversarial Robustness of Instruction-Tuned Large Language Models for Code
by: Hossen, Md Imran, et al.
Published: (2024)
by: Hossen, Md Imran, et al.
Published: (2024)
Invisible Prompts, Visible Threats: Malicious Font Injection in External Resources for Large Language Models
by: Xiong, Junjie, et al.
Published: (2025)
by: Xiong, Junjie, et al.
Published: (2025)
Demo: SGCode: A Flexible Prompt-Optimizing System for Secure Generation of Code
by: Ton, Khiem, et al.
Published: (2024)
by: Ton, Khiem, et al.
Published: (2024)
Malicious GenAI Chrome Extensions: Unpacking Data Exfiltration and Malicious Behaviours
by: Seetharam, Shresta B., et al.
Published: (2025)
by: Seetharam, Shresta B., et al.
Published: (2025)
Similar Items
-
PurpCode: Reasoning for Safer Code Generation
by: Liu, Jiawei, et al.
Published: (2025) -
CALICO: Part-Focused Semantic Co-Segmentation with Large Vision-Language Models
by: Nguyen, Kiet A., et al.
Published: (2024) -
Tiny but Trusted: Efficient Vision-Language Reasoning for Time-Series Anomaly Detection
by: Zhou, Xiaona, et al.
Published: (2026) -
Beyond BeautifulSoup: Benchmarking LLM-Powered Web Scraping for Everyday Users
by: Bhardwaj, Arth, et al.
Published: (2026) -
Part$^{2}$GS: Part-aware Modeling of Articulated Objects using 3D Gaussian Splatting
by: Yu, Tianjiao, et al.
Published: (2025)