Can LLMs Deeply Detect Complex Malicious Queries? A Framework for Jailbreaking via Obfuscating Intent
Fuente:
arXiv
Saved in:
| Main Authors: | Shang, Shang, Zhao, Xinqiang, Yao, Zhongjiang, Yao, Yepeng, Su, Liya, Fan, Zijing, Zhang, Xiaodan, Jiang, Zhengwei |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PhishSigma++: Malicious Email Detection with Typed Entity Relations
by: Shang, Shang, et al.
Published: (2026)
by: Shang, Shang, et al.
Published: (2026)
Detecting Malicious Intents in Smart Contracts with Pre-trained Programming Language Models
by: Huang, Youwei, et al.
Published: (2025)
by: Huang, Youwei, et al.
Published: (2025)
Mitigating Jailbreaks with Intent-Aware LLMs
by: Yeo, Wei Jie, et al.
Published: (2025)
by: Yeo, Wei Jie, et al.
Published: (2025)
Evolving Security in LLMs: A Study of Jailbreak Attacks and Defenses
by: Shang, Zhengchun, et al.
Published: (2025)
by: Shang, Zhengchun, et al.
Published: (2025)
Exploring Jailbreak Attacks on LLMs through Intent Concealment and Diversion
by: Cui, Tiehan, et al.
Published: (2025)
by: Cui, Tiehan, et al.
Published: (2025)
On Feasibility of Intent Obfuscating Attacks
by: Li, Zhaobin, et al.
Published: (2024)
by: Li, Zhaobin, et al.
Published: (2024)
Breaking Obfuscation: Cluster-Aware Graph with LLM-Aided Recovery for Malicious JavaScript Detection
by: Liang, Zhihong, et al.
Published: (2025)
by: Liang, Zhihong, et al.
Published: (2025)
Overlooked Safety Vulnerability in LLMs: Malicious Intelligent Optimization Algorithm Request and its Jailbreak
by: Gu, Haoran, et al.
Published: (2026)
by: Gu, Haoran, et al.
Published: (2026)
Adjacent Words, Divergent Intents: Jailbreaking Large Language Models via Task Concurrency
by: Jiang, Yukun, et al.
Published: (2025)
by: Jiang, Yukun, et al.
Published: (2025)
Why Neural Structural Obfuscation Can't Kill White-Box Watermarks for Good!
by: Jiang, Yanna, et al.
Published: (2026)
by: Jiang, Yanna, et al.
Published: (2026)
SecureNT: Smart Topology Obfuscation for Privacy-Aware Network Monitoring
by: Du, Chengze, et al.
Published: (2024)
by: Du, Chengze, et al.
Published: (2024)
Beyond Surface-Level Patterns: An Essence-Driven Defense Framework Against Jailbreak Attacks in LLMs
by: Xiang, Shiyu, et al.
Published: (2025)
by: Xiang, Shiyu, et al.
Published: (2025)
Babel: Jailbreaking Safety Attention via Obfuscation Distribution Optimized Sampling
by: Wang, Ziwei, et al.
Published: (2026)
by: Wang, Ziwei, et al.
Published: (2026)
Taint-Based Code Slicing for LLMs-based Malicious NPM Package Detection
by: Nguyen, Dang-Khoa, et al.
Published: (2025)
by: Nguyen, Dang-Khoa, et al.
Published: (2025)
ShallowJail: Steering Jailbreaks against Large Language Models
by: Liu, Shang, et al.
Published: (2026)
by: Liu, Shang, et al.
Published: (2026)
Malicious Package Detection using Metadata Information
by: Halder, S., et al.
Published: (2024)
by: Halder, S., et al.
Published: (2024)
Can LLMs Obfuscate Code? A Systematic Analysis of Large Language Models into Assembly Code Obfuscation
by: Mohseni, Seyedreza, et al.
Published: (2024)
by: Mohseni, Seyedreza, et al.
Published: (2024)
Jailbreaking with Universal Multi-Prompts
by: Hsu, Yu-Ling, et al.
Published: (2025)
by: Hsu, Yu-Ling, et al.
Published: (2025)
RealVul: Can We Detect Vulnerabilities in Web Applications with LLM?
by: Cao, Di, et al.
Published: (2024)
by: Cao, Di, et al.
Published: (2024)
You Can't Eat Your Cake and Have It Too: The Performance Degradation of LLMs with Jailbreak Defense
by: Mai, Wuyuao, et al.
Published: (2025)
by: Mai, Wuyuao, et al.
Published: (2025)
FuzzLLM: A Novel and Universal Fuzzing Framework for Proactively Discovering Jailbreak Vulnerabilities in Large Language Models
by: Yao, Dongyu, et al.
Published: (2023)
by: Yao, Dongyu, et al.
Published: (2023)
Adversarial Distilled Retrieval-Augmented Guarding Model for Online Malicious Intent Detection
by: Guo, Yihao, et al.
Published: (2025)
by: Guo, Yihao, et al.
Published: (2025)
Automated Hardware Logic Obfuscation Framework Using GPT
by: Latibari, Banafsheh Saber, et al.
Published: (2024)
by: Latibari, Banafsheh Saber, et al.
Published: (2024)
ASTRA: An Automated Framework for Strategy Discovery, Retrieval, and Evolution for Jailbreaking LLMs
by: Liu, Xu, et al.
Published: (2025)
by: Liu, Xu, et al.
Published: (2025)
PolyJailbreak: Cross-Modal Jailbreaking Attacks on Black-Box Multimodal LLMs
by: Wang, Xinkai, et al.
Published: (2025)
by: Wang, Xinkai, et al.
Published: (2025)
Model-Driven Security Analysis of Self-Sovereign Identity Systems
by: Ding, Yepeng, et al.
Published: (2024)
by: Ding, Yepeng, et al.
Published: (2024)
SecureFed: A Two-Phase Framework for Detecting Malicious Clients in Federated Learning
by: Kavuri, Likhitha Annapurna, et al.
Published: (2025)
by: Kavuri, Likhitha Annapurna, et al.
Published: (2025)
Preventing Jailbreak Prompts as Malicious Tools for Cybercriminals: A Cyber Defense Perspective
by: Tshimula, Jean Marie, et al.
Published: (2024)
by: Tshimula, Jean Marie, et al.
Published: (2024)
DMLDroid: Deep Multimodal Fusion Framework for Android Malware Detection with Resilience to Code Obfuscation and Adversarial Perturbations
by: Trung, Doan Minh, et al.
Published: (2025)
by: Trung, Doan Minh, et al.
Published: (2025)
Obfuscated Memory Malware Detection
by: P, Sharmila S, et al.
Published: (2024)
by: P, Sharmila S, et al.
Published: (2024)
Can Agents Secure Hardware? Evaluating Agentic LLM-Driven Obfuscation for IP Protection
by: Ghimire, Sujan, et al.
Published: (2026)
by: Ghimire, Sujan, et al.
Published: (2026)
Cryptic Bytes: WebAssembly Obfuscation for Evading Cryptojacking Detection
by: Harnes, Håkon, et al.
Published: (2024)
by: Harnes, Håkon, et al.
Published: (2024)
PII Jailbreaking in LLMs via Activation Steering Reveals Personal Information Leakage
by: Nakka, Krishna Kanth, et al.
Published: (2025)
by: Nakka, Krishna Kanth, et al.
Published: (2025)
FedAdOb: Privacy-Preserving Federated Deep Learning with Adaptive Obfuscation
by: Gu, Hanlin, et al.
Published: (2024)
by: Gu, Hanlin, et al.
Published: (2024)
Can Differentially Private Fine-tuning LLMs Protect Against Privacy Attacks?
by: Du, Hao, et al.
Published: (2025)
by: Du, Hao, et al.
Published: (2025)
ModelObfuscator: Obfuscating Model Information to Protect Deployed ML-based Systems
by: Zhou, Mingyi, et al.
Published: (2023)
by: Zhou, Mingyi, et al.
Published: (2023)
Re-Triggering Safeguards within LLMs for Jailbreak Detection
by: Lin, Zheng, et al.
Published: (2026)
by: Lin, Zheng, et al.
Published: (2026)
Assessing LLMs in Malicious Code Deobfuscation of Real-world Malware Campaigns
by: Patsakis, Constantinos, et al.
Published: (2024)
by: Patsakis, Constantinos, et al.
Published: (2024)
JailbreaksOverTime: Detecting Jailbreak Attacks Under Distribution Shift
by: Piet, Julien, et al.
Published: (2025)
by: Piet, Julien, et al.
Published: (2025)
SAFE: Self-Supervised Anomaly Detection Framework for Intrusion Detection
by: Li, Elvin, et al.
Published: (2025)
by: Li, Elvin, et al.
Published: (2025)
Similar Items
-
PhishSigma++: Malicious Email Detection with Typed Entity Relations
by: Shang, Shang, et al.
Published: (2026) -
Detecting Malicious Intents in Smart Contracts with Pre-trained Programming Language Models
by: Huang, Youwei, et al.
Published: (2025) -
Mitigating Jailbreaks with Intent-Aware LLMs
by: Yeo, Wei Jie, et al.
Published: (2025) -
Evolving Security in LLMs: A Study of Jailbreak Attacks and Defenses
by: Shang, Zhengchun, et al.
Published: (2025) -
Exploring Jailbreak Attacks on LLMs through Intent Concealment and Diversion
by: Cui, Tiehan, et al.
Published: (2025)