Not All Tokens Are Created Equal: Query-Efficient Jailbreak Fuzzing for LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Wenyu, Meng, Xiangtao, Zang, Chuanchao, Wang, Li, Gao, Xinyu, Wang, Jianing, Zhan, Peng, Li, Zheng, Guo, Shanqing |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Defenses at Odds: Measuring and Explaining Defense Conflicts in Large Language Models
by: Meng, Xiangtao, et al.
Published: (2026)
by: Meng, Xiangtao, et al.
Published: (2026)
Fuzz-Testing Meets LLM-Based Agents: An Automated and Efficient Framework for Jailbreaking Text-To-Image Generation Models
by: Dong, Yingkai, et al.
Published: (2024)
by: Dong, Yingkai, et al.
Published: (2024)
From Defender to Devil? Unintended Risk Interactions Induced by LLM Defenses
by: Meng, Xiangtao, et al.
Published: (2025)
by: Meng, Xiangtao, et al.
Published: (2025)
Beyond Known Fakes: Generalized Detection of AI-Generated Images via Post-hoc Distribution Alignment
by: Wang, Li, et al.
Published: (2025)
by: Wang, Li, et al.
Published: (2025)
DCMI: A Differential Calibration Membership Inference Attack Against Retrieval-Augmented Generation
by: Gao, Xinyu, et al.
Published: (2025)
by: Gao, Xinyu, et al.
Published: (2025)
PAPILLON: Efficient and Stealthy Fuzz Testing-Powered Jailbreaks for LLMs
by: Gong, Xueluan, et al.
Published: (2024)
by: Gong, Xueluan, et al.
Published: (2024)
Not All Entities are Created Equal: A Dynamic Anonymization Framework for Privacy-Preserving RAG
by: Zhu, Xinyuan, et al.
Published: (2026)
by: Zhu, Xinyuan, et al.
Published: (2026)
AVA: Inconspicuous Attribute Variation-based Adversarial Attack bypassing DeepFake Detection
by: Meng, Xiangtao, et al.
Published: (2023)
by: Meng, Xiangtao, et al.
Published: (2023)
Beyond the Safety Tax: Mitigating Unsafe Text-to-Image Generation via External Safety Rectification
by: Meng, Xiangtao, et al.
Published: (2025)
by: Meng, Xiangtao, et al.
Published: (2025)
VidLeaks: Membership Inference Attacks Against Text-to-Video Models
by: Wang, Li, et al.
Published: (2026)
by: Wang, Li, et al.
Published: (2026)
PolyJailbreak: Cross-Modal Jailbreaking Attacks on Black-Box Multimodal LLMs
by: Wang, Xinkai, et al.
Published: (2025)
by: Wang, Xinkai, et al.
Published: (2025)
Poisoning Attacks to Local Differential Privacy for Ranking Estimation
by: Zhan, Pei, et al.
Published: (2025)
by: Zhan, Pei, et al.
Published: (2025)
FuzzLLM: A Novel and Universal Fuzzing Framework for Proactively Discovering Jailbreak Vulnerabilities in Large Language Models
by: Yao, Dongyu, et al.
Published: (2023)
by: Yao, Dongyu, et al.
Published: (2023)
SafePTR: Token-Level Jailbreak Defense in Multimodal LLMs via Prune-then-Restore Mechanism
by: Chen, Beitao, et al.
Published: (2025)
by: Chen, Beitao, et al.
Published: (2025)
One Model Transfer to All: On Robust Jailbreak Prompts Generation against LLMs
by: Li, Linbao, et al.
Published: (2025)
by: Li, Linbao, et al.
Published: (2025)
A Simple and Efficient Jailbreak Method Exploiting LLMs' Helpfulness
by: Luo, Xuan, et al.
Published: (2025)
by: Luo, Xuan, et al.
Published: (2025)
BandFuzz: An ML-powered Collaborative Fuzzing Framework
by: Shi, Wenxuan, et al.
Published: (2025)
by: Shi, Wenxuan, et al.
Published: (2025)
LLMs Caught in the Crossfire: Malware Requests and Jailbreak Challenges
by: Li, Haoyang, et al.
Published: (2025)
by: Li, Haoyang, et al.
Published: (2025)
G-Fuzz: A Directed Fuzzing Framework for gVisor
by: Li, Yuwei, et al.
Published: (2024)
by: Li, Yuwei, et al.
Published: (2024)
Efficient and Stealthy Jailbreak Attacks via Adversarial Prompt Distillation from LLMs to SLMs
by: Li, Xiang, et al.
Published: (2025)
by: Li, Xiang, et al.
Published: (2025)
ICL-EVADER: Zero-Query Black-Box Evasion Attacks on In-Context Learning and Their Defenses
by: He, Ningyuan, et al.
Published: (2026)
by: He, Ningyuan, et al.
Published: (2026)
QueryAttack: Jailbreaking Aligned Large Language Models Using Structured Non-natural Query Language
by: Zou, Qingsong, et al.
Published: (2025)
by: Zou, Qingsong, et al.
Published: (2025)
TEMPLATEFUZZ: Fine-Grained Chat Template Fuzzing for Jailbreaking and Red Teaming LLMs
by: Shen, Qingchao, et al.
Published: (2026)
by: Shen, Qingchao, et al.
Published: (2026)
ProphetFuzz: Fully Automated Prediction and Fuzzing of High-Risk Option Combinations with Only Documentation via Large Language Model
by: Wang, Dawei, et al.
Published: (2024)
by: Wang, Dawei, et al.
Published: (2024)
Jailbreaking LLMs & VLMs: Mechanisms, Evaluation, and Unified Defense
by: Chen, Zejian, et al.
Published: (2026)
by: Chen, Zejian, et al.
Published: (2026)
GoldenFuzz: Generative Golden Reference Hardware Fuzzing
by: Wu, Lichao, et al.
Published: (2025)
by: Wu, Lichao, et al.
Published: (2025)
Tracker Installations Are Not Created Equal: Understanding Tracker Configuration of Form Data Collection
by: Kieserman, Julia B., et al.
Published: (2025)
by: Kieserman, Julia B., et al.
Published: (2025)
Evolve the Method, Not the Prompts: Evolutionary Synthesis of Jailbreak Attacks on LLMs
by: Chen, Yunhao, et al.
Published: (2025)
by: Chen, Yunhao, et al.
Published: (2025)
JailbreakLens: Interpreting Jailbreak Mechanism in the Lens of Representation and Circuit
by: He, Zeqing, et al.
Published: (2024)
by: He, Zeqing, et al.
Published: (2024)
Enhancing Jailbreak Attacks on LLMs via Persona Prompts
by: Zhang, Zheng, et al.
Published: (2025)
by: Zhang, Zheng, et al.
Published: (2025)
ChatHTTPFuzz: Large Language Model-Assisted IoT HTTP Fuzzing
by: Yang, Zhe, et al.
Published: (2024)
by: Yang, Zhe, et al.
Published: (2024)
EmbedFuzz: High Speed Fuzzing Through Transplantation
by: Hofhammer, Florian, et al.
Published: (2024)
by: Hofhammer, Florian, et al.
Published: (2024)
Re-Triggering Safeguards within LLMs for Jailbreak Detection
by: Lin, Zheng, et al.
Published: (2026)
by: Lin, Zheng, et al.
Published: (2026)
OrchJail: Jailbreaking Tool-Calling Text-to-Image Agents by Orchestration-Guided Fuzzing
by: Chen, Jianming, et al.
Published: (2026)
by: Chen, Jianming, et al.
Published: (2026)
TASO: Jailbreak LLMs via Alternative Template and Suffix Optimization
by: Wang, Yanting, et al.
Published: (2025)
by: Wang, Yanting, et al.
Published: (2025)
Prompt Fuzzing for Fuzz Driver Generation
by: Lyu, Yunlong, et al.
Published: (2023)
by: Lyu, Yunlong, et al.
Published: (2023)
Untargeted Jailbreak Attack
by: Huang, Xinzhe, et al.
Published: (2025)
by: Huang, Xinzhe, et al.
Published: (2025)
Virtual Context: Enhancing Jailbreak Attacks with Special Token Injection
by: Zhou, Yuqi, et al.
Published: (2024)
by: Zhou, Yuqi, et al.
Published: (2024)
DeepInception: Hypnotize Large Language Model to Be Jailbreaker
by: Li, Xuan, et al.
Published: (2023)
by: Li, Xuan, et al.
Published: (2023)
Fuzz on the Beach: Fuzzing Solana Smart Contracts
by: Smolka, Sven, et al.
Published: (2023)
by: Smolka, Sven, et al.
Published: (2023)
Similar Items
-
Defenses at Odds: Measuring and Explaining Defense Conflicts in Large Language Models
by: Meng, Xiangtao, et al.
Published: (2026) -
Fuzz-Testing Meets LLM-Based Agents: An Automated and Efficient Framework for Jailbreaking Text-To-Image Generation Models
by: Dong, Yingkai, et al.
Published: (2024) -
From Defender to Devil? Unintended Risk Interactions Induced by LLM Defenses
by: Meng, Xiangtao, et al.
Published: (2025) -
Beyond Known Fakes: Generalized Detection of AI-Generated Images via Post-hoc Distribution Alignment
by: Wang, Li, et al.
Published: (2025) -
DCMI: A Differential Calibration Membership Inference Attack Against Retrieval-Augmented Generation
by: Gao, Xinyu, et al.
Published: (2025)