GeneBreaker: Jailbreak Attacks against DNA Language Models with Pathogenicity Guidance
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Zaixi, Zhou, Zhenghong, Jin, Ruofan, Cong, Le, Wang, Mengdi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Securing the Language of Life: Inheritable Watermarks from DNA Language Models to Proteins
by: Zhang, Zaixi, et al.
Published: (2025)
by: Zhang, Zaixi, et al.
Published: (2025)
SafeProtein: Red-Teaming Framework and Benchmark for Protein Foundation Models
by: Fan, Jigang, et al.
Published: (2025)
by: Fan, Jigang, et al.
Published: (2025)
FoldMark: Protecting Protein Generative Models with Watermarking
by: Zhang, Zaixi, et al.
Published: (2024)
by: Zhang, Zaixi, et al.
Published: (2024)
FIMBA: Evaluating the Robustness of AI in Genomics via Feature Importance Adversarial Attacks
by: Skovorodnikov, Heorhii, et al.
Published: (2024)
by: Skovorodnikov, Heorhii, et al.
Published: (2024)
Validating GWAS Findings through Reverse Engineering of Contingency Tables
by: Jiang, Yuzhou, et al.
Published: (2024)
by: Jiang, Yuzhou, et al.
Published: (2024)
From Exponential to Polynomial Complexity: Efficient Permutation Counting with Subword Constraints
by: Mathew, Martin, et al.
Published: (2024)
by: Mathew, Martin, et al.
Published: (2024)
Quantifying Memorization and Privacy Risks in Genomic Language Models
by: Nemecek, Alexander, et al.
Published: (2026)
by: Nemecek, Alexander, et al.
Published: (2026)
DP-SNP-TIHMM: Differentially Private, Time-Inhomogeneous Hidden Markov Models for Synthesizing Genome-Wide Association Datasets
by: Rahimian, Shadi, et al.
Published: (2025)
by: Rahimian, Shadi, et al.
Published: (2025)
dabih -- encrypted data storage and sharing platform
by: Huttner, Michael, et al.
Published: (2024)
by: Huttner, Michael, et al.
Published: (2024)
Federated Learning in Genetics: Extended Analysis of Accuracy, Performance and Privacy Trade-offs
by: Hannemann, Anika, et al.
Published: (2024)
by: Hannemann, Anika, et al.
Published: (2024)
SafeAligner: Safety Alignment against Jailbreak Attacks via Response Disparity Guidance
by: Huang, Caishuang, et al.
Published: (2024)
by: Huang, Caishuang, et al.
Published: (2024)
White-box Membership Inference Attacks against Diffusion Models
by: Pang, Yan, et al.
Published: (2023)
by: Pang, Yan, et al.
Published: (2023)
SoK: Robustness in Large Language Models against Jailbreak Attacks
by: Xu, Feiyue, et al.
Published: (2026)
by: Xu, Feiyue, et al.
Published: (2026)
Align is not Enough: Multimodal Universal Jailbreak Attack against Multimodal Large Language Models
by: Wang, Youze, et al.
Published: (2025)
by: Wang, Youze, et al.
Published: (2025)
Persona Attack: Incremental Memory Injection Jailbreak Attack against Large Language Models
by: Park, Junyoung, et al.
Published: (2026)
by: Park, Junyoung, et al.
Published: (2026)
BaThe: Defense against the Jailbreak Attack in Multimodal Large Language Models by Treating Harmful Instruction as Backdoor Trigger
by: Chen, Yulin, et al.
Published: (2024)
by: Chen, Yulin, et al.
Published: (2024)
JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation
by: Zhang, Shenyi, et al.
Published: (2025)
by: Zhang, Shenyi, et al.
Published: (2025)
Jailbreaking Prompt Attack: A Controllable Adversarial Attack against Diffusion Models
by: Ma, Jiachen, et al.
Published: (2024)
by: Ma, Jiachen, et al.
Published: (2024)
RTL-Breaker: Assessing the Security of LLMs against Backdoor Attacks on HDL Code Generation
by: Mankali, Lakshmi Likhitha, et al.
Published: (2024)
by: Mankali, Lakshmi Likhitha, et al.
Published: (2024)
GateBreaker: Gate-Guided Attacks on Mixture-of-Expert LLMs
by: Wu, Lichao, et al.
Published: (2025)
by: Wu, Lichao, et al.
Published: (2025)
Jailbreaking Attack against Multimodal Large Language Model
by: Niu, Zhenxing, et al.
Published: (2024)
by: Niu, Zhenxing, et al.
Published: (2024)
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks
by: Yin, Ziyi, et al.
Published: (2025)
by: Yin, Ziyi, et al.
Published: (2025)
AISA: Awakening Intrinsic Safety Awareness in Large Language Models against Jailbreak Attacks
by: Song, Weiming, et al.
Published: (2026)
by: Song, Weiming, et al.
Published: (2026)
A Biosecurity Agent for Lifecycle LLM Biosecurity Alignment
by: Meng, Meiyin, et al.
Published: (2025)
by: Meng, Meiyin, et al.
Published: (2025)
StructuralSleight: Automated Jailbreak Attacks on Large Language Models Utilizing Uncommon Text-Organization Structures
by: Li, Bangxin, et al.
Published: (2024)
by: Li, Bangxin, et al.
Published: (2024)
RL-JACK: Reinforcement Learning-powered Black-box Jailbreaking Attack against LLMs
by: Chen, Xuan, et al.
Published: (2024)
by: Chen, Xuan, et al.
Published: (2024)
Prefix Guidance: A Steering Wheel for Large Language Models to Defend Against Jailbreak Attacks
by: Zhao, Jiawei, et al.
Published: (2024)
by: Zhao, Jiawei, et al.
Published: (2024)
Model-Editing-Based Jailbreak against Safety-aligned Large Language Models
by: Li, Yuxi, et al.
Published: (2024)
by: Li, Yuxi, et al.
Published: (2024)
EnJa: Ensemble Jailbreak on Large Language Models
by: Zhang, Jiahao, et al.
Published: (2024)
by: Zhang, Jiahao, et al.
Published: (2024)
LineBreaker: Finding Token-Inconsistency Bugs with Large Language Models
by: Chen, Hongbo, et al.
Published: (2024)
by: Chen, Hongbo, et al.
Published: (2024)
Prompt Inversion Attack against Collaborative Inference of Large Language Models
by: Qu, Wenjie, et al.
Published: (2025)
by: Qu, Wenjie, et al.
Published: (2025)
RobustKV: Defending Large Language Models against Jailbreak Attacks via KV Eviction
by: Jiang, Tanqiu, et al.
Published: (2024)
by: Jiang, Tanqiu, et al.
Published: (2024)
Defensive Prompt Patch: A Robust and Interpretable Defense of LLMs against Jailbreak Attacks
by: Xiong, Chen, et al.
Published: (2024)
by: Xiong, Chen, et al.
Published: (2024)
AutoJailbreak: Exploring Jailbreak Attacks and Defenses through a Dependency Lens
by: Lu, Lin, et al.
Published: (2024)
by: Lu, Lin, et al.
Published: (2024)
Systematic Scaling Analysis of Jailbreak Attacks in Large Language Models
by: Wang, Xiangwen, et al.
Published: (2026)
by: Wang, Xiangwen, et al.
Published: (2026)
Chain-of-Scrutiny: Detecting Backdoor Attacks for Large Language Models
by: Li, Xi, et al.
Published: (2024)
by: Li, Xi, et al.
Published: (2024)
Knowledge-to-Jailbreak: Investigating Knowledge-driven Jailbreaking Attacks for Large Language Models
by: Tu, Shangqing, et al.
Published: (2024)
by: Tu, Shangqing, et al.
Published: (2024)
PandaGuard: Systematic Evaluation of LLM Safety against Jailbreaking Attacks
by: Shen, Guobin, et al.
Published: (2025)
by: Shen, Guobin, et al.
Published: (2025)
Virtual Context: Enhancing Jailbreak Attacks with Special Token Injection
by: Zhou, Yuqi, et al.
Published: (2024)
by: Zhou, Yuqi, et al.
Published: (2024)
MacPrompt: Maraconic-guided Jailbreak against Text-to-Image Models
by: Ye, Xi, et al.
Published: (2026)
by: Ye, Xi, et al.
Published: (2026)
Similar Items
-
Securing the Language of Life: Inheritable Watermarks from DNA Language Models to Proteins
by: Zhang, Zaixi, et al.
Published: (2025) -
SafeProtein: Red-Teaming Framework and Benchmark for Protein Foundation Models
by: Fan, Jigang, et al.
Published: (2025) -
FoldMark: Protecting Protein Generative Models with Watermarking
by: Zhang, Zaixi, et al.
Published: (2024) -
FIMBA: Evaluating the Robustness of AI in Genomics via Feature Importance Adversarial Attacks
by: Skovorodnikov, Heorhii, et al.
Published: (2024) -
Validating GWAS Findings through Reverse Engineering of Contingency Tables
by: Jiang, Yuzhou, et al.
Published: (2024)