SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Xiaodong, Li, Xiangman, Li, Qi, Liu, Lingshuang, Ni, Jianbing |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Robustness of Watermarking on Text-to-Image Diffusion Models
by: Wu, Xiaodong, et al.
Published: (2024)
by: Wu, Xiaodong, et al.
Published: (2024)
SoK: Evaluating Jailbreak Guardrails for Large Language Models
by: Wang, Xunguang, et al.
Published: (2025)
by: Wang, Xunguang, et al.
Published: (2025)
SoK: a Comprehensive Causality Analysis Framework for Large Language Model Security
by: Zhao, Wei, et al.
Published: (2025)
by: Zhao, Wei, et al.
Published: (2025)
Security in LLM-as-a-Judge: A Comprehensive SoK
by: Masoud, Aiman Al, et al.
Published: (2026)
by: Masoud, Aiman Al, et al.
Published: (2026)
SoK: Robustness in Large Language Models against Jailbreak Attacks
by: Xu, Feiyue, et al.
Published: (2026)
by: Xu, Feiyue, et al.
Published: (2026)
SoK: Semantic Privacy in Large Language Models
by: Ma, Baihe, et al.
Published: (2025)
by: Ma, Baihe, et al.
Published: (2025)
SecureT2I: No More Unauthorized Manipulation on AI Generated Images from Prompts
by: Wu, Xiaodong, et al.
Published: (2025)
by: Wu, Xiaodong, et al.
Published: (2025)
SoK: Security Analysis of Blockchain-based Cryptocurrency
by: Liu, Zekai, et al.
Published: (2025)
by: Liu, Zekai, et al.
Published: (2025)
PDLRecover: Privacy-preserving Decentralized Model Recovery with Machine Unlearning
by: Li, Xiangman, et al.
Published: (2025)
by: Li, Xiangman, et al.
Published: (2025)
When There Is No Decoder: Removing Watermarks from Stable Diffusion Models in a No-box Setting
by: Wu, Xiaodong, et al.
Published: (2025)
by: Wu, Xiaodong, et al.
Published: (2025)
SoK: On the Semantic AI Security in Autonomous Driving
by: Shen, Junjie, et al.
Published: (2022)
by: Shen, Junjie, et al.
Published: (2022)
From AI-Generated Content to Agentic Action: Security and Safety Threats in Generative AI
by: Zhang, Zelin, et al.
Published: (2026)
by: Zhang, Zelin, et al.
Published: (2026)
SoK: Taxonomy and Evaluation of Prompt Security in Large Language Models
by: Hong, Hanbin, et al.
Published: (2025)
by: Hong, Hanbin, et al.
Published: (2025)
SoK: Security and Privacy of AI Agents for Blockchain
by: Romandini, Nicolò, et al.
Published: (2025)
by: Romandini, Nicolò, et al.
Published: (2025)
SoK: Towards Security and Safety of Edge AI
by: Wingarz, Tatjana, et al.
Published: (2024)
by: Wingarz, Tatjana, et al.
Published: (2024)
SoK: Understanding (New) Security Issues Across AI4Code Use Cases
by: Wu, Qilong, et al.
Published: (2025)
by: Wu, Qilong, et al.
Published: (2025)
MelShield: Robust Mel-Domain Audio Watermarking for Provenance Attribution of AI Generated Synthesized Speech
by: Jin, Yutong, et al.
Published: (2026)
by: Jin, Yutong, et al.
Published: (2026)
SoK: An Introspective Analysis of RPKI Security
by: Mirdita, Donika, et al.
Published: (2024)
by: Mirdita, Donika, et al.
Published: (2024)
SoK: Security and Privacy Risks of Healthcare AI
by: Chang, Yuanhaur, et al.
Published: (2024)
by: Chang, Yuanhaur, et al.
Published: (2024)
SoK: On Gradient Leakage in Federated Learning
by: Du, Jiacheng, et al.
Published: (2024)
by: Du, Jiacheng, et al.
Published: (2024)
SoK: Security of Programmable Logic Controllers
by: López-Morales, Efrén, et al.
Published: (2024)
by: López-Morales, Efrén, et al.
Published: (2024)
SoK: Unlearnability and Unlearning for Model Dememorization
by: Zhang, Mengying, et al.
Published: (2026)
by: Zhang, Mengying, et al.
Published: (2026)
SoK: Trust-Authorization Mismatch in LLM Agent Interactions
by: Shi, Guanquan, et al.
Published: (2025)
by: Shi, Guanquan, et al.
Published: (2025)
SoK: The Last Line of Defense: On Backdoor Defense Evaluation
by: Abad, Gorka, et al.
Published: (2025)
by: Abad, Gorka, et al.
Published: (2025)
SoK: Security of EMV Contactless Payment Systems
by: Nezhad, Mahshid Mehr, et al.
Published: (2025)
by: Nezhad, Mahshid Mehr, et al.
Published: (2025)
SoK: Watermarking for AI-Generated Content
by: Zhao, Xuandong, et al.
Published: (2024)
by: Zhao, Xuandong, et al.
Published: (2024)
SoK: The Security-Safety Continuum of Multimodal Foundation Models through Information Flow and Global Game-Theoretic Analysis of Asymmetric Threats
by: Sun, Ruoxi, et al.
Published: (2024)
by: Sun, Ruoxi, et al.
Published: (2024)
SoK: Benchmarking Poisoning Attacks and Defenses in Federated Learning
by: Zhang, Heyi, et al.
Published: (2025)
by: Zhang, Heyi, et al.
Published: (2025)
SoK: Security of Autonomous LLM Agents in Agentic Commerce
by: Mao, Qian'ang, et al.
Published: (2026)
by: Mao, Qian'ang, et al.
Published: (2026)
SoK: Large Language Model Copyright Auditing via Fingerprinting
by: Shao, Shuo, et al.
Published: (2025)
by: Shao, Shuo, et al.
Published: (2025)
SoK: Public Blockchain Sharding
by: Barat, Md Mohaimin Al, et al.
Published: (2024)
by: Barat, Md Mohaimin Al, et al.
Published: (2024)
SoK: Analysis techniques for WebAssembly
by: Harnes, Håkon, et al.
Published: (2024)
by: Harnes, Håkon, et al.
Published: (2024)
SoK: Leveraging Transformers for Malware Analysis
by: Kunwar, Pradip, et al.
Published: (2024)
by: Kunwar, Pradip, et al.
Published: (2024)
SoK: Evolution, Security, and Fundamental Properties of Transactional Systems
by: Waterpeace, Sky Pelletier, et al.
Published: (2026)
by: Waterpeace, Sky Pelletier, et al.
Published: (2026)
SoK: Verifiable Cross-Silo FL
by: Korneev, Aleksei, et al.
Published: (2024)
by: Korneev, Aleksei, et al.
Published: (2024)
SoK: Prompt Hacking of Large Language Models
by: Rababah, Baha, et al.
Published: (2024)
by: Rababah, Baha, et al.
Published: (2024)
SoK: Runtime Integrity
by: Ammar, Mahmoud, et al.
Published: (2024)
by: Ammar, Mahmoud, et al.
Published: (2024)
SoK: How Robust is Audio Watermarking in Generative AI models?
by: Wen, Yizhu, et al.
Published: (2025)
by: Wen, Yizhu, et al.
Published: (2025)
Towards Understanding the Safety Boundaries of DeepSeek Models: Evaluation and Findings
by: Ying, Zonghao, et al.
Published: (2025)
by: Ying, Zonghao, et al.
Published: (2025)
SoK: Speedy Secure Finality
by: Saraswat, Yash, et al.
Published: (2025)
by: Saraswat, Yash, et al.
Published: (2025)
Similar Items
-
Robustness of Watermarking on Text-to-Image Diffusion Models
by: Wu, Xiaodong, et al.
Published: (2024) -
SoK: Evaluating Jailbreak Guardrails for Large Language Models
by: Wang, Xunguang, et al.
Published: (2025) -
SoK: a Comprehensive Causality Analysis Framework for Large Language Model Security
by: Zhao, Wei, et al.
Published: (2025) -
Security in LLM-as-a-Judge: A Comprehensive SoK
by: Masoud, Aiman Al, et al.
Published: (2026) -
SoK: Robustness in Large Language Models against Jailbreak Attacks
by: Xu, Feiyue, et al.
Published: (2026)