Can LLMs Threaten Human Survival? Benchmarking Potential Existential Threats from LLMs via Prefix Completion
Fuente:
arXiv
Saved in:
| Main Authors: | Cui, Yu, Liu, Yifei, Fu, Hang, Pan, Sicheng, Zhang, Haibin, Zuo, Cong, Wang, Licheng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VortexPIA: Indirect Prompt Injection Attack against LLMs for Efficient Extraction of User Privacy
by: Cui, Yu, et al.
Published: (2025)
by: Cui, Yu, et al.
Published: (2025)
Spore: Efficient and Training-Free Privacy Extraction Attack on LLMs via Inference-Time Hybrid Probing
by: Cui, Yu, et al.
Published: (2026)
by: Cui, Yu, et al.
Published: (2026)
Free-MAD: Consensus-Free Multi-Agent Debate
by: Cui, Yu, et al.
Published: (2025)
by: Cui, Yu, et al.
Published: (2025)
Towards Provably Secure Generative AI: Reliable Consensus Sampling
by: Cui, Yu, et al.
Published: (2025)
by: Cui, Yu, et al.
Published: (2025)
Ramp Up NTT in Record Time using GPU-Accelerated Algorithms and LLM-based Code Generation
by: Cui, Yu, et al.
Published: (2025)
by: Cui, Yu, et al.
Published: (2025)
Benchmarking LLMs in an Embodied Environment for Blue Team Threat Hunting
by: Liu, Xiaoqun, et al.
Published: (2025)
by: Liu, Xiaoqun, et al.
Published: (2025)
Practical Reasoning Interruption Attacks on Reasoning Large Language Models
by: Cui, Yu, et al.
Published: (2025)
by: Cui, Yu, et al.
Published: (2025)
CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence
by: Alam, Md Tanvirul, et al.
Published: (2024)
by: Alam, Md Tanvirul, et al.
Published: (2024)
Using LLMs to Automate Threat Intelligence Analysis Workflows in Security Operation Centers
by: Tseng, PeiYu, et al.
Published: (2024)
by: Tseng, PeiYu, et al.
Published: (2024)
CyberSOCEval: Benchmarking LLMs Capabilities for Malware Analysis and Threat Intelligence Reasoning
by: Deason, Lauren, et al.
Published: (2025)
by: Deason, Lauren, et al.
Published: (2025)
AthenaBench: A Dynamic Benchmark for Evaluating LLMs in Cyber Threat Intelligence
by: Alam, Md Tanvirul, et al.
Published: (2025)
by: Alam, Md Tanvirul, et al.
Published: (2025)
Merger-as-a-Stealer: Stealing Targeted PII from Aligned LLMs with Model Merging
by: Lu, Lin, et al.
Published: (2025)
by: Lu, Lin, et al.
Published: (2025)
When LLMs Go Online: The Emerging Threat of Web-Enabled LLMs
by: Kim, Hanna, et al.
Published: (2024)
by: Kim, Hanna, et al.
Published: (2024)
Can LLMs Classify CVEs? Investigating LLMs Capabilities in Computing CVSS Vectors
by: Marchiori, Francesco, et al.
Published: (2025)
by: Marchiori, Francesco, et al.
Published: (2025)
SafeRedirect: Defeating Internal Safety Collapse via Task-Completion Redirection in Frontier LLMs
by: Pan, Chao, et al.
Published: (2026)
by: Pan, Chao, et al.
Published: (2026)
Can Watermarked LLMs be Identified by Users via Crafted Prompts?
by: Liu, Aiwei, et al.
Published: (2024)
by: Liu, Aiwei, et al.
Published: (2024)
Can LLMs be Fooled? Investigating Vulnerabilities in LLMs
by: Abdali, Sara, et al.
Published: (2024)
by: Abdali, Sara, et al.
Published: (2024)
Can LLMs be Scammed? A Baseline Measurement Study
by: Sehwag, Udari Madhushani, et al.
Published: (2024)
by: Sehwag, Udari Madhushani, et al.
Published: (2024)
You Can't Eat Your Cake and Have It Too: The Performance Degradation of LLMs with Jailbreak Defense
by: Mai, Wuyuao, et al.
Published: (2025)
by: Mai, Wuyuao, et al.
Published: (2025)
Can Developers rely on LLMs for Secure IaC Development?
by: Firouzi, Ehsan, et al.
Published: (2026)
by: Firouzi, Ehsan, et al.
Published: (2026)
Proactively Detecting Threats: A Novel Approach Using LLMs
by: Chawla, Aniesh, et al.
Published: (2026)
by: Chawla, Aniesh, et al.
Published: (2026)
"To Survive, I Must Defect": Jailbreaking LLMs via the Game-Theory Scenarios
by: Sun, Zhen, et al.
Published: (2025)
by: Sun, Zhen, et al.
Published: (2025)
LLMs Can Unlearn Refusal with Only 1,000 Benign Samples
by: Guo, Yangyang, et al.
Published: (2026)
by: Guo, Yangyang, et al.
Published: (2026)
Confusion is the Final Barrier: Rethinking Jailbreak Evaluation and Investigating the Real Misuse Threat of LLMs
by: Yan, Yu, et al.
Published: (2025)
by: Yan, Yu, et al.
Published: (2025)
GridSE: Towards Practical Secure Geographic Search via Prefix Symmetric Searchable Encryption (Full Version)
by: Guo, Ruoyang, et al.
Published: (2024)
by: Guo, Ruoyang, et al.
Published: (2024)
Advancing Autonomous Incident Response: Leveraging LLMs and Cyber Threat Intelligence
by: Tellache, Amine, et al.
Published: (2025)
by: Tellache, Amine, et al.
Published: (2025)
Can LLMs Patch Security Issues?
by: Alrashedy, Kamel, et al.
Published: (2023)
by: Alrashedy, Kamel, et al.
Published: (2023)
LLMs in Cybersecurity: Friend or Foe in the Human Decision Loop?
by: Pekaric, Irdin, et al.
Published: (2025)
by: Pekaric, Irdin, et al.
Published: (2025)
Knowledge Transfer from LLMs to Provenance Analysis: A Semantic-Augmented Method for APT Detection
by: Zuo, Fei, et al.
Published: (2025)
by: Zuo, Fei, et al.
Published: (2025)
Threat Me Right: A Human HARMS Threat Model for Technical Systems
by: Turk, Kieron Ivy, et al.
Published: (2025)
by: Turk, Kieron Ivy, et al.
Published: (2025)
PREE: Towards Harmless and Adaptive Fingerprint Editing in Large Language Models via Knowledge Prefix Enhancement
by: Yue, Xubin, et al.
Published: (2025)
by: Yue, Xubin, et al.
Published: (2025)
The Ethics of Interaction: Mitigating Security Threats in LLMs
by: Kumar, Ashutosh, et al.
Published: (2024)
by: Kumar, Ashutosh, et al.
Published: (2024)
AdvPrefix: An Objective for Nuanced LLM Jailbreaks
by: Zhu, Sicheng, et al.
Published: (2024)
by: Zhu, Sicheng, et al.
Published: (2024)
Can LLMs Hack Enterprise Networks? -- Replicated Computational Results (RCR) Report
by: Happe, Andreas, et al.
Published: (2026)
by: Happe, Andreas, et al.
Published: (2026)
CIMemories: A Compositional Benchmark for Contextual Integrity of Persistent Memory in LLMs
by: Mireshghallah, Niloofar, et al.
Published: (2025)
by: Mireshghallah, Niloofar, et al.
Published: (2025)
Fine-Tuning LLMs for Code Mutation: A New Era of Cyber Threats
by: Setak, Mohammad, et al.
Published: (2024)
by: Setak, Mohammad, et al.
Published: (2024)
Chimera: Harnessing Multi-Agent LLMs for Automatic Insider Threat Simulation
by: Yu, Jiongchi, et al.
Published: (2025)
by: Yu, Jiongchi, et al.
Published: (2025)
Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting
by: Meng, Yuqiao, et al.
Published: (2025)
by: Meng, Yuqiao, et al.
Published: (2025)
CyLens: Towards Reinventing Cyber Threat Intelligence in the Paradigm of Agentic Large Language Models
by: Liu, Xiaoqun, et al.
Published: (2025)
by: Liu, Xiaoqun, et al.
Published: (2025)
Agentic Misalignment: How LLMs Could Be Insider Threats
by: Lynch, Aengus, et al.
Published: (2025)
by: Lynch, Aengus, et al.
Published: (2025)
Similar Items
-
VortexPIA: Indirect Prompt Injection Attack against LLMs for Efficient Extraction of User Privacy
by: Cui, Yu, et al.
Published: (2025) -
Spore: Efficient and Training-Free Privacy Extraction Attack on LLMs via Inference-Time Hybrid Probing
by: Cui, Yu, et al.
Published: (2026) -
Free-MAD: Consensus-Free Multi-Agent Debate
by: Cui, Yu, et al.
Published: (2025) -
Towards Provably Secure Generative AI: Reliable Consensus Sampling
by: Cui, Yu, et al.
Published: (2025) -
Ramp Up NTT in Record Time using GPU-Accelerated Algorithms and LLM-based Code Generation
by: Cui, Yu, et al.
Published: (2025)