Language Models Can Autonomously Hack and Self-Replicate
Fuente:
arXiv
Saved in:
| Main Authors: | Air, Alena, Reworr, Kotov, Nikolaj, Volkov, Dmitrii, Steidley, John, Ladish, Jeffrey |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LLM Agent Honeypot: Monitoring AI Hacking Agents in the Wild
by: Reworr, et al.
Published: (2024)
by: Reworr, et al.
Published: (2024)
GPT-5 at CTFs: Case Studies From Top-Tier Cybersecurity Events
by: Reworr, et al.
Published: (2025)
by: Reworr, et al.
Published: (2025)
Hacking CTFs with Plain Agents
by: Turtayev, Rustem, et al.
Published: (2024)
by: Turtayev, Rustem, et al.
Published: (2024)
Can LLMs Hack Enterprise Networks? -- Replicated Computational Results (RCR) Report
by: Happe, Andreas, et al.
Published: (2026)
by: Happe, Andreas, et al.
Published: (2026)
Evaluating AI cyber capabilities with crowdsourced elicitation
by: Petrov, Artem, et al.
Published: (2025)
by: Petrov, Artem, et al.
Published: (2025)
BadGPT-4o: stripping safety finetuning from GPT models
by: Krupkina, Ekaterina, et al.
Published: (2024)
by: Krupkina, Ekaterina, et al.
Published: (2024)
Can LLMs Hack Enterprise Networks? Autonomous Assumed Breach Penetration-Testing Active Directory Networks
by: Happe, Andreas, et al.
Published: (2025)
by: Happe, Andreas, et al.
Published: (2025)
Badllama 3: removing safety finetuning from Llama 3 in minutes
by: Volkov, Dmitrii
Published: (2024)
by: Volkov, Dmitrii
Published: (2024)
Unelicitable Backdoors in Language Models via Cryptographic Transformer Circuits
by: Draguns, Andis, et al.
Published: (2024)
by: Draguns, Andis, et al.
Published: (2024)
LLM Agents can Autonomously Hack Websites
by: Fang, Richard, et al.
Published: (2024)
by: Fang, Richard, et al.
Published: (2024)
RepliBench: Evaluating the Autonomous Replication Capabilities of Language Model Agents
by: Black, Sid, et al.
Published: (2025)
by: Black, Sid, et al.
Published: (2025)
A Diamond Model Analysis on Twitter's Biggest Hack
by: Rahalkar, Chaitanya
Published: (2023)
by: Rahalkar, Chaitanya
Published: (2023)
Ethical Hacking and its role in Cybersecurity
by: Asif, Fatima, et al.
Published: (2024)
by: Asif, Fatima, et al.
Published: (2024)
Hacked in Translation -- from Subtitles to Complete Takeover
by: Herscovici, Omri, et al.
Published: (2024)
by: Herscovici, Omri, et al.
Published: (2024)
Cybersecurity AI: Hacking Consumer Robots in the AI Era
by: Mayoral-Vilches, Víctor, et al.
Published: (2026)
by: Mayoral-Vilches, Víctor, et al.
Published: (2026)
Cybersecurity AI: Hacking the AI Hackers via Prompt Injection
by: Mayoral-Vilches, Víctor, et al.
Published: (2025)
by: Mayoral-Vilches, Víctor, et al.
Published: (2025)
RedShell: A Generative AI-Based Approach to Ethical Hacking
by: Bessa, Ricardo, et al.
Published: (2026)
by: Bessa, Ricardo, et al.
Published: (2026)
SoK: A Review of Cross-Chain Bridge Hacks in 2023
by: Belenkov, Nikita, et al.
Published: (2025)
by: Belenkov, Nikita, et al.
Published: (2025)
The Postman: A Journey of Ethical Hacking in PosteID/SPID Borderland
by: Costa, Gabriele
Published: (2025)
by: Costa, Gabriele
Published: (2025)
Hacking the Fabric: Targeting Partial Reconfiguration for Fault Injection in FPGA Fabrics
by: Chaudhuri, Jayeeta, et al.
Published: (2024)
by: Chaudhuri, Jayeeta, et al.
Published: (2024)
ChipmunkRing: A Practical Post-Quantum Ring Signature Scheme for Blockchain Applications
by: Gerasimov, Dmitrii A.
Published: (2025)
by: Gerasimov, Dmitrii A.
Published: (2025)
Marginal sets in semigroups and semirings
by: Buchinskiy, I., et al.
Published: (2025)
by: Buchinskiy, I., et al.
Published: (2025)
Hacking Predictors Means Hacking Cars: Using Sensitivity Analysis to Identify Trajectory Prediction Vulnerabilities for Autonomous Driving Security
by: Gibson, Marsalis, et al.
Published: (2024)
by: Gibson, Marsalis, et al.
Published: (2024)
Bridging the Gap: A Survey and Classification of Research-Informed Ethical Hacking Tools
by: Modesti, Paolo, et al.
Published: (2024)
by: Modesti, Paolo, et al.
Published: (2024)
Passive Hack-Back Strategies for Cyber Attribution: Covert Vectors in Denied Environment
by: Weinberg, Abraham Itzhak
Published: (2025)
by: Weinberg, Abraham Itzhak
Published: (2025)
Hack Me If You Can: Aggregating AutoEncoders for Countering Persistent Access Threats Within Highly Imbalanced Data
by: Benabderrahmane, Sidahmed, et al.
Published: (2024)
by: Benabderrahmane, Sidahmed, et al.
Published: (2024)
A Dual-Level Cancelable Framework for Palmprint Verification and Hack-Proof Data Storage
by: Yang, Ziyuan, et al.
Published: (2024)
by: Yang, Ziyuan, et al.
Published: (2024)
Recipe: Hardware-Accelerated Replication Protocols
by: Giantsidi, Dimitra, et al.
Published: (2025)
by: Giantsidi, Dimitra, et al.
Published: (2025)
HackCar: a test platform for attacks and defenses on a cost-contained automotive architecture
by: Stabili, Dario, et al.
Published: (2024)
by: Stabili, Dario, et al.
Published: (2024)
X Hacking: The Threat of Misguided AutoML
by: Sharma, Rahul, et al.
Published: (2024)
by: Sharma, Rahul, et al.
Published: (2024)
SoK: Prompt Hacking of Large Language Models
by: Rababah, Baha, et al.
Published: (2024)
by: Rababah, Baha, et al.
Published: (2024)
ScamFerret: Detecting Scam Websites Autonomously with Large Language Models
by: Nakano, Hiroki, et al.
Published: (2025)
by: Nakano, Hiroki, et al.
Published: (2025)
Can Large Language Models Automate the Refinement of Cellular Network Specifications?
by: Dong, Jianshuo, et al.
Published: (2025)
by: Dong, Jianshuo, et al.
Published: (2025)
PenTest++: Elevating Ethical Hacking with AI and Automation
by: Al-Sinani, Haitham S., et al.
Published: (2025)
by: Al-Sinani, Haitham S., et al.
Published: (2025)
Work-in-Progress: Crash Course: Can (Under Attack) Autonomous Driving Beat Human Drivers?
by: Marchiori, Francesco, et al.
Published: (2024)
by: Marchiori, Francesco, et al.
Published: (2024)
Agentic Knowledge Distillation: Autonomous Training of Small Language Models for SMS Threat Detection
by: ElZemity, Adel, et al.
Published: (2026)
by: ElZemity, Adel, et al.
Published: (2026)
AI-Enhanced Ethical Hacking: A Linux-Focused Experiment
by: Al-Sinani, Haitham S., et al.
Published: (2024)
by: Al-Sinani, Haitham S., et al.
Published: (2024)
Trusted-Execution Environment (TEE) for Solving the Replication Crisis in Academia
by: Li, Jiasun, et al.
Published: (2026)
by: Li, Jiasun, et al.
Published: (2026)
SPARE: Securing Progressive Web Applications Against Unauthorized Replications
by: Talukder, Sajib, et al.
Published: (2025)
by: Talukder, Sajib, et al.
Published: (2025)
Fair Ordering in Replicated Systems via Streaming Social Choice
by: Ramseyer, Geoffrey, et al.
Published: (2023)
by: Ramseyer, Geoffrey, et al.
Published: (2023)
Similar Items
-
LLM Agent Honeypot: Monitoring AI Hacking Agents in the Wild
by: Reworr, et al.
Published: (2024) -
GPT-5 at CTFs: Case Studies From Top-Tier Cybersecurity Events
by: Reworr, et al.
Published: (2025) -
Hacking CTFs with Plain Agents
by: Turtayev, Rustem, et al.
Published: (2024) -
Can LLMs Hack Enterprise Networks? -- Replicated Computational Results (RCR) Report
by: Happe, Andreas, et al.
Published: (2026) -
Evaluating AI cyber capabilities with crowdsourced elicitation
by: Petrov, Artem, et al.
Published: (2025)