Out of the Cage: How Stochastic Parrots Win in Cyber Security Environments
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rigaki, Maria, Lukáš, Ondřej, Catania, Carlos A., Garcia, Sebastian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Hackphyr: A Local Fine-Tuned LLM Agent for Network Security Environments
von: Rigaki, Maria, et al.
Veröffentlicht: (2024)
von: Rigaki, Maria, et al.
Veröffentlicht: (2024)
The Power of MEME: Adversarial Malware Creation with Model-Based Reinforcement Learning
von: Rigaki, Maria, et al.
Veröffentlicht: (2023)
von: Rigaki, Maria, et al.
Veröffentlicht: (2023)
LLM in the Shell: Generative Honeypots
von: Sladić, Muris, et al.
Veröffentlicht: (2023)
von: Sladić, Muris, et al.
Veröffentlicht: (2023)
VelLMes: A high-interaction AI-based deception framework
von: Sladić, Muris, et al.
Veröffentlicht: (2025)
von: Sladić, Muris, et al.
Veröffentlicht: (2025)
Personal Information Parroting in Language Models
von: Subramani, Nishant, et al.
Veröffentlicht: (2026)
von: Subramani, Nishant, et al.
Veröffentlicht: (2026)
Evaluating Generalization Mechanisms in Autonomous Cyber Attack Agents
von: Lukáš, Ondřej, et al.
Veröffentlicht: (2026)
von: Lukáš, Ondřej, et al.
Veröffentlicht: (2026)
Actionable Cyber Threat Intelligence using Knowledge Graphs and Large Language Models
von: Fieblinger, Romy, et al.
Veröffentlicht: (2024)
von: Fieblinger, Romy, et al.
Veröffentlicht: (2024)
ExCyTIn-Bench: Evaluating LLM agents on Cyber Threat Investigation
von: Wu, Yiran, et al.
Veröffentlicht: (2025)
von: Wu, Yiran, et al.
Veröffentlicht: (2025)
The Use of Large Language Models (LLM) for Cyber Threat Intelligence (CTI) in Cybercrime Forums
von: Clairoux-Trepanier, Vanessa, et al.
Veröffentlicht: (2024)
von: Clairoux-Trepanier, Vanessa, et al.
Veröffentlicht: (2024)
Towards a scalable AI-driven framework for data-independent Cyber Threat Intelligence Information Extraction
von: Sorokoletova, Olga, et al.
Veröffentlicht: (2025)
von: Sorokoletova, Olga, et al.
Veröffentlicht: (2025)
When RAG Chatbots Expose Their Backend: An Anonymized Case Study of Privacy and Security Risks in Patient-Facing Medical AI
von: Madrid-García, Alfredo, et al.
Veröffentlicht: (2026)
von: Madrid-García, Alfredo, et al.
Veröffentlicht: (2026)
Watch Out for Your Agents! Investigating Backdoor Threats to LLM-Based Agents
von: Yang, Wenkai, et al.
Veröffentlicht: (2024)
von: Yang, Wenkai, et al.
Veröffentlicht: (2024)
AnnoCTR: A Dataset for Detecting and Linking Entities, Tactics, and Techniques in Cyber Threat Reports
von: Lange, Lukas, et al.
Veröffentlicht: (2024)
von: Lange, Lukas, et al.
Veröffentlicht: (2024)
AVISE: Framework for Evaluating the Security of AI Systems
von: Lempinen, Mikko, et al.
Veröffentlicht: (2026)
von: Lempinen, Mikko, et al.
Veröffentlicht: (2026)
The Ethics of Interaction: Mitigating Security Threats in LLMs
von: Kumar, Ashutosh, et al.
Veröffentlicht: (2024)
von: Kumar, Ashutosh, et al.
Veröffentlicht: (2024)
Institutional Platform for Secure Self-Service Large Language Model Exploration
von: Bumgardner, V. K. Cody, et al.
Veröffentlicht: (2024)
von: Bumgardner, V. K. Cody, et al.
Veröffentlicht: (2024)
LLM for SoC Security: A Paradigm Shift
von: Saha, Dipayan, et al.
Veröffentlicht: (2023)
von: Saha, Dipayan, et al.
Veröffentlicht: (2023)
How Well Can LLM Agents Simulate End-User Security and Privacy Attitudes and Behaviors?
von: Li, Yuxuan, et al.
Veröffentlicht: (2026)
von: Li, Yuxuan, et al.
Veröffentlicht: (2026)
A Survey on Agentic Security: Applications, Threats and Defenses
von: Shahriar, Asif, et al.
Veröffentlicht: (2025)
von: Shahriar, Asif, et al.
Veröffentlicht: (2025)
Text Embedding Inversion Security for Multilingual Language Models
von: Chen, Yiyi, et al.
Veröffentlicht: (2024)
von: Chen, Yiyi, et al.
Veröffentlicht: (2024)
Cheating Automatic LLM Benchmarks: Null Models Achieve High Win Rates
von: Zheng, Xiaosen, et al.
Veröffentlicht: (2024)
von: Zheng, Xiaosen, et al.
Veröffentlicht: (2024)
Security and Privacy Challenges of Large Language Models: A Survey
von: Das, Badhan Chandra, et al.
Veröffentlicht: (2024)
von: Das, Badhan Chandra, et al.
Veröffentlicht: (2024)
Towards Reliable and Practical LLM Security Evaluations via Bayesian Modelling
von: Llewellyn, Mary, et al.
Veröffentlicht: (2025)
von: Llewellyn, Mary, et al.
Veröffentlicht: (2025)
A High-Capacity and Secure Disambiguation Algorithm for Neural Linguistic Steganography
von: Feng, Yapei, et al.
Veröffentlicht: (2025)
von: Feng, Yapei, et al.
Veröffentlicht: (2025)
GradingAttack: Exposing Security Vulnerabilities in LLM Based Educational Grading Agents
von: Li, Xueyi, et al.
Veröffentlicht: (2026)
von: Li, Xueyi, et al.
Veröffentlicht: (2026)
From Vulnerabilities to Remediation: A Systematic Literature Review of LLMs in Code Security
von: Basic, Enna, et al.
Veröffentlicht: (2024)
von: Basic, Enna, et al.
Veröffentlicht: (2024)
Balancing Innovation and Privacy: Data Security Strategies in Natural Language Processing Applications
von: Liu, Shaobo, et al.
Veröffentlicht: (2024)
von: Liu, Shaobo, et al.
Veröffentlicht: (2024)
Secure Code Generation via Online Reinforcement Learning with Vulnerability Reward Model
von: Wu, Tianyi, et al.
Veröffentlicht: (2026)
von: Wu, Tianyi, et al.
Veröffentlicht: (2026)
Supporting Artifact Evaluation with LLMs: A Study with Published Security Research Papers
von: Heye, David, et al.
Veröffentlicht: (2026)
von: Heye, David, et al.
Veröffentlicht: (2026)
AgentSOC: A Multi-Layer Agentic AI Framework for Security Operations Automation
von: Roy, Joyjit, et al.
Veröffentlicht: (2026)
von: Roy, Joyjit, et al.
Veröffentlicht: (2026)
Swiss-Bench 003: Evaluating LLM Reliability and Adversarial Security for Swiss Regulatory Contexts
von: Uenal, Fatih
Veröffentlicht: (2026)
von: Uenal, Fatih
Veröffentlicht: (2026)
EventHunter: Dynamic Clustering and Ranking of Security Events from Hacker Forum Discussions
von: Ech-Chammakhy, Yasir, et al.
Veröffentlicht: (2025)
von: Ech-Chammakhy, Yasir, et al.
Veröffentlicht: (2025)
Guarding Your Conversations: Privacy Gatekeepers for Secure Interactions with Cloud-Based AI Models
von: Uzor, GodsGift, et al.
Veröffentlicht: (2025)
von: Uzor, GodsGift, et al.
Veröffentlicht: (2025)
Attestable Audits: Verifiable AI Safety Benchmarks Using Trusted Execution Environments
von: Schnabl, Christoph, et al.
Veröffentlicht: (2025)
von: Schnabl, Christoph, et al.
Veröffentlicht: (2025)
PyRIT: A Framework for Security Risk Identification and Red Teaming in Generative AI System
von: Munoz, Gary D. Lopez, et al.
Veröffentlicht: (2024)
von: Munoz, Gary D. Lopez, et al.
Veröffentlicht: (2024)
TITAN: Graph-Executable Reasoning for Cyber Threat Intelligence
von: Simoni, Marco, et al.
Veröffentlicht: (2025)
von: Simoni, Marco, et al.
Veröffentlicht: (2025)
How Jailbreak Defenses Work and Ensemble? A Mechanistic Investigation
von: Long, Zhuohang, et al.
Veröffentlicht: (2025)
von: Long, Zhuohang, et al.
Veröffentlicht: (2025)
LLM Cyber Evaluations Don't Capture Real-World Risk
von: Lukošiūtė, Kamilė, et al.
Veröffentlicht: (2025)
von: Lukošiūtė, Kamilė, et al.
Veröffentlicht: (2025)
Counteracting Concept Drift by Learning with Future Malware Predictions
von: Bosansky, Branislav, et al.
Veröffentlicht: (2024)
von: Bosansky, Branislav, et al.
Veröffentlicht: (2024)
Relevance as a Vulnerability: How Web Retrieval Degrades Safety Alignment in LLM Agents
von: Nawal, Aditya, et al.
Veröffentlicht: (2026)
von: Nawal, Aditya, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Hackphyr: A Local Fine-Tuned LLM Agent for Network Security Environments
von: Rigaki, Maria, et al.
Veröffentlicht: (2024) -
The Power of MEME: Adversarial Malware Creation with Model-Based Reinforcement Learning
von: Rigaki, Maria, et al.
Veröffentlicht: (2023) -
LLM in the Shell: Generative Honeypots
von: Sladić, Muris, et al.
Veröffentlicht: (2023) -
VelLMes: A high-interaction AI-based deception framework
von: Sladić, Muris, et al.
Veröffentlicht: (2025) -
Personal Information Parroting in Language Models
von: Subramani, Nishant, et al.
Veröffentlicht: (2026)