Agentic Misalignment: How LLMs Could Be Insider Threats
Fuente:
arXiv
Saved in:
| Main Authors: | Lynch, Aengus, Wright, Benjamin, Larson, Caleb, Ritchie, Stuart J., Mindermann, Soren, Hubinger, Evan, Perez, Ethan, Troy, Kevin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Insider Threats Mitigation: Role of Penetration Testing
by: Chauhan, Krutarth
Published: (2024)
by: Chauhan, Krutarth
Published: (2024)
TabSec: A Collaborative Framework for Novel Insider Threat Detection
by: Huang, Zilin, et al.
Published: (2024)
by: Huang, Zilin, et al.
Published: (2024)
The Danger Within: Insider Threat Modeling Using Business Process Models
by: von der Assen, Jan, et al.
Published: (2024)
by: von der Assen, Jan, et al.
Published: (2024)
Chimera: Harnessing Multi-Agent LLMs for Automatic Insider Threat Simulation
by: Yu, Jiongchi, et al.
Published: (2025)
by: Yu, Jiongchi, et al.
Published: (2025)
Insight-LLM: LLM-enhanced Multi-view Fusion in Insider Threat Detection
by: Song, Chengyu, et al.
Published: (2025)
by: Song, Chengyu, et al.
Published: (2025)
Signal Decomposition Reveals Structure in Insider Threat Detection under Sparse Temporal Data
by: Beadles, Hayden, et al.
Published: (2026)
by: Beadles, Hayden, et al.
Published: (2026)
Real-Time Detection of Insider Threats Using Behavioral Analytics and Deep Evidential Clustering
by: Ali, Anas, et al.
Published: (2025)
by: Ali, Anas, et al.
Published: (2025)
Facade: High-Precision Insider Threat Detection Using Deep Contextual Anomaly Detection
by: Kantchelian, Alex, et al.
Published: (2024)
by: Kantchelian, Alex, et al.
Published: (2024)
What Do We Know About the Psychology of Insider Threats?
by: Ruohonen, Jukka, et al.
Published: (2024)
by: Ruohonen, Jukka, et al.
Published: (2024)
"Blockchain-Enabled Zero Trust Framework for Securing FinTech Ecosystems Against Insider Threats and Cyber Attacks"
by: Singh, Avinash, et al.
Published: (2025)
by: Singh, Avinash, et al.
Published: (2025)
MambaITD: An Efficient Cross-Modal Mamba Network for Insider Threat Detection
by: Kong, Kaichuan, et al.
Published: (2025)
by: Kong, Kaichuan, et al.
Published: (2025)
Audit-LLM: Multi-Agent Collaboration for Log-based Insider Threat Detection
by: Song, Chengyu, et al.
Published: (2024)
by: Song, Chengyu, et al.
Published: (2024)
A Proactive Insider Threat Management Framework Using Explainable Machine Learning
by: Shikonde, Selma, et al.
Published: (2025)
by: Shikonde, Selma, et al.
Published: (2025)
Insider Threat Detection Using GCN and Bi-LSTM with Explicit and Implicit Graph Representations
by: Yumlembam, Rahul, et al.
Published: (2025)
by: Yumlembam, Rahul, et al.
Published: (2025)
OrgForge-IT: A Verifiable Synthetic Benchmark for LLM-Based Insider Threat Detection
by: Flynt, Jeffrey
Published: (2026)
by: Flynt, Jeffrey
Published: (2026)
Behavioral Analytics for Continuous Insider Threat Detection in Zero-Trust Architectures
by: Sarraf, Gaurav
Published: (2026)
by: Sarraf, Gaurav
Published: (2026)
LAN: Learning Adaptive Neighbors for Real-Time Insider Threat Detection
by: Cai, Xiangrui, et al.
Published: (2024)
by: Cai, Xiangrui, et al.
Published: (2024)
From Secure Agentic AI to Secure Agentic Web: Challenges, Threats, and Future Directions
by: Deng, Zhihang, et al.
Published: (2026)
by: Deng, Zhihang, et al.
Published: (2026)
Log2Sig: Frequency-Aware Insider Threat Detection via Multivariate Behavioral Signal Decomposition
by: Kong, Kaichuan, et al.
Published: (2025)
by: Kong, Kaichuan, et al.
Published: (2025)
RMSL: Weakly-Supervised Insider Threat Detection with Robust Multi-sphere Learning
by: Wang, Yang, et al.
Published: (2025)
by: Wang, Yang, et al.
Published: (2025)
RedChronos: A Large Language Model-Based Log Analysis System for Insider Threat Detection in Enterprises
by: Li, Chenyu, et al.
Published: (2025)
by: Li, Chenyu, et al.
Published: (2025)
Integrating Multi-Agent Simulation, Behavioral Forensics, and Trust-Aware Machine Learning for Adaptive Insider Threat Detection
by: Kausar, Firdous, et al.
Published: (2026)
by: Kausar, Firdous, et al.
Published: (2026)
FedAT: Federated Adversarial Training for Distributed Insider Threat Detection
by: Gayathri, R G, et al.
Published: (2024)
by: Gayathri, R G, et al.
Published: (2024)
The Art of (Mis)alignment: How Fine-Tuning Methods Effectively Misalign and Realign LLMs in Post-Training
by: Zhang, Rui, et al.
Published: (2026)
by: Zhang, Rui, et al.
Published: (2026)
Agentic Knowledge Distillation: Autonomous Training of Small Language Models for SMS Threat Detection
by: ElZemity, Adel, et al.
Published: (2026)
by: ElZemity, Adel, et al.
Published: (2026)
Transforming Cyber Defense: Harnessing Agentic and Frontier AI for Proactive, Ethical Threat Intelligence
by: Tallam, Krti
Published: (2025)
by: Tallam, Krti
Published: (2025)
From AI-Generated Content to Agentic Action: Security and Safety Threats in Generative AI
by: Zhang, Zelin, et al.
Published: (2026)
by: Zhang, Zelin, et al.
Published: (2026)
ThreatGPT: An Agentic AI Framework for Enhancing Public Safety through Threat Modeling
by: Zisad, Sharif Noor, et al.
Published: (2025)
by: Zisad, Sharif Noor, et al.
Published: (2025)
Benchmarking LLMs in an Embodied Environment for Blue Team Threat Hunting
by: Liu, Xiaoqun, et al.
Published: (2025)
by: Liu, Xiaoqun, et al.
Published: (2025)
CyLens: Towards Reinventing Cyber Threat Intelligence in the Paradigm of Agentic Large Language Models
by: Liu, Xiaoqun, et al.
Published: (2025)
by: Liu, Xiaoqun, et al.
Published: (2025)
Fine Grained Insider Risk Detection
by: Huber, Birkett, et al.
Published: (2024)
by: Huber, Birkett, et al.
Published: (2024)
Securing Agentic AI: Threat Modeling and Risk Analysis for Network Monitoring Agentic AI System
by: Zambare, Pallavi, et al.
Published: (2025)
by: Zambare, Pallavi, et al.
Published: (2025)
Federated Graph AGI for Cross-Border Insider Threat Intelligence in Government Financial Schemes
by: Nayak, Srikumar, et al.
Published: (2026)
by: Nayak, Srikumar, et al.
Published: (2026)
Using LLMs to Automate Threat Intelligence Analysis Workflows in Security Operation Centers
by: Tseng, PeiYu, et al.
Published: (2024)
by: Tseng, PeiYu, et al.
Published: (2024)
Can LLMs Threaten Human Survival? Benchmarking Potential Existential Threats from LLMs via Prefix Completion
by: Cui, Yu, et al.
Published: (2025)
by: Cui, Yu, et al.
Published: (2025)
Security Threats in Agentic AI System
by: Khan, Raihan, et al.
Published: (2024)
by: Khan, Raihan, et al.
Published: (2024)
PMU-Data: Data Traces Could be Distinguished
by: Li, Zhouyang, et al.
Published: (2025)
by: Li, Zhouyang, et al.
Published: (2025)
DMFI: A Dual-Modality Log Analysis Framework for Insider Threat Detection with LoRA-Tuned Language Models
by: Kong, Kaichuan, et al.
Published: (2025)
by: Kong, Kaichuan, et al.
Published: (2025)
ASTRIDE: A Security Threat Modeling Platform for Agentic-AI Applications
by: Bandara, Eranga, et al.
Published: (2025)
by: Bandara, Eranga, et al.
Published: (2025)
Toward Risk Thresholds for AI-Enabled Cyber Threats: Enhancing Decision-Making Under Uncertainty with Bayesian Networks
by: Jackson, Krystal, et al.
Published: (2026)
by: Jackson, Krystal, et al.
Published: (2026)
Similar Items
-
Insider Threats Mitigation: Role of Penetration Testing
by: Chauhan, Krutarth
Published: (2024) -
TabSec: A Collaborative Framework for Novel Insider Threat Detection
by: Huang, Zilin, et al.
Published: (2024) -
The Danger Within: Insider Threat Modeling Using Business Process Models
by: von der Assen, Jan, et al.
Published: (2024) -
Chimera: Harnessing Multi-Agent LLMs for Automatic Insider Threat Simulation
by: Yu, Jiongchi, et al.
Published: (2025) -
Insight-LLM: LLM-enhanced Multi-view Fusion in Insider Threat Detection
by: Song, Chengyu, et al.
Published: (2025)