Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting
Fuente:
arXiv
Saved in:
| Main Authors: | Meng, Yuqiao, Tang, Luoxi, Yu, Feiyang, Li, Xi, Yan, Guanhua, Yang, Ping, Xi, Zhaohan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Uncovering Vulnerabilities of LLM-Assisted Cyber Threat Intelligence
by: Meng, Yuqiao, et al.
Published: (2025)
by: Meng, Yuqiao, et al.
Published: (2025)
Benchmarking LLMs in an Embodied Environment for Blue Team Threat Hunting
by: Liu, Xiaoqun, et al.
Published: (2025)
by: Liu, Xiaoqun, et al.
Published: (2025)
SafeGPT: Preventing Data Leakage and Unethical Outputs in Enterprise LLM Use
by: Desai, Pratyush, et al.
Published: (2026)
by: Desai, Pratyush, et al.
Published: (2026)
Smart Privacy Policy Assistant: An LLM-Powered System for Transparent and Actionable Privacy Notices
by: Kalvakuntla, Sriharshini, et al.
Published: (2026)
by: Kalvakuntla, Sriharshini, et al.
Published: (2026)
POLAR: Automating Cyber Threat Prioritization through LLM-Powered Assessment
by: Tang, Luoxi, et al.
Published: (2025)
by: Tang, Luoxi, et al.
Published: (2025)
Data to Defense: The Role of Curation in Customizing LLMs Against Jailbreaking Attacks
by: Liu, Xiaoqun, et al.
Published: (2024)
by: Liu, Xiaoqun, et al.
Published: (2024)
Policy-Guided Threat Hunting: An LLM enabled Framework with Splunk SOC Triage
by: Sahay, Rishikesh, et al.
Published: (2026)
by: Sahay, Rishikesh, et al.
Published: (2026)
Robustifying Safety-Aligned Large Language Models through Clean Data Curation
by: Liu, Xiaoqun, et al.
Published: (2024)
by: Liu, Xiaoqun, et al.
Published: (2024)
LLM-Assisted Proactive Threat Intelligence for Automated Reasoning
by: Paul, Shuva, et al.
Published: (2025)
by: Paul, Shuva, et al.
Published: (2025)
Blue Teaming Function-Calling Agents
by: Dolcetti, Greta, et al.
Published: (2026)
by: Dolcetti, Greta, et al.
Published: (2026)
Technique Inference Engine: A Recommender Model to Support Cyber Threat Hunting
by: Turner, Matthew J., et al.
Published: (2025)
by: Turner, Matthew J., et al.
Published: (2025)
Unsupervised Threat Hunting using Continuous Bag-of-Terms-and-Time (CBoTT)
by: Kayhan, Varol, et al.
Published: (2024)
by: Kayhan, Varol, et al.
Published: (2024)
CTIArena: Benchmarking LLM Knowledge and Reasoning Across Heterogeneous Cyber Threat Intelligence
by: Cheng, Yutong, et al.
Published: (2025)
by: Cheng, Yutong, et al.
Published: (2025)
All Your Knowledge Belongs to Us: Stealing Knowledge Graphs via Reasoning APIs
by: Xi, Zhaohan
Published: (2025)
by: Xi, Zhaohan
Published: (2025)
Towards Secure Retrieval-Augmented Generation: A Comprehensive Review of Threats, Defenses and Benchmarks
by: Mu, Yanming, et al.
Published: (2026)
by: Mu, Yanming, et al.
Published: (2026)
Towards Explainable and Lightweight AI for Real-Time Cyber Threat Hunting in Edge Networks
by: Rahmati, Milad
Published: (2025)
by: Rahmati, Milad
Published: (2025)
Adversarial Network Imagination: Causal LLMs and Digital Twins for Proactive Telecom Mitigation
by: Sriram, Vignesh, et al.
Published: (2026)
by: Sriram, Vignesh, et al.
Published: (2026)
Securing the Future: Proactive Threat Hunting for Sustainable IoT Ecosystems
by: Ghasemshirazi, Saeid, et al.
Published: (2024)
by: Ghasemshirazi, Saeid, et al.
Published: (2024)
From Legacy to Standard: LLM-Assisted Transformation of Cybersecurity Playbooks into CACAO Format
by: Gurabi, Mehdi Akbari, et al.
Published: (2025)
by: Gurabi, Mehdi Akbari, et al.
Published: (2025)
ThreatModeling-LLM: Automating Threat Modeling using Large Language Models for Banking System
by: Wu, Tingmin, et al.
Published: (2024)
by: Wu, Tingmin, et al.
Published: (2024)
Distillability of LLM Security Logic: Predicting Attack Success Rate of Outline Filling Attack via Ranking Regression
by: Zhang, Tianyu, et al.
Published: (2025)
by: Zhang, Tianyu, et al.
Published: (2025)
Taming OpenClaw: Security Analysis and Mitigation of Autonomous LLM Agent Threats
by: Deng, Xinhao, et al.
Published: (2026)
by: Deng, Xinhao, et al.
Published: (2026)
Red Teaming Large Reasoning Models
by: Chen, Jiawei, et al.
Published: (2025)
by: Chen, Jiawei, et al.
Published: (2025)
Breaking Minds, Breaking Systems: Jailbreaking Large Language Models via Human-like Psychological Manipulation
by: Liu, Zehao, et al.
Published: (2025)
by: Liu, Zehao, et al.
Published: (2025)
CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence
by: Alam, Md Tanvirul, et al.
Published: (2024)
by: Alam, Md Tanvirul, et al.
Published: (2024)
Audit-LLM: Multi-Agent Collaboration for Log-based Insider Threat Detection
by: Song, Chengyu, et al.
Published: (2024)
by: Song, Chengyu, et al.
Published: (2024)
ATLAS: AI-Assisted Threat-to-Assertion Learning for System-on-Chip Security Verification
by: Tashdid, Ishraq, et al.
Published: (2026)
by: Tashdid, Ishraq, et al.
Published: (2026)
Towards Effective Offensive Security LLM Agents: Hyperparameter Tuning, LLM as a Judge, and a Lightweight CTF Benchmark
by: Shao, Minghao, et al.
Published: (2025)
by: Shao, Minghao, et al.
Published: (2025)
Cyber Defense Benchmark: Agentic Threat Hunting Evaluation for LLMs in SecOps
by: Chona, Alankrit, et al.
Published: (2026)
by: Chona, Alankrit, et al.
Published: (2026)
CyberSOCEval: Benchmarking LLMs Capabilities for Malware Analysis and Threat Intelligence Reasoning
by: Deason, Lauren, et al.
Published: (2025)
by: Deason, Lauren, et al.
Published: (2025)
AthenaBench: A Dynamic Benchmark for Evaluating LLMs in Cyber Threat Intelligence
by: Alam, Md Tanvirul, et al.
Published: (2025)
by: Alam, Md Tanvirul, et al.
Published: (2025)
WIPI: A New Web Threat for LLM-Driven Web Agents
by: Wu, Fangzhou, et al.
Published: (2024)
by: Wu, Fangzhou, et al.
Published: (2024)
BARTPredict: Empowering IoT Security with LLM-Driven Cyber Threat Prediction
by: Diaf, Alaeddine, et al.
Published: (2025)
by: Diaf, Alaeddine, et al.
Published: (2025)
Benchmarking Safety Risks of Knowledge-Intensive Reasoning under Malicious Knowledge Editing
by: Mao, Qinghua, et al.
Published: (2026)
by: Mao, Qinghua, et al.
Published: (2026)
RAG-targeted Adversarial Attack on LLM-based Threat Detection and Mitigation Framework
by: Ikbarieh, Seif, et al.
Published: (2025)
by: Ikbarieh, Seif, et al.
Published: (2025)
Enabling Trustworthy Federated Learning via Remote Attestation for Mitigating Byzantine Threats
by: Zhang, Chaoyu, et al.
Published: (2025)
by: Zhang, Chaoyu, et al.
Published: (2025)
When the Manual Lies: A Realistic Benchmark to Evaluate MCP Poisoning Attacks for LLM Agents
by: Liu, Shi, et al.
Published: (2026)
by: Liu, Shi, et al.
Published: (2026)
Automatic Red Teaming LLM-based Agents with Model Context Protocol Tools
by: He, Ping, et al.
Published: (2025)
by: He, Ping, et al.
Published: (2025)
Incalmo: An Autonomous LLM-assisted System for Red Teaming Multi-Host Networks
by: Singer, Brian, et al.
Published: (2025)
by: Singer, Brian, et al.
Published: (2025)
Trojan Horses in Recruiting: A Red-Teaming Case Study on Indirect Prompt Injection in Standard vs. Reasoning Models
by: Wirth, Manuel
Published: (2026)
by: Wirth, Manuel
Published: (2026)
Similar Items
-
Uncovering Vulnerabilities of LLM-Assisted Cyber Threat Intelligence
by: Meng, Yuqiao, et al.
Published: (2025) -
Benchmarking LLMs in an Embodied Environment for Blue Team Threat Hunting
by: Liu, Xiaoqun, et al.
Published: (2025) -
SafeGPT: Preventing Data Leakage and Unethical Outputs in Enterprise LLM Use
by: Desai, Pratyush, et al.
Published: (2026) -
Smart Privacy Policy Assistant: An LLM-Powered System for Transparent and Actionable Privacy Notices
by: Kalvakuntla, Sriharshini, et al.
Published: (2026) -
POLAR: Automating Cyber Threat Prioritization through LLM-Powered Assessment
by: Tang, Luoxi, et al.
Published: (2025)