Send to which account? Evaluation of an LLM-based Scambaiting System
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Siadati, Hossein, Jafarian, Haadi, Jafarikhah, Sima |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Cyber Defense Benchmark: Agentic Threat Hunting Evaluation for LLMs in SecOps
von: Chona, Alankrit, et al.
Veröffentlicht: (2026)
von: Chona, Alankrit, et al.
Veröffentlicht: (2026)
Measuring Harmfulness of Computer-Using Agents
von: Tian, Aaron Xuxiang, et al.
Veröffentlicht: (2025)
von: Tian, Aaron Xuxiang, et al.
Veröffentlicht: (2025)
Countermind: A Multi-Layered Security Architecture for Large Language Models
von: Schwarz, Dominik
Veröffentlicht: (2025)
von: Schwarz, Dominik
Veröffentlicht: (2025)
AgentSentry: Mitigating Indirect Prompt Injection in LLM Agents via Temporal Causal Diagnostics and Context Purification
von: Zhang, Tian, et al.
Veröffentlicht: (2026)
von: Zhang, Tian, et al.
Veröffentlicht: (2026)
Evaluating the Reliability of Digital Forensic Evidence Discovered by Large Language Model: A Case Study
von: Khatiwala, Jeel Piyushkumar, et al.
Veröffentlicht: (2026)
von: Khatiwala, Jeel Piyushkumar, et al.
Veröffentlicht: (2026)
CritBench: A Framework for Evaluating Cybersecurity Capabilities of Large Language Models in IEC 61850 Digital Substation Environments
von: Keppler, Gustav, et al.
Veröffentlicht: (2026)
von: Keppler, Gustav, et al.
Veröffentlicht: (2026)
MASH: Evading Black-Box AI-Generated Text Detectors via Style Humanization
von: Gu, Yongtong, et al.
Veröffentlicht: (2026)
von: Gu, Yongtong, et al.
Veröffentlicht: (2026)
Benchmarking Large Language Models for IoC Recovery under Adversarial Code Obfuscation and Encryption
von: Morales, Jaime, et al.
Veröffentlicht: (2026)
von: Morales, Jaime, et al.
Veröffentlicht: (2026)
AegisShield: Democratizing Cyber Threat Modeling with Generative AI
von: Grofsky, Matthew
Veröffentlicht: (2025)
von: Grofsky, Matthew
Veröffentlicht: (2025)
Multilingual AI-Driven Password Strength Estimation with Similarity-Based Detection
von: Palaniappan, Nikitha M., et al.
Veröffentlicht: (2026)
von: Palaniappan, Nikitha M., et al.
Veröffentlicht: (2026)
Towards Agentic Investigation of Security Alerts
von: Eilertsen, Even, et al.
Veröffentlicht: (2026)
von: Eilertsen, Even, et al.
Veröffentlicht: (2026)
Adversarial Feeds Steer LLM Agent Decisions Against Their Defaults
von: Usman, Rana Muhammad
Veröffentlicht: (2026)
von: Usman, Rana Muhammad
Veröffentlicht: (2026)
Kill-Chain Canaries: Stage-Level Tracking of Prompt Injection Across Attack Surfaces and Model Safety Tiers
von: Wang, Haochuan Kevin, et al.
Veröffentlicht: (2026)
von: Wang, Haochuan Kevin, et al.
Veröffentlicht: (2026)
Governance Architecture for Autonomous Agent Systems: Threats, Framework, and Engineering Practice
von: Ge, Yuxu
Veröffentlicht: (2026)
von: Ge, Yuxu
Veröffentlicht: (2026)
LLM Detectors Still Fall Short of Real World: Case of LLM-Generated Short News-Like Posts
von: Gameiro, Henrique Da Silva, et al.
Veröffentlicht: (2024)
von: Gameiro, Henrique Da Silva, et al.
Veröffentlicht: (2024)
SALLIE: Safeguarding Against Latent Language & Image Exploits
von: Azov, Guy, et al.
Veröffentlicht: (2026)
von: Azov, Guy, et al.
Veröffentlicht: (2026)
Adaptive Defense Orchestration for RAG: A Sentinel-Strategist Architecture against Multi-Vector Attacks
von: Pallerla, Pranav, et al.
Veröffentlicht: (2026)
von: Pallerla, Pranav, et al.
Veröffentlicht: (2026)
Sola-Visibility-ISPM: Benchmarking Agentic AI for Identity Security Posture Management Visibility
von: Engelberg, Gal, et al.
Veröffentlicht: (2026)
von: Engelberg, Gal, et al.
Veröffentlicht: (2026)
A Validated Prompt Bank for Malicious Code Generation: Separating Executable Weapons from Security Knowledge in 1,554 Consensus-Labeled Prompts
von: Young, Richard J., et al.
Veröffentlicht: (2026)
von: Young, Richard J., et al.
Veröffentlicht: (2026)
AIRTBench: Measuring Autonomous AI Red Teaming Capabilities in Language Models
von: Dawson, Ads, et al.
Veröffentlicht: (2025)
von: Dawson, Ads, et al.
Veröffentlicht: (2025)
The Automation Advantage in AI Red Teaming
von: Mulla, Rob, et al.
Veröffentlicht: (2025)
von: Mulla, Rob, et al.
Veröffentlicht: (2025)
Token-Level Generalization in LoRA Adapter Backdoors: Attack Characterization and Behavioral Detection
von: Lelle, Travis
Veröffentlicht: (2026)
von: Lelle, Travis
Veröffentlicht: (2026)
Predicting Known Vulnerabilities from Attack Descriptions Using Sentence Transformers
von: Othman, Refat
Veröffentlicht: (2026)
von: Othman, Refat
Veröffentlicht: (2026)
Semantic Superiority vs. Forensic Efficiency: A Comparative Analysis of Deep Learning and Psycholinguistics for Business Email Compromise Detection
von: Adjei, Yaw Osei, et al.
Veröffentlicht: (2025)
von: Adjei, Yaw Osei, et al.
Veröffentlicht: (2025)
JPS: Jailbreak Multimodal Large Language Models with Collaborative Visual Perturbation and Textual Steering
von: Chen, Renmiao, et al.
Veröffentlicht: (2025)
von: Chen, Renmiao, et al.
Veröffentlicht: (2025)
Whisper Leak: a side-channel attack on Large Language Models
von: McDonald, Geoff, et al.
Veröffentlicht: (2025)
von: McDonald, Geoff, et al.
Veröffentlicht: (2025)
Towards Modeling Cybersecurity Behavior of Humans in Organizations
von: Kürtz, Klaas Ole
Veröffentlicht: (2026)
von: Kürtz, Klaas Ole
Veröffentlicht: (2026)
Binary BPE: A Family of Cross-Platform Tokenizers for Binary Analysis
von: Bommarito II, Michael J.
Veröffentlicht: (2025)
von: Bommarito II, Michael J.
Veröffentlicht: (2025)
$δ$-STEAL: LLM Stealing Attack with Local Differential Privacy
von: Dang, Kieu, et al.
Veröffentlicht: (2025)
von: Dang, Kieu, et al.
Veröffentlicht: (2025)
Security Considerations for Multi-agent Systems
von: Nguyen, Tam, et al.
Veröffentlicht: (2026)
von: Nguyen, Tam, et al.
Veröffentlicht: (2026)
Retrieval Augmented Classification for Confidential Documents
von: Chang, Yeseul E., et al.
Veröffentlicht: (2026)
von: Chang, Yeseul E., et al.
Veröffentlicht: (2026)
Jailbreak Mimicry: Automated Discovery of Narrative-Based Jailbreaks for Large Language Models
von: Ntais, Pavlos
Veröffentlicht: (2025)
von: Ntais, Pavlos
Veröffentlicht: (2025)
Research on Security Enhancement Methods for Adversarial Robust Large Language Model Intelligent Agents for Medical Decision-Making Tasks
von: Hu, Saisai
Veröffentlicht: (2026)
von: Hu, Saisai
Veröffentlicht: (2026)
SBASH: a Framework for Designing and Evaluating RAG vs. Prompt-Tuned LLM Honeypots
von: Adebimpe, Adetayo, et al.
Veröffentlicht: (2025)
von: Adebimpe, Adetayo, et al.
Veröffentlicht: (2025)
Refusal Evaluation in Coding LLMs and Code Agents: A Systematic Review of Thirteen Malicious-Code Prompt Corpora (2023-2025)
von: Young, Richard J., et al.
Veröffentlicht: (2026)
von: Young, Richard J., et al.
Veröffentlicht: (2026)
Same Payload, Different Channel: Measuring Trust Asymmetry in Tool-Using Language Models
von: Syed, Mohammed Sameer, et al.
Veröffentlicht: (2026)
von: Syed, Mohammed Sameer, et al.
Veröffentlicht: (2026)
Amplifying Training Data Exposure through Fine-Tuning with Pseudo-Labeled Memberships
von: Oh, Myung Gyo, et al.
Veröffentlicht: (2024)
von: Oh, Myung Gyo, et al.
Veröffentlicht: (2024)
Impact of Phonetics on Speaker Identity in Adversarial Voice Attack
von: Dar, Daniyal Kabir, et al.
Veröffentlicht: (2025)
von: Dar, Daniyal Kabir, et al.
Veröffentlicht: (2025)
From nuclear safety to LLM security: Applying non-probabilistic risk management strategies to build safe and secure LLM-powered systems
von: Gutfraind, Alexander, et al.
Veröffentlicht: (2025)
von: Gutfraind, Alexander, et al.
Veröffentlicht: (2025)
CEKER: A Generalizable LLM Framework for Literature Analysis with a Case Study in Unikernel Security
von: Wollman, Alex, et al.
Veröffentlicht: (2024)
von: Wollman, Alex, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Cyber Defense Benchmark: Agentic Threat Hunting Evaluation for LLMs in SecOps
von: Chona, Alankrit, et al.
Veröffentlicht: (2026) -
Measuring Harmfulness of Computer-Using Agents
von: Tian, Aaron Xuxiang, et al.
Veröffentlicht: (2025) -
Countermind: A Multi-Layered Security Architecture for Large Language Models
von: Schwarz, Dominik
Veröffentlicht: (2025) -
AgentSentry: Mitigating Indirect Prompt Injection in LLM Agents via Temporal Causal Diagnostics and Context Purification
von: Zhang, Tian, et al.
Veröffentlicht: (2026) -
Evaluating the Reliability of Digital Forensic Evidence Discovered by Large Language Model: A Case Study
von: Khatiwala, Jeel Piyushkumar, et al.
Veröffentlicht: (2026)