Benchmarking Large Language Models for IoC Recovery under Adversarial Code Obfuscation and Encryption
Fuente:
arXiv
Guardado en:
| Autores principales: | Morales, Jaime, Pastrana, Sergio, Tapiador, Juan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MASH: Evading Black-Box AI-Generated Text Detectors via Style Humanization
por: Gu, Yongtong, et al.
Publicado: (2026)
por: Gu, Yongtong, et al.
Publicado: (2026)
Countermind: A Multi-Layered Security Architecture for Large Language Models
por: Schwarz, Dominik
Publicado: (2025)
por: Schwarz, Dominik
Publicado: (2025)
Cyber Defense Benchmark: Agentic Threat Hunting Evaluation for LLMs in SecOps
por: Chona, Alankrit, et al.
Publicado: (2026)
por: Chona, Alankrit, et al.
Publicado: (2026)
A Validated Prompt Bank for Malicious Code Generation: Separating Executable Weapons from Security Knowledge in 1,554 Consensus-Labeled Prompts
por: Young, Richard J., et al.
Publicado: (2026)
por: Young, Richard J., et al.
Publicado: (2026)
AIRTBench: Measuring Autonomous AI Red Teaming Capabilities in Language Models
por: Dawson, Ads, et al.
Publicado: (2025)
por: Dawson, Ads, et al.
Publicado: (2025)
The Automation Advantage in AI Red Teaming
por: Mulla, Rob, et al.
Publicado: (2025)
por: Mulla, Rob, et al.
Publicado: (2025)
Send to which account? Evaluation of an LLM-based Scambaiting System
por: Siadati, Hossein, et al.
Publicado: (2025)
por: Siadati, Hossein, et al.
Publicado: (2025)
Measuring Harmfulness of Computer-Using Agents
por: Tian, Aaron Xuxiang, et al.
Publicado: (2025)
por: Tian, Aaron Xuxiang, et al.
Publicado: (2025)
Predicting Known Vulnerabilities from Attack Descriptions Using Sentence Transformers
por: Othman, Refat
Publicado: (2026)
por: Othman, Refat
Publicado: (2026)
Evaluating the Reliability of Digital Forensic Evidence Discovered by Large Language Model: A Case Study
por: Khatiwala, Jeel Piyushkumar, et al.
Publicado: (2026)
por: Khatiwala, Jeel Piyushkumar, et al.
Publicado: (2026)
Refusal Evaluation in Coding LLMs and Code Agents: A Systematic Review of Thirteen Malicious-Code Prompt Corpora (2023-2025)
por: Young, Richard J., et al.
Publicado: (2026)
por: Young, Richard J., et al.
Publicado: (2026)
CritBench: A Framework for Evaluating Cybersecurity Capabilities of Large Language Models in IEC 61850 Digital Substation Environments
por: Keppler, Gustav, et al.
Publicado: (2026)
por: Keppler, Gustav, et al.
Publicado: (2026)
Towards Modeling Cybersecurity Behavior of Humans in Organizations
por: Kürtz, Klaas Ole
Publicado: (2026)
por: Kürtz, Klaas Ole
Publicado: (2026)
SALLIE: Safeguarding Against Latent Language & Image Exploits
por: Azov, Guy, et al.
Publicado: (2026)
por: Azov, Guy, et al.
Publicado: (2026)
Same Payload, Different Channel: Measuring Trust Asymmetry in Tool-Using Language Models
por: Syed, Mohammed Sameer, et al.
Publicado: (2026)
por: Syed, Mohammed Sameer, et al.
Publicado: (2026)
Adversarial Feeds Steer LLM Agent Decisions Against Their Defaults
por: Usman, Rana Muhammad
Publicado: (2026)
por: Usman, Rana Muhammad
Publicado: (2026)
Semantic Superiority vs. Forensic Efficiency: A Comparative Analysis of Deep Learning and Psycholinguistics for Business Email Compromise Detection
por: Adjei, Yaw Osei, et al.
Publicado: (2025)
por: Adjei, Yaw Osei, et al.
Publicado: (2025)
Code as a Weapon: A Consensus-Labeled Prompt Bank for Measuring Coding-Model Compliance with Malicious-Code Requests
por: Young, Richard J., et al.
Publicado: (2026)
por: Young, Richard J., et al.
Publicado: (2026)
Sola-Visibility-ISPM: Benchmarking Agentic AI for Identity Security Posture Management Visibility
por: Engelberg, Gal, et al.
Publicado: (2026)
por: Engelberg, Gal, et al.
Publicado: (2026)
AegisShield: Democratizing Cyber Threat Modeling with Generative AI
por: Grofsky, Matthew
Publicado: (2025)
por: Grofsky, Matthew
Publicado: (2025)
Multilingual AI-Driven Password Strength Estimation with Similarity-Based Detection
por: Palaniappan, Nikitha M., et al.
Publicado: (2026)
por: Palaniappan, Nikitha M., et al.
Publicado: (2026)
Towards Agentic Investigation of Security Alerts
por: Eilertsen, Even, et al.
Publicado: (2026)
por: Eilertsen, Even, et al.
Publicado: (2026)
AgentSentry: Mitigating Indirect Prompt Injection in LLM Agents via Temporal Causal Diagnostics and Context Purification
por: Zhang, Tian, et al.
Publicado: (2026)
por: Zhang, Tian, et al.
Publicado: (2026)
Research on Security Enhancement Methods for Adversarial Robust Large Language Model Intelligent Agents for Medical Decision-Making Tasks
por: Hu, Saisai
Publicado: (2026)
por: Hu, Saisai
Publicado: (2026)
Amplifying Training Data Exposure through Fine-Tuning with Pseudo-Labeled Memberships
por: Oh, Myung Gyo, et al.
Publicado: (2024)
por: Oh, Myung Gyo, et al.
Publicado: (2024)
Kill-Chain Canaries: Stage-Level Tracking of Prompt Injection Across Attack Surfaces and Model Safety Tiers
por: Wang, Haochuan Kevin, et al.
Publicado: (2026)
por: Wang, Haochuan Kevin, et al.
Publicado: (2026)
$δ$-STEAL: LLM Stealing Attack with Local Differential Privacy
por: Dang, Kieu, et al.
Publicado: (2025)
por: Dang, Kieu, et al.
Publicado: (2025)
Whisper Leak: a side-channel attack on Large Language Models
por: McDonald, Geoff, et al.
Publicado: (2025)
por: McDonald, Geoff, et al.
Publicado: (2025)
JPS: Jailbreak Multimodal Large Language Models with Collaborative Visual Perturbation and Textual Steering
por: Chen, Renmiao, et al.
Publicado: (2025)
por: Chen, Renmiao, et al.
Publicado: (2025)
Governance Architecture for Autonomous Agent Systems: Threats, Framework, and Engineering Practice
por: Ge, Yuxu
Publicado: (2026)
por: Ge, Yuxu
Publicado: (2026)
Adaptive Defense Orchestration for RAG: A Sentinel-Strategist Architecture against Multi-Vector Attacks
por: Pallerla, Pranav, et al.
Publicado: (2026)
por: Pallerla, Pranav, et al.
Publicado: (2026)
Jailbreak Mimicry: Automated Discovery of Narrative-Based Jailbreaks for Large Language Models
por: Ntais, Pavlos
Publicado: (2025)
por: Ntais, Pavlos
Publicado: (2025)
Toward Secure and Compliant AI: Organizational Standards and Protocols for NLP Model Lifecycle Management
por: Arora, Sunil, et al.
Publicado: (2025)
por: Arora, Sunil, et al.
Publicado: (2025)
SBASH: a Framework for Designing and Evaluating RAG vs. Prompt-Tuned LLM Honeypots
por: Adebimpe, Adetayo, et al.
Publicado: (2025)
por: Adebimpe, Adetayo, et al.
Publicado: (2025)
Token-Level Generalization in LoRA Adapter Backdoors: Attack Characterization and Behavioral Detection
por: Lelle, Travis
Publicado: (2026)
por: Lelle, Travis
Publicado: (2026)
DWFS-Obfuscation: Dynamic Weighted Feature Selection for Robust Malware Familial Classification under Obfuscation
por: Wei, Xingyuan, et al.
Publicado: (2025)
por: Wei, Xingyuan, et al.
Publicado: (2025)
Systematic Capability Benchmarking of Frontier Large Language Models for Offensive Cyber Tasks
por: Merves, Tyler H., et al.
Publicado: (2026)
por: Merves, Tyler H., et al.
Publicado: (2026)
RouteScan: A Non-Intrusive Approach to Auditing MoE LLMs Safety via Expert Routing Telemetry
por: Lv, Bo, et al.
Publicado: (2026)
por: Lv, Bo, et al.
Publicado: (2026)
LLM Detectors Still Fall Short of Real World: Case of LLM-Generated Short News-Like Posts
por: Gameiro, Henrique Da Silva, et al.
Publicado: (2024)
por: Gameiro, Henrique Da Silva, et al.
Publicado: (2024)
Adversarial Feature Alignment: Balancing Robustness and Accuracy in Deep Learning via Adversarial Training
por: Park, Leo Hyun, et al.
Publicado: (2024)
por: Park, Leo Hyun, et al.
Publicado: (2024)
Ejemplares similares
-
MASH: Evading Black-Box AI-Generated Text Detectors via Style Humanization
por: Gu, Yongtong, et al.
Publicado: (2026) -
Countermind: A Multi-Layered Security Architecture for Large Language Models
por: Schwarz, Dominik
Publicado: (2025) -
Cyber Defense Benchmark: Agentic Threat Hunting Evaluation for LLMs in SecOps
por: Chona, Alankrit, et al.
Publicado: (2026) -
A Validated Prompt Bank for Malicious Code Generation: Separating Executable Weapons from Security Knowledge in 1,554 Consensus-Labeled Prompts
por: Young, Richard J., et al.
Publicado: (2026) -
AIRTBench: Measuring Autonomous AI Red Teaming Capabilities in Language Models
por: Dawson, Ads, et al.
Publicado: (2025)