Prompted Contextual Vectors for Spear-Phishing Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Nahmias, Daniel, Engelberg, Gal, Klein, Dan, Shabtai, Asaf |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Ignore Me But Don't Replace Me: Utilizing Non-Linguistic Elements for Pretraining on the Cybersecurity Domain
by: Jang, Eugene, et al.
Published: (2024)
by: Jang, Eugene, et al.
Published: (2024)
Watermarking Degrades Alignment in Language Models: Analysis and Mitigation
by: Verma, Apurv, et al.
Published: (2025)
by: Verma, Apurv, et al.
Published: (2025)
Seeing the Forest through the Trees: Data Leakage from Partial Transformer Gradients
by: Li, Weijun, et al.
Published: (2024)
by: Li, Weijun, et al.
Published: (2024)
GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing
by: Zhang, Peiyan, et al.
Published: (2025)
by: Zhang, Peiyan, et al.
Published: (2025)
Protection Is (Nearly) All You Need: Structural Protection Dominates Scoring in Globally Capped KV Eviction
by: Garcia, Gabriel
Published: (2026)
by: Garcia, Gabriel
Published: (2026)
Scaling Trends in Language Model Robustness
by: Howe, Nikolaus, et al.
Published: (2024)
by: Howe, Nikolaus, et al.
Published: (2024)
Adversarial Attacks on Large Language Models Using Regularized Relaxation
by: Chacko, Samuel Jacob, et al.
Published: (2024)
by: Chacko, Samuel Jacob, et al.
Published: (2024)
Unlocking the Potential of Large Language Models for Clinical Text Anonymization: A Comparative Study
by: Pissarra, David, et al.
Published: (2024)
by: Pissarra, David, et al.
Published: (2024)
Exploiting Novel GPT-4 APIs
by: Pelrine, Kellin, et al.
Published: (2023)
by: Pelrine, Kellin, et al.
Published: (2023)
Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems
by: Hackett, William, et al.
Published: (2025)
by: Hackett, William, et al.
Published: (2025)
Operationalizing a Threat Model for Red-Teaming Large Language Models (LLMs)
by: Verma, Apurv, et al.
Published: (2024)
by: Verma, Apurv, et al.
Published: (2024)
sudoLLM: On Multi-role Alignment of Language Models
by: Saha, Soumadeep, et al.
Published: (2025)
by: Saha, Soumadeep, et al.
Published: (2025)
UniC-RAG: Universal Knowledge Corruption Attacks to Retrieval-Augmented Generation
by: Geng, Runpeng, et al.
Published: (2025)
by: Geng, Runpeng, et al.
Published: (2025)
Shortcuts Arising from Contrast: Effective and Covert Clean-Label Attacks in Prompt-Based Learning
by: Xie, Xiaopeng, et al.
Published: (2024)
by: Xie, Xiaopeng, et al.
Published: (2024)
PoTS: Proof-of-Training-Steps for Backdoor Detection in Large Language Models
by: Seddik, Issam, et al.
Published: (2025)
by: Seddik, Issam, et al.
Published: (2025)
Can Watermarked LLMs be Identified by Users via Crafted Prompts?
by: Liu, Aiwei, et al.
Published: (2024)
by: Liu, Aiwei, et al.
Published: (2024)
Learning Obfuscations Of LLM Embedding Sequences: Stained Glass Transform
by: Roberts, Jay, et al.
Published: (2025)
by: Roberts, Jay, et al.
Published: (2025)
Amplifying Training Data Exposure through Fine-Tuning with Pseudo-Labeled Memberships
by: Oh, Myung Gyo, et al.
Published: (2024)
by: Oh, Myung Gyo, et al.
Published: (2024)
Same Payload, Different Channel: Measuring Trust Asymmetry in Tool-Using Language Models
by: Syed, Mohammed Sameer, et al.
Published: (2026)
by: Syed, Mohammed Sameer, et al.
Published: (2026)
PromptSAM+: Malware Detection based on Prompt Segment Anything Model
by: Wei, Xingyuan, et al.
Published: (2024)
by: Wei, Xingyuan, et al.
Published: (2024)
Defending against Backdoor Attacks via Module Switching
by: Li, Weijun, et al.
Published: (2025)
by: Li, Weijun, et al.
Published: (2025)
Anomaly Detection in IEC-61850 GOOSE Networks: Evaluating Unsupervised and Temporal Learning for Real-Time Intrusion Detection
by: Moore, Joseph
Published: (2026)
by: Moore, Joseph
Published: (2026)
Blind Spots in the Guard: How Domain-Camouflaged Injection Attacks Evade Detection in Multi-Agent LLM Systems
by: Pai, Aaditya
Published: (2026)
by: Pai, Aaditya
Published: (2026)
Logits of API-Protected LLMs Leak Proprietary Information
by: Finlayson, Matthew, et al.
Published: (2024)
by: Finlayson, Matthew, et al.
Published: (2024)
Token-Level Generalization in LoRA Adapter Backdoors: Attack Characterization and Behavioral Detection
by: Lelle, Travis
Published: (2026)
by: Lelle, Travis
Published: (2026)
SBASH: a Framework for Designing and Evaluating RAG vs. Prompt-Tuned LLM Honeypots
by: Adebimpe, Adetayo, et al.
Published: (2025)
by: Adebimpe, Adetayo, et al.
Published: (2025)
Code as a Weapon: A Consensus-Labeled Prompt Bank for Measuring Coding-Model Compliance with Malicious-Code Requests
by: Young, Richard J., et al.
Published: (2026)
by: Young, Richard J., et al.
Published: (2026)
Evaluating the efficacy of LLM Safety Solutions : The Palit Benchmark Dataset
by: Palit, Sayon, et al.
Published: (2025)
by: Palit, Sayon, et al.
Published: (2025)
MarkLLM: An Open-Source Toolkit for LLM Watermarking
by: Pan, Leyi, et al.
Published: (2024)
by: Pan, Leyi, et al.
Published: (2024)
Large Language Models are Advanced Anonymizers
by: Staab, Robin, et al.
Published: (2024)
by: Staab, Robin, et al.
Published: (2024)
Jailbreaking Attacks vs. Content Safety Filters: How Far Are We in the LLM Safety Arms Race?
by: Xin, Yuan, et al.
Published: (2025)
by: Xin, Yuan, et al.
Published: (2025)
A Semantic Invariant Robust Watermark for Large Language Models
by: Liu, Aiwei, et al.
Published: (2023)
by: Liu, Aiwei, et al.
Published: (2023)
Power-Softmax: Towards Secure LLM Inference over Encrypted Data
by: Zimerman, Itamar, et al.
Published: (2024)
by: Zimerman, Itamar, et al.
Published: (2024)
JavelinGuard: Low-Cost Transformer Architectures for LLM Security
by: Datta, Yash, et al.
Published: (2025)
by: Datta, Yash, et al.
Published: (2025)
Breaking to Build: A Threat Model of Prompt-Based Attacks for Securing LLMs
by: Hill, Brennen, et al.
Published: (2025)
by: Hill, Brennen, et al.
Published: (2025)
DP-2Stage: Adapting Language Models as Differentially Private Tabular Data Generators
by: Afonja, Tejumade, et al.
Published: (2024)
by: Afonja, Tejumade, et al.
Published: (2024)
LLM Detectors Still Fall Short of Real World: Case of LLM-Generated Short News-Like Posts
by: Gameiro, Henrique Da Silva, et al.
Published: (2024)
by: Gameiro, Henrique Da Silva, et al.
Published: (2024)
RouteScan: A Non-Intrusive Approach to Auditing MoE LLMs Safety via Expert Routing Telemetry
by: Lv, Bo, et al.
Published: (2026)
by: Lv, Bo, et al.
Published: (2026)
Siren -- Advancing Cybersecurity through Deception and Adaptive Analysis
by: Ananthanarayanan, Samhruth, et al.
Published: (2024)
by: Ananthanarayanan, Samhruth, et al.
Published: (2024)
Jailbreak Mimicry: Automated Discovery of Narrative-Based Jailbreaks for Large Language Models
by: Ntais, Pavlos
Published: (2025)
by: Ntais, Pavlos
Published: (2025)
Similar Items
-
Ignore Me But Don't Replace Me: Utilizing Non-Linguistic Elements for Pretraining on the Cybersecurity Domain
by: Jang, Eugene, et al.
Published: (2024) -
Watermarking Degrades Alignment in Language Models: Analysis and Mitigation
by: Verma, Apurv, et al.
Published: (2025) -
Seeing the Forest through the Trees: Data Leakage from Partial Transformer Gradients
by: Li, Weijun, et al.
Published: (2024) -
GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing
by: Zhang, Peiyan, et al.
Published: (2025) -
Protection Is (Nearly) All You Need: Structural Protection Dominates Scoring in Globally Capped KV Eviction
by: Garcia, Gabriel
Published: (2026)