LLMs for Domain Generation Algorithm Detection
Fuente:
arXiv
Guardado en:
| Autores principales: | La O, Reynier Leyva, Catania, Carlos A., Parlanti, Tatiana |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Temporal Analysis Framework for Intrusion Detection Systems: A Novel Taxonomy for Time-Aware Cybersecurity
por: Parlanti, Tatiana S., et al.
Publicado: (2025)
por: Parlanti, Tatiana S., et al.
Publicado: (2025)
LLM in the Shell: Generative Honeypots
por: Sladić, Muris, et al.
Publicado: (2023)
por: Sladić, Muris, et al.
Publicado: (2023)
Out of the Cage: How Stochastic Parrots Win in Cyber Security Environments
por: Rigaki, Maria, et al.
Publicado: (2023)
por: Rigaki, Maria, et al.
Publicado: (2023)
VelLMes: A high-interaction AI-based deception framework
por: Sladić, Muris, et al.
Publicado: (2025)
por: Sladić, Muris, et al.
Publicado: (2025)
Overlooked Safety Vulnerability in LLMs: Malicious Intelligent Optimization Algorithm Request and its Jailbreak
por: Gu, Haoran, et al.
Publicado: (2026)
por: Gu, Haoran, et al.
Publicado: (2026)
Beyond Pattern Matching: Seven Cross-Domain Techniques for Prompt Injection Detection
por: Munirathinam, Thamilvendhan
Publicado: (2026)
por: Munirathinam, Thamilvendhan
Publicado: (2026)
HLPD: Aligning LLMs to Human Language Preference for Machine-Revised Text Detection
por: Dai, Fangqi, et al.
Publicado: (2025)
por: Dai, Fangqi, et al.
Publicado: (2025)
GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis
por: Xie, Yueqi, et al.
Publicado: (2024)
por: Xie, Yueqi, et al.
Publicado: (2024)
A General Pseudonymization Framework for Cloud-Based LLMs: Replacing Privacy Information in Controlled Text Generation
por: Hou, Shilong, et al.
Publicado: (2025)
por: Hou, Shilong, et al.
Publicado: (2025)
Zero-Shot Embedding Drift Detection: A Lightweight Defense Against Prompt Injections in LLMs
por: Sekar, Anirudh, et al.
Publicado: (2026)
por: Sekar, Anirudh, et al.
Publicado: (2026)
A Character-based Diffusion Embedding Algorithm for Enhancing the Generation Quality of Generative Linguistic Steganographic Texts
por: Chen, Yingquan, et al.
Publicado: (2025)
por: Chen, Yingquan, et al.
Publicado: (2025)
One Model Transfer to All: On Robust Jailbreak Prompts Generation against LLMs
por: Li, Linbao, et al.
Publicado: (2025)
por: Li, Linbao, et al.
Publicado: (2025)
Chain-of-Code Collapse: Reasoning Failures in LLMs via Adversarial Prompting in Code Generation
por: Roh, Jaechul, et al.
Publicado: (2025)
por: Roh, Jaechul, et al.
Publicado: (2025)
Real-time and Zero-footprint Bag of Synthetic Syllables Algorithm for E-mail Spam Detection Using Subject Line and Short Text Fields
por: Selitskiy, Stanislav
Publicado: (2025)
por: Selitskiy, Stanislav
Publicado: (2025)
Fingerprinting LLMs via Prompt Injection
por: Hu, Yuepeng, et al.
Publicado: (2025)
por: Hu, Yuepeng, et al.
Publicado: (2025)
Mitigating Jailbreaks with Intent-Aware LLMs
por: Yeo, Wei Jie, et al.
Publicado: (2025)
por: Yeo, Wei Jie, et al.
Publicado: (2025)
Iron Sharpens Iron: Defending Against Attacks in Machine-Generated Text Detection with Adversarial Training
por: Li, Yuanfan, et al.
Publicado: (2025)
por: Li, Yuanfan, et al.
Publicado: (2025)
SoK: Are Watermarks in LLMs Ready for Deployment?
por: Dang, Kieu, et al.
Publicado: (2025)
por: Dang, Kieu, et al.
Publicado: (2025)
Watermarking Conditional Text Generation for AI Detection: Unveiling Challenges and a Semantic-Aware Watermark Remedy
por: Fu, Yu, et al.
Publicado: (2023)
por: Fu, Yu, et al.
Publicado: (2023)
Fight Poison with Poison: Enhancing Robustness in Few-shot Machine-Generated Text Detection with Adversarial Training
por: Duan, Wenjing, et al.
Publicado: (2026)
por: Duan, Wenjing, et al.
Publicado: (2026)
Fast-MIA: Efficient and Scalable Membership Inference for LLMs
por: Takahashi, Hiromu, et al.
Publicado: (2025)
por: Takahashi, Hiromu, et al.
Publicado: (2025)
Cross-Task Defense: Instruction-Tuning LLMs for Content Safety
por: Fu, Yu, et al.
Publicado: (2024)
por: Fu, Yu, et al.
Publicado: (2024)
A Simple and Efficient Jailbreak Method Exploiting LLMs' Helpfulness
por: Luo, Xuan, et al.
Publicado: (2025)
por: Luo, Xuan, et al.
Publicado: (2025)
Dataset Protection via Watermarked Canaries in Retrieval-Augmented LLMs
por: Liu, Yepeng, et al.
Publicado: (2025)
por: Liu, Yepeng, et al.
Publicado: (2025)
Jailbreaking Commercial Black-Box LLMs with Explicitly Harmful Prompts
por: Zhang, Chiyu, et al.
Publicado: (2025)
por: Zhang, Chiyu, et al.
Publicado: (2025)
Segment-Level Coherence for Robust Harmful Intent Probing in LLMs
por: He, Xuanli, et al.
Publicado: (2026)
por: He, Xuanli, et al.
Publicado: (2026)
Do Reasoning LLMs Refuse What They Infer in Long Contexts?
por: Fu, Yu, et al.
Publicado: (2026)
por: Fu, Yu, et al.
Publicado: (2026)
Continual Pretraining on Encrypted Synthetic Data for Privacy-Preserving LLMs
por: Liu, Honghao, et al.
Publicado: (2026)
por: Liu, Honghao, et al.
Publicado: (2026)
ShadowCoT: Cognitive Hijacking for Stealthy Reasoning Backdoors in LLMs
por: Zhao, Gejian, et al.
Publicado: (2025)
por: Zhao, Gejian, et al.
Publicado: (2025)
Evolve the Method, Not the Prompts: Evolutionary Synthesis of Jailbreak Attacks on LLMs
por: Chen, Yunhao, et al.
Publicado: (2025)
por: Chen, Yunhao, et al.
Publicado: (2025)
The Model's Language Matters: A Comparative Privacy Analysis of LLMs
por: Mishra, Abhishek K., et al.
Publicado: (2025)
por: Mishra, Abhishek K., et al.
Publicado: (2025)
Ingest-And-Ground: Dispelling Hallucinations from Continually-Pretrained LLMs with RAG
por: Fang, Chenhao, et al.
Publicado: (2024)
por: Fang, Chenhao, et al.
Publicado: (2024)
TRUCE: Private Benchmarking to Prevent Contamination and Improve Comparative Evaluation of LLMs
por: Rajore, Tanmay, et al.
Publicado: (2024)
por: Rajore, Tanmay, et al.
Publicado: (2024)
FFT: Towards Harmlessness Evaluation and Analysis for LLMs with Factuality, Fairness, Toxicity
por: Cui, Shiyao, et al.
Publicado: (2023)
por: Cui, Shiyao, et al.
Publicado: (2023)
Federated Domain-Specific Knowledge Transfer on Large Language Models Using Synthetic Data
por: Li, Haoran, et al.
Publicado: (2024)
por: Li, Haoran, et al.
Publicado: (2024)
Self and Cross-Model Distillation for LLMs: Effective Methods for Refusal Pattern Alignment
por: Li, Jie, et al.
Publicado: (2024)
por: Li, Jie, et al.
Publicado: (2024)
TuBA: Cross-Lingual Transferability of Backdoor Attacks in LLMs with Instruction Tuning
por: He, Xuanli, et al.
Publicado: (2024)
por: He, Xuanli, et al.
Publicado: (2024)
MetaBackdoor: Exploiting Positional Encoding as a Backdoor Attack Surface in LLMs
por: Wen, Rui, et al.
Publicado: (2026)
por: Wen, Rui, et al.
Publicado: (2026)
Jailbreaking LLMs via Semantically Relevant Nested Scenarios with Targeted Toxic Knowledge
por: Xu, Ning, et al.
Publicado: (2025)
por: Xu, Ning, et al.
Publicado: (2025)
Semantic-Preserving Adversarial Attacks on LLMs: An Adaptive Greedy Binary Search Approach
por: Zhang, Chong, et al.
Publicado: (2025)
por: Zhang, Chong, et al.
Publicado: (2025)
Ejemplares similares
-
Temporal Analysis Framework for Intrusion Detection Systems: A Novel Taxonomy for Time-Aware Cybersecurity
por: Parlanti, Tatiana S., et al.
Publicado: (2025) -
LLM in the Shell: Generative Honeypots
por: Sladić, Muris, et al.
Publicado: (2023) -
Out of the Cage: How Stochastic Parrots Win in Cyber Security Environments
por: Rigaki, Maria, et al.
Publicado: (2023) -
VelLMes: A high-interaction AI-based deception framework
por: Sladić, Muris, et al.
Publicado: (2025) -
Overlooked Safety Vulnerability in LLMs: Malicious Intelligent Optimization Algorithm Request and its Jailbreak
por: Gu, Haoran, et al.
Publicado: (2026)