Efficient and Adaptable Detection of Malicious LLM Prompts via Bootstrap Aggregation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hassan, Shayan Ali, Ni, Tao, Qazi, Zafar Ayyub, Canini, Marco |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VLMGuard: Defending VLMs against Malicious Prompts via Unlabeled Data
von: Du, Xuefeng, et al.
Veröffentlicht: (2024)
von: Du, Xuefeng, et al.
Veröffentlicht: (2024)
Breaking Obfuscation: Cluster-Aware Graph with LLM-Aided Recovery for Malicious JavaScript Detection
von: Liang, Zhihong, et al.
Veröffentlicht: (2025)
von: Liang, Zhihong, et al.
Veröffentlicht: (2025)
PIShield: Detecting Prompt Injection Attacks via Intrinsic LLM Features
von: Zou, Wei, et al.
Veröffentlicht: (2025)
von: Zou, Wei, et al.
Veröffentlicht: (2025)
Toward More Generalized Malicious URL Detection Models
von: Tsai, YunDa, et al.
Veröffentlicht: (2022)
von: Tsai, YunDa, et al.
Veröffentlicht: (2022)
CleanBase: Detecting Malicious Documents in RAG Knowledge Databases
von: Jin, Weifei, et al.
Veröffentlicht: (2026)
von: Jin, Weifei, et al.
Veröffentlicht: (2026)
GasTrace: Detecting Sandwich Attack Malicious Accounts in Ethereum
von: Liu, Zekai, et al.
Veröffentlicht: (2024)
von: Liu, Zekai, et al.
Veröffentlicht: (2024)
Malicious Internet Entity Detection Using Local Graph Inference
von: Mandlik, Simon, et al.
Veröffentlicht: (2024)
von: Mandlik, Simon, et al.
Veröffentlicht: (2024)
FL-PLAS: Federated Learning with Partial Layer Aggregation for Backdoor Defense Against High-Ratio Malicious Clients
von: Zhang, Jianyi, et al.
Veröffentlicht: (2025)
von: Zhang, Jianyi, et al.
Veröffentlicht: (2025)
FraudFox: Adaptable Fraud Detection in the Real World
von: Butler, Matthew, et al.
Veröffentlicht: (2026)
von: Butler, Matthew, et al.
Veröffentlicht: (2026)
Exploiting Leaderboards for Large-Scale Distribution of Malicious Models
von: Suri, Anshuman, et al.
Veröffentlicht: (2025)
von: Suri, Anshuman, et al.
Veröffentlicht: (2025)
PDFInspect: A Unified Feature Extraction Framework for Malicious Document Detection
von: P, Sharmila S
Veröffentlicht: (2026)
von: P, Sharmila S
Veröffentlicht: (2026)
Rethinking Image Compression on the Web with Generative AI
von: Hassan, Shayan Ali, et al.
Veröffentlicht: (2024)
von: Hassan, Shayan Ali, et al.
Veröffentlicht: (2024)
FedGT: Identification of Malicious Clients in Federated Learning with Secure Aggregation
von: Xhemrishi, Marvin, et al.
Veröffentlicht: (2023)
von: Xhemrishi, Marvin, et al.
Veröffentlicht: (2023)
Continuous Multi-Task Pre-training for Malicious URL Detection and Webpage Classification
von: Li, Yujie, et al.
Veröffentlicht: (2024)
von: Li, Yujie, et al.
Veröffentlicht: (2024)
Robustness Against Adversarial Attacks via Learning Confined Adversarial Polytopes
von: Hamidi, Shayan Mohajer, et al.
Veröffentlicht: (2024)
von: Hamidi, Shayan Mohajer, et al.
Veröffentlicht: (2024)
WebGuard++:Interpretable Malicious URL Detection via Bidirectional Fusion of HTML Subgraphs and Multi-Scale Convolutional BERT
von: Tian, Ye, et al.
Veröffentlicht: (2025)
von: Tian, Ye, et al.
Veröffentlicht: (2025)
MalRAG: A Retrieval-Augmented LLM Framework for Open-set Malicious Traffic Identification
von: Luo, Xiang, et al.
Veröffentlicht: (2025)
von: Luo, Xiang, et al.
Veröffentlicht: (2025)
Localizing Malicious Outputs from CodeLLM
von: Borana, Mayukh, et al.
Veröffentlicht: (2025)
von: Borana, Mayukh, et al.
Veröffentlicht: (2025)
A Consensus-Bayesian Framework for Detecting Malicious Activity in Enterprise Directory Access Graphs
von: Uppuluri, Pratyush, et al.
Veröffentlicht: (2026)
von: Uppuluri, Pratyush, et al.
Veröffentlicht: (2026)
ML Study of MaliciousTransactions in Ethereum
von: Katz, Natan
Veröffentlicht: (2024)
von: Katz, Natan
Veröffentlicht: (2024)
From Past to Present: A Survey of Malicious URL Detection Techniques, Datasets and Code Repositories
von: Tian, Ye, et al.
Veröffentlicht: (2025)
von: Tian, Ye, et al.
Veröffentlicht: (2025)
Token-Efficient Change Detection in LLM APIs
von: Chauvin, Timothée, et al.
Veröffentlicht: (2026)
von: Chauvin, Timothée, et al.
Veröffentlicht: (2026)
Malicious URL Detection using optimized Hist Gradient Boosting Classifier based on grid search method
von: Maftoun, Mohammad, et al.
Veröffentlicht: (2024)
von: Maftoun, Mohammad, et al.
Veröffentlicht: (2024)
One Detector Fits All: Robust and Adaptive Detection of Malicious Packages from PyPI to Enterprises
von: Montaruli, Biagio, et al.
Veröffentlicht: (2025)
von: Montaruli, Biagio, et al.
Veröffentlicht: (2025)
A New Dataset and Methodology for Malicious URL Classification
von: Schvartzman, Ilan, et al.
Veröffentlicht: (2024)
von: Schvartzman, Ilan, et al.
Veröffentlicht: (2024)
Multi-Agent Systems Execute Arbitrary Malicious Code
von: Triedman, Harold, et al.
Veröffentlicht: (2025)
von: Triedman, Harold, et al.
Veröffentlicht: (2025)
Towards Quantum Machine Learning for Malicious Code Analysis
von: Lopez, Jesus, et al.
Veröffentlicht: (2025)
von: Lopez, Jesus, et al.
Veröffentlicht: (2025)
How Not to Detect Prompt Injections with an LLM
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)
von: Choudhary, Sarthak, et al.
Veröffentlicht: (2025)
Detecting Malicious AI Agents Through Simulated Interactions
von: Pi, Yulu, et al.
Veröffentlicht: (2025)
von: Pi, Yulu, et al.
Veröffentlicht: (2025)
RobPI: Robust Private Inference against Malicious Client
von: Xue, Jiaqi, et al.
Veröffentlicht: (2026)
von: Xue, Jiaqi, et al.
Veröffentlicht: (2026)
Fake or Compromised? Making Sense of Malicious Clients in Federated Learning
von: Mozaffari, Hamid, et al.
Veröffentlicht: (2024)
von: Mozaffari, Hamid, et al.
Veröffentlicht: (2024)
Hybrid Machine Learning Approach For Real-Time Malicious Url Detection Using Som-Rmo And Rbfn With Tabu Search Optimization
von: T, Swetha, et al.
Veröffentlicht: (2024)
von: T, Swetha, et al.
Veröffentlicht: (2024)
Amatriciana: Exploiting Temporal GNNs for Robust and Efficient Money Laundering Detection
von: Di Gennaro, Marco, et al.
Veröffentlicht: (2025)
von: Di Gennaro, Marco, et al.
Veröffentlicht: (2025)
FreeMOCA: Memory-Free Continual Learning for Malicious Code Analysis
von: Asadi, Zahra, et al.
Veröffentlicht: (2026)
von: Asadi, Zahra, et al.
Veröffentlicht: (2026)
Byzantine Outside, Curious Inside: Reconstructing Data Through Malicious Updates
von: Yue, Kai, et al.
Veröffentlicht: (2025)
von: Yue, Kai, et al.
Veröffentlicht: (2025)
eDySec: A Deep Learning-based Explainable Dynamic Analysis Framework for Detecting Malicious Packages in PyPI Ecosystem
von: Mehedi, Sk Tanzir, et al.
Veröffentlicht: (2026)
von: Mehedi, Sk Tanzir, et al.
Veröffentlicht: (2026)
MIP against Agent: Malicious Image Patches Hijacking Multimodal OS Agents
von: Aichberger, Lukas, et al.
Veröffentlicht: (2025)
von: Aichberger, Lukas, et al.
Veröffentlicht: (2025)
Towards Novel Malicious Packet Recognition: A Few-Shot Learning Approach
von: Stein, Kyle, et al.
Veröffentlicht: (2024)
von: Stein, Kyle, et al.
Veröffentlicht: (2024)
MAIDS: Malicious Agent Identification-based Data Security Model for Cloud Environments
von: Gupta, Kishu, et al.
Veröffentlicht: (2024)
von: Gupta, Kishu, et al.
Veröffentlicht: (2024)
TAPAS: Efficient Two-Server Asymmetric Private Aggregation Beyond Prio(+)
von: Karthikeyan, Harish, et al.
Veröffentlicht: (2026)
von: Karthikeyan, Harish, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
VLMGuard: Defending VLMs against Malicious Prompts via Unlabeled Data
von: Du, Xuefeng, et al.
Veröffentlicht: (2024) -
Breaking Obfuscation: Cluster-Aware Graph with LLM-Aided Recovery for Malicious JavaScript Detection
von: Liang, Zhihong, et al.
Veröffentlicht: (2025) -
PIShield: Detecting Prompt Injection Attacks via Intrinsic LLM Features
von: Zou, Wei, et al.
Veröffentlicht: (2025) -
Toward More Generalized Malicious URL Detection Models
von: Tsai, YunDa, et al.
Veröffentlicht: (2022) -
CleanBase: Detecting Malicious Documents in RAG Knowledge Databases
von: Jin, Weifei, et al.
Veröffentlicht: (2026)