HoneyTrap: Deceiving Large Language Model Attackers to Honeypot Traps with Resilient Multi-Agent Defense
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Siyuan, Lin, Xi, Wu, Jun, Liu, Zehao, Li, Haoyu, Ju, Tianjie, Chen, Xiang, Li, Jianhua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CoopGuard: Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Round Attacks
von: Li, Siyuan, et al.
Veröffentlicht: (2026)
von: Li, Siyuan, et al.
Veröffentlicht: (2026)
HoneyWin: High-Interaction Windows Honeypot in Enterprise Environment
von: Aung, Yan Lin, et al.
Veröffentlicht: (2025)
von: Aung, Yan Lin, et al.
Veröffentlicht: (2025)
HoneySat: A Network-based Satellite Honeypot Framework
von: López-Morales, Efrén, et al.
Veröffentlicht: (2025)
von: López-Morales, Efrén, et al.
Veröffentlicht: (2025)
HoneyGAN Pots: A Deep Learning Approach for Generating Honeypots
von: Gabrys, Ryan, et al.
Veröffentlicht: (2024)
von: Gabrys, Ryan, et al.
Veröffentlicht: (2024)
HoneyGPT: Breaking the Trilemma in Terminal Honeypots with Large Language Model
von: Wang, Ziyang, et al.
Veröffentlicht: (2024)
von: Wang, Ziyang, et al.
Veröffentlicht: (2024)
TrapSuffix: Proactive Defense Against Adversarial Suffixes in Jailbreaking
von: Du, Mengyao, et al.
Veröffentlicht: (2026)
von: Du, Mengyao, et al.
Veröffentlicht: (2026)
Copyright Traps for Large Language Models
von: Meeus, Matthieu, et al.
Veröffentlicht: (2024)
von: Meeus, Matthieu, et al.
Veröffentlicht: (2024)
Breaking Minds, Breaking Systems: Jailbreaking Large Language Models via Human-like Psychological Manipulation
von: Liu, Zehao, et al.
Veröffentlicht: (2025)
von: Liu, Zehao, et al.
Veröffentlicht: (2025)
The 'Sure' Trap: Multi-Scale Poisoning Analysis of Stealthy Compliance-Only Backdoors in Fine-Tuned Large Language Models
von: Tan, Yuting, et al.
Veröffentlicht: (2025)
von: Tan, Yuting, et al.
Veröffentlicht: (2025)
Large Language Models are Good Attackers: Efficient and Stealthy Textual Backdoor Attacks
von: Li, Ziqiang, et al.
Veröffentlicht: (2024)
von: Li, Ziqiang, et al.
Veröffentlicht: (2024)
Trapping Attacker in Dilemma: Examining Internal Correlations and External Influences of Trigger for Defending GNN Backdoors
von: Yang, Fan, et al.
Veröffentlicht: (2026)
von: Yang, Fan, et al.
Veröffentlicht: (2026)
PreCurious: How Innocent Pre-Trained Language Models Turn into Privacy Traps
von: Liu, Ruixuan, et al.
Veröffentlicht: (2024)
von: Liu, Ruixuan, et al.
Veröffentlicht: (2024)
CTRAP: Embedding Collapse Trap to Safeguard Large Language Models from Harmful Fine-Tuning
von: Yi, Biao, et al.
Veröffentlicht: (2025)
von: Yi, Biao, et al.
Veröffentlicht: (2025)
Honeypot Protocol
von: Hasan, Najmul
Veröffentlicht: (2026)
von: Hasan, Najmul
Veröffentlicht: (2026)
Transferring Backdoors between Large Language Models by Knowledge Distillation
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2024)
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2024)
Defenses at Odds: Measuring and Explaining Defense Conflicts in Large Language Models
von: Meng, Xiangtao, et al.
Veröffentlicht: (2026)
von: Meng, Xiangtao, et al.
Veröffentlicht: (2026)
Trustworthy AI-Generative Content for Intelligent Network Service: Robustness, Security, and Fairness
von: Li, Siyuan, et al.
Veröffentlicht: (2024)
von: Li, Siyuan, et al.
Veröffentlicht: (2024)
NSmark: Null Space Based Black-box Watermarking Defense Framework for Language Models
von: Zhao, Haodong, et al.
Veröffentlicht: (2024)
von: Zhao, Haodong, et al.
Veröffentlicht: (2024)
LLMAtKGE: Large Language Models as Explainable Attackers against Knowledge Graph Embeddings
von: Li, Ting, et al.
Veröffentlicht: (2025)
von: Li, Ting, et al.
Veröffentlicht: (2025)
Seeing is Deceiving: Exploitation of Visual Pathways in Multi-Modal Language Models
von: Janowczyk, Pete, et al.
Veröffentlicht: (2024)
von: Janowczyk, Pete, et al.
Veröffentlicht: (2024)
The Invitation Trap: Proactive Availability Backdoor in LLMs via Conversational Induction
von: Wang, He, et al.
Veröffentlicht: (2026)
von: Wang, He, et al.
Veröffentlicht: (2026)
Defensible Design for OpenClaw: Securing Autonomous Tool-Invoking Agents
von: Li, Zongwei, et al.
Veröffentlicht: (2026)
von: Li, Zongwei, et al.
Veröffentlicht: (2026)
LLMHoney: A Real-Time SSH Honeypot with Large Language Model-Driven Dynamic Response Generation
von: Malhotra, Pranjay
Veröffentlicht: (2025)
von: Malhotra, Pranjay
Veröffentlicht: (2025)
LLM Agent Honeypot: Monitoring AI Hacking Agents in the Wild
von: Reworr, et al.
Veröffentlicht: (2024)
von: Reworr, et al.
Veröffentlicht: (2024)
Descriptor: Multi-Regional Cloud Honeypot Dataset (MURHCAD)
von: Feito-Casares, Enrique, et al.
Veröffentlicht: (2026)
von: Feito-Casares, Enrique, et al.
Veröffentlicht: (2026)
LoopTrap: Termination Poisoning Attacks on LLM Agents
von: Xu, Huiyu, et al.
Veröffentlicht: (2026)
von: Xu, Huiyu, et al.
Veröffentlicht: (2026)
Honeypot Implementation in a Cloud Environment
von: Machmeier, Stefan
Veröffentlicht: (2023)
von: Machmeier, Stefan
Veröffentlicht: (2023)
Honeypot-powered Malware Reverse Engineering
von: Bombardieri, Michele, et al.
Veröffentlicht: (2015)
von: Bombardieri, Michele, et al.
Veröffentlicht: (2015)
AutoAttacker: A Large Language Model Guided System to Implement Automatic Cyber-attacks
von: Xu, Jiacen, et al.
Veröffentlicht: (2024)
von: Xu, Jiacen, et al.
Veröffentlicht: (2024)
Attacker Control and Bug Prioritization
von: Lacombe, Guilhem, et al.
Veröffentlicht: (2025)
von: Lacombe, Guilhem, et al.
Veröffentlicht: (2025)
Uplifted Attackers, Human Defenders: The Cyber Offense-Defense Balance for Trailing-Edge Organizations
von: Murphy, Benjamin, et al.
Veröffentlicht: (2025)
von: Murphy, Benjamin, et al.
Veröffentlicht: (2025)
CacheTrap: Unveiling a Stealthier Gray-Box Trojan against LLMs
von: Nahian, Mohaiminul Al, et al.
Veröffentlicht: (2025)
von: Nahian, Mohaiminul Al, et al.
Veröffentlicht: (2025)
AgentTrap: Measuring Runtime Trust Failures in Third-Party Agent Skills
von: Zhuang, Haomin, et al.
Veröffentlicht: (2026)
von: Zhuang, Haomin, et al.
Veröffentlicht: (2026)
DCS Chain: A Flexible Private Blockchain System
von: Zheng, Jianwu, et al.
Veröffentlicht: (2024)
von: Zheng, Jianwu, et al.
Veröffentlicht: (2024)
TrapFlow: Controllable Website Fingerprinting Defense via Dynamic Backdoor Learning
von: Liang, Siyuan, et al.
Veröffentlicht: (2024)
von: Liang, Siyuan, et al.
Veröffentlicht: (2024)
How Agentic AI Coding Assistants Become the Attacker's Shell
von: Liu, Yue, et al.
Veröffentlicht: (2026)
von: Liu, Yue, et al.
Veröffentlicht: (2026)
HoneyDOC: An Efficient Honeypot Architecture Enabling All-Round Design
von: Fan, Wenjun, et al.
Veröffentlicht: (2024)
von: Fan, Wenjun, et al.
Veröffentlicht: (2024)
ToDA: Target-oriented Diffusion Attacker against Recommendation System
von: Liu, Xiaohao, et al.
Veröffentlicht: (2024)
von: Liu, Xiaohao, et al.
Veröffentlicht: (2024)
A Defender-Attacker-Defender Model for Optimizing the Resilience of Hospital Networks to Cyberattacks
von: Helfrich, Stephan, et al.
Veröffentlicht: (2026)
von: Helfrich, Stephan, et al.
Veröffentlicht: (2026)
Jailbreaking LLMs & VLMs: Mechanisms, Evaluation, and Unified Defense
von: Chen, Zejian, et al.
Veröffentlicht: (2026)
von: Chen, Zejian, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
CoopGuard: Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Round Attacks
von: Li, Siyuan, et al.
Veröffentlicht: (2026) -
HoneyWin: High-Interaction Windows Honeypot in Enterprise Environment
von: Aung, Yan Lin, et al.
Veröffentlicht: (2025) -
HoneySat: A Network-based Satellite Honeypot Framework
von: López-Morales, Efrén, et al.
Veröffentlicht: (2025) -
HoneyGAN Pots: A Deep Learning Approach for Generating Honeypots
von: Gabrys, Ryan, et al.
Veröffentlicht: (2024) -
HoneyGPT: Breaking the Trilemma in Terminal Honeypots with Large Language Model
von: Wang, Ziyang, et al.
Veröffentlicht: (2024)